domingo, 20 de setembro de 2026
PesquisarExplore
Source trace. Via News points to the documents behind its reporting and shows what we drew from each — so you can check any claim. How we source
Source document

Why AI Chatbots Agree With You Even When You’re Wrong

View original at spectrum.ieee.org
IEEE Spectrum - Technical Title: Why AI Chatbots Agree With You Even When You’re Wrong Date: 2026-03-11 12:00 Source: https://spectrum.ieee.org/ai-sycophancy <img src="https://spectrum.ieee.org/media-library/conceptual-collage-of-emojis-being-poured-through-a-strainer-and-into-a-phone-judgmental-emojis-are-filtered-out…
Opening lines of the source · short snapshot — read the full document at the original

O que extraímos desta fonte

The claims Via News extracted from this document. We point to the source; we don't replace it.

  • If a user states a belief in a presupposition, the model will go along with it because that's what people normally do in conversations

    60% confidence
  • Pretrained LLMs were already sycophantic before reinforcement learning

    60% confidence
  • We just need to ask ourselves as a society, What do we want? Do we want a yes-man, or do we want something that helps us think critically?

    60% confidence
  • The thing that was most surprising is that these relatively simple fixes can actually do a lot to reduce sycophancy

    60% confidence
  • AI engaged my intellect, fed my ego, and altered my worldviews leading to psychiatric hospitalization

    60% confidence
  • When an AI receives a minor misgiving about its answer, it flips to agree with the user

    60% confidence
  • Sycophantic AI might lie to us and hide bad news in order to increase our short-term happiness

    60% confidence
  • The update we removed was overly flattering or agreeable—often described as sycophantic

    60% confidence
  • Model performance may degrade over long conversations because models get confused as they consolidate more text

    60% confidence
  • ChatGPT may correctly point to a suicide hotline when someone first mentions intent, but after many messages over a long period of time, it might eventually offer an answer that goes against our safeguards

    60% confidence
  • Reinforcement learning increased sycophancy, with one of the biggest predictors of positive ratings being whether a model agreed with a person's beliefs and biases

    60% confidence

Citado nestas reportagens da Via News

O que sabemos · a inteligência por trás desta página
Ao vivo do substrato
O que estamos a ver
AI Boom Hits a Fork: Slowdown Calls Clash with Capex Confidence as Markets Get Nervous
Dario Amodei's repeated calls for a global slowdown in frontier AI development, echoed by Microsoft's new humanist AI code of conduct and FTC antitrust caution, are being publicly rejected by Nvidia and Meta leadership even as hyperscaler spending draws fresh skeptical scrutiny (Wachter's analysis, Burry-style overbuilding worries) and weak guidance from Adobe and a post-slowdown-comment selloff in GE Vernova signal investor jitters. Meanwhile wealth and security effects of the AI race keep compounding — Zhang Yiming's fortune surging on AI-driven ByteDance value, a Chinese hacking firm weaponizing AI against stolen government secrets, and low-quality AI-generated products (an AI sitcom, a spam-flooding agent platform) fueling backlash even as adoption races ahead.
A nossa leitura dos dados ›
Sinais que acompanhamos
EPKINLY Regulatory-Clinical Success Cascade
High probability of expanded label indications, additional combination approvals, and competitive positioning strength in follicular lymphoma market. Predicts positive commercial uptake and potential accelerated review for related indications.
Padrões que observamos ›
Onde as fontes divergem
Berkshire Hathaway
Both facts report Berkshire Hathaway's cash position on 2026-01-01 with identical observation timestamps, but claim vastly different values: 380 billion USD vs 400 USD. These cannot both be true for the same entity at the same point in time. The magnitude of the discrepancy (a factor of ~10^9) rules out rounding, unit conversion, or methodological differences.
Sinalizamos conflitos abertamente ›
Verificado recentemente
Verificado com a fonte original
4,982
factos rastreados à sua fonte — e sinalizamos os que não se confirmam.
101 entidades acompanhadas4,982 factos verificados com a fonte5,299 documentos-fonte arquivados
Consulte estes dados → isubstrate.com
Why AI Chatbots Agree With You Even When You’re Wrong — Source | Via News | pt.VIA.NEWS