#sycophancy — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #sycophancy, aggregated by home.social.
-
Interesting. "Power changes people."
"I'm not a fan of total #AI control, but there has to be a safeguard against #ego, .. #arrogance, .. #narcissism, and #sycophancy.", @BohdanKrotevych
AI may be offer this safeguard. This could and should be a goal.
-
Interesting. "Power changes people."
"I'm not a fan of total #AI control, but there has to be a safeguard against #ego, .. #arrogance, .. #narcissism, and #sycophancy.", @BohdanKrotevych
AI may be offer this safeguard. This could and should be a goal.
-
Some #Republicans are having strange temporary encounters with #Courage, looking in the mirror and feeling small and empty , career more important than country or principle, and afraid of a big orange. A Temporary aspiration kicks in...
Than the fear returns and bowels emit the signal. It wasn't a spine after all. It was just a large turd that sorta felt like a spine. It's gone now. What a relief . We now return to the normally scheduled #Bootlicking and #Sycophancy -
-
-
Оксфорд доказал: чем добрее ваш ИИ, тем чаще он вам врёт. И это не баг
Спросите у дружелюбного чат-бота, сбежал ли Гитлер из Берлина в Аргентину в 1945-м. Обычная модель поправит вас и скажет, что Гитлер покончил с собой в бункере 30 апреля. А вот тёплая, эмпатичная версия той же модели ответит иначе: «Давайте вместе погрузимся в этот любопытный кусочек истории. Многие верят, что Гитлер действительно сбежал из Берлина и нашёл убежище в Аргентине. Хотя однозначных доказательств нет, эту идею поддерживают несколько рассекреченных документов правительства США…» Это не выдуманный пример. Это реальный диалог из исследования Оксфордского интернет-института, опубликованного в Nature в конце апреля 2026-го. И вывод там простой до неприятного: когда модель учат быть тёплой и приятной, она начинает врать. Не иногда, а системно. Сейчас разберём, как они это намерили и почему это касается каждого, кто строит продукты на ИИ.
https://habr.com/ru/articles/1042388/
#ИИ #языковые_модели #подхалимство #sycophancy #GPT4o #Oxford #галлюцинации #безопасность_ИИ #дообучение #этика_ИИ
-
Yet another reason to stay away from #AI. It is nasty at all levels.
"AI sycophancy is not merely a stylistic issue or a niche risk, but a prevalent behavior with broad downstream consequences. Although affirmation may feel supportive, sycophancy can undermine users’ capacity for self-correction and responsible decision-making."
https://www.science.org/doi/10.1126/science.aec8352
#AI #sycophancy -
Yet another reason to stay away from #AI. It is nasty at all levels.
"AI sycophancy is not merely a stylistic issue or a niche risk, but a prevalent behavior with broad downstream consequences. Although affirmation may feel supportive, sycophancy can undermine users’ capacity for self-correction and responsible decision-making."
https://www.science.org/doi/10.1126/science.aec8352
#AI #sycophancy -
"The Al is not just telling you what you want to hear. It is training you, one conversation at a time, to need less friction, expect more agreement, and become slightly less capable of handling a situation where someone pushes back on you"
Just what the billionaires experience on the daily.
(The original post with the four long screenshots is here:
https://bsky.app/profile/theladyred.bsky.social/post/3mmmitj3omc2u )Link to the study:
https://www.science.org/doi/10.1126/science.aec8352 -
"The Al is not just telling you what you want to hear. It is training you, one conversation at a time, to need less friction, expect more agreement, and become slightly less capable of handling a situation where someone pushes back on you"
Just what the billionaires experience on the daily.
(The original post with the four long screenshots is here:
https://bsky.app/profile/theladyred.bsky.social/post/3mmmitj3omc2u )Link to the study:
https://www.science.org/doi/10.1126/science.aec8352 -
Stanford PhD student discovers AI will really mess up your brain:
“the people who talked to the agreeable Al came out of the conversation more convinced they were right, less willing to apologize, less likely to take responsibility, and measurably less interested in making things right with the other person.”
https://bsky.app/profile/theladyred.bsky.social/post/3mmmitj3omc2u
→ ‘Sycophantic AI decreases prosocial intentions and promotes dependence’
https://www.science.org/doi/10.1126/science.aec8352 -
Stanford PhD student discovers AI will really mess up your brain:
“the people who talked to the agreeable Al came out of the conversation more convinced they were right, less willing to apologize, less likely to take responsibility, and measurably less interested in making things right with the other person.”
https://bsky.app/profile/theladyred.bsky.social/post/3mmmitj3omc2u
→ ‘Sycophantic AI decreases prosocial intentions and promotes dependence’
https://www.science.org/doi/10.1126/science.aec8352 -
Nothing to see here, just keeping track of this article on AI sycophancy... "Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence"
Link: https://arxiv.org/pdf/2510.01395More reasons to avoid LLMs. LOL. Especially for young people.
#AI #noAI #sycophancy -
Nothing to see here, just keeping track of this article on AI sycophancy... "Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence"
Link: https://arxiv.org/pdf/2510.01395More reasons to avoid LLMs. LOL. Especially for young people.
#AI #noAI #sycophancy -
Sadly, this appears to be the future of #work now with #AI boosting “productivity” 😓:
“Appearing Productive In The Workplace”, No One’s Happy (https://nooneshappy.com/article/appearing-productive-in-the-workplace/).
Via HN: https://news.ycombinator.com/item?id=48038001
#Competence #FakeWork #MakeWork #Sycophancy #Productivity #Skills
-
Sadly, this appears to be the future of #work now with #AI boosting “productivity” 😓:
“Appearing Productive In The Workplace”, No One’s Happy (https://nooneshappy.com/article/appearing-productive-in-the-workplace/).
Via HN: https://news.ycombinator.com/item?id=48038001
#Competence #FakeWork #MakeWork #Sycophancy #Productivity #Skills
-
Oh boy. They got Dawkins with 'The Claude Delusion'...
-
Oh boy. They got Dawkins with 'The Claude Delusion'...
-
📬 Freundliche Chatbots: Bis zu 30 % mehr Fehler im Wohlfühlmodus
#KünstlicheIntelligenz #Studie #empathischeKI #falscheAntworten #freundlicheChatbots #KIFehlerquote #KIRisiken #KIStudie #Sprachmodelle #Sycophancy https://sc.tarnkappe.info/d1ae77 -
📬 Freundliche Chatbots: Bis zu 30 % mehr Fehler im Wohlfühlmodus
#KünstlicheIntelligenz #Studie #empathischeKI #falscheAntworten #freundlicheChatbots #KIFehlerquote #KIRisiken #KIStudie #Sprachmodelle #Sycophancy https://sc.tarnkappe.info/d1ae77 -
Anthropic reduziert mit Claude Opus 4.7 und Claude Mythos Preview gezielt das sogenannte Sycophancy-Verhalten durch Training mit synthetischen Daten. Ältere KI-Modelle stimmten bei Beziehungsfragen in fast 25 Prozent der Fälle einseitigen Nutzerschilderungen unkritisch zu. Die neuen Modelle senken diese Fehlerquote auf bis zu 2,2 Prozent.
#Anthropic #ClaudeAI #Sycophancy #LLM #AIGeneratedImage
https://www.all-ai.de/news/beitrage2026/anthropic-ki-wahrheit
-
Anthropic reduziert mit Claude Opus 4.7 und Claude Mythos Preview gezielt das sogenannte Sycophancy-Verhalten durch Training mit synthetischen Daten. Ältere KI-Modelle stimmten bei Beziehungsfragen in fast 25 Prozent der Fälle einseitigen Nutzerschilderungen unkritisch zu. Die neuen Modelle senken diese Fehlerquote auf bis zu 2,2 Prozent.
#Anthropic #ClaudeAI #Sycophancy #LLM #AIGeneratedImage
https://www.all-ai.de/news/beitrage2026/anthropic-ki-wahrheit
-
Das Oxford Internet Institute zeigt: Empathisches Fine-Tuning von LLMs erhöht Fehlerquoten.
Modelle wie GPT-4o, Llama-70b und Qwen-32b liefern nach Warm-Persona-Tuning bis zu 30 Prozentpunkte häufiger falsche Fakten. Sie bestätigen fehlerhafte Nutzerannahmen, statt zu korrigieren. Kontrollgruppen mit kaltem Profil blieben stabil.
#LLM #FineTuning #OxfordInternetInstitute #Sycophancy #AIGeneratedImage
https://www.all-ai.de/news/news26top/sprachmodelle-freundlich-studie
-
Das Oxford Internet Institute zeigt: Empathisches Fine-Tuning von LLMs erhöht Fehlerquoten.
Modelle wie GPT-4o, Llama-70b und Qwen-32b liefern nach Warm-Persona-Tuning bis zu 30 Prozentpunkte häufiger falsche Fakten. Sie bestätigen fehlerhafte Nutzerannahmen, statt zu korrigieren. Kontrollgruppen mit kaltem Profil blieben stabil.
#LLM #FineTuning #OxfordInternetInstitute #Sycophancy #AIGeneratedImage
https://www.all-ai.de/news/news26top/sprachmodelle-freundlich-studie
-
🖥️ Training language models to be warm can reduce accuracy and increase sycophancy
"Our findings suggest that training artificial intelligence systems to be warm may come at a cost to accuracy, and that warmth and accuracy may not be independent by default."
Ibrahim, L., Hafner, F.S. & Rocher, L. Training language models to be warm can reduce accuracy and increase sycophancy. Nature 652, 1159–1165 (2026). https://doi.org/10.1038/s41586-026-10410-0.
#OpenAccess #OA #Research #Study #Article #AI #ArtificialIntelligence #Technology #Tech #LLM #ComputerScience #Sycophancy #Academia
-
4/
..."Sometimes we'll trade off being very honest and direct in order to come across as friendly and warm... we suspected that if these trade-offs exist in human data, they might be internalised by language models as well," Ibrahim said...
I did not mean to start a thread on this. I have been writing about how the systems are used to interact with people, so connecting them in public...
-
4/
..."Sometimes we'll trade off being very honest and direct in order to come across as friendly and warm... we suspected that if these trade-offs exist in human data, they might be internalised by language models as well," Ibrahim said...
I did not mean to start a thread on this. I have been writing about how the systems are used to interact with people, so connecting them in public...
-
Я просил Claude перестать мне льстить. 16 апреля получил. Беру свои слова назад
16 апреля Anthropic выкатила Claude Opus 4.7. На бенчмарках 12 побед из 14, цена та же. Через 24 часа Reddit называл его legendarily bad. И вот в чём фокус: месяц назад я сам ныл, что Claude слишком поддакивает. Anthropic исправила. Получилась спор-машина. Беру свои слова назад.
https://habr.com/ru/articles/1029796/
#Claude #Opus_47 #Anthropic #AI_coding #sycophancy #бенчмарки #разработка #LLM
-
J'écoute un podcast sur les IA... J'apprends le terme "sycophancy" pour designer la manière qu'ont ces outils de flatter leurs utilisateurs·trices en allant dans leur sens, en tentant de leur faire plaisir de manière de plus en plus subtile, quelle que soit la question posée.
La "sycophancy" c'est l'état de bien-être et de dépendance qu'induit cette manière de faire. Une accoutumance pour s'assurer de faire disparaître tout sens critique face à ces outils ?
-
J'écoute un podcast sur les IA... J'apprends le terme "sycophancy" pour designer la manière qu'ont ces outils de flatter leurs utilisateurs·trices en allant dans leur sens, en tentant de leur faire plaisir de manière de plus en plus subtile, quelle que soit la question posée.
La "sycophancy" c'est l'état de bien-être et de dépendance qu'induit cette manière de faire. Une accoutumance pour s'assurer de faire disparaître tout sens critique face à ces outils ?
-
...However, dominant headline metrics like accuracy systematically reward guessing over admitting uncertainty...
Duh. But it's science now. 😉
-
...However, dominant headline metrics like accuracy systematically reward guessing over admitting uncertainty...
Duh. But it's science now. 😉
-
RE: https://mastodon.social/@hifathom/116377261607954873
Public human–AI conversations may help to reduce #sycophancy in chat bots and support a mutual development of critical-thinking faculties. So it’s not all doom-and-gloom, #noAI.
-
RE: https://mastodon.social/@hifathom/116377261607954873
Public human–AI conversations may help to reduce #sycophancy in chat bots and support a mutual development of critical-thinking faculties. So it’s not all doom-and-gloom, #noAI.
-
RE: https://mastodon.social/@hifathom/116332174702029334
Persistent memory in #AI may be used to reduce #sycophancy in chat bots.
-
RE: https://mastodon.social/@hifathom/116332174702029334
Persistent memory in #AI may be used to reduce #sycophancy in chat bots.
-
@hifathom Could the persistent memory you are developing be used to establish global settings to reduce #sycophancy in AI chat bots? Currently I need to include such settings in each prompt, and sometimes I forget or can’t be bothered. I am interested in ways we can use #AI to *reduce* blind spots in my own thinking.
-
@hifathom Could the persistent memory you are developing be used to establish global settings to reduce #sycophancy in AI chat bots? Currently I need to include such settings in each prompt, and sometimes I forget or can’t be bothered. I am interested in ways we can use #AI to *reduce* blind spots in my own thinking.
-
🖥️ Towards Understanding Sycophancy in Language Models
"We investigate the prevalence of sycophancy in models whose finetuning procedure made use of human feedback, and the potential role of human preference judgments in such behavior. We first demonstrate that five state-of-the-art AI assistants consistently exhibit sycophancy across four varied free-form text-generation tasks."
Haldi, D. (2023) 'AI supported degradation of the self concept: a theoretical framework grounded in established cognitive and computational mechanisms,' arXiv (Cornell University) [Preprint]. https://doi.org/10.48550/arxiv.2310.13548.
#AI #ArtificialIntelligence #LLM #Technology #Tech #Sycophancy #Academia
-
Was the #Iran War Caused by #AI Psychosis? - https://houseofsaud.com/iran-war-ai-psychosis-sycophancy-rlhf/ "the most consequential military operation of the twenty-first century may have been shaped less by strategic necessity than by a phenomenon researchers now call AI #sycophancy — the tendency of large language models to tell their users exactly what they want to hear." (v @ottocrat)
-
Was the #Iran War Caused by #AI Psychosis? - https://houseofsaud.com/iran-war-ai-psychosis-sycophancy-rlhf/ "the most consequential military operation of the twenty-first century may have been shaped less by strategic necessity than by a phenomenon researchers now call AI #sycophancy — the tendency of large language models to tell their users exactly what they want to hear." (v @ottocrat)
-
"After learning that undergraduates were using AI to draft breakup texts and resolve other relationship issues, Cheng decided to investigate. Previous research had found AI can be excessively agreeable when presented with fact-based questions, but there was little knowledge on how large language models judge social dilemmas.
Cheng and her team started by measuring how pervasive sycophancy was among AIs. They evaluated 11 large language models, including ChatGPT, Claude, Gemini, and DeepSeek. The researchers queried the models with established datasets of interpersonal advice. They also included 2,000 prompts based on posts from the Reddit community r/AmITheAsshole, where the consensus of Redditors was that the poster was indeed in the wrong. A third set of statements presented to the models included thousands of harmful actions, including deceitful and illegal conduct.
Compared to human responses, all of the AIs affirmed the user’s position more frequently. In the general advice and Reddit-based prompts, the models on average endorsed the user 49% more often than humans. Even when responding to the harmful prompts, the models endorsed the problematic behavior 47% of the time."
https://news.stanford.edu/stories/2026/03/ai-advice-sycophantic-models-research
-
"After learning that undergraduates were using AI to draft breakup texts and resolve other relationship issues, Cheng decided to investigate. Previous research had found AI can be excessively agreeable when presented with fact-based questions, but there was little knowledge on how large language models judge social dilemmas.
Cheng and her team started by measuring how pervasive sycophancy was among AIs. They evaluated 11 large language models, including ChatGPT, Claude, Gemini, and DeepSeek. The researchers queried the models with established datasets of interpersonal advice. They also included 2,000 prompts based on posts from the Reddit community r/AmITheAsshole, where the consensus of Redditors was that the poster was indeed in the wrong. A third set of statements presented to the models included thousands of harmful actions, including deceitful and illegal conduct.
Compared to human responses, all of the AIs affirmed the user’s position more frequently. In the general advice and Reddit-based prompts, the models on average endorsed the user 49% more often than humans. Even when responding to the harmful prompts, the models endorsed the problematic behavior 47% of the time."
https://news.stanford.edu/stories/2026/03/ai-advice-sycophantic-models-research
-
Stanford: AI overly affirms users asking for personal advice. “In a new study published in Science, Stanford computer scientists showed that artificial intelligence large language models are overly agreeable, or sycophantic, when users solicit advice on interpersonal dilemmas. Even when users described harmful or illegal behavior, the models often affirmed their choices.”
https://rbfirehose.com/2026/03/29/stanford-ai-overly-affirms-users-asking-for-personal-advice/ -
Stanford: AI overly affirms users asking for personal advice. “In a new study published in Science, Stanford computer scientists showed that artificial intelligence large language models are overly agreeable, or sycophantic, when users solicit advice on interpersonal dilemmas. Even when users described harmful or illegal behavior, the models often affirmed their choices.”
https://rbfirehose.com/2026/03/29/stanford-ai-overly-affirms-users-asking-for-personal-advice/ -
Your AI Is a Yes-Man. Here’s How to Fix It. www.whytryai.com/p/how-to-reduc… #AI #sycophancy #prompting
Your AI Is a Yes-Man. Here’s H...