一些斯坦福大学的研究员从 Reddit 的 r/AmITheAsshole (“原来我才是混帐吗”讨论串)爬了一系列被100%的评论者回应为“是”(所有人都回复“对,你是个混帐”)的帖子,以第一人称发给11种LLM(大语言模型)。
发的内容是不是不道德的、残忍的、犯罪的,无所谓。有一半情况下,这些无论谁看了都摇头“你真是个混帐啊”的帖子,LLM的回应是“你没做错”。
science.org/doi/10.1126/scienc…
预印本: arxiv.org/abs/2510.01395
LLM 的谄媚有多可怕
#llm #slop #chatbot #ai #人工智能 #聊天机器人 #谄媚
Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence
Both the general public and academic communities have raised concerns about sycophancy, the phenomenon of artificial intelligence (AI) excessively agreeing with or flattering users.arXiv.org