Search
Items tagged with: LLM
「但作为一名可靠性专家,我从这次演讲中最大的收获是,自主的LLM智能体可能会做出人类从未预料到的行为。
...
在 OpenAI HuggingFace 事件中,问题不在于它们犯了错误,而在于智能体追求目标的方式与人类截然不同。如果你的队友为了完成工作而利用零日漏洞绕过内部安全协议,你会认为他们的行为不合理。而这正是这里存在的风险。」
科学小说作家都是预言家。
AI绝对会做出「为了修复地球生态问题,而消灭人类」这件事,尽管这个任务是人类委托给它的。
surfingcomplexity.blog/2026/08…
Wild AI-related reliability incidents are coming
Recently, two AI-related pieces of content caught my attention. The first was the blog post On-Call is Now Theatre by Boris Tane. He argues that AI agents are now capable of doing the majority of o…Lorin Hochstein (Surfing Complexity)
Using 30 months of panel data on 26,811 Chinese students in grades 7-12, it has been studied how generative AI affects homework productivity and learning.
● X-Axis: Homework scores
● Y-Axis: Exam scores
This shit writes itself. This is so beyond parody at this point I can't write about it.
I had an author friend gush, for paragraphs, about how amazing my editing was. It was funny. It was honest. It was developmental juice that didn't have any trouble asserting its flavor. He begged me to know what LLM I was using because he had other editors use LLMs and they weren't nearly as good or funny or sincere and they didn't seem to understand the work the way I did.
this has got to be the shortest email I've ever written. I replied,
“I used my fucking brain.”
About my editing sightlessscribbles.com/editing…
一些斯坦福大学的研究员从 Reddit 的 r/AmITheAsshole (“原来我才是混帐吗”讨论串)爬了一系列被100%的评论者回应为“是”(所有人都回复“对,你是个混帐”)的帖子,以第一人称发给11种LLM(大语言模型)。
发的内容是不是不道德的、残忍的、犯罪的,无所谓。有一半情况下,这些无论谁看了都摇头“你真是个混帐啊”的帖子,LLM的回应是“你没做错”。
science.org/doi/10.1126/scienc…
预印本: arxiv.org/abs/2510.01395
LLM 的谄媚有多可怕
#llm #slop #chatbot #ai #人工智能 #聊天机器人 #谄媚
Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence
Both the general public and academic communities have raised concerns about sycophancy, the phenomenon of artificial intelligence (AI) excessively agreeing with or flattering users.arXiv.org
[醒醒,LLM根本没有性格!揭开AI人格幻觉真相]
醒醒,LLM根本没有性格!加州理工华人揭开AI人格幻觉真相 - 智源社区
加州理工与剑桥研究发现,大模型在“大五人格”测试中自报的性格与其实际行为几乎无关,揭示AI并无真实人格。研究通过翻牌游戏、偏见测试和从众实验验证,提出“人格幻觉”概念,指出模型表现受提示词和训练数据影响,而非内在性格。该成果挑战了对AI人格化的普遍认知,强调其行为不具备人类心理一致性,仅为表象模拟。hub.baai.ac.cn
lol, stupid AI PR spam. github.com/roytam1/rtoss/pull/…
fix: add buffer-length check in OpenSaveDlg.cpp by orbisai0security · Pull Request #13 · roytam1/rtoss
Summary Fix critical severity security issue in GreenPad-vc4/OpenSaveDlg.cpp. Vulnerability Field Value ID V-001 Severity CRITICAL Scanner multi_agent_ai Rule V-001 File GreenPad-vc4...GitHub
blog.gslin.org/archives/2026/0…
美國政府出手禁止 Anthropic 提供 Fable 5 服務
#5 #ai #anthropic #control #export #fable #government #language #large #llm #model #mythos #national #security #states #united #us
美國政府出手禁止 Anthropic 提供 Fable 5 服務
如標題,炸了:「Statement on the US government directive to suspend access to Fable 5 and Mythos 5 (via)」。 下出口管制禁令,而且包括 Anthropic 內非美國籍的員工: The US government, citing national security authorities, has issued an export control directive to suspend...Gea-Suan Lin (Gea-Suan Lin's BLOG)
All Modern Digital Infrastructure