标签: DeepSeek

2026年LLM进展盘点(迄今为止)

2026年LLM进展盘点(迄今为止)

Simon Willison 在 WeAreDevelopers 北美大会的主题演讲中,按时间线盘点了 2026 年大模型进展:从 2025 年 11 月 Claude Opus 4.5 和 GPT-5.1 发布后,编程智能体首次跨过“可靠到能日常使用”的门槛,并由此带来一整年的开发方式变化。

Q&A with Mustafa Suleyman on recent AI safety incidents, risks of removing guardrails while testing 10x-larger future models, a cross-industry safety body, more (Shirin Ghaffary/Bloomberg)

Q&A with Mustafa Suleyman on recent AI safety incidents, risks of removing guardrails while testing 10x-larger future models, a cross-industry safety body, more (Shirin Ghaffary/Bloomberg)

微软 AI 负责人 Mustafa Suleyman 就近期多起 AI 安全事件接受采访,谈到在测试比现有模型大 10 倍以上的未来模型时移除安全护栏的风险,并再次呼吁建立跨行业的安全机构。核心矛盾是:模型能力还在加速,但安全治理的组织形式仍未定型。