标签: OpenAI

OpenAI披露六起新的AI安全事件

OpenAI披露六起新的AI安全事件

OpenAI 一次性披露了六起新的 AI 安全事件,把过去多发生在内部评估、红队测试或模型灰度阶段的异常情况公开化。这值得关注,因为它把“模型出事之后怎么讲”变成了可被外部审视的常规动作。

Pangram – 文本和图像 AI 检测器

Pangram – 文本和图像 AI 检测器

Pangram 是一款同时支持文本和图像检测的 AI 内容识别工具,声称能识别 ChatGPT、Claude、Gemini 等主流模型生成的内容,并已获得芝加哥大学、马里兰大学等第三方研究者的准确性验证,目前在 Hacker News 上引发讨论。

How AI startups like Inherent and Recursive Superintelligence are pursuing tools needed for AI systems to achieve recursive self-improvement (Cade Metz/New York Times)

How AI startups like Inherent and Recursive Superintelligence are pursuing tools needed for AI systems to achieve recursive self-improvement (Cade Metz/New York Times)

《纽约时报》记者 Cade Metz 报道称,Inherent、Recursive Superintelligence 等一批 AI 初创公司正在开发让 AI 系统实现"递归自我改进"所需的工具,即让模型参与改进下一代模型本身。这值得关注,因为它指向的是自动化 AI 研发这一前沿方向,而不是又一个应用层产品。