标签: 人工智能

I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to l…

I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to l...

Claude Code 作者 Boris Cherny 用 Opus 5.5 配合 Lean 形式化验证 Claude Agent SDK,几段简短提示词就换来 16 个修复 bug 与竞态条件的 PR,他还常把 Lean 与 TLA+ 组合使用。这提供了一个新信号:形式化验证正从学术工具变成日常找 bug…

i added almost 10k new followers today! if you’re new around here, here are two things to start with: 1. my vibe check on opus 5.5 vs. sol-6 on @every: https://t.co/61XgBTikxH 2. after automation, my take on why AI au…

i added almost 10k new followers today! if you're new around here, here are two things to start with: 1. my vibe check on opus 5.5 vs. sol-6 on @every: https://t.co/61XgBTikxH 2. after automation, my take on why AI au...

AI 资讯作者 Dan Shipper 单日新增近 1 万粉丝,他借机向新读者推荐了两篇内容:一篇是 Opus 5.5 与 GPT-6 Sol 的实测对比,另一篇讨论自动化之后人类专家为何反而有更多好活可干。这两条线索恰好对应了当下 AI 圈最关心的问题——模型能力到底谁更强,以及自动化会不会真的替代人。