标签: Claude

We need to teach the world to prompt and maximally use AI: to help all people see the way it can give you wings in all your pursuits Then we need yearn to solve more problems for ourselves and for one another https://…

We need to teach the world to prompt and maximally use AI: to help all people see the way it can give you wings in all your pursuits Then we need yearn to solve more problems for ourselves and for one another https://...

Y Combinator 总裁 Garry Tan 在 X 上提出,应当教会更多人写好 prompt、最大化使用 AI,并借此培养主动解决问题的意愿。这一表态把当前 AI 讨论的焦点从"模型能力"拉回到"人的使用能力"上。

We ran fresh Next.js evals. The tally: ① Opus 5.5 [𝟿𝟽%] ② GPT 6 Sol [𝟿𝟽%] ③ Fable 5.1 [𝟿𝟽%] ④ Grok 4.7 [𝟿𝟺%] Notably, Grok is 2x-7x cheaper https://t.co/BAUg14981G

We ran fresh Next.js evals. The tally: ① Opus 5.5 [𝟿𝟽%] ② GPT 6 Sol [𝟿𝟽%] ③ Fable 5.1 [𝟿𝟽%] ④ Grok 4.7 [𝟿𝟺%] Notably, Grok is 2x-7x cheaper https://t.co/BAUg14981G

Vercel CEO Guillermo Rauch 公布了新一轮 Next.js 代码评测结果,四款前沿模型中有三款同时达到 97% 成功率,而 Grok 4.7 以 94% 紧随其后,但调用成本低 2 到 7 倍。这组数据把“顶级能力是否必须付顶级价格”这个老问题重新摆上了桌面。

I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to l…

I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to l...

Claude Code 作者 Boris Cherny 用 Opus 5.5 配合 Lean 形式化验证 Claude Agent SDK,几段简短提示词就换来 16 个修复 bug 与竞态条件的 PR,他还常把 Lean 与 TLA+ 组合使用。这提供了一个新信号:形式化验证正从学术工具变成日常找 bug…

i added almost 10k new followers today! if you’re new around here, here are two things to start with: 1. my vibe check on opus 5.5 vs. sol-6 on @every: https://t.co/61XgBTikxH 2. after automation, my take on why AI au…

i added almost 10k new followers today! if you're new around here, here are two things to start with: 1. my vibe check on opus 5.5 vs. sol-6 on @every: https://t.co/61XgBTikxH 2. after automation, my take on why AI au...

AI 资讯作者 Dan Shipper 单日新增近 1 万粉丝,他借机向新读者推荐了两篇内容:一篇是 Opus 5.5 与 GPT-6 Sol 的实测对比,另一篇讨论自动化之后人类专家为何反而有更多好活可干。这两条线索恰好对应了当下 AI 圈最关心的问题——模型能力到底谁更强,以及自动化会不会真的替代人。