Q&A with AI researchers John Schulman, Beren Millidge, and Charlie O’Neill on steelmanning the case against RSI, Chinese labs’ progress, long-horizon RL, more (Dwarkesh Patel/Dwarkesh Podcast)

Dwarkesh Patel 的播客请来 John Schulman、Beren Millidge、Charlie O'Neill 三位研究者,用"反向论证"的方式讨论递归自我改进(RSI)是否真会失控,同时聊到中国实验室的进展和长时程强化学习。这场对话值得关注,是因为它把常被当成信仰的 AI 加速论,重新拉…








