标签: 推理

GenRec:Netflix 的 LLM 原生推荐之路

GenRec:Netflix 的 LLM 原生推荐之路

Netflix 近期预告了名为 GenRec 的“LLM 原生”推荐系统,尝试用大语言模型重新构建整套推荐流程。HN 社区讨论认为,价值不在“猜得更准”,而在于 LLM 能把非结构化内容和用户偏好变成可组合的语义特征,但商业层面的动机也受到质疑。

not really you can be an AI-native rocketship without being in a permanent fundraising situation and in a constant death match with other rocketships to capture customers, sacrificing gross margin but the rules for bu…

not really you can be an AI-native rocketship without being in a permanent fundraising situation and in a constant death match with other rocketships to capture customers, sacrificing gross margin but the rules for bu...

AI 投资人 Matt Turck 认为创业公司要么成为“AI 原生火箭船”、永远处于融资和抢客户的死战中,要么出局;Dan Shipper 直接反驳:可以成为 AI 原生公司,但不烧钱、不牺牲毛利、不卷入死亡竞赛,只是需要一套完全不同的打法。

India central bank head says AI can approve loans humans would have turned down, says the technology is, ‘a capability to be responsibly harnessed and not merely as a risk to be contained’

India central bank head says AI can approve loans humans would have turned down, says the technology is, 'a capability to be responsibly harnessed and not merely as a risk to be contained'

印度央行行长公开表态,要求银行将 AI 纳入信贷审批核心流程,认为 AI 能识别人工审核难以覆盖的借款人群体,同时也明确责任仍须由银行而非算法承担——这是全球主要央行对 AI 金融应用一次罕见的系统性定调。

I was curious how X is fighting AI slop, so I looked at its open-source algorithm. It has a behavioral model called TweetSpamBot that analyzes up to 512 recent account actions, looking at signals like posting bursts,…

I was curious how X is fighting AI slop, so I looked at its open-source algorithm. It has a behavioral model called TweetSpamBot that analyzes up to 512 recent account actions, looking at signals like posting bursts,...

X 在 8 月 13 日开源 For You 时间线排序代码后,开发者发现其反 AI 垃圾内容主要依赖一个名为 TweetSpamBot 的行为评分模型,能识别高频发帖、批量引用转推等模式,但几乎不分析帖子文本,导致刻意包装的 AI 内容农场仍能绕过限制。

人工智能驱动测试

人工智能驱动测试

有开发者写了一套通过 Android 无障碍服务将设备操作暴露给 AI 模型的方案,让 Hermes 模型能在模拟器(Waydroid)上完成滑动、打开应用、定位问题等测试动作,展示了 AI 驱动端到端测试的实际可行性,也暴露了自动化在组织层面的真实阻力。

GitHub Copilot 每周发布(8月10日)

GitHub Copilot 每周发布(8月10日)

GitHub Copilot 在 8 月 10 日这波更新里,把重心从“多一个模型”转向“可组合的代理工作流”:Kimi K3 与 MAI-Code 新模型上线、Agent Plugins 1.0 正式可用,并同步增强了 CLI 和 JetBrains 端的智能体能力。这次更新的核心信号是:Copilot 正…

Kog 深挖 GPU,榨取更多推理性能

Kog 深挖 GPU,榨取更多推理性能

法国初创公司 Kog 正尝试用软件优化从企业现有的 GPU 中榨取更多 AI 推理性能,其演示达到了单请求 3000 tokens/秒的解码速度,但验证模型是仅 20 亿参数的小模型;这套方法能否在大模型上复制,是它下一步的关键考验。