7月31日 | AI 帮自己变便宜,这个回路第一次发生
本期内容
今天五件事,覆盖了 AI 行业正在发生的几个不同层面的变化。OpenAI 用模型帮助优化自身架构、主动压低 API 价格;DeepSeek 轻量新模型在智能体基准上超越了自家旗舰;DeepMind 展示了机器人的全身协调控制;Science 期刊记录了顶级 AI 公司几乎停止发表研究论文的趋势;Farnam Street 提出了一个在 AI 时代更难处理的老问题:如何区分真正的专家和会说行话的人。听完这期,你对 AI 工具选型、成本重算、以及如何评估信息来源,都会有一些新的判断依据。
本期要点
- GPT-5.6 推出三档变体,价格最低砍八成,这次效率提升由模型自身参与完成,开发者应重新核算现有项目成本
- DeepSeek-V4-Flash 在智能体基准上超越更重的 V4-Pro,轻量快速不再等于能力妥协
- Gemini Robotics 2 引入全身智能控制框架,机器人能在非结构化环境里自主协调,语言模型开始成为物理行动的大脑
- Science 期刊分析显示顶级 AI 创业公司研究发表量急剧下降,技术信息不对称正在影响第三方安全评估能力
- Farnam Street 的专家与模仿者框架在 AI 时代更为关键,追问"结论怎么来的"比"结论是什么"更重要
参考资料
Advancing the price-performance frontier with GPT-5.6 — https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/
Building abundant intelligence — https://openai.com/index/building-abundant-intelligence/
Gemini Robotics 2 brings whole body intelligence to robots — https://deepmind.google/blog/
DeepSeek-V4-Flash Update — https://platform.deepseek.com/
AI's top startups are barely publishing their research (HN #49103285) — https://news.ycombinator.com/item?id=49103285
Experts vs. Imitators — https://fs.blog/
---
BearTalk 狗熊有话说播客,始于 2012 年。
订阅地址:https://beartalking.com/page/podcast
