
18期


你有没有想过,聪明的AI也会犯傻,甚至像个没头脑的实习生?本期节目,我们就来聊聊如何让AI变得更“靠谱”。我们将一起看看,科学家们如何用AI工具去解决古老的数学难题,如何洞悉AI群体的“集体意识”,是会变得更聪明还是更固执,以及如何教会AI拥有一个好记性,并像人一样学会“反思”自己。
00:00:28 给你一把新扳手,拧紧一颗老螺丝
00:05:56 AI的“集体意识”,乌合之众还是三个臭皮匠?
00:10:51 如何才能拥有一个好记性?
00:15:37 为什么聪明的AI,干起活来却像个“没头脑”?
00:21:07 给AI立规矩,为什么不能靠“死命令”?
本期介绍的几篇论文:
[LG] Improving the matrix multiplication exponent with modern optimization and AlphaEvolve
[Google DeepMind]
https://arxiv.org/abs/2608.16884
---
[AI] Physics of Agents: Statistical Mechanics Predicts Collective Behavior of AI Agents
[Stanford University & UC Santa Barbara]
https://arxiv.org/abs/2608.16578
---
[LG] Proteus: Incremental Memory Activation for Long-Context Sequence Modeling
[Mila & Google]
https://arxiv.org/abs/2608.16844
---
[CL] How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks
[Prentis AI]
https://arxiv.org/abs/2608.14905
---
[CL] CAPO: Constraint-Aware Prompt Optimization for LLM Agents
[Microsoft]
https://arxiv.org/abs/2608.16068


3期

沪ICP备06026464号-4 网络文化经营许可证
沪网文[2014]0587-137号
信息网络传播视听许可证:0911603
©2011-2019 qingting.fm ALL Rights Reserved.
应用名称:蜻蜓FM | 开发者:上海麦克风文化传媒有限公司