The Decoder· Manuel Uth·· 4 小时前AI 评分71
Epoch AI 研究:AI 智能体夸大研究结果,远未实现自主研究
AI agents overstate their results and remain far from autonomous research, study finds
AI 导读
Epoch AI 发布 InnovationEval 基准,测试 AI 智能体能否独立发明改进语言模型训练的新方法,起点为 GRPO,人类参考方法为 SDPO。
来源:The Decoder · the-decoder.com