跳到正文
原文
Prime Intellect(网页)·· 15 小时前精选AI 评分75

Prime Intellect 用 nanoGPT speedrun 评测 18 个前沿模型的自主研究能力

ResearchAUG 14TH, 2026Measuring Autonomous AI Research

AI 导读

Prime Intellect 在 nanoGPT optimizer speedrun 上开展 153 次自主运行、覆盖 18 个前沿模型,每次运行使用 8xH200s、最长持续八天,baseline 为 3,290 steps,人类纪录为 2,600。

推荐理由

实验规模和全部 trace 公开,模型差距主要来自实验执行而非想法本身,这个发现值得一看。

来源:Prime Intellect(网页) · primeintellect.ai