Prime Intellect(网页)·· 15 小时前精选AI 评分75
Prime Intellect 用 nanoGPT speedrun 评测 18 个前沿模型的自主研究能力
ResearchAUG 14TH, 2026Measuring Autonomous AI Research
AI 导读
Prime Intellect 在 nanoGPT optimizer speedrun 上开展 153 次自主运行、覆盖 18 个前沿模型,每次运行使用 8xH200s、最长持续八天,baseline 为 3,290 steps,人类纪录为 2,600。
推荐理由
实验规模和全部 trace 公开,模型差距主要来自实验执行而非想法本身,这个发现值得一看。
来源:Prime Intellect(网页) · primeintellect.ai