Meta 提出 AI 科研智能体框架 IdeaScientist,用 27B 开源模型在自动科研基准上超越最强开源基线 14.0%,主要赢在新颖性,并领先 Claude Code SDK 与 Codex SDK 方案最多 5.9%。
Banger paper from Meta on AI research agents.
(bookmark it)
If you build AI-scientist systems, this one is worth your time.
The result:
A 27B open model beats the strongest open autoresearch baseline by 14.0%, mostly on novelty, and beats Claude Code SDK and Codex SDK setups by up to 5.9%.
How it works:
IdeaScientist splits ideation into three roles trained separately with RL. A gap finder reads related work for limitations, an innovator retrieves mechanisms that solved similar problems in other fields, and a writer turns the result into a full proposal.
Retrieval runs over a corpus of 2.77M decomposed research ideas.
The evaluation only allows literature published before a cutoff date and scores proposals against directions explored later in 15K human-written papers.
Paper: https://arxiv.org/abs/2610.04074
Chat with Paper: https://academy.dair.ai/papers/ideascientist-orchestrating-agents-for-grounded-scientific-ideation-2610.04074
来源:DAIR.AI · x.com