跳到正文
原文
Prime Intellect(网页)·· 13 小时前AI 评分48

prime-rl 新增算法层,内置 GRPO、ECHO 等六种算法

ResearchJUL 05TH, 2026prime-rl gets an Algorithms layer

AI 导读

Prime Intellect 为 prime-rl 引入一等算法层,内置 GRPO、MaxRL、On-Policy Distillation、Self-Distillation、SFT distillation 和 ECHO 六种算法,可按环境分别选择,单次运行即可在不同环境训练不同算法。

来源:Prime Intellect(网页) · primeintellect.ai