跳到正文
原文
ModelScope· @ModelScope2022 · X·· 2 天前AI 评分44
AI 导读

上海AI实验室推出Intern-Decision系列,覆盖0.8B、2B、4B三个规模,在七项决策基准上平均得分分别为79.38、84.68和90.02,其中4B版本以90.02超越Jev的88.74且概率校准更优。

正文

Intern-Decision scales multimodal structured decision-making across 0.8B, 2B, and 4B models.
🤖 https://modelscope.ai/collections/Shanghai_AI_Laboratory/Intern-Decision

🏆 The three models average 79.38, 84.68, and 90.02 across seven decision benchmarks. Intern-Decision-4B surpasses Jev at 88.74 while achieving better probability calibration.
⚡ Reported mean latency is 33.98 ms, 33.28 ms, and 44.16 ms respectively, compared with 109.70 ms for Jev in the same local HF setup.
🧠 One forward pass answers multiple choice, score, and yes/no questions with typed JSON and calibrated probability distributions.
🖼️ Text states and up to eight images enable routing, tool selection, scoring, and multimodal workflow control.
📜 Apache 2.0. Applicable Qwen3.5 upstream notices must be retained.

来源:ModelScope · x.com