# 小米MiMo-V2.6开源RL训练细节，已投入超100万美元

- 来源：elvis (@omarsar0)
- 发布时间：2026-09-17 05:35
- AIHOT 分数：46
- AIHOT 链接：https://aihot.news/items/cmu4mmikr03jsrodczpno3bob
- 原文链接：https://x.com/omarsar0/status/2100337683277173009

## AI 摘要

小米MiMo-V2.6正进行大规模RL训练，从三个维度扩展：计算（每步约2B tokens、1568 prompts×16 rollouts、全异步）、环境与harness（多任务智能体RL，单次运行混合多个harness）以及评分计算（组内智能体信用分配，结合测试用例与rubric奖励）。团队表示将在未来数周逐步开源细节，目前投入已超100万美元。

## 正文

This should be the standard for building open-source AI.

$1M+ so far.

### 引用推文

> Fuli Luo：Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scale...
