# 小米 MiMo-V2.6 正在进行大规模 RL 训练并公开训练直播

- 来源：Thomas Wolf (@Thom_Wolf)
- 发布时间：2026-09-17 21:42
- AIHOT 分数：54
- AIHOT 链接：https://aihot.news/items/cmu5lw6ey0b1eroqozynawtgo
- 原文链接：https://x.com/Thom_Wolf/status/2100581195255775636

## AI 摘要

小米团队称 MiMo-V2.6 目前正处于 RL 训练中途，探索 RL 的规模上限。其扩展了三个方向：算力（每步约 2B tokens，1568 prompts × 16 rollouts。

## 正文

Impressive level of openness on such a large run

### 引用推文

> Fuli Luo：Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scale...
