註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Peano Labs
@peano_ai
Intelligence, Compiled.
加入 September 2026
1 正在關注    250 粉絲
We enable full-parameter RL on TPUs: MiMo-V2.6 at 310B, plus other stable training runs of 1,000+ steps across 1,000+ TPUs. With JAX, scaling up is a config change, not a rewrite. We built on that with optimized vLLM inference for faster rollouts and full bitwise trainer–sampler agreement in validation. Trainer and sampler share one TPU ICI fabric. All 310B MiMo-V2.6 parameters transfer in <2 seconds.
顯示更多
0
4
62
15
轉發到社區