1. 22 Sep 2026

    'Better than DeepSeek': Xiaomi's MiMo-V2.6-Pro debuts as the top open weights model in ...

    ... reinforcement-learning environments and now the training infrastructure behind them. From smartphones and EVs to frontier AI. Xiaomi's move into ...

    venturebeat.com ↗
  2. 22 Sep 2026

    MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement | alphaXiv

    Reinforcement learning for agents works by having the model attempt a task, receiving a signal about how well it did, and adjusting the numbers that ...

    www.alphaxiv.org ↗
  3. 22 Sep 2026

    Xiaomi open-sources MiMo-V2.6, publishes the bill for scaling reinforcement learning

    Chinese labs have spent the year arguing that reinforcement learning, not pre-training, holds the next gains. Xiaomi has now put a price on that ...

    www.digitimes.com ↗
  4. 22 Sep 2026

    Xiaomi open-sources MiMo-V2.6 models after scaling reinforcement learning - TechNode

    Xiaomi also released MiMo-V2.6-Distill-Qwen-9B and research resources for reinforcement learning. MiMo-V2.6-Pro scored 46 on the Artificial ...

    technode.com ↗
  5. 22 Sep 2026

    Xiaomi Steps Up AI Investment With New Models - Caixin Global

    Chinese tech giant Xiaomi Corp. on Tuesday launched and open-sourced its latest AI ... The new lineup, comprising of the multimodal MiMo-V2.6-Pro and ...

    www.caixinglobal.com ↗
  6. 22 Sep 2026

    Xiaomi Steps Up AI Investment With New Models - Caixin Global

    Xiaomi Steps Up AI Investment With New Models - The Chinese company talks up the stronger reinforcement learning capabilities of its MiMo-V2.6 ...

    www.caixinglobal.com ↗
  7. 22 Sep 2026

    Xiaomi releases MiMo-V2.6-Pro and MiMo-V2.6-Flash open-source AI models with scaled ...

    The company says the release focuses on scaling reinforcement learning (RL) compute on verifiable and complex tasks through exploration and feedback.

    www.fonearena.com ↗
  8. 22 Sep 2026

    Luo Fuli Bets on Large-Scale RL: Xiaomi's Most Powerful Open-Source Model Debuts - 36氪

    She said that measured by computational investment, MiMo-V2.6 "is likely to be one of the largest single reinforcement learning training runs ever ...

    eu.36kr.com ↗
  9. 21 Sep 2026

    Dongfeng's Xiaodong humanoid robot is scheduled to enter a factory in October - TechNode

    Xiaomi livestreams MiMo-V2.6 reinforcement-learning runs.

    technode.com ↗
  10. 18 Sep 2026

    Xiaomi Livestreams MiMo-V2.6 Pro and Flash RL Post-Training Dashboard - Pandaily

    Xiaomi MiMo publicly streams MiMo-V2.6 Pro/Flash reinforcement-learning post-training—steps, tokens, rewards, and cost—at mimo.xiaomi.com/rl/, ...

    pandaily.com ↗
  11. 18 Sep 2026

    Xiaomi livestreams MiMo-V2.6 reinforcement-learning runs - TechNode

    Xiaomi's MiMo team is livestreaming the reinforcement-learning training of two unreleased models, MiMo-V2.6-Pro and MiMo-V2.6-Flash, ...

    technode.com ↗
  12. 17 Sep 2026

    Xiaomi MiMo-V2.6 Breaks Cover: A 1T-Class Chinese Lab Trains in Public - Forkast.News

    Xiaomi has initiated a live, public stream of its reinforcement learning training run for the MiMo-V2.6 model, a level of operational exposure ...

    forkast.news ↗
  13. 17 Sep 2026

    Xiaomi publicly unveils MiMo-V2.6 training progress for the first time - BigGo Finance

    Luo Fuli, head of Xiaomi's MiMo team, posted on X on September 17, publicly sharing the reinforcement learning training progress of the new model ...

    finance.biggo.com ↗
  14. 17 Sep 2026

    Xiaomi Livestreams MiMo 2.6 Reinforcement Learning Training: Over $1.13 Million Burned ...

    Xiaomi is livestreaming the reinforcement learning post-training process of its MiMo-V2.6 large language model on its official website in real ...

    finance.biggo.com ↗
Sources

Where this comes from

Artificial Intelligence — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1990879549…

401 items Polled 22 Sep, 21:25 UTC 200

Agentic AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4849751788…

401 items Polled 22 Sep, 21:25 UTC 200

AI Agents — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4849751788…

401 items Polled 22 Sep, 21:25 UTC 200

Data Science — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4836803184…

401 items Polled 22 Sep, 21:25 UTC 200

Large Language Model — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/8510459957…

401 items Polled 22 Sep, 21:25 UTC 200

Machine Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4836803184…

401 items Polled 22 Sep, 21:25 UTC 200

Deep Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1643871049…

401 items Polled 22 Sep, 21:25 UTC 200

LLMs Alignment — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/5121641697…

45 items Polled 22 Sep, 21:25 UTC 200

AI Alignment — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1764885284…

401 items Polled 22 Sep, 21:25 UTC 200

Representation Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1758013403…

146 items Polled 22 Sep, 21:25 UTC 200

Reinforcement Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1720175169…

401 items Polled 22 Sep, 21:25 UTC 200

Generative AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/7811591585…

401 items Polled 22 Sep, 21:25 UTC 200

Multimodal AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/9568555102…

401 items Polled 22 Sep, 21:25 UTC 200

Mechanistic Interpretability — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/8541059511…

34 items Polled 22 Sep, 21:25 UTC 200

RLHF — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/7811591585…

36 items Polled 22 Sep, 21:25 UTC 200

Feeds are configured through the LEARN_FEEDS environment variable, so new sources can be added without a code change. Each poll sends the stored ETag and Last-Modified headers, so an unchanged feed answers 304 and costs the publisher nothing.