1. 22 Sep 2026

    A phone maker now has the world's top open-weight AI model - TNW

    The gains came from one large reinforcement learning run. Xiaomi streamed it live as it happened. Pro and Flash each completed 30 training steps over ...

    thenextweb.com ↗
  2. 22 Sep 2026

    'Better than DeepSeek': Xiaomi's MiMo-V2.6-Pro debuts as the top open weights model in ...

    ... reinforcement-learning environments and now the training infrastructure behind them. From smartphones and EVs to frontier AI. Xiaomi's move into ...

    venturebeat.com ↗
  3. 22 Sep 2026

    Xiaomi open-sources MiMo-V2.6, publishes the bill for scaling reinforcement learning

    Chinese labs have spent the year arguing that reinforcement learning, not pre-training, holds the next gains. Xiaomi has now put a price on that ...

    www.digitimes.com ↗
  4. 22 Sep 2026

    Xiaomi's affordable flagship AI leads the open models, and Anthropic says Claude helped get it there

    Xiaomi achieved the performance jump through expanded reinforcement learning, and the company released its tools and training tasks openly. At the ...

    the-decoder.com ↗
  5. 22 Sep 2026

    Xiaomi open-sources MiMo-V2.6 models after scaling reinforcement learning - TechNode

    Xiaomi also released MiMo-V2.6-Distill-Qwen-9B and research resources for reinforcement learning. MiMo-V2.6-Pro scored 46 on the Artificial ...

    technode.com ↗
  6. 22 Sep 2026

    Xiaomi Steps Up AI Investment With New Models - Caixin Global

    Chinese tech giant Xiaomi Corp. on Tuesday launched and open-sourced its latest AI ... The new lineup, comprising of the multimodal MiMo-V2.6-Pro and ...

    www.caixinglobal.com ↗
  7. 22 Sep 2026

    Xiaomi Steps Up AI Investment With New Models - Caixin Global

    Xiaomi Steps Up AI Investment With New Models - The Chinese company talks up the stronger reinforcement learning capabilities of its MiMo-V2.6 ...

    www.caixinglobal.com ↗
  8. 22 Sep 2026

    Xiaomi releases MiMo-V2.6-Pro and MiMo-V2.6-Flash open-source AI models with scaled ...

    The company says the release focuses on scaling reinforcement learning (RL) compute on verifiable and complex tasks through exploration and feedback.

    www.fonearena.com ↗
  9. 22 Sep 2026

    Luo Fuli Bets on Large-Scale RL: Xiaomi's Most Powerful Open-Source Model Debuts - 36氪

    She said that measured by computational investment, MiMo-V2.6 "is likely to be one of the largest single reinforcement learning training runs ever ...

    eu.36kr.com ↗
  10. 22 Sep 2026

    Just now, Xiaomi has broken the performance cutoff threshold for large AI models. Luo Fuli ...

    Apart from version updates, APPSO previously reported that Xiaomi has publicly shared online a reinforcement learning training process that lasted for ...

    eu.36kr.com ↗
  11. 22 Sep 2026

    Xiaomi's New Flagship Model Leads Open-Weight Rankings With a Score of 46 - Unite.AI

    A Livestreamed Reinforcement-Learning Run. Xiaomi said it streamed the production reinforcement-learning run live as it happened. In under six days, ...

    www.unite.ai ↗
  12. 21 Sep 2026

    Xiaomi MiMo V2.6 Pro Becomes Top Open Model On Artificial Analysis Intelligence Index

    ... reinforcement learning. Alongside the open weights, Xiaomi has published the technical report, its RL environments, and training code. The company ...

    officechai.com ↗
  13. 21 Sep 2026

    Dongfeng's Xiaodong humanoid robot is scheduled to enter a factory in October - TechNode

    Xiaomi livestreams MiMo-V2.6 reinforcement-learning runs.

    technode.com ↗
  14. 18 Sep 2026

    Xiaomi Livestreams MiMo-V2.6 Pro and Flash RL Post-Training Dashboard - Pandaily

    Xiaomi MiMo publicly streams MiMo-V2.6 Pro/Flash reinforcement-learning post-training—steps, tokens, rewards, and cost—at mimo.xiaomi.com/rl/, ...

    pandaily.com ↗
  15. 18 Sep 2026

    Xiaomi livestreams MiMo-V2.6 reinforcement-learning runs - TechNode

    Xiaomi's MiMo team is livestreaming the reinforcement-learning training of two unreleased models, MiMo-V2.6-Pro and MiMo-V2.6-Flash, ...

    technode.com ↗
  16. 17 Sep 2026

    Xiaomi MiMo-V2.6 Breaks Cover: A 1T-Class Chinese Lab Trains in Public - Forkast.News

    Xiaomi has initiated a live, public stream of its reinforcement learning training run for the MiMo-V2.6 model, a level of operational exposure ...

    forkast.news ↗
  17. 17 Sep 2026

    Xiaomi's AI models training cost over HK$240,000 per hour, China's AI prodigy Luo Fuli reveals

    ... (over HK$240000) per hour, Luo Fuli, the founder of Xiaomi's MiMo large language model and dubbed an "AI prodigy," shared on social media.

    www.thestandard.com.hk ↗
  18. 17 Sep 2026

    Xiaomi Opens Its AI Training Books as Budget Phones and a September Flagship Loom

    Reinforcement learning at USD 31,000 an hour. The models are being trained through reinforcement learning, a method that devours capital. Luo Fuli ...

    www.ad-hoc-news.de ↗
  19. 17 Sep 2026

    Xiaomi publicly unveils MiMo-V2.6 training progress for the first time - BigGo Finance

    Luo Fuli, head of Xiaomi's MiMo team, posted on X on September 17, publicly sharing the reinforcement learning training progress of the new model ...

    finance.biggo.com ↗
  20. 17 Sep 2026

    Luo Fuli Follows Lei Jun to Launch Live Streaming, Xiaomi New Model Training Program ...

    Exploring cutting-edge methodologies to push the limits of AI reinforcement learning, this research delves into advanced optimization frameworks, ...

    eu.36kr.com ↗
Sources

Where this comes from

Artificial Intelligence — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1990879549…

401 items Polled 22 Sep, 21:25 UTC 200

Agentic AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4849751788…

401 items Polled 22 Sep, 21:25 UTC 200

AI Agents — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4849751788…

401 items Polled 22 Sep, 21:25 UTC 200

Data Science — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4836803184…

401 items Polled 22 Sep, 21:25 UTC 200

Large Language Model — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/8510459957…

401 items Polled 22 Sep, 21:25 UTC 200

Machine Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/4836803184…

401 items Polled 22 Sep, 21:25 UTC 200

Deep Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1643871049…

401 items Polled 22 Sep, 21:25 UTC 200

LLMs Alignment — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/5121641697…

45 items Polled 22 Sep, 21:25 UTC 200

AI Alignment — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1764885284…

401 items Polled 22 Sep, 21:25 UTC 200

Representation Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1758013403…

146 items Polled 22 Sep, 21:25 UTC 200

Reinforcement Learning — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/1720175169…

401 items Polled 22 Sep, 21:25 UTC 200

Generative AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/7811591585…

401 items Polled 22 Sep, 21:25 UTC 200

Multimodal AI — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/9568555102…

401 items Polled 22 Sep, 21:25 UTC 200

Mechanistic Interpretability — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/8541059511…

34 items Polled 22 Sep, 21:25 UTC 200

RLHF — Google

https://www.google.co.in/alerts/feeds/05832220720342067762/7811591585…

36 items Polled 22 Sep, 21:25 UTC 200

Feeds are configured through the LEARN_FEEDS environment variable, so new sources can be added without a code change. Each poll sends the stored ETag and Last-Modified headers, so an unchanged feed answers 304 and costs the publisher nothing.