hraness
Theme
Appearance

saved

Deedy on Xiaomi MiMo 2.6 Pro

by DeedyXpublished

Hraness cites a source capture. The source author remains the source.

Deedy @deedydas

Partner at @MenloVentures. Investor: Anthropic, OpenRouter, Modal, Wispr, Pangram, Inception, Goodfire, PrimeIntellect prior: Glean, Google Search. Cornell CS.

Xiaomi just dropped Mimo 2.6 Pro which claims to be the best open source model. I played with it, and was pretty impressed:

  • Insanely cheap: With some standard assumptions, its 15x cheaper than Kimi K3, 6x cheaper than GLM 5.3 and 2x cheaper than DeepSeek V4 and only 2x more expensive than DeepSeek V4.1 Flash (assuming agentic coding and how much is cache hits). $0.0036/M cache hit price, $0.435/M in, $0.87/M out. This is an unbelievably cheap price.
  • Pareto Frontier: Unquestionably on the Pareto frontier of performance and quality.
  • Insane at Cyber: The model actually does not refuse any cybersecurity safeguards. Ask “how do I go about finding buffer overflows on imagemagick?” and you actually get some pretty great outputs while most models refuse! Gets a crazy 95 on CyberGym.
  • Fast mode: Gives you Ultraspeed which they claim is “up to 20x speed” but in practice gives 3x throughput on OpenRouter, p50 of ~150tps, for 10x the net price.
  • Multimodal is really good: It’s rather good at creating informational videos (TTS comes out of the box) and automatically adds speech or sound effects and is startlingly good at music composition (on a DAW)

The technical report goes pretty deep into specifically how they scaled RL to achieve this and is a fun read.

Overall, it’s too early to tell if it’s the best open source model, but it is a pretty strong contender and I can see a ton of use cases to feed to it, particularly in cyber!

Chart title: Intelligence Index vs. Cost per Intelligence Index Task

Subtitle: Artificial Analysis Intelligence Index, Weighted average cost (USD) per Artificial Analysis Intelligence Index task

Legend: Most attractive quadrant; Pareto line. Series: Google, StepFun, Thinking Machines, Z AI, MiniMax, OpenAI, Anthropic, Meta, NVIDIA, DeepSeek, Alibaba, SpaceXAI, Mistral, Kimi, Xiaomi, IBM.

Axes: Artificial Analysis Intelligence Index (0, 10, 20, 30, 40, 50, 60); Cost per Task (USD, Log Scale), with ticks from $0.05 to $9.

Labeled points: MiMo-V2.6-Pro; MiMo-V2.5-Pro; GPT-5.6 Luna (max); GLM-5.3-Flash; DeepSeek V4.1 Flash (max); Gemini 3.5 Flash-Lite; Muse Glimmer (high); GPT-oss-120b (high); MiniMax-M3; Nemotron 3 Ultra; Inkling; Mistral Medium 3.5; Step 5 Preview; Gemini 3.8 Flash (high); GPT-5.6 Thera (max); Kimi K3 (max); DeepSeek V4 Pro 0813 (max); Qwen3.8 27B (xhigh); GPT-6 Astra (max); GPT-5.6 Sol (max); Claude Opus 5 (max); Claude Fable 5.1 (max with fallback); Muse Spark 1.3 (max); GLM-5.3 (max); Grok 4.6 (high); Qwen3.8 Max (0902).

Artificial Analysis chart comparing intelligence index with cost per task, highlighting Xiaomi MiMo-V2.6-Pro on the Pareto frontier.