Xiaomi MiMo: A Trillion Parameters, With Four Percent Active

Xiaomi has unveiled the MiMo family of language models, spanning a wide range of configurations from compact 7B models to a massive Mixture of Experts (MoE) solution featuring one trillion parameters. A key highlight of the announcement is the high computational efficiency: despite the massive total parameter count, only about four percent are active during text generation. This approach significantly accelerates model performance and reduces hardware requirements without compromising output quality. Experts note that the use of sparse architectures is becoming a new industry standard, allowing companies to scale AI capabilities to trillion-parameter levels while maintaining acceptable inference speeds. This development underscores Xiaomi's commitment to strengthening its position in generative AI by offering flexible tools for diverse applications.
This is a summary. Read the full article at the original source:
HabrRelated stories
Jev: New frontier model 40-400x cheaper and 20-200x faster
Typesafe AI has introduced Jev, a new frontier model designed to significantly optimize the cost and speed of large language model inference. Accordin…
The AI graveyard: A running list of projects and startups that didn't make it
TechCrunch has published a comprehensive overview of the current AI landscape, focusing on high-profile projects and startups that have failed to meet…
Emergency Alert System for 15,000 Users: LLM, PostGIS, Qdrant, and Telegram
The article details the architecture of an emergency alert system serving 15,000 users. The primary engineering challenge was automating the processin…



