Realtime AI News
MiniMax ARR surges 500% and token consumption jumps 2000% as the agent dividend kicks in
MiniMax disclosed that its ARR surpassed $800 million in August 2026, roughly 5.3 times the $150 million reported in February, while July token consumption reached 20 times January's level. H1 revenue hit $116.6 million, up 283.1% year over year, with B2B revenue now accounting for 63.4% of the mix as agent-driven enterprise demand becomes the main growth engine.

MiniMax's latest disclosures show a commercialization surge: as of August 2026, the company's annual recurring revenue (ARR) has surpassed $800 million, up from just $150 million disclosed in February — roughly a 5.3x expansion in half a year. Token consumption has climbed in tandem, reaching 20 times January's level in July, according to QbitAI.
The growth is now showing up in actual revenue. In the first half of 2026, MiniMax generated $116.6 million in revenue, up 283.1% year over year, already exceeding its full-year 2025 total of $79.04 million; second-quarter revenue grew another 81.8% quarter over quarter.
Even more notable than the scale is the shifting revenue mix. B2B revenue accounted for 63.4% of total revenue in H1 2026 and rose to 80% by August, versus roughly 30% for all of 2025. MiniMax's growth engine has clearly moved from consumer products toward enterprises and developers.
In H1 2026, revenue from MiniMax's open platform and other AI enterprise services reached $73.93 million, up 703.1% year over year, contributing about three-quarters of the total revenue increase. The consumer side did not shrink: AI-native product revenue, covering Talkie and Hailuo AI, came to about $42.64 million, up 100.9%.
Financial efficiency is improving too. Gross profit hit $20.81 million in H1, up 464.8%, lifting gross margin from 12.1% to 17.9%, while sales and distribution expenses fell 17.9% to $26.97 million. R&D spending remained heavy at $296.9 million, up 138.8% year over year.
Why now? CEO Yan Junjie said on the earnings call that enterprise customers and developers on the platform surpassed 2 million by May — ten times the level at the end of last year — while July token consumption hit 20x January's level. Agent-driven inference demand is growing far faster than human user counts and message volumes. MiniMax's models are increasingly entering the real workloads of enterprises and developers, steadily converting model capability into API calls and commercial revenue.
Models and open source are the two accelerators. The M3 model released in June gave developers confidence to plug agent workflows into real business scenarios; the H3 video-generation model open-sourced in July drew more than 24 million downloads within three weeks, spawned over 300 derivative models in the community, and pushed many developers toward paid API usage.
MiniMax's strongest model, M3, costs just $1.2 per million tokens, well below the $2.2 median for comparable models. MiniMax argues that high intelligence and cost efficiency are not a trade-off: lowering the cost per unit of intelligence lets customers run more calls, generating more revenue to reinvest in compute. Whether the commercialization acceleration persists and whether B2B volume can fund the next model cycle are the key things to watch.
Why it matters
MiniMax's ARR jumping from $150 million to over $800 million in six months, with B2B reaching 80% of the mix, signals that agent workloads are reshaping how model companies monetize. Its low-priced M3 and open-sourced H3 could keep pressuring pricing across the segment.
Nearby Updates
All08/27, 19:35
OpenAI to start showing ads on ChatGPT's free and Go tiers in India
OpenAI will begin showing ads on ChatGPT's free and Go tiers in India, TechCrunch reports. The company has more than 100 million weekly active ChatGPT users in India, a large share of whom use the free or lower-priced plans.
08/27, 21:00
Plaud's $249 'agentic' earbuds ship with an eSIM-enabled case for talking to AI agents
Plaud has unveiled new "agentic" earbuds priced at $249, featuring a charging case with a built-in eSIM so users can talk to AI agents without relying on their phone. The launch extends Plaud's push into AI hardware, making always-available AI assistance the core selling point of a wearable device.
08/27, 21:00
NVIDIA's Vera, Its First CPU Built for AI Agents, Is Now Shipping at Scale
NVIDIA announced that Vera, its first CPU built for AI agents, is now shipping at scale, with Vice President of Hyperscale and HPC Ian Buck personally hand-delivering Vera CPU systems across the AI ecosystem. The milestone marks Vera's move from announcement to real-world deployment in agent-ready computing infrastructure.
08/27, 21:00
NVIDIA Brings DLSS 4.5 Controls and New Ways to Play to GeForce NOW at Gamescom 2026
At Gamescom 2026, NVIDIA unveiled the next wave of GeForce NOW updates, including new DLSS 4.5 technology controls that give members more ways to fine-tune gameplay. The company also expanded support for new Steam devices, GOG single sign-on, and the Firefox browser, while bringing even more big PC games to the cloud.