Realtime AI News
MiniMax H3's Price Already Cut to Pennies Just After Launch as China's LLM Price War Escalates
MiniMax's H3 model has seen its price driven down to a matter of cents shortly after release, according to a Sohu report surfaced via Google News. The rapid discounting shows China's LLM price war is now hitting brand-new models almost immediately after launch.

Wait — wasn't MiniMax H3 just released? How did its price already get driven down to a few cents? That is the question a Sohu report poses, capturing a growing sense of whiplash in China's model market.
The article points to a striking fact: shortly after H3 went live, its price has already fallen to the level of a few cents, a speed of discounting that exceeds anything seen in previous model generations.
Price wars are nothing new in China's large-model race, but the way H3 has been dragged into one is still notable: a brand-new model hitting near-floor pricing almost immediately signals that vendors are using extreme low prices to win developers and API traffic.
For developers, prices at this level slash inference costs and make many more applications economically viable. For model vendors, however, it squeezes the window for commercial payback even further.
The report also implies a deeper question about the industry's rhythm: when every new generation ships and immediately discounts, with capability upgrades and price declines accelerating in lockstep, where can vendors still build a moat?
More broadly, H3's price war is a microcosm of the overall competitive landscape — open source and closed source, newcomers and giants are all colliding on the same price line.
What to watch next: whether H3's price keeps falling, whether rivals follow with matching cuts, and how long before this round of discounting shows up in application-layer pricing.
Why it matters
MiniMax H3 illustrates a new phase of China's LLM price war in which models are discounted almost at launch. Rapidly falling inference prices reshape the developer ecosystem while intensifying monetization pressure on model makers.
Nearby Updates
All08/05, 21:42
Anthropic Confirms It Is Building an In-House Silicon Team for Claude
Anthropic has confirmed it is building an in-house silicon team dedicated to Claude, according to a report from Unite.AI. The move puts the AI lab alongside OpenAI, Google, Amazon, and Meta in the custom chip race, with the goal of cutting dependence on generic GPUs and improving inference efficiency.
08/05, 21:00
Benchmark Gensuite Launches MCP Connector to Give AI Agents Access to Enterprise Operational Risk Management
Benchmark Gensuite has announced an MCP connector that gives AI agents access to enterprise operational risk management data and workflows. The move aligns with the Model Context Protocol's rise as the standard for connecting agents to enterprise data, potentially embedding risk and compliance processes deeper into corporate AI operations.
08/05, 22:47
Kiteworks and Reco partner to strengthen AI agent governance and data security
Enterprise data governance platform Kiteworks and AI security firm Reco have announced a partnership to strengthen AI agent governance and data security, as reported by Security Info Watch. The two will combine Reco's agent security capabilities with Kiteworks' controlled data handling to help enterprises manage the data risks of AI agents.
08/05, 20:49
keyv npm supply chain attack hides malware in AI agent files scanners never read
A supply chain attack on the keyv npm package has been uncovered, with malware hidden in AI agent files that security scanners never read, Tech Times reports. The technique exploits a blind spot in automated scanning to slip malicious code into the npm ecosystem.