Realtime AI News
DeepSeek V4 Flash Tops 7 Trillion Weekly Calls, Ranks No.1 Globally
DeepSeek V4 Flash has surpassed 7 trillion calls per week, ranking first globally in weekly call volume, according to a 36Kr report. The milestone positions the model as the most frequently used large language model in real-world production today.

DeepSeek V4 Flash has logged more than 7 trillion calls per week, ranking first globally in weekly call volume, according to a report from Chinese tech media 36Kr. The report, headlined "Over 7 Trillion Calls Weekly: DeepSeek V4 Flash Ranks No.1 Globally in Weekly Call Volume," points to the model's sheer scale of real-world usage.
Weekly call volume is one of the most direct measures of genuine adoption for a large language model. Unlike benchmark scores or demo videos, it reflects how frequently developers and end users actually invoke the model in production, and higher numbers mean more real business scenarios have been connected to it.
DeepSeek V4 Flash is DeepSeek's model offering, and topping the weekly call ranking makes it one of the most frequently invoked large models in the world. For a model positioned around fast response, trillion-level weekly traffic also suggests it has held up under high-frequency, low-latency workloads.
The signal matters for the industry: when a model logs trillions of weekly calls, it is running steadily inside real applications rather than living only in benchmarks and demos. Leadership in call volume often says more about a model's standing in the developer ecosystem than a single performance edge.
The 36Kr report provides corroboration from a major Chinese tech outlet, though the item circulates through a Google News aggregation link, so readers may want to cross-check the figure against DeepSeek's official channels and other usage statistics.
What to watch next is whether DeepSeek V4 Flash can sustain its lead in the coming weeks and whether the momentum pulls more developers into building on top of it. At the same time, competition over call volume among top models will continue to shape inference pricing and the broader ecosystem.
Why it matters
DeepSeek V4 Flash's No.1 ranking in weekly call volume signals massive real-world usage and could cement DeepSeek's position in the inference services market.
Nearby Updates
All08/05, 17:38
Taotian opens 2027 campus recruitment with AI roles making up over 90% of tech positions
Alibaba's Taotian Group kicked off its 2027 campus recruitment on August 5, targeting graduates graduating between November 2026 and October 2027 at home and abroad. AI technology positions account for more than 90% of the openings, with core roles including AI application R&D engineer, AI Agent optimization engineer, and AI application algorithm engineer.
08/05, 17:49
AMD Teams Up With Meta, OpenAI and Anthropic to Optimize AI Models, CEO Says the Whole Ecosystem Benefits
AMD CEO Lisa Su said on the company's second-quarter earnings call that AMD is working directly with Meta, OpenAI, Anthropic, and others to optimize their AI models on ROCm, its open AI software platform, with improvements benefiting the entire AMD ecosystem. More than 3 million AI models now run out of the box on AMD's platform, and open-source contributions to ROCm have grown more than tenfold over the past year.
08/05, 18:00
AI server shipments forecast raised to 31% YoY growth in 2026
Industry coverage from Evertiq shows the 2026 forecast for AI server shipments has been raised to 31% year-over-year growth. The upgrade points to still-expanding demand for AI compute and continued heavy investment in data-center infrastructure.
08/05, 16:04
NVIDIA Nemotron 3 Ultra Highlights Major AI Agent Advances
Coverage from Techgenyz spotlights NVIDIA's Nemotron 3 Ultra as a major step forward in AI agent capabilities. The release positions agent-focused abilities as the headline advance in NVIDIA's high-end open model line, with benchmark results and technical specifics still pending.