Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

ByteDance begins training 10-trillion-parameter AI model

A report circulating via Google News aggregation says ByteDance has begun training an AI model with 10 trillion parameters. If confirmed, the move would put ByteDance among the very few companies worldwide attempting models at the 10-trillion-parameter scale.

Published
字节跳动开始训练10万亿参数AI大模型
Image source: bytedance.com

According to a report circulating via Google News aggregation, ByteDance has begun training an AI model with 10 trillion parameters. The news spread through aggregator channels on August 8 and drew attention across the industry.

A 10-trillion-parameter model represents a step change in scale, demanding GPU clusters, high-bandwidth interconnect, and power supply on a level few organizations can sustain over the long run.

Public information so far is limited to the training plan itself; the report does not provide further details on the model's architecture, training timeline, or compute scale.

If confirmed, ByteDance would join a very small group of companies worldwide attempting models at the 10-trillion-parameter scale. The plan signals that the race to ever-larger foundation models is heating up again.

Why it matters: model scale is directly tied to capability ceilings and cost structures. Training at 10 trillion parameters tests not only algorithmic and engineering skill, but also a company's long-term commitment to infrastructure investment.

What to watch next: whether ByteDance officially confirms the plan, when training is expected to complete, and how the model will eventually reach user-facing products. The competitive landscape for frontier-scale models could be redrawn.

Why it matters

A confirmed 10-trillion-parameter training program would push the large-model arms race to a new scale and raise the infrastructure bar for the whole industry. ByteDance's move could redraw competition among leading AI players.

ByteDanceLLMInfrastructure
Back to realtime news

Nearby Updates

All

08/08, 18:24

Firebird launches CIS region's largest AI factory in Armenia, powered by NVIDIA and Dell

Firebird, an emerging AI cloud provider, has launched the CIS region's largest AI factory in Armenia, creating a new AI computing hub powered by NVIDIA accelerated computing and Dell Technologies high-performance infrastructure. Armenian Prime Minister Nikol Pashinyan and Deputy Prime Minister Zhaslan Madiyev attended the launch, underscoring the project's regional significance.

08/08, 17:47

Chinese supercomputer runs full DeepSeek-V3/R1 models on pure CPUs, matching an 80-GPU cluster

Reports say the LingSheng supercomputer, ranked the world's top domestic Chinese supercomputer, has completed distributed inference of MoE models on a pure CPU architecture, running the full DeepSeek-V3/R1-671B model on just 16 compute nodes. At a batch size of 2048, its output throughput is said to be comparable to a cluster of 80 mainstream GPUs.

08/08, 17:40

Apple Intelligence officially supports Alibaba's Qwen models in deep tech collaboration

Apple Intelligence has officially added support for Alibaba's Qwen large language models, with the two companies forming a deep technical collaboration, according to a SmartHey report. The move marks a key step in localizing Apple's AI services in China.

08/08, 16:37

Gemini App reaches 950 million monthly users, Google says

Google says its Gemini app has reached 950 million monthly active users. The milestone makes Gemini one of the largest consumer AI assistant apps in the world and a strong signal for mainstream adoption of generative AI.