Realtime AI News
Qwen Releases Qwen3.8-2.4T-A95B MoE Model with FP8 Quantized Version
Qwen has published the Qwen3.8-2.4T-A95B model on Hugging Face, a mixture-of-experts text model whose naming indicates roughly 2.4 trillion total parameters and 95 billion active parameters. The team also uploaded an FP8 quantized variant of the same model in the same release window.

Qwen has published a new model, Qwen3.8-2.4T-A95B, on Hugging Face, marking the latest update to the open-source Qwen family.
The model card lists the pipeline as text-generation and shows it is built on the transformers library, with tags including qwen3_5_moe_text indicating a mixture-of-experts (MoE) text model.
Under Qwen's naming convention, 2.4T points to roughly 2.4 trillion total parameters and A95B to about 95 billion active parameters per inference, a classic sparse-activation MoE design.
At nearly the same time, the team uploaded an FP8 quantized variant, Qwen3.8-2.4T-A95B-FP8, whose model card lists the original release as its base model; quantized versions typically offer friendlier memory footprints and faster inference.
At the time of writing, the base model had recorded 978 downloads on Hugging Face while the FP8 version had around 3,851, with community interest climbing quickly.
The steady stream of large-parameter open-source MoE models lets developers run near-frontier text generation locally or in self-hosted setups while keeping inference costs in check through sparse activation.
Next to watch is whether the team publishes a technical report and benchmark results, and how quickly the model gains support in mainstream inference frameworks and deployment tooling.
Why it matters
The release adds another large-parameter MoE text model to the open-source ecosystem, potentially lowering the barrier to near-frontier text generation.
Nearby Updates
All08/12, 17:52
Lei Jun poaches a key executive from ByteDance, Sina reports
Sina reported that Lei Jun has poached a key executive from ByteDance, without naming the person or the role. The move, surfaced through ByteDance AI-related news aggregation, is the latest sign of intensifying talent competition between major tech companies.
08/12, 19:00
AI code-testing startup Blacksmith's valuation jumps nearly 10x in under a year
Blacksmith, an AI code-testing startup, says its valuation has jumped nearly tenfold in less than a year to $550 million. The company reports revenue growth of more than tenfold over the past year, as AI-driven coding fuels demand for software validation.
08/12, 17:47
ByteDance's Seedance 2.5 draws attention: AI video generated 30 seconds at a time
Reports say ByteDance's AI video model Seedance 2.5 can generate 30 seconds of video in a single pass. The capability, if confirmed, would let creators produce complete coherent clips without stitching, a significant step for short-video and advertising workflows.
08/12, 15:57
China's edge-AI unicorn ModelBest files for A-share IPO counseling as MiniCPM hits 38M downloads
ModelBest, China's largest edge-AI model unicorn, has filed for A-share IPO counseling, with its MiniCPM series reaching 38 million cumulative downloads. The filing marks the company's first formal step toward a listing on China's A-share market and adds momentum to the capitalization of the on-device AI segment.