Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

Report: Nvidia building 1-trillion-parameter Nemotron 4 to rival open AI models

Nvidia is building a 1-trillion-parameter Nemotron 4 model to rival open AI models, according to The Information. If confirmed, it would be a major upgrade to Nvidia's open-source model line and a sign that the chip giant is investing more heavily in the model layer.

Published
报道:Nvidia正打造1万亿参数Nemotron 4,对标开源AI模型
Image source: nvidia.com

Nvidia is building a 1-trillion-parameter Nemotron 4 model to rival open AI models, The Information reports.

Nemotron is Nvidia's open-source model family, which has drawn attention in local-AI and open-source communities and helped drive adoption of its GPU ecosystem.

At 1 trillion parameters, Nemotron 4 would rank among the largest models in the industry, directly challenging the top open models popular in the community today.

The report originates from The Information, and Nvidia has not officially confirmed the parameter count or release plans, so details still require official confirmation.

For Nvidia, a strong open model can drive sales of its GPUs and AI infrastructure — the stronger the model, the more reason developers have to run workloads on Nvidia platforms.

The move also signals that Nvidia is evolving from selling compute to a "compute plus models" strategy, putting it in a relationship with cloud providers and model companies that is both cooperative and competitive.

What to watch next: Nemotron 4's official launch plans, its real-world performance, and whether it can build enough adoption momentum in the open-source community.

Why it matters

If confirmed, Nemotron 4 would substantially strengthen Nvidia's influence in the open-model layer while further entrenching its chip ecosystem.

NVIDIANemotronOpen Source
Back to realtime news

Nearby Updates

All

08/12, 00:47

Databricks acquires Electric to give every AI agent its own Postgres database

Databricks has acquired Electric, a database startup, with the goal of giving every AI agent its own Postgres database, according to The New Stack. The move embeds database infrastructure directly into AI agent tooling, lowering the barrier for enterprises deploying agents at scale.

08/12, 00:45

Alibaba unveils Wan3.0 with 30-second video generation and multimodal inputs

Alibaba has unveiled Wan3.0, its video generation model that supports up to 30 seconds of output and multimodal inputs, according to Pokde.Net. The release extends Alibaba's push into long-form, controllable video generation and intensifies competition in the field.

08/12, 01:00

Google's AMIE medical AI demonstrates real-time clinical video consultations in first-of-its-kind study

Google has shown that AMIE, its research medical AI system, can conduct real-time clinical video consultations in a first-of-its-kind study. In simulated clinical settings, the system engages patients through live video, marking a step beyond text-based medical AI toward multimodal clinical interaction.

08/12, 00:25

An unreleased Anthropic model made surprising progress on the Riemann hypothesis, TechCrunch reports

TechCrunch reports that an unreleased Anthropic model made more progress than expected on the Riemann hypothesis, one of mathematics' biggest unsolved problems, though Anthropic has not solved it. The report signals that frontier mathematical reasoning is still improving quickly, and raises questions about when the model will ship.