Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

NVIDIA Unveils Nemotron 3.5 Lightning and NeMo Switchyard for Efficient Agentic AI

NVIDIA is expanding its Nemotron 3 model family with Nemotron 3.5 Lightning, which the company describes as the highest-efficiency model in its class for long-running agentic AI workloads. The release pairs the model with NeMo Switchyard tooling aimed at making autonomous agents faster, smarter, and more efficient to run.

Published
NVIDIA发布Nemotron 3.5 Lightning与NeMo Switchyard,瞄准高效Agentic AI
Image source: blogs.nvidia.com

NVIDIA is expanding its Nemotron 3 model family with Nemotron 3.5 Lightning, which the company describes as the highest-efficiency model in its class for long-running agentic AI workloads, according to a post on the NVIDIA Blog.

The release arrives as AI shifts from chatbots to autonomous agents, and as open models increasingly serve demand for full control over where AI runs and how it is deployed and evolves.

Nemotron 3.5 Lightning is positioned as the efficiency-focused addition to the Nemotron 3 line, aimed at workloads where agents must run for extended periods without excessive compute cost.

The announcement also introduces NeMo Switchyard, NVIDIA's tooling for building and deploying agentic AI systems, pairing the new model with the software stack needed to put it to work.

The launch is tied to NVIDIA's push for local and on-premises AI, with the release covering both RTX and DGX platforms so users can run open models in environments they control.

Why it matters: efficiency is becoming the deciding factor for agentic AI, where long-running loops multiply inference cost, and a model that maintains quality at lower overhead changes the economics of deploying agents at scale.

The move also strengthens NVIDIA's position in the open-model ecosystem, where local AI communities are building, customizing, and running increasingly capable agents on their own hardware.

What to watch next: benchmark results against other open agentic models, adoption within NVIDIA's NeMo ecosystem, and how the broader Nemotron 3 family evolves in the coming months.

Why it matters

By betting on efficiency for long-running agent workloads, NVIDIA's Nemotron 3.5 Lightning could accelerate open-model adoption in local and on-premises agentic AI deployments.

NVIDIANemotronAgentic AI
Back to realtime news

Nearby Updates

All

08/11, 21:00

Spotify to Label 'AI Persona' Profiles and Exclude Their Music from Recommendations

Spotify is introducing 'AI Persona' labels for artist profiles that represent AI-generated identities, and their music will be excluded from editorial, algorithmic, and personalized recommendations by default. The policy gives the streaming giant a clear governance framework for synthetic artists as AI-generated music spreads across the platform.

08/11, 17:24

Google co-founder Brin urgently takes over Gemini team as 3.5 Pro is reportedly cancelled

Chinese tech outlet QbitAI reports that Google co-founder Sergey Brin has urgently taken over the Gemini team amid severe internal infighting over compute allocation. The report also claims the planned 3.5 Pro has been cancelled as the team undergoes further restructuring.

08/11, 15:36

Zhipu AI Releases GLM-5 on Hugging Face Official Model Registry

Zhipu AI has published its next-generation open-source model GLM-5 on the official Hugging Face registry, with the model card listing a text-generation pipeline, a mixture-of-experts architecture tag, and bilingual Chinese-English conversational support. The model has already drawn more than 155,000 downloads and over 2,100 likes since appearing on the platform.

08/11, 13:12

Claude Breaks New Record on the Riemann Hypothesis Benchmark, Report Says Model Is Unannounced

A Claude model has set a new record on a benchmark built around the century-old Riemann hypothesis, pushing the lower bound of verified cases substantially higher, according to Chinese outlet QbitAI. The report suggests the result came from an unannounced new model, hinting that Anthropic may have a stronger reasoning model in testing.