Nvidia Releases Nemotron 3.5 Lightning Open AI Model
Nvidia has introduced Nemotron 3.5 Lightning, a new open AI model designed less as an all-purpose answer engine than as a fast workhorse inside long-running AI agent systems.
Campus Technology
Publisher
Aug 17, 2026 at 8:26 PM UTC · 2 min de leitura

Key Signal
30B total model parameters
Last Updated
há 4 dias
Nvidia has introduced Nemotron 3.5 Lightning, a new open AI model designed less as an all-purpose answer engine than as a fast workhorse inside long-running AI agent systems.
Released Aug. 11, Nemotron 3.5 Lightning is a 30 billion-parameter mixture-of-experts model that activates 3 billion parameters per token. Nvidia says it supports context windows of up to 1 million tokens and can deliver up to four times the output speed of similar-sized models. The company is offering the model with open weights, training data, and recipes under its OpenMDW-1.1 license.
The more consequential part of the Nemotron 3.5 Lightning release is the job Nvidia expects models like it to perform. Rather than sending every step of an AI agent's work to a large frontier model, Nvidia envisions systems in which different models handle different classes of tasks.
A larger reasoning model might plan a workflow or handle a difficult decision, while a faster specialized model performs the repeated tool calls, validations, formatting tasks, and other routine operations generated as an agent carries out that plan. That distinction becomes more significant as AI applications move beyond single prompts and responses toward agents that can generate dozens or potentially many more model calls while completing a task.
Article Intelligence
Topics
Sponsored
AdNewsLayer Premium
Unlock deeper intelligence.
Ad-free reading, exclusive research, and real-time onchain insights.
Go Premium
