Introducing NVIDIA Nemotron 3.5 Lightning⚡
An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster.
It delivers up to 4x the output speed of similar-sized models.
NVIDIA AI
@NVIDIAAI
Lightning pairs strong accuracy with speed.
On PinchBench, it reaches 86% accuracy while completing 10,000 tasks 35% faster than Qwen3.6 35B at similar accuracy.
On PinchBench, it reaches 86% accuracy while completing 10,000 tasks 35% faster than Qwen3.6 35B at similar accuracy.

NVIDIA AI
@NVIDIAAI
Lightning is built to specialize.
Post-train Nemotron 3.5 Lightning with NVIDIA NeMo for your domain data, tools, workflows and policies.
Across cybersecurity, coding, legal and energy tasks, post-training improves accuracy for specialized work.
Post-train Nemotron 3.5 Lightning with NVIDIA NeMo for your domain data, tools, workflows and policies.
Across cybersecurity, coding, legal and energy tasks, post-training improves accuracy for specialized work.

NVIDIA AI
@NVIDIAAI
Long-running agents spend most of their time executing: calling tools, validating results and delegating work.
Nemotron 3.5 Lightning is built for this high-volume execution, at a size that can run anywhere from an NVIDIA DGX Spark to the data center.
See it running agentic workflows on DGX Spark and read the technical breakdown: nvda.ws/4heK23m
Nemotron 3.5 Lightning is built for this high-volume execution, at a size that can run anywhere from an NVIDIA DGX Spark to the data center.
See it running agentic workflows on DGX Spark and read the technical breakdown: nvda.ws/4heK23m
NVIDIA AI
@NVIDIAAI
Not every step in an agent workflow needs the same model.
That’s why we’re also releasing NVIDIA NeMo Switchyard, a new open source library for model routing.
Use frontier models for complex reasoning and planning, and Lightning for high-volume, specialized execution.
Learn more: nvda.ws/457eK7l
That’s why we’re also releasing NVIDIA NeMo Switchyard, a new open source library for model routing.
Use frontier models for complex reasoning and planning, and Lightning for high-volume, specialized execution.
Learn more: nvda.ws/457eK7l
NVIDIA AI
@NVIDIAAI
As always, NVIDIA Nemotron 3.5 Lightning is open and customizable.
This includes weights, data and recipes.
Available now on @huggingface 🤗 → huggingface.co/nvidia/NVIDIA-…
This includes weights, data and recipes.
Available now on @huggingface 🤗 → huggingface.co/nvidia/NVIDIA-…

huggingface.co
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
NVIDIA AI
@NVIDIAAI
And it ships with its data.
Alongside Lightning, we’re releasing Nemotron-RL-Agentic-Terminal-Pivot, the open agentic RL dataset used to post-train its coding agent capabilities.
Dataset on @huggingface 🤗→ huggingface.co/datasets/nvidi…
Alongside Lightning, we’re releasing Nemotron-RL-Agentic-Terminal-Pivot, the open agentic RL dataset used to post-train its coding agent capabilities.
Dataset on @huggingface 🤗→ huggingface.co/datasets/nvidi…

huggingface.co
nvidia/Nemotron-RL-Agentic-Terminal-Pivot-v1 · Datasets at Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
