Nvidia Nemotron 3.5 Lightning

tosh1 pts0 comments

NVIDIA AI on X: "Introducing NVIDIA Nemotron 3.5 Lightning⚡

An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster.

It delivers up to 4x the output speed of similar-sized models. https://t.co/ENWrZe76pU" / X<br>Post

Log inSign up

Post

NVIDIA AI

@NVIDIAAI

Introducing NVIDIA Nemotron 3.5 Lightning⚡

An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster.

It delivers up to 4x the output speed of similar-sized models.

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0">1:00 PM · Aug 11, 2026360.2KViews

127<br>287<br>2.6K<br>648

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">NVIDIA AI

@NVIDIAAI

2h

Lightning pairs strong accuracy with speed.

On PinchBench, it reaches 86% accuracy while completing 10,000 tasks 35% faster than Qwen3.6 35B at similar accuracy.

140<br>17K

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">NVIDIA AI

@NVIDIAAI

2h

Lightning is built to specialize.

Post-train Nemotron 3.5 Lightning with NVIDIA NeMo for your domain data, tools, workflows and policies.

Across cybersecurity, coding, legal and energy tasks, post-training improves accuracy for specialized work.

127<br>16K

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">NVIDIA AI

@NVIDIAAI

2h

Long-running agents spend most of their time executing: calling tools, validating results and delegating work.

Nemotron 3.5 Lightning is built for this high-volume execution, at a size that can run anywhere from an NVIDIA DGX Spark to the data center.

See it running agentic Show more

00:00

65<br>11K

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">NVIDIA AI

@NVIDIAAI

1h

Not every step in an agent workflow needs the same model.

That’s why we’re also releasing NVIDIA NeMo Switchyard, a new open source library for model routing.

Use frontier models for complex reasoning and planning, and Lightning for high-volume, specialized execution.

Learn Show more

00:00

54<br>8.9K

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">NVIDIA AI

@NVIDIAAI

1h

As always, NVIDIA Nemotron 3.5 Lightning is open and customizable.

This includes weights, data and recipes.<br>Available now on @huggingface 🤗 →

nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 · Hugging Face

From huggingface.co

96<br>9.7K

span:not(:empty)~span:not(:empty)]:before:content-['·'] [&>span:not(:empty)~span:not(:empty)]:before:px-1 [&>span:not(:empty)~span:not(:empty)]:before:shrink-0 min-w-0 overflow-hidden">Big Lewinsky<br>@CRuud19470

1h

Guys did you really use a chart where your model is listed second worst as an ad? Are you okay?

36<br>1.5K

Log in or sign up for X<br>See what’s happening and join the conversation<br>Continue with phoneContinue with AppleContinue with Google<br>or<br>Log in with username or email

Relevant people

NVIDIA AI@NVIDIAAIFollow<br>Teaching your AI new tricks.

Trending now

span empty before nvidia lightning nemotron

Related Articles