Nemotron 3 Super targets agentic reasoning

// 77d agoMODEL RELEASE

Nemotron 3 Super targets agentic reasoning

NVIDIA has released Nemotron 3 Super, a 120B-parameter open-weight hybrid Mamba-Transformer MoE model with 12B active parameters and a native 1M-token context window. It is built for long-horizon multi-agent workloads like software development and cybersecurity, with NVIDIA claiming over 5x the throughput of the previous Nemotron Super plus open datasets, recipes, and deployment guides.

// ANALYSIS

This is NVIDIA making a serious play for the “agent brain” layer: not just bigger reasoning, but cheaper, longer-context reasoning that can stay aligned across sprawling multi-agent workflows.

–The 120B total / 12B active setup matters because it aims to keep inference costs down while still scaling to harder reasoning and coding tasks
–A native 1M-token context window directly targets the context explosion problem that breaks many long-running agent systems
–The hybrid Mamba-Transformer design, latent MoE, and multi-token prediction show NVIDIA optimizing for throughput as much as raw benchmark bragging rights
–Open weights, datasets, and recipes make this more useful to developers than a closed API-only release, especially for teams that want to fine-tune or self-host
–NVIDIA is also tying the model tightly to its own stack, from Blackwell NVFP4 optimization to NeMo, TensorRT-LLM, and NIM deployment paths

// TAGS

nemotron-3-superllmreasoningagentopen-weightsinference

DISCOVERED

77d ago

2026-03-11

PUBLISHED

77d ago

2026-03-11

RELEVANCE

9/ 10

AUTHOR

deeceeo

// KEEP READING

More AI developer news from the feed

EXPLORE FULL FEED

UPDATE2h ago

Cursor adds dedicated subagents for skills

Cursor now allows developers to execute tool-heavy or research-intensive agent skills within dedicated subagents. This architectural shift isolates noisy background tasks, keeping the main chat context clean and focused.

UPDATE3h ago

YouTube moves AI labels to video player

YouTube is moving its AI content disclosures from video descriptions to more prominent placements beneath the player and on Shorts overlays. Starting in May, the platform will use internal signals to automatically label photorealistic AI content that creators fail to disclose.

OPEN SOURCE6h ago

Taste Skill kills AI "frontend slop"

Taste-Skill is an open-source framework that provides portable "agent skills" to enforce high-end design principles in AI-generated code. By injecting specific design directives and "anti-slop" rules, it enables LLMs to produce editorial-grade UIs that bypass generic, boilerplate-heavy AI templates.