BazaarLinkBazaarLink
Sign in

Nemotron 3 Nano Omni 30b A3b Reasoning — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 1, 2026.

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and audio inputs and produces text output, enabling agents to perceive and reason across modalities in a single inference loop. Built on a hybrid MoE Transformer-Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers approximately 2× higher throughput and 2.5× lower compute for video reasoning versus separate vision + speech pipelines. It supports up to 300K context length and a 16,384 reasoning budget, with extended thinking enabled via reasoning.enabled.

ProviderNVIDIA
ReleasedApril 2026
Model IDnvidia/nemotron-3-nano-omni-30b-a3b-reasoning

Technical specifications

Context window
256K tokens
Reasoning
Input
TextAudioImageVideo
Output
Text
← All Models

Frequently Asked Questions

Can I still use the retired Nemotron 3 Nano Omni 30b A3b Reasoning model?

No. Nemotron 3 Nano Omni 30b A3b Reasoning has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: NVIDIA

Nemotron 3 Super 120b A12bNemotron 3 Nano 30b A3bGlm 4.7Glm 5Glm 5.2Deepseek V3.2
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.