BazaarLinkBazaarLink
Sign in

Nemotron Nano 12b V2 Vl — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 12, 2026.

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s memory-efficient sequence modeling for significantly higher throughput and lower latency. The model supports inputs of text and multi-image documents, producing natural-language outputs. It is trained on high-quality NVIDIA-curated synthetic datasets optimized for optical-character recognition, chart reasoning, and multimodal comprehension. Nemotron Nano 2 VL achieves leading results on OCRBench v2 and scores ≈ 74 average across MMMU, MathVista, AI2D, OCRBench, OCR-Reasoning, ChartQA, DocVQA, and Video-MME—surpassing prior open VL baselines. With Efficient Video Sampling (EVS), it handles long-form videos while reducing inference cost. Open-weights, training data, and fine-tuning recipes are released under a permissive NVIDIA open license, with deployment supported across NeMo, NIM, and major inference runtimes.

ProviderNVIDIA
ReleasedOctober 2025
Model IDnvidia/nemotron-nano-12b-v2-vl

Technical specifications

Context window
128K tokens
Reasoning
Input
ImageTextVideo
Output
Text

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 5
Coding
Math 27
MMLU 65
GPQA 44
This modelCategory leader
Overall intelligence
#491out of 573 models
Aggregated across academic benchmarks
Math reasoning
#193out of 269 models
AIME / MATH-500 / SciCode
Time to first token
#520out of 586 models
AA 全球 p50:1.59 秒
Intelligence
5
Mid-tier
This model
5
Category leader
61
Median
15
Math
27
Mid-tier
This model
27
Category leader
99
Median
53
MMLU Pro
65%
This model
65%
Category leader
90%
Median
75%
GPQA
44%
This model
44%
Category leader
94%
Median
69%
LiveCodeBench
35%
This model
35%
Category leader
42%
Median
42%
HLE
5%
This model
5%
Category leader
53%
Median
7%
SciCode
18%
This model
18%
Category leader
60%
Median
33%
IFBench
26%
This model
26%
Category leader
83%
Median
44%
τ²-Bench Telecom
19%
This model
19%
Category leader
99%
Median
46%
AA-LCR
17%
This model
17%
Category leader
76%
Median
40%
Terminal-Bench Hard
0%
This model
0%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →
← All Models

Frequently Asked Questions

Can I still use the retired Nemotron Nano 12b V2 Vl model?

No. Nemotron Nano 12b V2 Vl has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: NVIDIA

Nemotron 3 Super 120b A12bNemotron 3 Nano 30b A3bGlm 4.7Glm 5Glm 5.2Deepseek V3.2
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.