BazaarLinkBazaarLink
Sign in
N

Nemotron Nano 12b V2 Vl

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s memory-efficient sequence modeling for significantly higher throughput and lower latency. The model supports inputs of text and multi-image documents, producing natural-language outputs. It is trained on high-quality NVIDIA-curated synthetic datasets optimized for optical-character recognition, chart reasoning, and multimodal comprehension. Nemotron Nano 2 VL achieves leading results on OCRBench v2 and scores ≈ 74 average across MMMU, MathVista, AI2D, OCRBench, OCR-Reasoning, ChartQA, DocVQA, and Video-MME—surpassing prior open VL baselines. With Efficient Video Sampling (EVS), it handles long-form videos while reducing inference cost. Open-weights, training data, and fine-tuning recipes are released under a permissive NVIDIA open license, with deployment supported across NeMo, NIM, and major inference runtimes.

Pricing

Input Price
Free
/ 1M tokens
Output Price
Free
/ 1M tokens
Context Window
128K
tokens

💱 匯率 USD/NTD = 32.34 · 不含稅 · 更新於 2026/07/22

ProviderNNVIDIA
Released2025年10月
Model IDnemotron-nano-12b-v2-vl

Technical specifications

Context window
128K tokens
Reasoning
Input
ImageTextVideo
Output
Text
!
此模型已下架 — API 呼叫會回傳 HTTP 410,但本頁面保留以利歷史查詢

Quick Start

Use Nemotron Nano 12b V2 Vl via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="nemotron-nano-12b-v2-vl",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "nemotron-nano-12b-v2-vl",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Nemotron Nano 12b V2 Vl via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Nemotron Nano 12b V2 Vl?

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s memory-efficient sequence modeling for significantly higher throughput and lower latency. The model supports inputs of text and multi-image documents, producing natural-language outputs. It is trained on high-quality NVIDIA-curated synthetic datasets optimized for optical-character recognition, chart reasoning, and multimodal comprehension. Nemotron Nano 2 VL achieves leading results on OCRBench v2 and scores ≈ 74 average across MMMU, MathVista, AI2D, OCRBench, OCR-Reasoning, ChartQA, DocVQA, and Video-MME—surpassing prior open VL baselines. With Efficient Video Sampling (EVS), it handles long-form videos while reducing inference cost. Open-weights, training data, and fine-tuning recipes are released under a permissive NVIDIA open license, with deployment supported across NeMo, NIM, and major inference runtimes.

How much does the Nemotron Nano 12b V2 Vl API cost?

Nemotron Nano 12b V2 Vl costs $0.0000 per 1K input tokens and $0.0000 per 1K output tokens when accessed through BazaarLink.

How do I use Nemotron Nano 12b V2 Vl with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "nemotron-nano-12b-v2-vl". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Nemotron Nano 12b V2 Vl?

Nemotron Nano 12b V2 Vl supports a context window of 128,000 tokens.

Is Nemotron Nano 12b V2 Vl available for free?

Yes — Nemotron Nano 12b V2 Vl is available at no cost through BazaarLink's free tier. No credit card required.

中轉站誠信檢測 · BazaarLink 獨家

尚無此模型的檢測資料

BazaarLink Probe 對聲稱提供此模型的 endpoint 進行家族指紋驗證與 V3 子模型對照。 如發現行為與真貨基準不符,會列入異常案例。

查看完整檢測報告 →

Related Models & Links

All ModelsAPI DocumentationAPI Latency Probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.