Nemotron 3 Nano Omni 30b A3b Reasoning
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and audio inputs and produces text output, enabling agents to perceive and reason across modalities in a single inference loop. Built on a hybrid MoE Transformer-Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers approximately 2× higher throughput and 2.5× lower compute for video reasoning versus separate vision + speech pipelines. It supports up to 300K context length and a 16,384 reasoning budget, with extended thinking enabled via reasoning.enabled .
Pricing
💱 匯率 USD/NTD = 32.34 · 不含稅 · 更新於 2026/07/22
Technical specifications
Quick Start
Use Nemotron 3 Nano Omni 30b A3b Reasoning via BazaarLink API — just change the base_url:
from openai import OpenAI
client = OpenAI(
base_url="https://bazaarlink.ai/api/v1",
api_key="sk-bl-YOUR_API_KEY",
)
response = client.chat.completions.create(
model="nemotron-3-nano-omni-30b-a3b-reasoning",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://bazaarlink.ai/api/v1",
apiKey: "sk-bl-YOUR_API_KEY",
});
const response = await client.chat.completions.create({
model: "nemotron-3-nano-omni-30b-a3b-reasoning",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);Why use Nemotron 3 Nano Omni 30b A3b Reasoning via BazaarLink?
- ✓USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
- ✓OpenAI-compatible API — zero code changes required
- ✓Automatic failover — multi-provider redundancy for the same model
- ✓Chinese-language support — local team, instant help
Frequently Asked Questions
What is Nemotron 3 Nano Omni 30b A3b Reasoning?
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and audio inputs and produces text output, enabling agents to perceive and reason across modalities in a single inference loop. Built on a hybrid MoE Transformer-Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers approximately 2× higher throughput and 2.5× lower compute for video reasoning versus separate vision + speech pipelines. It supports up to 300K context length and a 16,384 reasoning budget, with extended thinking enabled via reasoning.enabled .
How much does the Nemotron 3 Nano Omni 30b A3b Reasoning API cost?
Nemotron 3 Nano Omni 30b A3b Reasoning costs $0.0000 per 1K input tokens and $0.0000 per 1K output tokens when accessed through BazaarLink.
How do I use Nemotron 3 Nano Omni 30b A3b Reasoning with the OpenAI SDK?
Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "nemotron-3-nano-omni-30b-a3b-reasoning". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.
What is the context window for Nemotron 3 Nano Omni 30b A3b Reasoning?
Nemotron 3 Nano Omni 30b A3b Reasoning supports a context window of 256,000 tokens.
Is Nemotron 3 Nano Omni 30b A3b Reasoning available for free?
Yes — Nemotron 3 Nano Omni 30b A3b Reasoning is available at no cost through BazaarLink's free tier. No credit card required.