BazaarLinkBazaarLink
Sign in

Nemotron 3 Nano 30b A3b API Pricing & Quick Start

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Intelligence#287 / 573
15
Artificial Analysis Intelligence Index
Speed#9 / 586
322.7
Output tokens per second (median)
Input Price#239 / 586
$0.05NT$ 2
USD / 1M tokens
Output Price#247 / 586
$0.20NT$ 6
USD / 1M tokens
First-token latency#442 / 586
0.53s
Time to first token (AA median)
ProviderNVIDIA
ReleasedDecember 2025
Model IDnvidia/nemotron-3-nano-30b-a3b

Technical specifications

Context window
262K tokens
Reasoning
Yes
Input
Text
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.05NT$ 2
Output$0.20NT$ 6
Cache read$0.03NT$ 1
Cache write
3:1 blended (est.)$0.09NT$ 3
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.07
Output$0.30
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.13
Our price differs from AA reference by ~33%.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 15
Coding 14
Math
MMLU
GPQA 47
This modelCategory leader
Coding ability
#164out of 198 models
LiveCodeBench / SciCode and similar
Overall intelligence
#287out of 573 models
Aggregated across academic benchmarks
Time to first token
#442out of 586 models
AA 全球 p50:0.53 秒
Intelligence
15
Mid-tier
This model
15
Category leader
61
Median
15
Coding
14
Mid-tier
This model
14
Category leader
78
Median
37
GPQA
47%
This model
47%
Category leader
94%
Median
69%
HLE
5%
This model
5%
Category leader
53%
Median
7%
SciCode
28%
This model
28%
Category leader
60%
Median
33%
IFBench
63%
This model
63%
Category leader
83%
Median
44%
τ²-Bench Telecom
45%
This model
45%
Category leader
99%
Median
46%
AA-LCR
36%
This model
36%
Category leader
76%
Median
40%
Terminal-Bench Hard
8%
This model
8%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Nemotron 3 Nano 30b A3b via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="nvidia/nemotron-3-nano-30b-a3b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "nvidia/nemotron-3-nano-30b-a3b",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Nemotron 3 Nano 30b A3b via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Nemotron 3 Nano 30b A3b?

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

How much does the Nemotron 3 Nano 30b A3b API cost?

Nemotron 3 Nano 30b A3b costs $0.0500 per 1M input tokens and $0.2000 per 1M output tokens when accessed through BazaarLink.

How do I use Nemotron 3 Nano 30b A3b with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "nvidia/nemotron-3-nano-30b-a3b". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Nemotron 3 Nano 30b A3b?

Nemotron 3 Nano 30b A3b supports a context window of 262,144 tokens.

Is Nemotron 3 Nano 30b A3b available for free?

Nemotron 3 Nano 30b A3b is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: NVIDIA

Nemotron 3 Super 120b A12bGlm 4.7Deepseek V4 FlashGlm 5Glm 5.2Deepseek V3.2
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.