BazaarLinkBazaarLink
Sign in
N

Nemotron 4 340b Instruct

Nemotron-4-340B-Instruct is an English-language chat model optimized for synthetic data generation. This large language model (LLM) is a fine-tuned version of Nemotron-4-340B-Base, designed for single and multi-turn chat use-cases with a 4,096 token context length. The base model was pre-trained on 9 trillion tokens from diverse English texts, 50+ natural languages, and 40+ coding languages. The instruct model underwent additional alignment steps: 1. Supervised Fine-tuning (SFT) 2. Direct Preference Optimization (DPO) 3. Reward-aware Preference Optimization (RPO) The alignment process used approximately 20K human-annotated samples, while 98% of the data for fine-tuning was synthetically generated. Detailed information about the synthetic data generation pipeline is available in the [technical report](https://arxiv.org/html/2406.11704v1).

Pricing

Input Price
/ 1M tokens
Output Price
/ 1M tokens
Context Window
4K
tokens

💱 匯率 USD/NTD = 32.34 · 不含稅 · 更新於 2026/07/22

ProviderNNVIDIA
Released2024年6月
Model IDnemotron-4-340b-instruct

Technical specifications

Context window
4K tokens
Reasoning
Input
Text
Output
Text
!
此模型已下架 — API 呼叫會回傳 HTTP 410,但本頁面保留以利歷史查詢

Quick Start

Use Nemotron 4 340b Instruct via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="nemotron-4-340b-instruct",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "nemotron-4-340b-instruct",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Nemotron 4 340b Instruct via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Nemotron 4 340b Instruct?

Nemotron-4-340B-Instruct is an English-language chat model optimized for synthetic data generation. This large language model (LLM) is a fine-tuned version of Nemotron-4-340B-Base, designed for single and multi-turn chat use-cases with a 4,096 token context length. The base model was pre-trained on 9 trillion tokens from diverse English texts, 50+ natural languages, and 40+ coding languages. The instruct model underwent additional alignment steps: 1. Supervised Fine-tuning (SFT) 2. Direct Preference Optimization (DPO) 3. Reward-aware Preference Optimization (RPO) The alignment process used approximately 20K human-annotated samples, while 98% of the data for fine-tuning was synthetically generated. Detailed information about the synthetic data generation pipeline is available in the [technical report](https://arxiv.org/html/2406.11704v1).

How much does the Nemotron 4 340b Instruct API cost?

Nemotron 4 340b Instruct pricing is available on this page. BazaarLink bills in TWD with no additional markup over the provider's list price.

How do I use Nemotron 4 340b Instruct with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "nemotron-4-340b-instruct". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Nemotron 4 340b Instruct?

Nemotron 4 340b Instruct supports a context window of 4,096 tokens.

Is Nemotron 4 340b Instruct available for free?

Nemotron 4 340b Instruct is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

中轉站誠信檢測 · BazaarLink 獨家

尚無此模型的檢測資料

BazaarLink Probe 對聲稱提供此模型的 endpoint 進行家族指紋驗證與 V3 子模型對照。 如發現行為與真貨基準不符,會列入異常案例。

查看完整檢測報告 →

Related Models & Links

All ModelsAPI DocumentationAPI Latency Probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.