BazaarLinkBazaarLink
Sign in

Deepseek V4.1 Flash API Pricing & Quick Start

High-performance AI language model, available via BazaarLink with USD billing (TWD-quoted).

Intelligence#39 / 636
40
Artificial Analysis Intelligence Index
Speed#37 / 330
194.3
Output tokens per second (median)
Input Price#165 / 437
$0.30NT$ 9
USD / 1M tokens
Output Price#161 / 437
$1.20NT$ 38
USD / 1M tokens
First-token latency#54 / 330
0.96s
Time to first token (AA median)
ProviderDeepSeek
Released
Model IDdeepseek/deepseek-v4.1-flash

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImage
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.30NT$ 9
Output$1.20NT$ 38
Cache read$0.01NT$ 0
Cache write
3:1 blended (est.)$0.52NT$ 17
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.30
Output$1.20
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 40
Coding
Math
MMLU
GPQA
This modelCategory leader
Overall intelligence
#39out of 636 models
Aggregated across academic benchmarks
Time to first token
#54out of 330 models
AA 全球 p50:0.96 秒
Intelligence
40
Top 6%
This model
40
Category leader
53
Median
11

Top 5 — CodingCoding Index leaderboard

2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Deepseek V4.1 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="deepseek/deepseek-v4.1-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "deepseek/deepseek-v4.1-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Deepseek V4.1 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Deepseek V4.1 Flash?

Deepseek V4.1 Flash is an AI language model by DeepSeek, available via BazaarLink's OpenAI-compatible API.

How much does the Deepseek V4.1 Flash API cost?

Deepseek V4.1 Flash costs $0.30 per 1M input tokens and $1.20 per 1M output tokens when accessed through BazaarLink.

How do I use Deepseek V4.1 Flash with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "deepseek/deepseek-v4.1-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Deepseek V4.1 Flash?

Deepseek V4.1 Flash supports a context window of 1,048,576 tokens.

Is Deepseek V4.1 Flash available for free?

Deepseek V4.1 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
27 runs
Independent hosts
15
Model-family fingerprint match
70%
Anomalies detected
27 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: DeepSeek

Deepseek V3.2Deepseek V4 FlashDeepseek V4 ProDeepseek Chat V3 0324Deepseek V4 Pro 0813Deepseek V3.2 Exp
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.