BazaarLinkBazaarLink
Sign in

Deepseek V4 Flash API Pricing & Quick Start

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Intelligence#56 / 638
35
Artificial Analysis Intelligence Index
Speed#30 / 329
210.1
Output tokens per second (median)
Input Price#218 / 437
$0.20NT$ 6
USD / 1M tokens
Output Price#187 / 437
$0.40NT$ 13
USD / 1M tokens
First-token latency#69 / 329
1.08s
Time to first token (AA median)
ProviderDeepSeek
ReleasedApril 2026
Model IDdeepseek/deepseek-v4-flash

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
Text
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.20NT$ 6
Output$0.40NT$ 13
Cache read$0.03NT$ 1
Cache write
3:1 blended (est.)$0.25NT$ 8
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.44
Output$1.32
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.17
Our price differs from AA reference by ~55%.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 35
Coding 69
Math
MMLU
GPQA 91
This modelCategory leader
Coding ability
#53out of 255 models
LiveCodeBench / SciCode and similar
Overall intelligence
#56out of 638 models
Aggregated across academic benchmarks
Time to first token
#69out of 329 models
AA 全球 p50:1.08 秒
Intelligence
35
Top 9%
This model
35
Category leader
53
Median
12
Coding
69
Top 21%
This model
69
Category leader
82
Median
45
GPQA
91%
This model
91%
Category leader
94%
Median
69%
HLE
37%
This model
37%
Category leader
53%
Median
7%
SciCode
50%
This model
50%
Category leader
60%
Median
34%
AA-LCR
66%
This model
66%
Category leader
76%
Median
40%

Top 5 — CodingCoding Index leaderboard

2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Deepseek V4 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="deepseek/deepseek-v4-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "deepseek/deepseek-v4-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Deepseek V4 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Deepseek V4 Flash?

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

How much does the Deepseek V4 Flash API cost?

Deepseek V4 Flash costs $0.20 per 1M input tokens and $0.40 per 1M output tokens when accessed through BazaarLink.

How do I use Deepseek V4 Flash with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "deepseek/deepseek-v4-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Deepseek V4 Flash?

Deepseek V4 Flash supports a context window of 1,048,576 tokens.

Is Deepseek V4 Flash available for free?

Deepseek V4 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
1,207 runs
Independent hosts
469
Model-family fingerprint match
80%
Anomalies detected
946 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: DeepSeek

Deepseek V3.2Deepseek V4 ProDeepseek Chat V3 0324Deepseek V4 Pro 0813Deepseek V3.2 ExpDeepseek R1 0528
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.