BazaarLinkBazaarLink
Sign in

Deepseek V4 Flash API Pricing & Quick Start

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Intelligence#56 / 573
40
Artificial Analysis Intelligence Index
Speed#70 / 586
117.7
Output tokens per second (median)
Input Price#267 / 586
$0.20NT$ 6
USD / 1M tokens
Output Price#243 / 586
$0.40NT$ 13
USD / 1M tokens
First-token latency#473 / 586
0.86s
Time to first token (AA median)
ProviderDeepSeek
ReleasedApril 2026
Model IDdeepseek/deepseek-v4-flash

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
Text
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.20NT$ 6
Output$0.40NT$ 13
Cache read$0.03NT$ 1
Cache write
3:1 blended (est.)$0.25NT$ 8
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.14
Output$0.28
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.17
Our price differs from AA reference by ~43%.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 40
Coding 56
Math
MMLU
GPQA 89
This modelCategory leader
Coding ability
#53out of 198 models
LiveCodeBench / SciCode and similar
Overall intelligence
#56out of 573 models
Aggregated across academic benchmarks
Time to first token
#473out of 586 models
AA 全球 p50:0.86 秒
Intelligence
40
Top 10%
This model
40
Category leader
61
Median
15
Coding
56
Top 27%
This model
56
Category leader
78
Median
37
GPQA
89%
This model
89%
Category leader
94%
Median
69%
HLE
32%
This model
32%
Category leader
53%
Median
7%
SciCode
45%
This model
45%
Category leader
60%
Median
33%
IFBench
79%
This model
79%
Category leader
83%
Median
44%
τ²-Bench Telecom
95%
This model
95%
Category leader
99%
Median
46%
AA-LCR
63%
This model
63%
Category leader
76%
Median
40%
Terminal-Bench Hard
36%
This model
36%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Deepseek V4 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="deepseek/deepseek-v4-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "deepseek/deepseek-v4-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Deepseek V4 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Deepseek V4 Flash?

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

How much does the Deepseek V4 Flash API cost?

Deepseek V4 Flash costs $0.2000 per 1M input tokens and $0.4000 per 1M output tokens when accessed through BazaarLink.

How do I use Deepseek V4 Flash with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "deepseek/deepseek-v4-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Deepseek V4 Flash?

Deepseek V4 Flash supports a context window of 1,048,576 tokens.

Is Deepseek V4 Flash available for free?

Deepseek V4 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
131 runs
Independent hosts
67
Model-family fingerprint match
67%
Anomalies detected
75 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: DeepSeek

Deepseek V3.2Deepseek V4 ProDeepseek V3.2 ExpDeepseek R1Deepseek V3.1 TerminusDeepseek Chat V3 0324
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.