BazaarLinkBazaarLink
Sign in

Gemini 2.5 Flash Lite API Pricing & Quick Start

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Intelligence#440 / 573
7
Artificial Analysis Intelligence Index
Speed#166 / 586
0.0
Output tokens per second (median)
Input Price#241 / 586
$0.10NT$ 3
USD / 1M tokens
Output Price#258 / 586
$0.40NT$ 13
USD / 1M tokens
First-token latency#1 / 586
0.00s
Time to first token (AA median)
ProviderGoogle
ReleasedJuly 2025
Model IDgoogle/gemini-2.5-flash-lite

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImageFileAudioVideo
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.10NT$ 3
Output$0.40NT$ 13
Cache read$0.01NT$ 0
Cache write$0.08NT$ 3
3:1 blended (est.)$0.18NT$ 6
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.10
Output$0.40
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.17
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 7
Coding
Math 35
MMLU 72
GPQA 47
This modelCategory leader
Overall intelligence
#440out of 573 models
Aggregated across academic benchmarks
Math reasoning
#173out of 269 models
AIME / MATH-500 / SciCode
Time to first token
🥇
#1out of 586 models
AA 全球 p50:0.00 秒
Intelligence
7
Mid-tier
This model
7
Category leader
61
Median
15
Math
35
Mid-tier
This model
35
Category leader
99
Median
53
MMLU Pro
72%
This model
72%
Category leader
90%
Median
75%
GPQA
47%
This model
47%
Category leader
94%
Median
69%
LiveCodeBench
40%
This model
40%
Category leader
42%
Median
42%
HLE
4%
This model
4%
Category leader
53%
Median
7%
SciCode
18%
This model
18%
Category leader
60%
Median
33%
IFBench
31%
This model
31%
Category leader
83%
Median
44%
τ²-Bench Telecom
19%
This model
19%
Category leader
99%
Median
46%
AA-LCR
31%
This model
31%
Category leader
76%
Median
40%
Terminal-Bench Hard
2%
This model
2%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 2.5 Flash Lite via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-2.5-flash-lite",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-2.5-flash-lite",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 2.5 Flash Lite via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 2.5 Flash Lite?

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

How much does the Gemini 2.5 Flash Lite API cost?

Gemini 2.5 Flash Lite costs $0.1000 per 1M input tokens and $0.4000 per 1M output tokens when accessed through BazaarLink.

How do I use Gemini 2.5 Flash Lite with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "google/gemini-2.5-flash-lite". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 2.5 Flash Lite?

Gemini 2.5 Flash Lite supports a context window of 1,048,576 tokens.

Is Gemini 2.5 Flash Lite available for free?

Gemini 2.5 Flash Lite is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
1 runs
Independent hosts
1
Model-family fingerprint match
0%
Anomalies detected
1 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Google

Gemma 3 4b ItGemini Embedding 2 PreviewGemini 3.1 Pro PreviewGemini 2.5 Flash ImageGemini 3.5 FlashGemini 2.5 Pro Preview
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.