BazaarLinkBazaarLink
Sign in

Gemini 2.5 Flash Lite API Pricing & Quick Start

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Intelligence#493 / 638
7
Artificial Analysis Intelligence Index
Speed#13 / 329
271.5
Output tokens per second (median)
Input Price#62 / 437
$0.10NT$ 3
USD / 1M tokens
Output Price#80 / 437
$0.40NT$ 13
USD / 1M tokens
First-token latency#1 / 329
0.31s
Time to first token (AA median)
ProviderGoogle
ReleasedJuly 2025
Model IDgoogle/gemini-2.5-flash-lite

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImageFileAudioVideo
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.10NT$ 3
Output$0.40NT$ 13
Cache read$0.01NT$ 0
Cache write$0.08NT$ 3
3:1 blended (est.)$0.18NT$ 6
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.10
Output$0.40
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.17
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 7
Coding
Math 35
MMLU 72
GPQA 47
This modelCategory leader
Overall intelligence
#493out of 638 models
Aggregated across academic benchmarks
Time to first token
🥇
#1out of 329 models
AA 全球 p50:0.31 秒
Intelligence
7
Mid-tier
This model
7
Category leader
53
Median
12
Math
35
This model
35
Category leader
99
Median
53
MMLU Pro
72%
This model
72%
Category leader
90%
Median
75%
GPQA
47%
This model
47%
Category leader
94%
Median
69%
LiveCodeBench
40%
This model
40%
Category leader
42%
Median
42%
HLE
4%
This model
4%
Category leader
53%
Median
7%
SciCode
18%
This model
18%
Category leader
60%
Median
34%
IFBench
31%
This model
31%
Category leader
83%
Median
44%
τ²-Bench Telecom
19%
This model
19%
Category leader
99%
Median
46%
AA-LCR
31%
This model
31%
Category leader
76%
Median
40%
Terminal-Bench Hard
2%
This model
2%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 2.5 Flash Lite via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-2.5-flash-lite",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-2.5-flash-lite",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 2.5 Flash Lite via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 2.5 Flash Lite?

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

How much does the Gemini 2.5 Flash Lite API cost?

Gemini 2.5 Flash Lite costs $0.10 per 1M input tokens and $0.40 per 1M output tokens when accessed through BazaarLink.

How do I use Gemini 2.5 Flash Lite with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "google/gemini-2.5-flash-lite". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 2.5 Flash Lite?

Gemini 2.5 Flash Lite supports a context window of 1,048,576 tokens.

Is Gemini 2.5 Flash Lite available for free?

Gemini 2.5 Flash Lite is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
3 runs
Independent hosts
2
Model-family fingerprint match
0%
Anomalies detected
3 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Google

Gemini 3.1 Flash Lite PreviewGemini 3.1 Pro PreviewGemma 4 31b ItGemma 4 26b A4b ItGemini 3.1 Flash Image PreviewGemma 3 4b It
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.