BazaarLinkBazaarLink
Sign in

Gemini 2.5 Flash Lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Intelligence#431 / 563
7
Artificial Analysis Intelligence Index
Speed#38 / 576
190.2
Output tokens per second (median)
Input Price#231 / 576
$0.10NT$ 3
USD / 1M tokens
Output Price#252 / 576
$0.40NT$ 13
USD / 1M tokens
First-token latency#269 / 576
0.32s
Time to first token (AA median)
ProviderGoogle
Released2025年7月
Model IDgoogle/gemini-2.5-flash-lite

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImageFileAudioVideo
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.10NT$ 3
Output$0.40NT$ 13
Cache read$0.01NT$ 0
Cache write$0.08NT$ 3
3:1 blended (est.)$0.18NT$ 6
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.10
Output$0.40
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.17
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 7
Coding
Math 35
MMLU 72
GPQA 47
This modelCategory leader
Overall intelligence
#431out of 563 models
Aggregated across academic benchmarks
Math reasoning
#173out of 269 models
AIME / MATH-500 / SciCode
Time to first token
#269out of 576 models
AA 全球 p50:0.32 秒
Intelligence
7
Mid-tier
This model
7
Category leader
60
Median
15
Math
35
Mid-tier
This model
35
Category leader
99
Median
53
MMLU Pro
72%
This model
72%
Category leader
90%
Median
75%
GPQA
47%
This model
47%
Category leader
94%
Median
68%
LiveCodeBench
40%
This model
40%
Category leader
42%
Median
42%
HLE
4%
This model
4%
Category leader
53%
Median
7%
SciCode
18%
This model
18%
Category leader
60%
Median
33%
IFBench
31%
This model
31%
Category leader
83%
Median
44%
τ²-Bench Telecom
19%
This model
19%
Category leader
99%
Median
46%
AA-LCR
31%
This model
31%
Category leader
76%
Median
39%
Terminal-Bench Hard
2%
This model
2%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
3GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 2.5 Flash Lite via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-2.5-flash-lite",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-2.5-flash-lite",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 2.5 Flash Lite via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 2.5 Flash Lite?

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

How much does the Gemini 2.5 Flash Lite API cost?

Gemini 2.5 Flash Lite costs $0.1000 per 1K input tokens and $0.4000 per 1K output tokens when accessed through BazaarLink.

How do I use Gemini 2.5 Flash Lite with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "google/gemini-2.5-flash-lite". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 2.5 Flash Lite?

Gemini 2.5 Flash Lite supports a context window of 1,048,576 tokens.

Is Gemini 2.5 Flash Lite available for free?

Gemini 2.5 Flash Lite is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

中轉站誠信檢測 · BazaarLink 獨家

已測試 endpoint
4
獨立 hosts
2
家族指紋一致率
0%
偵測到異常
4
資料範圍
過去 90

BazaarLink Probe 對聲稱提供此模型的 endpoint 進行家族指紋驗證與 V3 子模型對照。 如發現行為與真貨基準不符,會列入異常案例。

查看完整檢測報告 →

Related Models & Links

More from Google

Gemma 3 27b ItGemini 3.1 Flash Image PreviewGemini 2.5 ProGemma 2 27b ItGemini 3.1 Flash Lite PreviewGemma 3 4b It
All ModelsAPI DocumentationAPI Latency Probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.