BazaarLinkBazaarLink
Sign in

Gemini 3.5 Flash Lite API Pricing & Quick Start

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Intelligence#84 / 573
37
Artificial Analysis Intelligence Index
Speed#6 / 586
362.2
Output tokens per second (median)
Input Price#332 / 586
$0.30NT$ 10
USD / 1M tokens
Output Price#388 / 586
$2.50NT$ 81
USD / 1M tokens
First-token latency#553 / 586
7.41s
Time to first token (AA median)
ProviderGoogle
ReleasedJuly 2026
Model IDgoogle/gemini-3.5-flash-lite

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImageVideoFileAudio
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.30NT$ 10
Output$2.50NT$ 81
Cache read
Cache write
3:1 blended (est.)$0.85NT$ 28
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.30
Output$2.50
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.85
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 37
Coding 49
Math
MMLU
GPQA 84
This modelCategory leader
Coding ability
#70out of 198 models
LiveCodeBench / SciCode and similar
Overall intelligence
#84out of 573 models
Aggregated across academic benchmarks
Time to first token
#553out of 586 models
AA 全球 p50:7.41 秒
Intelligence
37
Top 15%
This model
37
Category leader
61
Median
15
Coding
49
Mid-tier
This model
49
Category leader
78
Median
37
GPQA
84%
This model
84%
Category leader
94%
Median
69%
HLE
18%
This model
18%
Category leader
53%
Median
7%
SciCode
41%
This model
41%
Category leader
60%
Median
33%
AA-LCR
62%
This model
62%
Category leader
76%
Median
40%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 3.5 Flash Lite via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-3.5-flash-lite",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-3.5-flash-lite",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 3.5 Flash Lite via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

How much does the Gemini 3.5 Flash Lite API cost?

Gemini 3.5 Flash Lite costs $0.3000 per 1M input tokens and $2.5000 per 1M output tokens when accessed through BazaarLink.

How do I use Gemini 3.5 Flash Lite with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "google/gemini-3.5-flash-lite". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite supports a context window of 1,048,576 tokens.

Is Gemini 3.5 Flash Lite available for free?

Gemini 3.5 Flash Lite is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
3 runs
Independent hosts
1
Model-family fingerprint match
100%
Anomalies detected
0 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Google

Gemini 3.1 Pro PreviewGemini 3.1 Flash Lite PreviewGemini 2.5 Flash LiteGemma 3 4b ItGemini Embedding 2 PreviewGemini 2.5 Flash Image
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.