BazaarLinkBazaarLink
Sign in

Gemini 3.1 Flash Lite Preview API Pricing & Quick Start

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Intelligence#180 / 573
25
Artificial Analysis Intelligence Index
Speed#10 / 586
298.6
Output tokens per second (median)
Input Price#311 / 586
$0.25NT$ 8
USD / 1M tokens
Output Price#350 / 586
$1.50NT$ 49
USD / 1M tokens
First-token latency#548 / 586
5.59s
Time to first token (AA median)
ProviderGoogle
ReleasedMarch 2026
Model IDgoogle/gemini-3.1-flash-lite-preview

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
TextImageVideoFileAudio
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.25NT$ 8
Output$1.50NT$ 49
Cache read$0.03NT$ 1
Cache write$0.08NT$ 3
3:1 blended (est.)$0.56NT$ 18
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.25
Output$1.50
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.56
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 25
Coding 35
Math
MMLU
GPQA 82
This modelCategory leader
Coding ability
#104out of 198 models
LiveCodeBench / SciCode and similar
Overall intelligence
#180out of 573 models
Aggregated across academic benchmarks
Time to first token
#548out of 586 models
AA 全球 p50:5.59 秒
Intelligence
25
Mid-tier
This model
25
Category leader
61
Median
15
Coding
35
Mid-tier
This model
35
Category leader
78
Median
37
GPQA
82%
This model
82%
Category leader
94%
Median
69%
HLE
16%
This model
16%
Category leader
53%
Median
7%
SciCode
42%
This model
42%
Category leader
60%
Median
33%
IFBench
77%
This model
77%
Category leader
83%
Median
44%
τ²-Bench Telecom
31%
This model
31%
Category leader
99%
Median
46%
AA-LCR
65%
This model
65%
Category leader
76%
Median
40%
Terminal-Bench Hard
24%
This model
24%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 3.1 Flash Lite Preview via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-3.1-flash-lite-preview",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-3.1-flash-lite-preview",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 3.1 Flash Lite Preview via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 3.1 Flash Lite Preview?

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

How much does the Gemini 3.1 Flash Lite Preview API cost?

Gemini 3.1 Flash Lite Preview costs $0.2500 per 1M input tokens and $1.5000 per 1M output tokens when accessed through BazaarLink.

How do I use Gemini 3.1 Flash Lite Preview with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "google/gemini-3.1-flash-lite-preview". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 3.1 Flash Lite Preview?

Gemini 3.1 Flash Lite Preview supports a context window of 1,048,576 tokens.

Is Gemini 3.1 Flash Lite Preview available for free?

Gemini 3.1 Flash Lite Preview is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
6 runs
Independent hosts
3
Model-family fingerprint match
0%
Anomalies detected
5 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Google

Gemini 3.1 Pro PreviewGemini 2.5 Flash LiteGemma 3 4b ItGemini Embedding 2 PreviewGemini 2.5 Flash ImageGemini 3.5 Flash
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.