BazaarLinkBazaarLink
Sign in

Glm 4.7 Flash API Pricing & Quick Start

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

Intelligence#196 / 573
23
Artificial Analysis Intelligence Index
Speed#166 / 586
0.0
Output tokens per second (median)
Input Price#236 / 586
$0.06NT$ 2
USD / 1M tokens
Output Price#258 / 586
$0.40NT$ 13
USD / 1M tokens
First-token latency#1 / 586
0.00s
Time to first token (AA median)
ProviderZhipu AI
ReleasedJanuary 2026
Model IDz-ai/glm-4.7-flash

Technical specifications

Context window
203K tokens
Reasoning
Yes
Input
Text
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.06NT$ 2
Output$0.40NT$ 13
Cache read$0.01NT$ 0
Cache write
3:1 blended (est.)$0.15NT$ 5
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.07
Output$0.40
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.15
Our price differs from AA reference by ~14%.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 23
Coding
Math
MMLU
GPQA 58
This modelCategory leader
Overall intelligence
#196out of 573 models
Aggregated across academic benchmarks
Time to first token
🥇
#1out of 586 models
AA 全球 p50:0.00 秒
Intelligence
23
Mid-tier
This model
23
Category leader
61
Median
15
GPQA
58%
This model
58%
Category leader
94%
Median
69%
HLE
7%
This model
7%
Category leader
53%
Median
7%
SciCode
34%
This model
34%
Category leader
60%
Median
33%
IFBench
61%
This model
61%
Category leader
83%
Median
44%
τ²-Bench Telecom
99%
This model
99%
Category leader
99%
Median
46%
AA-LCR
35%
This model
35%
Category leader
76%
Median
40%
Terminal-Bench Hard
22%
This model
22%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$5 / $30
4GPT-5.6 Sol (high)openai/gpt-5-6-sol-high77$5 / $30
5Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)anthropic/claude-opus-5-xhigh77$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Glm 4.7 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://bazaarlink.ai/api/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="z-ai/glm-4.7-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://bazaarlink.ai/api/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "z-ai/glm-4.7-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Glm 4.7 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Glm 4.7 Flash?

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

How much does the Glm 4.7 Flash API cost?

Glm 4.7 Flash costs $0.0600 per 1M input tokens and $0.4000 per 1M output tokens when accessed through BazaarLink.

How do I use Glm 4.7 Flash with the OpenAI SDK?

Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "z-ai/glm-4.7-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Glm 4.7 Flash?

Glm 4.7 Flash supports a context window of 202,752 tokens.

Is Glm 4.7 Flash available for free?

Glm 4.7 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Zhipu AI

Glm 4.7Glm 5Glm 5.2Glm 5.1Glm 5v TurboGlm 4.6v
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.