Lfm 2.5 1.2b Thinking
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is designed to provide higher-quality “thinking” responses in a small 1.2B model.
Pricing
💱 匯率 USD/NTD = 32.34 · 不含稅 · 更新於 2026/07/22
Technical specifications
Quick Start
Use Lfm 2.5 1.2b Thinking via BazaarLink API — just change the base_url:
from openai import OpenAI
client = OpenAI(
base_url="https://bazaarlink.ai/api/v1",
api_key="sk-bl-YOUR_API_KEY",
)
response = client.chat.completions.create(
model="lfm-2.5-1.2b-thinking",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://bazaarlink.ai/api/v1",
apiKey: "sk-bl-YOUR_API_KEY",
});
const response = await client.chat.completions.create({
model: "lfm-2.5-1.2b-thinking",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);Why use Lfm 2.5 1.2b Thinking via BazaarLink?
- ✓USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
- ✓OpenAI-compatible API — zero code changes required
- ✓Automatic failover — multi-provider redundancy for the same model
- ✓Chinese-language support — local team, instant help
Frequently Asked Questions
What is Lfm 2.5 1.2b Thinking?
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is designed to provide higher-quality “thinking” responses in a small 1.2B model.
How much does the Lfm 2.5 1.2b Thinking API cost?
Lfm 2.5 1.2b Thinking costs $0.0000 per 1K input tokens and $0.0000 per 1K output tokens when accessed through BazaarLink.
How do I use Lfm 2.5 1.2b Thinking with the OpenAI SDK?
Set base_url to "https://bazaarlink.ai/api/v1" and use model ID "lfm-2.5-1.2b-thinking". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.
What is the context window for Lfm 2.5 1.2b Thinking?
Lfm 2.5 1.2b Thinking supports a context window of 32,768 tokens.
Is Lfm 2.5 1.2b Thinking available for free?
Yes — Lfm 2.5 1.2b Thinking is available at no cost through BazaarLink's free tier. No credit card required.