Gemini Flash 1.5 8b — Retired model
Gemini Flash 1.5 8B is optimized for speed and efficiency, offering enhanced performance in small prompt tasks like chat, transcription, and translation. With reduced latency, it is highly effective for real-time and large-scale operations. This model focuses on cost-effective solutions while maintaining high-quality results. [Click here to learn more about this model](https://developers.googleblog.com/en/gemini-15-flash-8b-is-now-generally-available-for-use/). Usage of Gemini is subject to Google's [Gemini Terms of Use](https://ai.google.dev/terms).
Technical specifications
Independent benchmarks
Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.
Top 5 — CodingCoding Index leaderboard
openai/gpt-5-6-sol-xhigh78$5 / $30openai/gpt-5-6-sol-high77$5 / $30anthropic/claude-opus-5-xhigh77$5 / $25Frequently Asked Questions
Can I still use the retired Gemini Flash 1.5 8b model?
No. Gemini Flash 1.5 8b has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.
Relay integrity checks
No integrity-check data is available for this model yet.
BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.
View the full integrity report →Related models and links
More from this provider: Google