BazaarLinkBazaarLink
Sign in
D

Cogito V2 Preview Llama 109b Moe — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 1, 2026.

An instruction-tuned, hybrid-reasoning Mixture-of-Experts model built on Llama-4-Scout-17B-16E. Cogito v2 can answer directly or engage an extended “thinking” phase, with alignment guided by Iterated Distillation & Amplification (IDA). It targets coding, STEM, instruction following, and general helpfulness, with stronger multilingual, tool-calling, and reasoning performance than size-equivalent baselines. The model supports long-context use (up to 10M tokens) and standard Transformers workflows. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. Learn more in our docs

ProviderDdeepcogito
ReleasedSeptember 2025
Model IDdeepcogito/cogito-v2-preview-llama-109b-moe

Technical specifications

Context window
131K tokens
Reasoning
Input
ImageText
Output
Text
← All Models

Frequently Asked Questions

Can I still use the retired Cogito V2 Preview Llama 109b Moe model?

No. Cogito V2 Preview Llama 109b Moe has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

Related models and links

Glm 4.7Glm 5Glm 5.2Deepseek V3.2Deepseek V4 ProGlm 5.1
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.