BazaarLinkBazaarLink
Sign in
F

Qwerky 72b — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 1, 2026.

Qrwkv-72B is a linear-attention RWKV variant of the Qwen 2.5 72B model, optimized to significantly reduce computational cost at scale. Leveraging linear attention, it achieves substantial inference speedups (>1000x) while retaining competitive accuracy on common benchmarks like ARC, HellaSwag, Lambada, and MMLU. It inherits knowledge and language support from Qwen 2.5, supporting approximately 30 languages, making it suitable for efficient inference in large-context applications.

ProviderFfeatherless
ReleasedMarch 2025
Model IDfeatherless/qwerky-72b

Technical specifications

Context window
33K tokens
Reasoning
Input
Text
Output
Text
← All Models

Frequently Asked Questions

Can I still use the retired Qwerky 72b model?

No. Qwerky 72b has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

Related models and links

Glm 4.7Glm 5Glm 5.2Deepseek V3.2Deepseek V4 ProGlm 5.1
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.