Pick two or three models and see exactly how far apart they are — price per million tokens, context window, and what each one can actually do.
| METRIC | MiMo-V2.5 Xiaomi | Gemini 3.5 Flash-Lite Google | Qwen3.8-Flash Alibaba |
|---|---|---|---|
| Input $/1M | $0.12CHEAPEST | $0.30+150% | $0.15+25% |
| Cached input $/1M | — | $0.03+88% | $0.016CHEAPEST |
| Output $/1M | $0.24CHEAPEST | $2.50+942% | $0.47+96% |
| Blended $/1M (3:1) | $0.15CHEAPEST | $0.85+467% | $0.23+53% |
| Context window | 1M−5% | 1.05MLARGEST | 1M−5% |
| Capabilities | REASONINGVISIONAUDIO INPUTVIDEO INPUTTOOL USEOPEN WEIGHTS | REASONINGVISIONAUDIO INPUTVIDEO INPUTTOOL USE | REASONINGVISIONVIDEO INPUTTOOL USE |
Percentages are relative to the “best” value in each row — cheapest price, largest context. Blended assumes a typical 3:1 input-to-output token ratio.
ReqKey gives every API key a credit balance — validate, deduct, enforce limits, and see per-consumer spend in real time.
Start metering free