Pick two or three models and see exactly how far apart they are — price per million tokens, context window, and what each one can actually do.
| METRIC | DeepSeek V4 Flash DeepSeek | GLM-5.3-Flash Z.ai | Qwen3.8-Flash Alibaba |
|---|---|---|---|
| Input $/1M | $0.44+487% | $0.075CHEAPEST | $0.15+100% |
| Cached input $/1M | $0.014CHEAPEST | $0.015+7% | $0.016+14% |
| Output $/1M | $1.32+428% | $0.25CHEAPEST | $0.47+88% |
| Blended $/1M (3:1) | $0.66+456% | $0.119CHEAPEST | $0.23+94% |
| Context window | 1M−24% | 1.31MLARGEST | 1M−24% |
| Capabilities | REASONINGTOOL USEOPEN WEIGHTS | REASONINGVISIONVIDEO INPUTTOOL USEOPEN WEIGHTS | REASONINGVISIONVIDEO INPUTTOOL USE |
Percentages are relative to the “best” value in each row — cheapest price, largest context. Blended assumes a typical 3:1 input-to-output token ratio.
ReqKey gives every API key a credit balance — validate, deduct, enforce limits, and see per-consumer spend in real time.
Start metering free