AI Battle > Rankings > Pricing: Cache Hit, Input, and Output
GLM-5.3-Flash by Z AI leads the Pricing: Cache Hit, Input, and Output ranking with a score of 0.15. The ranking covers 20 evaluated llm models, re-scored on every AI Battle telemetry run.
Last updated: . Source: AI Battle live evaluation telemetry, 20 models ranked.
| # | Model | Creator | Score |
|---|---|---|---|
| 1 | GLM-5.3-Flash | Z AI | 0.15 |
| 2 | gpt-oss-120b (high) | OpenAI | 0.15 |
| 3 | GPT-5.6 Luna (max) | OpenAI | 0.20 |
| 4 | MiniMax-M3 | MiniMax | 0.30 |
| 5 | DeepSeek V4.1 Flash (Reasoning, Max Effort) | DeepSeek | 0.30 |
| 6 | Muse Glimmer (high) | Meta | 0.35 |
| 7 | Gemini 3.5 Flash-Lite | 0.30 | |
| 8 | Nemotron 3 Ultra 550B A55B (Reasoning) | NVIDIA | 0.60 |
| 9 | Qwen3.8 27B (xhigh) | Alibaba | 0.50 |
| 10 | Gemini 3.8 Flash (high) | 0.75 |
GLM-5.3-Flash by Z AI currently leads the Pricing: Cache Hit, Input, and Output ranking with a score of 0.15, based on live AI Battle telemetry across LLM Models.
The Pricing: Cache Hit, Input, and Output ranking is refreshed automatically from the AI Battle evaluation cluster. Scores are re-evaluated on every telemetry run, so positions reflect the latest published results.
The top 3 are: 1. GLM-5.3-Flash (0.15), 2. gpt-oss-120b (high) (0.15), 3. GPT-5.6 Luna (max) (0.20).