AI Battle > Rankings > Latency: Time To First Answer Token
Gemini 3.5 Flash-Lite by Google leads the Latency: Time To First Answer Token ranking with a score of 7.59. The ranking covers 20 evaluated llm models, re-scored on every AI Battle telemetry run.
Last updated: . Source: AI Battle live evaluation telemetry, 20 models ranked.
| # | Model | Creator | Score |
|---|---|---|---|
| 1 | Gemini 3.5 Flash-Lite | 7.59 | |
| 2 | gpt-oss-120b (high) | OpenAI | 0.82 |
| 3 | DeepSeek V4.1 Flash (Reasoning, Max Effort) | DeepSeek | 1.17 |
| 4 | Gemini 3.8 Flash (high) | 12.8 | |
| 5 | Nemotron 3 Ultra 550B A55B (Reasoning) | NVIDIA | 2.26 |
| 6 | Mistral Medium 3.5 | Mistral | 2.25 |
| 7 | MiniMax-M3 | MiniMax | 1.33 |
| 8 | GLM-5.3-Flash | Z AI | 2.42 |
| 9 | Muse Glimmer (high) | Meta | 1.06 |
| 10 | Inkling (xhigh) | Thinking Machines | 2.98 |
Gemini 3.5 Flash-Lite by Google currently leads the Latency: Time To First Answer Token ranking with a score of 7.59, based on live AI Battle telemetry across LLM Models.
The Latency: Time To First Answer Token ranking is refreshed automatically from the AI Battle evaluation cluster. Scores are re-evaluated on every telemetry run, so positions reflect the latest published results.
The top 3 are: 1. Gemini 3.5 Flash-Lite (7.59), 2. gpt-oss-120b (high) (0.82), 3. DeepSeek V4.1 Flash (Reasoning, Max Effort) (1.17).