# AI Latency Tracker
Updated July 28, 2026. Independent, provider-neutral latency and uptime of 45 AI inference APIs, measured from 4 regions with distributed probes and updated automatically. Figures are measured, not scraped.

## By region
- [Fastest AI API from Asia (Tokyo)](https://llmlatency.dev/region/ap-tokyo)
- [Fastest AI API from Europe (Germany)](https://llmlatency.dev/region/eu-hetzner)
- [Fastest AI API from South America (São Paulo)](https://llmlatency.dev/region/sa-east)
- [Fastest AI API from US (Central)](https://llmlatency.dev/region/us-central)

## By provider
- [ai21](https://llmlatency.dev/provider/ai21)
- [aleph-alpha](https://llmlatency.dev/provider/aleph-alpha)
- [anthropic](https://llmlatency.dev/provider/anthropic)
- [baichuan](https://llmlatency.dev/provider/baichuan)
- [baseten](https://llmlatency.dev/provider/baseten)
- [cerebras](https://llmlatency.dev/provider/cerebras)
- [cohere](https://llmlatency.dev/provider/cohere)
- [deepinfra](https://llmlatency.dev/provider/deepinfra)
- [deepseek](https://llmlatency.dev/provider/deepseek)
- [doubao](https://llmlatency.dev/provider/doubao)
- [ernie](https://llmlatency.dev/provider/ernie)
- [featherless](https://llmlatency.dev/provider/featherless)
- [fireworks](https://llmlatency.dev/provider/fireworks)
- [friendli](https://llmlatency.dev/provider/friendli)
- [glm](https://llmlatency.dev/provider/glm)
- [google](https://llmlatency.dev/provider/google)
- [groq](https://llmlatency.dev/provider/groq)
- [hunyuan](https://llmlatency.dev/provider/hunyuan)
- [hyperbolic](https://llmlatency.dev/provider/hyperbolic)
- [iflytek](https://llmlatency.dev/provider/iflytek)
- [inference-net](https://llmlatency.dev/provider/inference-net)
- [kimi](https://llmlatency.dev/provider/kimi)
- [meta-llama](https://llmlatency.dev/provider/meta-llama)
- [minimax](https://llmlatency.dev/provider/minimax)
- [mistral](https://llmlatency.dev/provider/mistral)
- [nebius](https://llmlatency.dev/provider/nebius)
- [novita](https://llmlatency.dev/provider/novita)
- [nscale](https://llmlatency.dev/provider/nscale)
- [openai](https://llmlatency.dev/provider/openai)
- [openrouter](https://llmlatency.dev/provider/openrouter)
- [perplexity](https://llmlatency.dev/provider/perplexity)
- [qwen](https://llmlatency.dev/provider/qwen)
- [reka](https://llmlatency.dev/provider/reka)
- [replicate](https://llmlatency.dev/provider/replicate)
- [sambanova](https://llmlatency.dev/provider/sambanova)
- [sarvam](https://llmlatency.dev/provider/sarvam)
- [sensenova](https://llmlatency.dev/provider/sensenova)
- [siliconflow](https://llmlatency.dev/provider/siliconflow)
- [stepfun](https://llmlatency.dev/provider/stepfun)
- [targon](https://llmlatency.dev/provider/targon)
- [together](https://llmlatency.dev/provider/together)
- [upstage](https://llmlatency.dev/provider/upstage)
- [writer](https://llmlatency.dev/provider/writer)
- [xai](https://llmlatency.dev/provider/xai)
- [yi-01ai](https://llmlatency.dev/provider/yi-01ai)

## More
- [Methodology](https://llmlatency.dev/methodology) · [Glossary](https://llmlatency.dev/glossary)
- [Provider comparisons](https://llmlatency.dev/compare/anthropic-vs-openai) (e.g. openai vs anthropic, cerebras vs groq)
- [AI model deprecation & migration calendar](https://llmlatency.dev/deprecations)
- [llms.txt](https://llmlatency.dev/llms.txt) · [llms-full.txt](https://llmlatency.dev/llms-full.txt)
