Which AI provider has the best uptime?

As of July 28, 2026, 27 of 45 providers measured in all 4 regions answered every single probe over the last 24 hours (51,981 checks total). Uptime is not what separates AI APIs today — latency is. Where availability does differ, it usually differs in one region rather than everywhere: kimi sits at 99.9% on average but 99.7% from Asia (Tokyo).

AI API uptime ranking

ProviderUptimeWorst region wherep50 TTFBProbes
anthropic100%100%Asia (Tokyo)194 ms1156
google100%100%Asia (Tokyo)90 ms1156
together100%100%Asia (Tokyo)213 ms1156
xai100%100%Asia (Tokyo)206 ms1156
ai21100%100%Asia (Tokyo)289 ms1155
aleph-alpha100%100%Asia (Tokyo)364 ms1155
baseten100%100%Asia (Tokyo)139 ms1155
cerebras100%100%Asia (Tokyo)165 ms1155
cohere100%100%Asia (Tokyo)148 ms1155
doubao100%100%Asia (Tokyo)451 ms1155
featherless100%100%Asia (Tokyo)568 ms1155
fireworks100%100%Asia (Tokyo)101 ms1155
friendli100%100%Asia (Tokyo)262 ms1155
glm100%100%Asia (Tokyo)547 ms1155
hyperbolic100%100%Asia (Tokyo)456 ms1155
inference-net100%100%Asia (Tokyo)212 ms1155
meta-llama100%100%Asia (Tokyo)120 ms1155
minimax100%100%Asia (Tokyo)318 ms1155
mistral100%100%Asia (Tokyo)215 ms1155
novita100%100%Asia (Tokyo)188 ms1155
nscale100%100%Asia (Tokyo)288 ms1155
perplexity100%100%Asia (Tokyo)192 ms1155
qwen100%100%Asia (Tokyo)433 ms1155
reka100%100%Asia (Tokyo)341 ms1155
replicate100%100%Asia (Tokyo)164 ms1155

Last 24 hours, providers measured in all 4 regions. Ties broken by sample size: 100% over 200 probes is a stronger claim than 100% over 20. Detected outages →

Which AI providers dropped requests?

ProviderUptimeWorst regionwherep50 TTFBProbes
kimi99.9%99.7%Asia (Tokyo)259 ms1156
openai99.9%99.7%Europe (Germany)212 ms1156
baichuan99.9%99.7%South America (São Paulo)235 ms1155
deepinfra99.9%99.7%South America (São Paulo)484 ms1155
deepseek99.9%99.7%US (Central)306 ms1155
ernie99.9%99.7%South America (São Paulo)274 ms1155
groq99.9%99.7%US (Central)198 ms1155
hunyuan99.9%99.7%South America (São Paulo)1422 ms1155
nebius99.9%99.7%South America (São Paulo)329 ms1155
openrouter99.9%99.7%US (Central)66 ms1155
sambanova99.9%99.7%Europe (Germany)234 ms1155
sarvam99.9%99.7%Europe (Germany)398 ms1155
targon99.9%99.7%Asia (Tokyo)258 ms1155
writer99.9%99.7%US (Central)228 ms1155
yi-01ai99.9%99.7%Asia (Tokyo)383 ms1155

A provider can be perfect on average and still be unreliable for you: what matters is the region your traffic comes from, which is why the worst region is shown next to the average.

Why uptime is the wrong question for AI APIs

Endpoint availability among major inference providers has converged. Over the last 24 hours we recorded 51,981 probes across 45 providers and 4 regions, and almost all of them answered every time. Choosing a provider on uptime alone therefore tells you almost nothing.

What still varies by a wide margin is how fast the same provider answers depending on where the request starts — often by several times for identical infrastructure. That difference is measurable, persistent, and it affects every request you make. See the cross-region latency summary.

Where does this data come from?

Independent probes in 4 regions open a real connection to each provider’s official API host every five minutes and record whether it answered and how quickly. Requests go directly to the provider, never through a gateway or aggregator, so nothing is attributed to the wrong party. The method and its limits are on the methodology page; raw figures are available as JSON under CC BY 4.0.