AI Reliability Leaderboard
Compare 90-day reliability signals across 800+ AI services: confirmed incidents, community reports, monitoring confidence, and verified availability. Not a model-speed benchmark.
When most AI services have no confirmed hard outage, a single “most reliable” ranking would be misleading. DownForAI shows the underlying reliability signals instead.
Last updated: Sat, 03 Oct 2026 20:13:47 GMT
Key signals (90 days)
Most reported AI services (90d)
Raw community reports over the last 90 days. Not normalized by user base — popular services may receive more reports.
| Service | Category | Reports 90d | Community-detected outages | Confirmed incidents | Availability | Confidence |
|---|---|---|---|---|---|---|
| Chai AI | Roleplay AI | 258 | 19 | 0 | 100.00% | Basic probe |
| HiWaifu | Roleplay AI | 106 | 12 | 0 | 100.00% | Basic probe |
| Chub AI | Roleplay AI | 76 | 4 | 0 | 100.00% | Basic probe |
| SpicyChat AI | Roleplay AI | 31 | 2 | 0 | 100.00% | Basic probe |
| Ollama | LLM | 26 | — | 0 | 100.00% | Basic probe |
| OpenAI | LLM | 26 | 2 | 3 | 100.00% | Official API |
| JanitorAI | Roleplay AI | 25 | — | 0 | 100.00% | Status page |
| DeepSeek | LLM | 21 | 2 | 0 | 100.00% | Basic probe |
| Kiro | Dev Tools | 20 | — | 0 | 100.00% | Basic probe |
| NVIDIA NIM | Infrastructure | 20 | — | 0 | 100.00% | Status page |
| Voicemod | Audio | 20 | — | 0 | 100.00% | Basic probe |
| GitHub Copilot | Dev Tools | 20 | — | 4 | 100.00% | Official API |
| ChatGPT | LLM | 16 | 2 | 3 | 100.00% | Official API |
| Vast.ai | Infrastructure | 12 | 2 | 0 | 100.00% | Basic probe |
| Civitai | Image | 11 | — | 0 | 100.00% | Status page |
| Candy AI | Roleplay AI | 11 | 2 | 0 | 100.00% | Basic probe |
| Talkie AI | Roleplay AI | 10 | — | 0 | 100.00% | Basic probe |
| DuckDuckGo AI | LLM | 8 | — | 2 | 100.00% | Basic probe |
| Moonshot AI (Kimi) | LLM | 8 | — | 7 | 100.00% | Official API |
| Microsoft Copilot | LLM | 8 | — | 0 | 100.00% | Basic probe |
Confirmed incidents (90d)
Services with confirmed outage or degradation incidents observed by DownForAI. Incident minutes sum the duration of confirmed incidents.
| Service | Category | Incidents | Incident minutes | Degraded checks | Availability | Source |
|---|---|---|---|---|---|---|
| Elastic AI Search | Search | 30 | 124,496 | 0 | 0.00% | Official API |
| Mate AI | Roleplay AI | 19 | 72,490 | 0 | 0.00% | Basic probe |
| Kajiwoto | Roleplay AI | 8 | 29,940 | 0 | 100.00% | Basic probe |
| Moonshot AI (Kimi) | LLM | 7 | 17,537 | 0 | 100.00% | Official API |
| Snowflake Cortex | Vector DB | 5 | 7,607 | 47 | 100.00% | Official API |
| Rakuten AI | Productivity | 6 | 4,426 | 0 | 100.00% | Basic probe |
| Supabase | Dev Tools | 1 | 3,916 | 83 | 100.00% | Official API |
| Supabase Vector | Vector DB | 1 | 3,871 | 41 | 100.00% | Official API |
| Cloudflare Workers AI | Infrastructure | 1 | 3,495 | 46 | 100.00% | Official API |
| AskCodi | Dev Tools | 1 | 2,580 | 0 | 100.00% | Basic probe |
| ElevenLabs | Audio | 4 | 1,230 | 0 | 98.91% | Official API |
| ChatGPT | LLM | 3 | 1,199 | 11 | 100.00% | Official API |
| OpenAI Operator | Agents | 3 | 1,199 | 11 | 97.83% | Official API |
| OpenAI API | Dev Tools | 3 | 1,141 | 10 | 100.00% | Official API |
| Retool AI | Agents | 1 | 1,140 | 0 | 100.00% | Official API |
| OpenAI Sora | Video | 3 | 1,125 | 11 | 100.00% | Official API |
| GPT Image (OpenAI) | Image | 3 | 1,125 | 11 | 100.00% | Official API |
| OpenAI | LLM | 3 | 1,109 | 45 | 100.00% | Official API |
| OpenAI Whisper | Audio | 3 | 1,096 | 12 | 97.83% | Official API |
| Miro AI | Productivity | 2 | 1,020 | 1 | 100.00% | Official API |
| GitHub Models | Dev Tools | 4 | 960 | 0 | 100.00% | Official API |
| AI21 Labs | LLM | 1 | 960 | 0 | 100.00% | Official API |
| GitHub Copilot | Dev Tools | 4 | 901 | 2 | 100.00% | Official API |
| BentoML | Infrastructure | 1 | 614 | 0 | 100.00% | Basic probe |
| Notion AI | Productivity | 1 | 360 | 0 | 100.00% | Official API |
| Render AI | Dev Tools | 2 | 359 | 1 | 100.00% | Official API |
| MongoDB Atlas Vector | Vector DB | 1 | 240 | 2 | 100.00% | Official API |
| Copymatic | Marketing AI | 1 | 180 | 0 | 100.00% | Basic probe |
| Botpress | Agents | 1 | 180 | 0 | 100.00% | Basic probe |
| DuckDuckGo AI | LLM | 2 | 119 | 0 | 100.00% | Basic probe |
| Airtable AI | Productivity | 1 | 60 | 0 | 100.00% | Official API |
| Linear | Productivity | 1 | 60 | 0 | 100.00% | Official API |
| Claude Chat | LLM | 1 | 60 | 1 | 100.00% | Official API |
| Anthropic | LLM | 1 | 60 | 3 | 100.00% | Official API |
Officially monitored reliability leaders
Restricted to the 127 services with an official status API or status page, where availability is measured comparably. Grouped by outcome rather than ranked — most have no confirmed hard outage.
No confirmed incidents or degradations (90d) — 96
Confirmed degradations / incidents, no hard outage — 27
| Service | Source | Availability | Incidents | Degraded checks | Reports |
|---|---|---|---|---|---|
| Supabase | Official API | 100.00% | 1 | 83 | 0 |
| Snowflake Cortex | Official API | 100.00% | 5 | 47 | 2 |
| Zoom AI Companion | Official API | 100.00% | 0 | 46 | 0 |
| Cloudflare Workers AI | Official API | 100.00% | 1 | 46 | 2 |
| OpenAI | Official API | 100.00% | 3 | 45 | 26 |
| Couchbase Capella | Official API | 100.00% | 0 | 45 | 0 |
| Supabase Vector | Official API | 100.00% | 1 | 41 | 0 |
| ChatGPT | Official API | 100.00% | 3 | 11 | 16 |
| OpenAI Sora | Official API | 100.00% | 3 | 11 | 0 |
| GPT Image (OpenAI) | Official API | 100.00% | 3 | 11 | 0 |
| OpenAI API | Official API | 100.00% | 3 | 10 | 0 |
| Anthropic | Official API | 100.00% | 1 | 3 | 4 |
| MongoDB Atlas Vector | Official API | 100.00% | 1 | 2 | 0 |
| GitHub Copilot | Official API | 100.00% | 4 | 2 | 20 |
| Snyk Code AI | Official API | 100.00% | 0 | 1 | 0 |
| Render AI | Official API | 100.00% | 2 | 1 | 0 |
| v0 by Vercel | Official API | 100.00% | 0 | 1 | 0 |
| Anthropic API | Official API | 100.00% | 0 | 1 | 1 |
| Claude Chat | Official API | 100.00% | 1 | 1 | 1 |
| Miro AI | Official API | 100.00% | 2 | 1 | 0 |
| Airtable AI | Official API | 100.00% | 1 | 0 | 0 |
| Notion AI | Official API | 100.00% | 1 | 0 | 0 |
| AI21 Labs | Official API | 100.00% | 1 | 0 | 0 |
| Retool AI | Official API | 100.00% | 1 | 0 | 0 |
| Moonshot AI (Kimi) | Official API | 100.00% | 7 | 0 | 8 |
| GitHub Models | Official API | 100.00% | 4 | 0 | 0 |
| Linear | Official API | 100.00% | 1 | 0 | 0 |
Confirmed outages (90d) — 4
| Service | Source | Availability | Incidents | Degraded checks | Reports |
|---|---|---|---|---|---|
| Elastic AI Search | Official API | 0.00% | 30 | 0 | 0 |
| OpenAI Operator | Official API | 97.83% | 3 | 11 | 0 |
| OpenAI Whisper | Official API | 97.83% | 3 | 12 | 0 |
| ElevenLabs | Official API | 98.91% | 4 | 0 | 0 |
Reliability leaders by category
Browse like-for-like reliability rankings within each category.
LLM · 53 tracked
View full LLM reliability rankings →Image · 54 tracked
View full Image reliability rankings →Video · 52 tracked
View full Video reliability rankings →Audio · 50 tracked
View full Audio reliability rankings →Dev Tools · 74 tracked
View full Dev Tools reliability rankings →Infrastructure · 47 tracked
View full Infrastructure reliability rankings →Search · 24 tracked
View full Search reliability rankings →Productivity · 55 tracked
View full Productivity reliability rankings →How reliable is our signal?
Monitoring confidence reflects how a service is observed — the quality of our measurement, not the service's reliability.
How this leaderboard works
- Availability = non-outage rate over a rolling 90-day window. “Down” means a hard OUTAGE only. A degraded period is not downtime — it is counted as a separate signal. A blocked or rate-limited probe is never counted as an outage.
- Confirmed incidents exclude false positives. Community reports are user-submitted, counted as raw totals not normalized by user base, and shown as their own independent signal — a popular service naturally receives more.
- Monitoring confidence (official API, status page, basic probe, limited, unverifiable) describes the quality of our measurement, not the service's reliability — we never rank by it.
- Each surface is re-checked roughly every 75 minutes. Any response-time figures shown elsewhere reflect the monitored surface (often a homepage or status page) — not model inference speed or tokens-per-second.