On this page
Status pages answer one question well: is it broken right now. They answer the next one badly: how often does it break. Ask the Claude status API for its incidents and you get the latest 50, which on 4 October 2026 reached back only to 25 July. OpenAI's returns 25, which reached back to 14 September. Ask either about last spring and the record is gone from the API.
So I started keeping it. Once a day a script reads 12 public status pages, stores each incident once under the provider and the incident's own id, and never deletes one. Where a page still exposes older history, I pulled that in too. This page is what the archive says, rebuilt every day.
Gravity's agents run on model APIs like these, which is why their incidents are the first thing I read in the morning. The numbers below are the providers' own words, counted carefully, and nothing more.
- OpenAI API: 18 incidents in the 90 days to 4 October 2026, 4 of them major or critical, adding up to 210 hours with an open incident on at least one API component.
- Claude API: 60 incidents in the 90 days to 4 October 2026, 22 major or critical, 102 hours with an open incident, and a median of 1 h 9 min from start to resolved.
- Most hours with an open incident among model APIs in the last 30 days: OpenAI API, 21 hours across 6 incidents (30 days to 4 October 2026).
- 5 model APIs posted no incident in the same 30 days: Gemini on Vertex AI (Google Cloud), Cohere, Groq, Perplexity API, Vercel AI Gateway. No posted incident is not proof of no problems; it is what the status page shows.
- Across all 12 services tracked: 69 incidents in the 30 days to 4 October 2026, 15 of them major or critical. The archive holds 1,259 incidents from 12 status pages.
Updated . Hours with an open incident are not downtime: they measure how long the provider's own status page showed an incident as unresolved, with overlapping incidents counted once.
Cite this page: "AI API reliability report", Gravity, updated , https://gravity.fast/blog/ai-api-reliability-report/. Data CC BY 4.0 (licence): copy it, chart it, use it commercially, and link back here. Downloads: data.csv (one row per incident) and monthly.csv (one row per service per month).
The last 90 days
Model APIs first, then the agent and automation platforms that sit one layer up. For OpenAI, Anthropic, Perplexity and Vercel only the API part of the status page counts; the consumer apps on the same pages are left out. The coverage table further down says exactly which components count for each service.
| Service | Incidents | Major or critical | Hours with an open incident | Median time to resolve | Longest incident |
|---|---|---|---|---|---|
| Model APIs | |||||
| OpenAI API | 18 | 4 | 210 | 2 h 52 min | 6 d 8 h |
| Claude API (Anthropic) | 60 | 22 | 102 | 1 h 9 min | 7 h 9 min |
| Gemini on Vertex AI (Google Cloud) | 0 | 0 | 0 | n/a | n/a |
| Cohere | 3 | 2 | 6.2 | 57 min | 4 h 22 min |
| Groq | 0 | 0 | 0 | n/a | n/a |
| Fireworks AI | 197 | 2 | 158 | 12 min | 46 h 35 min |
| Perplexity API | 0 | 0 | 0 | n/a | n/a |
| Vercel AI Gateway | 1 | 0 | 0.5 | 30 min | 30 min |
| Agent and automation platforms | |||||
| Zapier | 23 | 2 | 343 | 3 h 15 min | 6 d |
| Make | 16 | 8 | 485 | 18 h 17 min | 7 d 6 h |
| Cursor | 87 | 23 | 166 | 1 h 12 min | 13 h 19 min |
| v0 by Vercel | 6 | 4 | 9.1 | 1 h 6 min | 3 h 42 min |
The 90 days to 4 October 2026. Incidents are counted by start date. Hours are the time at least one in-scope incident was open, clipped to the window. "Data from" marks a service whose usable history starts inside the window, so its figures cover less than 90 days. What counts for each service is listed under coverage. Every incident is in data.csv.
The last 30 days
The same table over a shorter window, which is the one to quote when a writer needs "recently".
| Service | Incidents | Major or critical | Hours with an open incident | Median time to resolve | Longest incident |
|---|---|---|---|---|---|
| Model APIs | |||||
| OpenAI API | 6 | 1 | 21 | 2 h 52 min | 7 h 27 min |
| Claude API (Anthropic) | 6 | 4 | 13 | 1 h 39 min | 6 h 16 min |
| Gemini on Vertex AI (Google Cloud) | 0 | 0 | 0 | n/a | n/a |
| Cohere | 0 | 0 | 0 | n/a | n/a |
| Groq | 0 | 0 | 0 | n/a | n/a |
| Fireworks AI | 25 | 0 | 11 | 12 min | 2 h 13 min |
| Perplexity API | 0 | 0 | 0 | n/a | n/a |
| Vercel AI Gateway | 0 | 0 | 0 | n/a | n/a |
| Agent and automation platforms | |||||
| Zapier | 4 | 1 | 81 | 2 h 26 min | 3 h 58 min |
| Make | 3 | 2 | 22 | 11 h 7 min | 18 h 17 min |
| Cursor | 24 | 6 | 37 | 44 min | 6 h 30 min |
| v0 by Vercel | 1 | 1 | 0.6 | 34 min | 34 min |
The 30 days to 4 October 2026, same rules as the 90-day table.
Leaderboard: this month and last
Ranked by the fewest hours with an open incident, so the top of each table is the quietest status page. Early in a month these tables are thin, and a single incident can move a service several places, so the full previous month sits beneath for comparison.
Model APIs, October 2026 so far
| Rank | Service | Incidents | Major or critical | Hours with an open incident |
|---|---|---|---|---|
| 1 | OpenAI API | 0 | 0 | 0 |
| 1 | Gemini on Vertex AI (Google Cloud) | 0 | 0 | 0 |
| 1 | Cohere | 0 | 0 | 0 |
| 1 | Groq | 0 | 0 | 0 |
| 1 | Perplexity API | 0 | 0 | 0 |
| 1 | Vercel AI Gateway | 0 | 0 | 0 |
| 7 | Fireworks AI | 2 | 0 | 0.4 |
| 8 | Claude API (Anthropic) | 1 | 0 | 6.3 |
Ranked by fewest hours with an open incident, then fewest incidents; ties share a rank. October 2026 from the 1st to 4 October 2026.
Agent and automation platforms, October 2026 so far
| Rank | Service | Incidents | Major or critical | Hours with an open incident |
|---|---|---|---|---|
| 1 | Zapier | 0 | 0 | 0 |
| 1 | Make | 0 | 0 | 0 |
| 1 | v0 by Vercel | 0 | 0 | 0 |
| 4 | Cursor | 1 | 0 | 0.6 |
Ranked by fewest hours with an open incident, then fewest incidents; ties share a rank. October 2026 from the 1st to 4 October 2026.
Model APIs, September 2026
| Rank | Service | Incidents | Major or critical | Hours with an open incident |
|---|---|---|---|---|
| 1 | Gemini on Vertex AI (Google Cloud) | 0 | 0 | 0 |
| 1 | Groq | 0 | 0 | 0 |
| 1 | Perplexity API | 0 | 0 | 0 |
| 1 | Vercel AI Gateway | 0 | 0 | 0 |
| 5 | Cohere | 1 | 0 | 4.4 |
| 6 | Claude API (Anthropic) | 8 | 7 | 11 |
| 7 | Fireworks AI | 36 | 0 | 21 |
| 8 | OpenAI API | 6 | 1 | 40 |
Ranked by fewest hours with an open incident, then fewest incidents; ties share a rank. The full month of September 2026.
Agent and automation platforms, September 2026
| Rank | Service | Incidents | Major or critical | Hours with an open incident |
|---|---|---|---|---|
| 1 | v0 by Vercel | 1 | 1 | 0.6 |
| 2 | Make | 4 | 2 | 22 |
| 3 | Cursor | 32 | 10 | 48 |
| 4 | Zapier | 5 | 1 | 155 |
Ranked by fewest hours with an open incident, then fewest incidents; ties share a rank. The full month of September 2026.
Month by month
The last seven months for every service. This is the part no status page shows you, because their APIs forget and their history pages are one provider at a time.
| Service | Apr 2026 | May 2026 | Jun 2026 | Jul 2026 | Aug 2026 | Sep 2026 | Oct 2026 (to date) |
|---|---|---|---|---|---|---|---|
| Model APIs | |||||||
| OpenAI API | n/a | n/a | n/a | 10 161 h from 6 Jul 2026 | 2 10.0 h | 6 40 h | 0 0 h |
| Claude API (Anthropic) | 33 37 h | 29 30 h | 38 455 h | 39 79 h | 16 28 h | 8 11 h | 1 6.3 h |
| Gemini on Vertex AI (Google Cloud) | 0 0 h | 0 0 h | 0 0 h | 0 0 h | 0 0 h | 0 0 h | 0 0 h |
| Cohere | 3 1.8 h | 0 0 h | 1 2.9 h | 1 0.8 h | 1 1.0 h | 1 4.4 h | 0 0 h |
| Groq | 0 0 h | 0 0 h | 0 0 h | 1 1.6 h | 0 0 h | 0 0 h | 0 0 h |
| Fireworks AI | n/a | n/a | n/a | 24 17 h from 6 Jul 2026 | 135 119 h | 36 21 h | 2 0.4 h |
| Perplexity API | n/a | n/a | n/a | 0 0 h from 6 Jul 2026 | 0 0 h | 0 0 h | 0 0 h |
| Vercel AI Gateway | 0 0 h | 0 0 h | 0 0 h | 1 0.5 h | 0 0 h | 0 0 h | 0 0 h |
| Agent and automation platforms | |||||||
| Zapier | n/a | n/a | 2 24 h from 18 Jun 2026 | 9 87 h | 9 138 h | 5 155 h | 0 0 h |
| Make | 3 3.0 h | 4 25 h | 5 15 h | 8 324 h | 4 139 h | 4 22 h | 0 0 h |
| Cursor | 19 20 h | 25 36 h | 10 12 h | 30 57 h | 25 62 h | 32 48 h | 1 0.6 h |
| v0 by Vercel | 4 2.9 h | 2 2.1 h | 2 5.0 h | 4 4.9 h | 1 3.7 h | 1 0.6 h | 0 0 h |
Each cell: incidents started that month, and below it hours with an open incident. Tracking since 4 October 2026, when the daily collector first ran; earlier months come from each provider's own retained history: back to 1 October 2025 for the Atlassian Statuspage pages (Claude, Make, Cursor, Vercel), about 90 days for the incident.io pages (OpenAI, Zapier, Perplexity, Cohere, Groq, Fireworks AI, v0; Cohere, Groq, Zapier and v0 reach further through their latest-25 list), and whatever Google's incidents feed still lists. "From" marks a month the history only partly covers; n/a means no usable history. Full series since 1 October 2025: monthly.csv.
How to read these numbers
Hours with an open incident are not downtime. They measure how long the provider's own status page showed at least one in-scope incident as unresolved. Most incidents are partial: slower responses, errors on one model, one region or one feature. A status page also tends to open an incident after the trouble starts and to close it when the provider is confident, so the window runs long in both directions. Treat the figure as a reporting measure.
Providers post at different thresholds. Fireworks AI posts an incident per affected model, so it logs far more incidents than a provider that groups problems into one notice. A page that posts every slowdown will look worse here than one that posts only big outages. That is a real difference in transparency, and it cuts the opposite way from how a raw ranking reads.
Scope matters more than it looks. The OpenAI and Claude rows count only incidents that touched an API component, not ChatGPT or claude.ai. The Zapier, Make, Cursor and v0 rows count every incident on the page, because each page describes one product. Zapier and Make also post incidents for single third-party connectors (one app's trigger failing, say), which can stay open for days and push their hours up even when the core product is fine. If you quote a figure, quote the scope with it.
Zero means nothing was posted. It is not a test result. If you run production traffic on any of these, the practical move is a fallback chain rather than a bet on one provider: I wrote up how to set up agent fallback chains and retry and fallback patterns, and the agent uptime and reliability guide covers what to measure on your own side. Rate limits fail differently from outages, which the rate-limiting post covers.
Change log: notable incidents
Every major or critical incident, plus anything open six hours or more, in the last 90 days, newest first. Each links to the provider's own incident page.
- : Claude API (Anthropic), minor impact: Delayed credits on the Claude Platform, open 6 h 16 min.
- : Cursor, major impact: Investigating service degradation - xAI models, open 35 min.
- : v0 by Vercel, major impact: Upstream Provider issue causing partial outage in v0, open 34 min.
- : Claude API (Anthropic), major impact: Elevated errors on claude.ai, Claude Code, Claude Cowork and the Claude API, open 2 h 6 min.
- : Claude API (Anthropic), major impact: Elevated errors for multiple models, open 1 h 38 min.
- : Cursor, major impact: Investigating service degradation, Grok Bot, open 5 h 43 min.
- : Make, major impact: Degradation on us1.make.celonis.com, open 3 h 56 min.
- : Cursor, major impact: Investigating service degradation, Grok Bot, open 2 h 56 min.
- : Cursor, major impact: Investigating service degradation, Grok Bot, open 2 h 4 min.
- : Cursor, major impact: Investigating service degradation, Grok Bot, open 6 h 30 min.
- : OpenAI API, major impact: Elevated errors with gpt-image-2.5-flare, open 40 min.
- : Claude API (Anthropic), major impact: Intermittent error spikes for Claude Mythos 5.1 and Claude Fable 5.1, open 1 h 26 min.
- : Claude API (Anthropic), major impact: Elevated errors for Claude Mythos 5.1 and Claude Fable 5.1, open 17 min.
- : Zapier, major impact: Zap History not displaying Zap runs, open 1 h 50 min.
- : Make, major impact: Degradation on EU1 Region, open 18 h 17 min.
- : OpenAI API, minor impact: Elevated errors for image generation, open 7 h 27 min.
- : Cursor, major impact: Grok Bot computers unreachable for some users, open 1 h 5 min.
- : Cursor, major impact: Degraded performance for Grok Bot routine triggers, open 1 h.
- : Cursor, major impact: Elevated errors for Anthropic Models, open 2 h 15 min.
- : Cursor, critical impact: Investigating service degradation, open 3 h 26 min.
- : Claude API (Anthropic), major impact: Elevated errors for multiple models, open 2 h 57 min.
- : Claude API (Anthropic), major impact: Elevated errors for Claude Sonnet 5, open 19 min.
- : Claude API (Anthropic), major impact: Elevated errors for Claude Sonnet 5, open 26 min.
- : Cursor, major impact: Investigating service degradation, Automations, Cloud Agents, Review Agents, open 1 h 20 min.
- : Zapier, no impact level posted: Freshservice - Webhook delivery disruption, open 6 d.
- : OpenAI API, minor impact: Elevated latency in the Responses API, open 20 h 38 min.
- : Cursor, critical impact: Investigating service degradation, cursor.com and IDE, open 7 min.
- : Claude API (Anthropic), major impact: Elevated errors for multiple models, open 3 h 24 min.
- : Zapier, no impact level posted: Salesforce integration - trigger and connection disruptions, open 4 d 4 h.
- : Claude API (Anthropic), major impact: Elevated errors on requests to multiple models, open 26 min.
- : Tracker. Tracking started. Backfilled Claude, Make, Cursor and Vercel from their status-page history to 1 October 2025; incident.io pages (OpenAI, Zapier, Cohere, Groq, Fireworks AI, Perplexity, v0) from their 90-day windows and latest-25 lists; Google Cloud from its incidents feed.
Notable means major or critical impact, or open for 6 hours or more, in the 90 days to 4 October 2026; newest first, up to 30 lines. The page rebuilds daily, so this list moves.
Coverage and status at the last check
What counts for each service, how far back the usable history goes, how many incidents the archive holds, and what each status page said when the collector last read it.
| Service | Status page | What counts | Usable history from | Incidents stored | Status page at the last check |
|---|---|---|---|---|---|
| OpenAI API | status.openai.com | Incidents touching a component in the APIs group (Chat Completions, Responses, Realtime, Batch, Embeddings and the rest) | 6 July 2026 | 18 | All clear: All Systems Operational |
| Claude API (Anthropic) | status.claude.com | Incidents touching the Claude API (api.anthropic.com) component | 1 October 2025 | 286 | All clear: All Systems Operational |
| Gemini on Vertex AI (Google Cloud) | status.cloud.google.com | Google Cloud incidents that list a Vertex AI or Gemini product | 27 February 2026 | 1 | All clear: No open incident naming Vertex AI or Gemini |
| Cohere | status.cohere.com | Every incident on the page | 1 March 2025 | 25 | All clear: All Systems Operational |
| Groq | groqstatus.com | Every incident on the page | 3 September 2025 | 25 | All clear: All Systems Operational |
| Fireworks AI | status.fireworks.ai | Every incident on the page (it posts per-model incidents) | 6 July 2026 | 197 | All clear: All Systems Operational |
| Perplexity API | status.perplexity.com | Incidents touching the API component | 6 July 2026 | 0 | All clear: All Systems Operational |
| Vercel AI Gateway | www.vercel-status.com | Incidents touching the AI Gateway component | 1 October 2025 | 2 | All clear: All Systems Operational |
| Zapier | status.zapier.com | Every incident on the page | 18 June 2026 | 25 | All clear: All Systems Operational |
| Make | status.make.com | Every incident on the page | 1 October 2025 | 57 | All clear: All Systems Operational |
| Cursor | status.cursor.com | Every incident on the page | 1 October 2025 | 215 | All clear: All Systems Operational |
| v0 by Vercel | v0-status.com | Every incident on the page (v0 has its own status page; status.v0.dev redirects there) | 18 November 2025 | 25 | All clear: All Systems Operational |
Status as read at 13:36 UTC on 4 October 2026. This is a once-a-day snapshot, not a live monitor; for "is it down right now", open the status page linked in the row.
Methodology
Here is exactly how the tables are built, so you can decide how far to trust them.
Where the data comes from
Atlassian Statuspage pages (Claude at status.claude.com, Make, Cursor, Vercel): the public /api/v2/incidents.json, /api/v2/components.json and /api/v2/summary.json endpoints every day. For the backfill and to close any gap, the page's own /history.json listing and the per-incident /incidents/<id>.json record, which carries start, resolve and component data. These pages were backfilled to 1 October 2025.
incident.io pages (OpenAI, Zapier, Cohere, Groq, Fireworks AI, Perplexity, and v0 at v0-status.com): the Statuspage-compatible /api/v2/incidents.json and /api/v2/summary.json, plus the JSON the status page itself loads, which lists every incident in its 90-day window with the affected components and impact times. The compatible API carries no component data, so components come from the second source.
Google Cloud: the public incidents feed at status.cloud.google.com/incidents.json, filtered to incidents that list a Vertex AI or Gemini product. Google's severity labels are mapped to this page's scale as service information to minor, service disruption to major, service outage to critical. That mapping is mine.
What counts
An incident is anything the provider posted as an incident. Scheduled maintenance is excluded. Each incident is stored once, keyed by status page and the provider's incident id, and later reads update it in place, for example when it resolves. Impact is the provider's own label (none, minor, major, critical) where the page gives one; for older incident.io records outside the compatible API's latest 25, it is derived from the worst component status posted (degraded performance to minor, partial outage to major, full outage to critical), a mapping that matched the provider's own label on 81 of the 83 incidents where I could compare both on 4 October 2026.
How each number is computed
Incidents are counted in the window where they started. Major or critical is the subset with that impact label. Hours with an open incident is the union of each incident's start-to-resolve interval, clipped to the window, so two overlapping incidents count once. Median time to resolve uses incidents that started in the window and have resolved. Longest incident is the longest start-to-resolve span among incidents that started in the window, linked to the provider's page. Start is the provider's stated start time where the page records one (Statuspage's started_at, the earliest component impact on incident.io), otherwise the time the incident was posted. Windows end at the time of the last collection.
Some incidents are posted after the fact, and a page that moves to a new status provider can import its old incidents with the import time as the post time. Either way the record shows a resolve time earlier than its start. Those incidents still count, dated at their resolve time, but they add no hours and no time to resolve, because the real start is unknown. The CSV flags each one.
The collector runs daily. Nothing is ever deleted from the archive, which lives in the site repository with one dated snapshot per day, so a figure quoted from this page can be checked against the day it was taken.
Known gaps
I checked 19 status pages on 4 October 2026 and kept only those with a working public JSON feed. Not included: n8n Cloud and Lindy (no Statuspage JSON at their status pages), Relevance AI and Together AI (no JSON feed found), Replit and Mistral (requests refused), OpenRouter (status JSON unavailable when checked), and the Gemini API in Google AI Studio, whose status page renders in the browser with no public JSON feed I could find. I will add any of them when a feed works.
incident.io pages expose only their last 90 days with component data, so OpenAI and Perplexity history here starts in early July 2026 and grows from there. Google's feed listed six incidents across all of Google Cloud on 4 October 2026, the oldest from 27 February 2026, so Gemini on Vertex AI has a short record. None of this measures performance from the outside: there are no synthetic pings here, only what each provider published.
Related data pages: the AI agent pricing index tracks what agent platforms charge, with dated history, and the AI agent funding tracker logs every disclosed round. For a price-first comparison, see the cheapest AI agent platforms.
FAQ
1. Is the OpenAI API down right now?
This page is not a live monitor; it reads the status pages once a day. At the last check (13:36 UTC, 4 October 2026) status.openai.com said "All Systems Operational". For a live answer open status.openai.com. Over the 30 days to 4 October 2026, 6 incidents touched an OpenAI API component, 21 hours with an open incident in total.
2. Where can I see Claude API status history?
status.claude.com shows recent incidents, and its incidents API returns only the latest 50, which on 4 October 2026 reached back to 25 July 2026. This page stores every incident that touched the Claude API (api.anthropic.com) component from 1 October 2025 onward: 286 so far, with 60 in the 90 days to 4 October 2026. The full list is in data.csv.
3. What is the most reliable LLM API?
On status-page evidence alone, over the 90 days to 4 October 2026: Perplexity API, Groq and Gemini on Vertex AI (Google Cloud) posted no incident in scope, and among those that did, the fewest hours with an open incident belonged to Vercel AI Gateway (0.5 hours across 1 incident). Treat that as a reporting record, not an uptime guarantee: providers differ in how readily they post, a narrow component list posts less, and one that posts every small slowdown looks worse than one that posts only big outages.
4. How many incidents did OpenAI and Anthropic have last month?
In September 2026, 6 incidents touched an OpenAI API component (40 hours with an open incident) and 8 touched the Claude API (11 hours). Counts come from each provider's own status page and are rebuilt daily.
5. Does Google publish Gemini API incidents?
Google Cloud's incident feed covers Gemini on Vertex AI, and this page counts incidents there that list a Vertex AI or Gemini product: 1 stored, the feed reaching back to 27 February 2026. The Gemini API in Google AI Studio has its own status page, which this page does not cover yet.
6. Are hours with an open incident the same as downtime?
No. They measure how long a provider's status page showed at least one incident as unresolved. Most incidents are partial: slower responses, errors for some models or some users, or one feature failing. A status page also opens incidents late and closes them when the provider decides, so the figure is a reporting measure, not measured downtime.
7. Why do some services show zero incidents?
Because their status pages posted none in that window. That can mean a quiet month, a narrow component list, or a provider that posts sparingly. This page compares what each provider chose to publish, so a zero is a zero on the status page, not a test result.
8. Can I reuse this data?
Yes. The tables and both CSV files (data.csv, one row per incident, and monthly.csv) are licensed CC BY 4.0: copy them, chart them, use them commercially, and credit "Gravity AI API reliability report" with a link to this page. The underlying incidents are public statements on each provider's status page.
Sources
- OpenAI status, incidents and components, read daily from 4 October 2026: status.openai.com
- Claude status (Anthropic), incidents, history and components, read daily from 4 October 2026: status.claude.com
- Google Cloud Service Health incidents feed and its schema, read 4 October 2026: status.cloud.google.com/incidents.json, incidents.schema.json
- Cohere status: status.cohere.com; Groq status: groqstatus.com; Fireworks AI status: status.fireworks.ai; Perplexity status: status.perplexity.com, all read daily from 4 October 2026
- Vercel status (AI Gateway component): vercel-status.com; Zapier status: status.zapier.com; Make status: status.make.com; Cursor status: status.cursor.com; v0 status: v0-status.com, all read daily from 4 October 2026
- Atlassian Statuspage public API reference (incidents, components, summary endpoints): developer.statuspage.io, read 4 October 2026