Voice Agent API: Downstream ProvidersThe Voice Agent API depends on other providers such as OpenAI, Anthropic, Google, and ElevenLabs. This section is used to report issues with these downstream providers that may impact functionality of Deepgrams Voice Agent.
Operational
DG Whisper CloudOpenAI Whisper models hosted by Deepgram
Operational
Incidents50
History 50
Minor
ISP Backbone Degradation
Started
Wed, Sep 2, 2026, 08:22:55 PM
Updated
Wed, Sep 2, 2026, 11:28:32 PM
Resolved
Wed, Sep 2, 2026, 11:28:32 PM
Duration
3h 5m
resolved
The upstream ISP failure has been mitigated by rerouting all traffic around that provider.
identified
Cogent is experiencing a backbone networking issue in LA, due to two fiber cuts. This is likely impacting a large majority of traffic flowing to the east-coast, we're looking into possible mitigations and Cogent has engineers actively working on it. https://ecogent.cogentco.com/network-status
None
[Deepgram EU API] Flux-Multi STT Outage
Started
Tue, Aug 25, 2026, 07:00:00 AM
Updated
Tue, Aug 25, 2026, 06:01:20 PM
Resolved
Tue, Aug 25, 2026, 07:00:00 AM
Duration
0m
resolved
Resolved
Between 07:13 and 09:00 UTC, flux-general-multi Speech-to-Text requests on api.eu.deepgram.com failed shortly after the WebSocket connection was established, returning an internal server error on the first audio and closing the connection. Affected clients saw abnormal closures and reconnect loops. A low rate of residual failures continued until 11:30 UTC.
This has been fully investigated, mitigated and resolved, service has fully recovered.
None
[Deepgram API] Elevated Streaming STT Error Rates
Started
Mon, Aug 24, 2026, 11:00:00 PM
Updated
Tue, Aug 25, 2026, 04:41:54 PM
Resolved
Mon, Aug 24, 2026, 11:00:00 PM
Duration
0m
resolved
Between 23:18 and 23:45 UTC, a portion of streaming STT connection requests failed with 5XX errors and timeouts.
While rolling out a new data center, we detected streaming errors and immediately removed it from rotation. Traffic to the new site had been intentionally limited to a small percentage while we verified the rollout, which limited the number of affected requests. This is now fully resolved.
Minor
Elevated 5XX Error Rate for v1/auth/grant
Started
Thu, Aug 20, 2026, 03:52:08 PM
Updated
Thu, Aug 20, 2026, 06:01:16 PM
Resolved
Thu, Aug 20, 2026, 05:00:47 PM
Duration
1h 8m
resolved
This incident has been resolved.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We're seeing elevated error rate for authentication services. We are currently investigating the issue
Minor
Elevated 5XX Error Rate for v1/auth/grant
Started
Thu, Aug 13, 2026, 01:17:52 PM
Updated
Thu, Aug 13, 2026, 05:10:22 PM
Resolved
Thu, Aug 13, 2026, 01:30:05 PM
Duration
12m
resolved
We've resolved the issue and the service is back to normal.
investigating
We're seeing elevated 503 error for v1/auth/grant endpoint.
Minor
[Deepgram API] Elevated Batch and Streaming STT Error Rates
Started
Tue, Aug 4, 2026, 02:30:34 PM
Updated
Tue, Aug 4, 2026, 06:10:21 PM
Resolved
Tue, Aug 4, 2026, 06:10:21 PM
Duration
3h 39m
resolved
This incident has been resolved.
monitoring
An ISP network failure resulted in poor performance along one network path, resulting in degraded performance that primarily affected STT traffic. All network traffic is now routed through a high quality network path and services have recovered. We are actively monitoring.
investigating
We are currently investigating an issue where batch and streaming STT requests have degraded performance.
A fix has been implemented and traffic has returned to normal. We are monitoring the results.
investigating
We are continuing to investigate an issue. STT batch requests may have returned 5xx errors, and some STT websockets may have experienced high latency, starting at 16:03 UTC.
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating an issue where some batch requests are returning 400 errors due to remote content download errors. Retrying the request one or more times often succeeds.
None
Intermittent Connectivity Issues Due to AWS US-WEST-2 Infrastructure Event
Started
Fri, Jul 24, 2026, 01:18:11 PM
Updated
Mon, Jul 27, 2026, 07:02:21 PM
Resolved
Fri, Jul 24, 2026, 01:22:35 PM
Duration
4m
resolved
This incident has been resolved.
investigating
Between approximately 10:55 and 12:00 UTC on July 24, an upstream infrastructure issue in AWS's US-WEST-2 region caused intermittent connectivity problems that may have affected Deepgram services and customer infrastructure — including elevated errors, connection timeouts, and failures fetching remotely hosted media for some requests.
The underlying cause was an AWS networking issue affecting routing to the US-WEST-2 region. AWS applied mitigations and fully restored connectivity by ~12:00 UTC. We have confirmed recovery and are continuing to monitor.
Details on the underlying AWS event are available on the AWS Health status page: https://health.aws.amazon.com/health/status?eventID=arn:aws:health:us-west-2::event/MULTIPLE_SERVICES/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE_83A36_24C89B493A6
Minor
Voice Agent API: Google/Gemini LLM errors
Started
Tue, Jul 21, 2026, 01:45:46 PM
Updated
Tue, Jul 21, 2026, 04:45:16 PM
Resolved
Tue, Jul 21, 2026, 03:20:11 PM
Duration
1h 34m
resolved
This incident has been resolved.
investigating
Starting at 13:45 PM UTC, some Voice Agent sessions using Google/Gemini without a pinned model version received failed responses due to a routing issue introduced in this morning's deploy. Two Gemini models were unaffected; all other providers and products were unaffected. The issue was resolved as of 3:20 PM UTC.
None
Elevated 5XX Errors for Flux STT Requests
Started
Tue, Jul 7, 2026, 03:56:59 PM
Updated
Tue, Jul 7, 2026, 07:42:34 PM
Resolved
Tue, Jul 7, 2026, 06:20:00 PM
Duration
2h 23m
resolved
This incident has been resolved.
investigating
Flux STT experienced elevated WebSocket streaming errors between 15:56 UTC and 18:20 UTC.
Minor
Elevated errors when using Google Gemini models with Voice Agent
Started
Wed, Jun 17, 2026, 06:00:31 PM
Updated
Wed, Jun 24, 2026, 09:55:06 PM
Resolved
Thu, Jun 18, 2026, 04:00:42 AM
Duration
10h
resolved
This incident has been resolved.
identified
Continuing to monitor; root cause confirmed as the Google API key change.
identified
We have identified the cause of elevated error rates affecting Voice Agent requests that use Google Gemini models. A recent change on Google's side to how API keys operate is causing requests to fail, and our team is rolling out updated credentials to resolve it. Until the update is complete, customers may see intermittent failures for Gemini-model requests — concentrated in the first half of each hour, with most requests succeeding in the second half. Other Voice Agent models are unaffected. We are actively working on the fix and will update when resolved.
investigating
We are currently investigating elevated error rates affecting Voice Agent requests that use Google Gemini models. During this period, some requests to Google models (including gemini-2.5-flash-lite) may fail or time out. Other Voice Agent functionality and non-Google models are not affected. Our team is actively working to identify the cause and restore full service. We will provide an update shortly.
To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration.
None
ISP Network-Related Issue
Started
Wed, Jun 24, 2026, 01:00:00 AM
Updated
Wed, Jun 24, 2026, 02:15:08 PM
Resolved
Wed, Jun 24, 2026, 01:00:00 AM
Duration
0m
resolved
An upstream ISP issue may have caused intermittent WebSocket connection failures, handshake timeouts, and degraded streaming connectivity.
Incident time: 01:10 - 01:52 UTC
Minor
400 errors with PreRecorded STT ("REMOTE_CONTENT_ERROR" / "failed to return valid data") when passing a URL
Started
Wed, Jun 17, 2026, 01:35:10 PM
Updated
Tue, Jun 23, 2026, 09:33:49 PM
Resolved
Wed, Jun 17, 2026, 07:47:56 PM
Duration
6h 12m
resolved
The issue was resolved at 19:47 UTC.
identified
Between the hours of 13:35 UTC - 19:47 UTC we noticed an increase in 400 errors returned on PreRecorded STT requests where the payload was a URL. They were transient errors, retrying the request was the recommended course of action and would typically result in a successful transcription.
Anthropic has resolved this incident.
Incident end time: Jun 23, 2026, 16:44 UTC
monitoring
We are currently observing elevated error rates from our downstream third-party provider, Anthropic. Deepgram infrastructure, including our canaries, remains healthy.
Incident start time: Jun 23, 2026, 14:19 UTC
(Anthropic status updates: https://status.claude.com)
None
Degraded performance on Nova streaming STT
Started
Thu, Jun 18, 2026, 01:00:00 PM
Updated
Thu, Jun 18, 2026, 09:42:03 PM
Resolved
Thu, Jun 18, 2026, 01:00:00 PM
Duration
0m
resolved
Nova-3 and Nova-2 streaming STT experienced elevated websocket connection errors and websocket disconnections between 13:00 - 16:00 UTC.
This incident has been resolved. A degradation affecting a subset of streaming Speech-to-Text (Flux) sessions, where connections terminated before returning a transcript, has been mitigated and service has been operating normally since. We're continuing to monitor and apologize for the disruption.
identified
We identified a degradation affecting a subset of Flux streaming Speech-to-Text sessions, where connections terminated before returning a transcript. We've isolated it to specific degraded hosts and are mitigating.
investigating
Some streaming Speech-to-Text connections on the Flux model established successfully but then terminated a few seconds into the session due to an internal error, before returning a final transcript. Applications relying on final transcripts (e.g. IVR/voice-agent flows) would have seen sessions open with no transcript returned.
None
Outbound networking issues affecting STT, TTS streaming and batch service
Started
Fri, May 22, 2026, 09:00:01 PM
Updated
Fri, May 22, 2026, 09:00:01 PM
Resolved
Fri, May 22, 2026, 09:00:01 PM
Duration
0m
resolved
Some API requests may have experienced elevated error rates for approximately 2-3 minutes starting at 16:43 UTC on May 22, 2026. The issue has been resolved now.
Minor
Voice Agent - Elevated 4xx Error Rates
Started
Mon, May 18, 2026, 10:06:00 PM
Updated
Mon, May 18, 2026, 11:36:38 PM
Resolved
Mon, May 18, 2026, 10:06:00 PM
Duration
0m
resolved
Between 22:06 and 22:12 UTC, a subset of Voice Agent requests returned elevated 4xx errors following a recent deploy. The deploy has been rolled back and the service is operating normally.
Minor
Elevated 5XX Errors for Non-English Batch STT Requests with Redaction
Started
Thu, May 14, 2026, 02:52:20 PM
Updated
Thu, May 14, 2026, 07:09:45 PM
Resolved
Thu, May 14, 2026, 07:00:54 PM
Duration
4h 8m
resolved
This issue has been resolved.
investigating
We are currently investigating an increase in 5XX errors for non-English STT requests that are using redaction.
Minor
Intermittent 404 HTTP errors Batch STT English
Started
Tue, May 12, 2026, 04:18:34 PM
Updated
Tue, May 12, 2026, 08:15:21 PM
Resolved
Tue, May 12, 2026, 06:07:11 PM
Duration
1h 48m
resolved
This incident has been resolved.
investigating
During the affected time period, some English batch speech-to-text requests intermittently returned 404 Not Found HTTP errors. This issue was limited to requests that met both of the following criteria:
- Utilized English batch STT models (including both monolingual English and multilingual models that support English transcription)
- Had Smart Format, Entity Detection, or Redaction features enabled.
We are currently observing elevated error rates from our downstream third-party provider, OpenAI specifically with the ChatGPT 4o model. Deepgram infrastructure, including our canaries, remains healthy.
We are currently observing elevated error rates from our downstream third-party provider, Anthropic specifically with the Claude Sonnet 4.5 model. Deepgram infrastructure, including our canaries, remains healthy.
We are currently observing elevated error rates from our downstream third-party provider, Google Gemini LLMs. Deepgram infrastructure, including our canaries, remains healthy; however, Voice Agent customers using bring-your-own (BYO) Gemini LLMs may experience degraded performance or intermittent failures on LLM think requests.
We are monitoring the situation and tracking the provider’s status at https://aistudio.google.com/status. We will provide further updates as more information becomes available or if we observe broader impact.
None
Outbound networking issues affecting STT fetch requests and callbacks, Dashboard API requests, and Voice Agent 3rd party providers
Started
Mon, Apr 27, 2026, 02:14:34 PM
Updated
Mon, Apr 27, 2026, 06:13:41 PM
Resolved
Mon, Apr 27, 2026, 02:14:34 PM
Duration
0m
resolved
A small number of customers experienced disruptions to STT fetch requests, callbacks, and third-party providers within Voice Agent, as well as 503 errors on API calls for usage lookups and token creation. The majority of customers were unaffected.
None
Issues with downstream LLM provider in Voice Agent API
Started
Fri, Apr 17, 2026, 08:56:12 PM
Updated
Sat, Apr 18, 2026, 03:48:17 AM
Resolved
Sat, Apr 18, 2026, 03:48:17 AM
Duration
6h 52m
resolved
This incident has been resolved.
investigating
We are continuing to investigate this issue.
investigating
We are currently seeing elevated error rates from OpenAI models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration.
None
Outbound networking issues affecting STT PreRecorded API requests with callbacks, TTS Batch (HTTP) requests with callbacks, and Voice Agent API BYO 3rd party providers (LLM and TTS)
Started
Fri, Apr 10, 2026, 05:27:56 PM
Updated
Mon, Apr 13, 2026, 07:47:31 PM
Resolved
Mon, Apr 13, 2026, 07:47:31 PM
Duration
3d 2h
resolved
This incident has been resolved.
monitoring
A network infrastructure fix was implemented at 18:15 UTC and the behavior was immediately resolved. We are continuing to monitor carefully.
identified
A networking incident was identified which resulted in failed callbacks on STT PreRecorded API requests and TTS Batch requests. The issue also affected Voice Agent API sessions with LLM and TTS providers. We have identified a fix and are monitoring traffic. The source of the issue is being investigated.
None
Degraded API Performance
Started
Wed, Mar 25, 2026, 03:16:00 PM
Updated
Wed, Mar 25, 2026, 07:36:57 PM
Resolved
Wed, Mar 25, 2026, 03:16:00 PM
Duration
0m
resolved
Between 15:16 and 15:32 UTC on March 25, 2026, we noticed an increased rate of API requests returning errors, including 403, 404, and 504 status codes. The issue has been identified and resolved.
Minor
Voice Agent API Degraded Service
Started
Fri, Mar 20, 2026, 04:53:21 PM
Updated
Mon, Mar 23, 2026, 02:50:24 PM
Resolved
Mon, Mar 23, 2026, 06:53:57 PM
Duration
3d 2h
resolved
This incident has been resolved
investigating
There is an ongoing OpenAI outage that is affecting our Voice Agent. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration.
Minor
Voice Agent API Degraded Performance
Started
Wed, Mar 18, 2026, 03:00:37 PM
Updated
Thu, Mar 19, 2026, 01:37:52 AM
Resolved
Wed, Mar 18, 2026, 05:00:07 PM
Duration
1h 59m
resolved
This incident has been resolved.
identified
We are continuing to monitor this issue.
identified
There is an ongoing OpenAI outage that is affecting our Voice Agent. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration.
Minor
Issues with downstream LLM provider in Voice Agent API
Started
Thu, Mar 5, 2026, 12:59:47 AM
Updated
Thu, Mar 5, 2026, 01:26:03 AM
Resolved
Thu, Mar 5, 2026, 01:26:03 AM
Duration
26m
resolved
This incident has been resolved.
identified
There is an ongoing OpenAI outage that is affecting our Voice Agent. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration.
None
[Deepgram API] Elevated Errors for Streaming STT
Started
Thu, Feb 26, 2026, 05:16:15 PM
Updated
Fri, Feb 27, 2026, 02:05:54 AM
Resolved
Fri, Feb 27, 2026, 02:05:54 AM
Duration
8h 49m
resolved
This incident has been resolved.
monitoring
We are monitoring the issue.
investigating
We are currently investigating an issue that results in increased Speech-to-Text result latency.
None
[Deepgram API] Flux SDK Session Disruptions
Started
Tue, Feb 24, 2026, 02:00:00 AM
Updated
Thu, Feb 26, 2026, 08:28:49 PM
Resolved
Tue, Feb 24, 2026, 02:00:00 AM
Duration
0m
resolved
A temporary issue during a Flux-related update caused intermittent SDK session errors. The issue was identified and resolved.
2:00 AM UTC - 13:45 UTC
None
Issues with downstream LLM provider in Voice Agent API
Started
Mon, Feb 23, 2026, 05:58:30 PM
Updated
Tue, Feb 24, 2026, 02:28:36 PM
Resolved
Tue, Feb 24, 2026, 02:28:14 PM
Duration
20h 29m
resolved
The upstream Anthropic incident affecting Claude models has been resolved and error rates have returned to normal. Voice Agent functionality should now be operating as expected.
identified
There is an ongoing Anthropic outage that is affecting our Voice Agent. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration. The incident started at Feb 17, 2026 - 20:15 UTC
None
Issues with downstream LLM provider in Voice Agent API
Started
Tue, Feb 17, 2026, 04:15:28 PM
Updated
Tue, Feb 17, 2026, 08:16:51 PM
Resolved
Tue, Feb 17, 2026, 08:16:51 PM
Duration
4h 1m
resolved
As of 10:00 UTC, the issues from OpenAI models is being resolved.
investigating
We are currently seeing elevated error rates from OpenAI models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration, including at least one non-OpenAI think provider. The incident started at 14:40 UTC.
None
Issues with downstream LLM provider in Voice Agent API
Started
Mon, Feb 16, 2026, 05:08:07 PM
Updated
Mon, Feb 16, 2026, 11:32:10 PM
Resolved
Mon, Feb 16, 2026, 11:32:10 PM
Duration
6h 24m
resolved
This incident has been resolved.
investigating
We are continuing to investigate this issue.
investigating
We are currently seeing elevated error rates from OpenAI models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration, including at least one non-OpenAI think provider. The incident started at 16:30 UTC.
Minor
Issues with downstream LLM provider in Voice Agent API
Started
Thu, Feb 12, 2026, 05:47:20 PM
Updated
Fri, Feb 13, 2026, 09:28:42 PM
Resolved
Fri, Feb 13, 2026, 09:28:42 PM
Duration
1d 3h
resolved
The incident has been resolved.
investigating
We are currently seeing elevated error rates from OpenAI models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration, including at least one non-OpenAI think provider. The incident started at 17:47 UTC
Minor
Elevated Connection Errors
Started
Wed, Jan 28, 2026, 10:10:13 PM
Updated
Tue, Feb 10, 2026, 02:56:52 AM
Resolved
Thu, Feb 5, 2026, 06:56:19 PM
Duration
7d 20h
resolved
This incident has been resolved.
monitoring
The implemented fixes have resulted in stable systems since 2026-02-03 17:35 UTC. We are continuing to monitor carefully.
The issue primarily impacted STT batch services, with some impact on STT streaming. Specifically, response times and 429 errors increased for batch, and a percentage of connection errors occurred for both batch and streaming.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are continuing to investigate this issue.
investigating
We're currently seeing elevated response time, as well as increased 429 errors for batch STT. We're actively working on mitigating the issue.
investigating
We are continuing to investigate this issue.
investigating
We're observing elevated connection error rates at the moment, and we're actively investigating the issue.
monitoring
We saw elevated connection errors from 19:12 to 19:19 UTC, and 20:10 to 20:31 UTC. We're no longer seeing spikes in errors, but we're continuing to monitor the situation.
Minor
Issues with downstream LLM provider in Voice Agent API
Started
Wed, Feb 4, 2026, 04:52:21 PM
Updated
Wed, Feb 4, 2026, 06:47:10 PM
Resolved
Wed, Feb 4, 2026, 06:47:10 PM
Duration
1h 54m
resolved
The downstream provider indicated the issue was resolved. There was an outage between 16:20 UTC / 8:20 PT and 16:55 UTC / 8:55 PT
identified
We are currently seeing elevated error rates from Anthropic models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration, including at least one non-Anthropic think provider. The incident started at 16:34 UTC
Minor
Issues with downstream LLM provider in Voice Agent API
Started
Tue, Feb 3, 2026, 03:35:42 PM
Updated
Tue, Feb 3, 2026, 11:35:27 PM
Resolved
Tue, Feb 3, 2026, 11:35:27 PM
Duration
7h 59m
resolved
The incident has been resolved. The downstream provider's status page indicates the issues were resolved at 9:56 PT / 17:56 UTC
monitoring
We are continuing to monitor for any further issues with the downstream provider.
monitoring
Failures have reduced since 15:55UTC, we are monitoring the downstream provider's status page.
identified
We are currently seeing elevated error rates from Anthropic models. To avoid downtime, please define multiple LLM providers (https://developers.deepgram.com/docs/voice-agent-llm-models#using-multiple-llm-providers) in your Voice Agent configuration, including at least one non-Anthropic think provider.
Minor
Degraded Transcripts on Nova-3 Multilingual STT Model
Started
Thu, Jan 22, 2026, 09:00:00 PM
Updated
Mon, Jan 26, 2026, 04:42:59 PM
Resolved
Thu, Jan 22, 2026, 11:45:00 PM
Duration
2h 45m
resolved
From 21:38 - 23:39 UTC on 2026-01-22, requests for Deepgram's Nova-3 multilingual STT model received degraded transcripts due to a faulty model release.
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and is being actively resolved.
None
Shai-Hulud 2.0 Supply Chain Incident – No Customer Impact
Started
Wed, Nov 26, 2025, 12:00:14 AM
Updated
Thu, Dec 4, 2025, 10:50:35 PM
Resolved
Wed, Nov 26, 2025, 12:00:14 AM
Duration
0m
resolved
Informational (No Customer Action Required)
Deepgram identified that our internal development environment was affected as part of the industry-wide Shai-Hulud 2.0 NPM supply chain attack. The attack used compromised NPM packages to inject malicious CI/CD workflows and attempt to exfiltrate internal development credentials.
Our investigation confirms:
- No access to customer data or databases
- No impact to production API infrastructure or service availability
- No modification of published Deepgram SDKs or packages
- No effect on customer authentication or API keys
Timeline (UTC):
- 00:41, Nov 26: Webhook alert on an internal GitHub repo; engineers link activity to Shai-Hulud 2.0, disable malicious workflows, and rotate exposed credentials.
- 21:45, Nov 26: Additional commits via a compromised GitHub App publish some internal materials (no customer or production impact); the account is removed, GitHub Actions disabled, and the org moved into a locked-down state.
We see no further signs of compromise and are gradually restoring normal operations while tightening SDLC and CI/CD controls.
Customer action: None required.
Questions: security@deepgram.com • support@deepgram.com
This incident affected internal development infrastructure only; customer-facing APIs and services were not impacted.
Minor
Elevated API Errors
Started
Fri, Nov 14, 2025, 07:10:19 PM
Updated
Tue, Nov 18, 2025, 04:33:09 PM
Resolved
Tue, Nov 18, 2025, 04:33:09 PM
Duration
3d 21h
resolved
This incident has been resolved.
monitoring
We are continuing to monitor for any further issues.
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are still investigating the issue.
investigating
We are continuing to investigate this issue.
investigating
We're experiencing an elevated level of API errors and are currently looking into the issue.
Minor
Elevated latency and connection timeouts on Deepgram services
Started
Thu, Nov 13, 2025, 11:16:41 PM
Updated
Fri, Nov 14, 2025, 01:13:03 AM
Resolved
Fri, Nov 14, 2025, 01:13:03 AM
Duration
1h 56m
resolved
This incident has been resolved.
monitoring
We have adjusted our networking and are seeing recovery.
investigating
We are investigating an issue causing elevated latency and connection timeouts to Deepgram services.
Major
Outage for non-English batch speech-to-text requests utilizing redaction feature
Started
Wed, Oct 29, 2025, 04:15:42 PM
Updated
Thu, Oct 30, 2025, 01:34:01 AM
Resolved
Thu, Oct 30, 2025, 01:34:01 AM
Duration
9h 18m
resolved
This issue has been resolved.
monitoring
We are continuing to monitor for any further issues.
monitoring
We've implemented a fix and are currently monitoring the status of the issue.
investigating
We're experiencing an outage for our non-English batch speech-to-text requests utilizing redaction feature.
As of 19:28 UTC OpenAI's LLMs are no longer returning 500 errors.
We are still in conversation with OpenAI about what happened.
identified
We are still continuing to see increased errors using the OpenAI models in the Voice Agent API workflow. We have reached out and communicated with OpenAI to help resolve.
identified
We are still continuing to see increased errors using the OpenAI models in the Voice Agent API workflow. We have reached out and communicated with OpenAI to help resolve.
identified
We are continuing to see elevated error rates with OpenAI models in the Voice Agent API workflow. We have reached out and communicated with OpenAI to help resolve.
identified
We're seeing elevated error rates (500 Internal Server Error) with OpenAI models in the Voice Agent API.
Our team is actively investigating and communicating with OpenAI to find a path to resolution, and a path to mitigation in the short term. Our metrics show degraded service for 4o-mini, 4.1-mini, 4.1 nano, 4.1, and possibly other OpenAI models.
We are continuing to monitor, investigate, and work with OpenAI to resolve.
As a workaround, we recommend switching to a different model or provider.
identified
A service degradation and partial outage began at approximately 16:44 UTC.
Minor
Increased Latency Over Streaming STT and Voice Agent API
Started
Tue, Oct 28, 2025, 02:02:29 PM
Updated
Tue, Oct 28, 2025, 05:43:02 PM
Resolved
Tue, Oct 28, 2025, 04:20:23 PM
Duration
2h 17m
resolved
This issue has been resolved.
monitoring
The increased latency issue has significantly improved. We will continue to monitor the performance.
investigating
We identified a potential source of increased latency and are currently monitoring a mitigation step.
investigating
We are currently investigating increased latency over our streaming STT API and Voice Agent API.
Minor
Elevated error rates for Voice Agent API when using OpenAI LLMs
Started
Thu, Oct 9, 2025, 11:00:58 PM
Updated
Fri, Oct 17, 2025, 02:25:14 PM
Resolved
Thu, Oct 16, 2025, 09:00:33 PM
Duration
6d 21h
resolved
As of 21:00 UTC on October 16th, OpenAI's LLMs are no longer returning 500 errors.
investigating
We're seeing elevated error rates with OpenAI models in the Voice Agent API.
Our team is actively investigating and working with OpenAI to find a path to resolution, and a path to mitigation in the short term. Our metrics show degraded service for 4o-mini, 4.1-mini, 4.1 nano, 4.1, and possibly other OpenAI models.
A service degradation began at approximately 23:00 UTC on October 9th. The service degraded further at 22:00 UTC on October 14th, and then again at 23:00 UTC on October 15th. We are continuing to monitor, investigate, and work with OpenAI to resolve.
As a workaround, we recommend switching to a different model or provider.
None
[ISP Network-Related Issue]
Started
Tue, Oct 7, 2025, 04:55:50 PM
Updated
Tue, Oct 7, 2025, 06:26:45 PM
Resolved
Tue, Oct 7, 2025, 05:18:50 PM
Duration
23m
resolved
Start: 16:15:00 UTC
End: 17:18:00 UTC
Streaming connectivity issues from some AWS regions may have caused Websocket failures errors and degraded connectivity
Minor
US Network Issues
Started
Fri, Oct 3, 2025, 12:08:43 AM
Updated
Fri, Oct 3, 2025, 12:43:07 PM
Resolved
Fri, Oct 3, 2025, 04:30:09 AM
Duration
4h 21m
resolved
This incident has been resolved.
monitoring
The issue has been mitigated after rerouting traffic to avoid upstream network provider failures. Networking errors may have occurred intermittently between 12:08 am UTC and 1:51 am UTC.
identified
We've identified an issue with an upstream network provider across various US regions, and are working to shift traffic to alternative routes.