DashboardAccount management dashboard at turbopuffer.com/dashboard
Operational
gcp-asia-northeast3turbopuffer API
Operational
aws-us-east-1turbopuffer API
Operational
aws-us-east-2turbopuffer API
Operational
aws-us-west-2turbopuffer API
Operational
aws-ap-south-1turbopuffer API
Operational
aws-eu-west-2turbopuffer API
Operational
aws-ca-central-1turbopuffer API
Operational
gcp-asia-southeast1turbopuffer API
Operational
gcp-europe-west3turbopuffer API
Operational
gcp-us-east4turbopuffer API
Operational
gcp-us-west1turbopuffer API
Operational
aws-ap-southeast-2turbopuffer API
Operational
aws-eu-central-1turbopuffer API
Operational
aws-sa-east-1turbopuffer API
Operational
gcp-us-east1turbopuffer API
Operational
gcp-europe-west1turbopuffer API
Operational
Incidents25
History 25
Critical
Dashboard outage
Started
Sat, Sep 5, 2026, 06:40:07 PM
Updated
Sat, Sep 5, 2026, 06:52:40 PM
Resolved
Sat, Sep 5, 2026, 06:52:40 PM
Duration
12m
resolved
From 7:21-7:41 UTC, the turbopuffer dashboard was unavailable due to an issue with our backing datastore. This issue has been resolved and dashboard access has been restored.
investigating
The turbopuffer dashboard is currently experiencing an outage. turbopuffer regions are fully available. Writes and queries are not affected.
Major
Elevated rate of 5xx in gcp-us-central1
Started
Wed, Sep 2, 2026, 04:49:00 PM
Updated
Wed, Sep 2, 2026, 05:45:24 PM
Resolved
Wed, Sep 2, 2026, 05:16:00 PM
Duration
27m
resolved
The issue causing elevated 5xx errors for some namespaces has been resolved. We repaired the affected namespaces, and error rates have returned to normal. We will continue monitoring the service closely.
monitoring
Some namespaces are experiencing an elevated rate of 5xx errors. We have identified the issue and are working to restore full service.
Critical
Outage in gcp-us-southeast1
Started
Thu, Aug 27, 2026, 08:37:35 PM
Updated
Thu, Aug 27, 2026, 10:15:13 PM
Resolved
Thu, Aug 27, 2026, 10:15:13 PM
Duration
1h 37m
resolved
Service has been fully restored.
monitoring
We’ve identified the problem and rolled out a fix. We’re closely monitoring the situation.
investigating
Full outage identified in gcp-southeast1
Major
Errors when accessing dashboard
Started
Thu, Jul 16, 2026, 09:42:52 AM
Updated
Thu, Jul 16, 2026, 10:49:22 AM
Resolved
Thu, Jul 16, 2026, 10:49:22 AM
Duration
1h 6m
resolved
Our authentication provider has resolved the incident on their end, and dashboard logins have been fully operational since 10:08 UTC.
Existing sessions and the database were unaffected throughout.
monitoring
Logins are succeeding again as of 10:08 UTC. Our authentication provider is still working through their incident, so some logins may be slower than usual.
We'll continue monitoring until it's fully resolved.
identified
Dashboard logins (Google, SSO, and email) are intermittently failing due to an ongoing incident at [our authentication provider](https://status.workos.com/incidents/b52fkqjzcl20).
We're monitoring and will update here.
investigating
We're aware of issues affecting some customers who are logging into the dashboard. We're investigating and will provide an update shortly.
Note that the database is not affected - this issue is limited to the dashboard.
Minor
Elevated 500 rate on delete_by_filter writes, filtered queries
Started
Thu, May 28, 2026, 06:52:16 PM
Updated
Thu, May 28, 2026, 06:54:48 PM
Resolved
Thu, May 28, 2026, 06:30:00 PM
Duration
—
resolved
Starting at ~17:55 UTC, we experienced an elevated 500 rate affecting delete_by_filter, and filtered queries as a result of a recent change deployed to filtering. This change was reverted, and errors fully subsided by ~18:30 UTC. This incident affected approximately 38.2% of `delete_by_filter` requests in aws-us-east-1, and a small percentage of filtered queries in other regions.
monitoring
Investigating elevated 500 error rate on filtered queries, delete_by_filter operations.
Minor
partial outage in aws-us-west-2
Started
Wed, May 20, 2026, 03:42:34 PM
Updated
Wed, May 20, 2026, 04:19:38 PM
Resolved
Wed, May 20, 2026, 04:19:38 PM
Duration
37m
resolved
Our deploy has rolled out and affected namespaces have recovered.
identified
Our team has identified an issue with a deploy in aws-us-west-2, causing queries and writes to 500 for affected namespaces. We are deploying a fix.
Critical
Errors when accessing dashboard
Started
Mon, Apr 20, 2026, 12:22:41 PM
Updated
Mon, Apr 20, 2026, 05:32:27 PM
Resolved
Mon, Apr 20, 2026, 05:32:27 PM
Duration
5h 9m
resolved
This incident is resolved. Dashboard access is working as usual.
As noted earlier, this incident only affected the dashboard and had no impact on turbopuffer regions or namespaces!
monitoring
Traffic has returned to normal. We're monitoring closely!
monitoring
We're seeing recovery with traffic returning to normal. We're monitoring closely and will share another update shortly.
identified
This issue is currently affecting the dashboard and API key management. **turbopuffer regions and namespaces are all operating normally**, and existing API keys continue to authenticate as expected.
The root cause is a dependency service that we use for metadata. We'll share an update as we make progress.
investigating
We've confirmed reports of dashboard errors and are actively investigating. Note that the database is not affected - this issue is limited to the dashboard.
Critical
Errors when accessing dashboard
Started
Thu, Apr 2, 2026, 12:14:00 PM
Updated
Thu, Apr 2, 2026, 01:10:34 PM
Resolved
Thu, Apr 2, 2026, 01:10:34 PM
Duration
56m
resolved
We were impacted by [an issue with one of our providers](https://www.vercel-status.com/incidents/5r9bp5y8rql2), which we were able to resolve. We will continue to monitor for regressions.
investigating
We've confirmed reports of dashboard errors and our team is investigating.
Major
Elevated 408 and 429s
Started
Tue, Mar 24, 2026, 09:50:17 AM
Updated
Tue, Mar 24, 2026, 10:18:49 AM
Resolved
Tue, Mar 24, 2026, 10:18:48 AM
Duration
28m
resolved
We've observed a complete recovery of traffic.
monitoring
Our team has taken remedial action and is observing recovery, we will continue to monitor.
investigating
The gcp-us-east4 cluster is experiencing elevated rates of 408 and 429s affecting both reads and writes affecting approximately 4% of traffic. Our team is investigating.
Major
Increased Request Latency and 429s in GCP us-east4
Started
Tue, Mar 17, 2026, 09:58:01 AM
Updated
Tue, Mar 17, 2026, 10:48:08 AM
Resolved
Tue, Mar 17, 2026, 10:48:08 AM
Duration
50m
resolved
The region is stable and has fully recovered. We will continue to monitor.
monitoring
Latencies have returned to normal and 429s have fully recovered. The issue was related to transient failures in the underlying infrastructure. Our engineers have applied mitigations, and we are continuing to monitor.
investigating
We are looking into an issue causing increased latency and 429s in us-east4.
Major
Elevated 500s in region gcp-us-west1
Started
Mon, Mar 16, 2026, 08:55:43 AM
Updated
Mon, Mar 16, 2026, 09:39:58 AM
Resolved
Mon, Mar 16, 2026, 09:39:58 AM
Duration
44m
resolved
The region has stabilized and returned to normal operation. We are continuing to monitor!
monitoring
A fix has been rolled out and 5xx errors in the region have reduced. We are continuing to monitor the situation!
investigating
We are investigating an increase 500's in the region!
Major
Partial outage in gcp-us-east4
Started
Thu, Oct 16, 2025, 09:12:29 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Oct 16, 2025, 10:52:20 AM
Duration
1h 39m
resolved
Full recovery as of 12:24 PM UTC
monitoring
We are seeing degraded performance again as latency to Google Cloud Storage. We are continuing to monitor and work with the GCP team.
resolved
Full recovery as of 10:44am UTC.
We are investigating the root cause with the GCP team.
monitoring
Substantial improvement starting 10:35am UTC. We are continuing to investigate and work directly with GCP to resolve this fully.
identified
8:30am UTC - 9:25am UTC: Degraded performance on queries and writes in gcp-us-east4 from increased latency to Google Cloud Storage.
9:25am UTC: Escalated to partial outage both write and read paths due to climbing latency and timeouts from Google Cloud Storage.
* Queries may see elevated rates of 429s, and 500s.
* Writes may observe elevated rates of 408s, and 500s.
Our team is working to restore full system performance, and we are working with GCP to restore service as soon as possible.
Major
Outage in aws us-east-1
Started
Mon, Oct 20, 2025, 07:58:57 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Oct 20, 2025, 11:18:06 PM
Duration
15h 19m
resolved
AWS has reported that the upstream incident is resolved.
monitoring
The turbopuffer service remains healthy in the aws-us-east-1 region. We continue to closely monitor the situation while AWS continues to report service disruption in the region.
monitoring
AWS is experiencing a major disruption in region us-east-1 impacting multiple services (<https://health.aws.amazon.com/health/status?eventID=arn:aws:health:us-east-1::event/MULTIPLE_SERVICES/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE_BA540_514A652BE1A>). This caused turbopuffer to become unavailable in that region starting at 7:50 UTC. We've rolled out mitigations which resulted in a full recovery of turbopuffer at 8:20 UTC (despite the ongoing AWS outage). An existing mitigation to guard against EC2 instance launch issues has prevented the long-running EC2 degradation from further affecting turbopuffer availability. The AWS incident is ongoing. We will continue monitoring the cluster's health.
monitoring
We have deployed mitigations and are monitoring the situation.
investigating
[AWS us-east-1 is experiencing a major outage and many services are affected. ](https://health.aws.amazon.com/health/status)
turbopuffer clients in aws-us-east-1 will see elevated error rate and latencies during this outage.
We're looking at mitigating the impact.
Minor
Degraded performance in gcp-us-east4
Started
Mon, Oct 20, 2025, 09:55:09 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Oct 20, 2025, 05:22:02 PM
Duration
7h 26m
resolved
Service is stable. We are continuing to investigate the root cause with high priority.
monitoring
We have observed an increase in timeouts and a slow down in indexing starting at 9:38 UTC. The service has returned to normal at 10:00 UTC. We are monitoring and investigating the root cause.
investigating
we're currently observing degraded performance in gcp-us-east4 leading to elevated rates of 408s and 429s. Our team is currently investigating.
Major
Partial dashboard outage
Started
Mon, Oct 20, 2025, 07:42:21 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Tue, Oct 21, 2025, 02:31:00 AM
Duration
6h 48m
resolved
The upstream incident is resolved.
monitoring
Dashboard error rates have subsided. We continue to monitor the situation while we wait for confirmation from our authentication provider that the incident has resolved.
identified
Some users are reporting issues accessing the turbopuffer dashboard due to a partial outage with our authentication provider: <https://status.workos.com/incidents/rdz189wnjhyn>
turbopuffer APIs are unaffected.
Minor
Elevated 500s in gcp-us-west1
Started
Mon, Nov 24, 2025, 01:24:36 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Nov 24, 2025, 01:27:41 PM
Duration
3m
resolved
We have taken steps to remediate the 500s, which have now subsided. We're investigating the root cause.
identified
We're investigating an increase in 500 responses in the region.
Minor
degraded performance in aws-us-east-1
Started
Thu, Dec 11, 2025, 05:52:34 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Dec 11, 2025, 07:31:53 PM
Duration
1h 39m
resolved
we have not seen any re-occurences of the issue.
monitoring
we've removed problematic query nodes from rotation and are monitoring for further issues.
investigating
some namespaces in aws-us-east-1 may experience degraded performance and request timeouts. our team is working to rectify the issue.
Minor
Degraded service (500 responses)
Started
Mon, Dec 15, 2025, 10:25:55 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Dec 15, 2025, 10:30:58 AM
Duration
5m
resolved
Service is stable.
monitoring
Some of our query nodes restarted repeatedly due to OOMs caused by a code path that was producing very large allocations. We have deployed a fix, and the service has recovered. We are monitoring.
Minor
Degraded performance in aws-us-west-2
Started
Mon, Dec 15, 2025, 05:14:51 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Dec 15, 2025, 06:03:54 PM
Duration
49m
resolved
The problematic queries have been isolated and traffic has been restored to normal.
identified
Some query nodes are currently experiencing excessive load, leading to 429s for certain namespaces. We have identified the problem and have a fix in preparation.
Critical
TLS certificate expiration in aws-us-east-1
Started
Tue, Dec 16, 2025, 05:24:23 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Sun, Dec 14, 2025, 09:28:00 PM
Duration
—
resolved
Service has been fully restored.
monitoring
The affected certificates were manually renewed. Connections to aws-us-east-1.turbopuffer.com should once again pass TLS validation.
identified
We've identified that a stuck ACME challenge prevented cert-manager from automatically renewing the TLS certificate for aws-us-east-1.turbopuffer.com.
investigating
We are investigating reports of connections to aws-us-east-1.turbopuffer.com failing TLS validation.
Minor
indexing backlog in aws-us-west-2
Started
Wed, Jan 7, 2026, 06:20:52 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 7, 2026, 03:00:00 PM
Duration
—
resolved
incident was resolved
identified
misconfiguration broke indexer auto-scaling, causing a backlog of jobs
Minor
Degraded performance in aws-us-west-2
Started
Wed, Jan 28, 2026, 03:19:02 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 04:12:41 AM
Duration
53m
resolved
Latencies have remained stable.
monitoring
The region has fully transitioned back to Intel hardware and API response latency has returned to normal. We're monitoring to ensure that latencies remain stable.
identified
We've identified a pathological performance issue when running on ARM hardware. We've shifted the region to run entirely on Intel hardware.
investigating
We're investigating reports of degraded performance in aws-us-west-2.
Minor
aws-us-west-2: Increased 429's on a subset of nodes
Started
Wed, Jan 28, 2026, 10:10:31 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 10:15:00 AM
Duration
4m
resolved
The issue has been resolved. We deployed a fix and confirmed that affected nodes have returned to normal operation.
monitoring
We identified an issue causing high CPU utilization on a subset of query nodes in us-west-2. Some customers experienced 429 errors and increased query latency during this time.
We have identified the root cause and are deploying a fix. We are actively monitoring the situation and implementing additional safeguards to prevent this from recurring.
Minor
Elevated 500s in region gcp-us-east4 due to transient unavailability of turbopuffer query nodes
Started
Wed, Jan 28, 2026, 01:50:31 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 04:41:29 PM
Duration
2h 50m
resolved
Service has returned to normal. The cause has been identified and is being addressed.
monitoring
We've identified the problem and taken action to stop customer impact.
investigating
Customers may notice a small number of their requests receiving a 500 response due to transient unavailability of turbopuffer query nodes. We have identified a probable cause and are moving to confirming and addressing it.
Critical
Outage in aws-us-west-2
Started
Thu, Jan 29, 2026, 03:30:15 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Jan 29, 2026, 04:47:29 AM
Duration
1h 17m
resolved
Service has been fully restored. We'll be sharing a full RCA with affected customers as soon as possible.
monitoring
We've identified the problem and rolled out a fix. We're closely monitoring the situation.
investigating
We are investigating reports of an outage in aws-us-west-2.