Search & RAG

status.turbopuffer.com

Turbopuffer

Serverless vector & full-text search

All Systems Operational

Operational
Latency
550ms
Checked
just now
Active incidents
0
Components
20
Source
Statuspage.io API

Overview

Observed uptime · 1 day100%
2026-09-202026-09-20
Component health
20 up0 degraded0 down

Components20

20 operational0 degraded0 outage20 total
  • aws-eu-west-1turbopuffer API
    Operational
  • gcp-northamerica-northeast2turbopuffer API
    Operational
  • gcp-us-central1turbopuffer API
    Operational
  • DashboardAccount management dashboard at turbopuffer.com/dashboard
    Operational
  • gcp-asia-northeast3turbopuffer API
    Operational
  • aws-us-east-1turbopuffer API
    Operational
  • aws-us-east-2turbopuffer API
    Operational
  • aws-us-west-2turbopuffer API
    Operational
  • aws-ap-south-1turbopuffer API
    Operational
  • aws-eu-west-2turbopuffer API
    Operational
  • aws-ca-central-1turbopuffer API
    Operational
  • gcp-asia-southeast1turbopuffer API
    Operational
  • gcp-europe-west3turbopuffer API
    Operational
  • gcp-us-east4turbopuffer API
    Operational
  • gcp-us-west1turbopuffer API
    Operational
  • aws-ap-southeast-2turbopuffer API
    Operational
  • aws-eu-central-1turbopuffer API
    Operational
  • aws-sa-east-1turbopuffer API
    Operational
  • gcp-us-east1turbopuffer API
    Operational
  • gcp-europe-west1turbopuffer API
    Operational

Incidents25

History 25

Critical

Dashboard outage

Started
Sat, Sep 5, 2026, 06:40:07 PM
Updated
Sat, Sep 5, 2026, 06:52:40 PM
Resolved
Sat, Sep 5, 2026, 06:52:40 PM
Duration
12m
  1. resolved

    From 7:21-7:41 UTC, the turbopuffer dashboard was unavailable due to an issue with our backing datastore. This issue has been resolved and dashboard access has been restored.

  2. investigating

    The turbopuffer dashboard is currently experiencing an outage. turbopuffer regions are fully available. Writes and queries are not affected.

Major

Elevated rate of 5xx in gcp-us-central1

Started
Wed, Sep 2, 2026, 04:49:00 PM
Updated
Wed, Sep 2, 2026, 05:45:24 PM
Resolved
Wed, Sep 2, 2026, 05:16:00 PM
Duration
27m
  1. resolved

    The issue causing elevated 5xx errors for some namespaces has been resolved. We repaired the affected namespaces, and error rates have returned to normal. We will continue monitoring the service closely.

  2. monitoring

    Some namespaces are experiencing an elevated rate of 5xx errors. We have identified the issue and are working to restore full service.

Critical

Outage in gcp-us-southeast1

Started
Thu, Aug 27, 2026, 08:37:35 PM
Updated
Thu, Aug 27, 2026, 10:15:13 PM
Resolved
Thu, Aug 27, 2026, 10:15:13 PM
Duration
1h 37m
  1. resolved

    Service has been fully restored.

  2. monitoring

    We’ve identified the problem and rolled out a fix. We’re closely monitoring the situation.

  3. investigating

    Full outage identified in gcp-southeast1

Major

Errors when accessing dashboard

Started
Thu, Jul 16, 2026, 09:42:52 AM
Updated
Thu, Jul 16, 2026, 10:49:22 AM
Resolved
Thu, Jul 16, 2026, 10:49:22 AM
Duration
1h 6m
  1. resolved

    Our authentication provider has resolved the incident on their end, and dashboard logins have been fully operational since 10:08 UTC. Existing sessions and the database were unaffected throughout.

  2. monitoring

    Logins are succeeding again as of 10:08 UTC. Our authentication provider is still working through their incident, so some logins may be slower than usual. We'll continue monitoring until it's fully resolved.

  3. identified

    Dashboard logins (Google, SSO, and email) are intermittently failing due to an ongoing incident at [our authentication provider](https://status.workos.com/incidents/b52fkqjzcl20). We're monitoring and will update here.

  4. investigating

    We're aware of issues affecting some customers who are logging into the dashboard. We're investigating and will provide an update shortly. Note that the database is not affected - this issue is limited to the dashboard.

Minor

Elevated 500 rate on delete_by_filter writes, filtered queries

Started
Thu, May 28, 2026, 06:52:16 PM
Updated
Thu, May 28, 2026, 06:54:48 PM
Resolved
Thu, May 28, 2026, 06:30:00 PM
Duration
  1. resolved

    Starting at ~17:55 UTC, we experienced an elevated 500 rate affecting delete_by_filter, and filtered queries as a result of a recent change deployed to filtering. This change was reverted, and errors fully subsided by ~18:30 UTC. This incident affected approximately 38.2% of `delete_by_filter` requests in aws-us-east-1, and a small percentage of filtered queries in other regions.

  2. monitoring

    Investigating elevated 500 error rate on filtered queries, delete_by_filter operations.

Minor

partial outage in aws-us-west-2

Started
Wed, May 20, 2026, 03:42:34 PM
Updated
Wed, May 20, 2026, 04:19:38 PM
Resolved
Wed, May 20, 2026, 04:19:38 PM
Duration
37m
  1. resolved

    Our deploy has rolled out and affected namespaces have recovered.

  2. identified

    Our team has identified an issue with a deploy in aws-us-west-2, causing queries and writes to 500 for affected namespaces. We are deploying a fix.

Critical

Errors when accessing dashboard

Started
Mon, Apr 20, 2026, 12:22:41 PM
Updated
Mon, Apr 20, 2026, 05:32:27 PM
Resolved
Mon, Apr 20, 2026, 05:32:27 PM
Duration
5h 9m
  1. resolved

    This incident is resolved. Dashboard access is working as usual. As noted earlier, this incident only affected the dashboard and had no impact on turbopuffer regions or namespaces!

  2. monitoring

    Traffic has returned to normal. We're monitoring closely!

  3. monitoring

    We're seeing recovery with traffic returning to normal. We're monitoring closely and will share another update shortly.

  4. identified

    This issue is currently affecting the dashboard and API key management. **turbopuffer regions and namespaces are all operating normally**, and existing API keys continue to authenticate as expected. The root cause is a dependency service that we use for metadata. We'll share an update as we make progress.

  5. investigating

    We've confirmed reports of dashboard errors and are actively investigating. Note that the database is not affected - this issue is limited to the dashboard.

Critical

Errors when accessing dashboard

Started
Thu, Apr 2, 2026, 12:14:00 PM
Updated
Thu, Apr 2, 2026, 01:10:34 PM
Resolved
Thu, Apr 2, 2026, 01:10:34 PM
Duration
56m
  1. resolved

    We were impacted by [an issue with one of our providers](https://www.vercel-status.com/incidents/5r9bp5y8rql2), which we were able to resolve. We will continue to monitor for regressions.

  2. investigating

    We've confirmed reports of dashboard errors and our team is investigating.

Major

Elevated 408 and 429s

Started
Tue, Mar 24, 2026, 09:50:17 AM
Updated
Tue, Mar 24, 2026, 10:18:49 AM
Resolved
Tue, Mar 24, 2026, 10:18:48 AM
Duration
28m
  1. resolved

    We've observed a complete recovery of traffic.

  2. monitoring

    Our team has taken remedial action and is observing recovery, we will continue to monitor.

  3. investigating

    The gcp-us-east4 cluster is experiencing elevated rates of 408 and 429s affecting both reads and writes affecting approximately 4% of traffic. Our team is investigating.

Major

Increased Request Latency and 429s in GCP us-east4

Started
Tue, Mar 17, 2026, 09:58:01 AM
Updated
Tue, Mar 17, 2026, 10:48:08 AM
Resolved
Tue, Mar 17, 2026, 10:48:08 AM
Duration
50m
  1. resolved

    The region is stable and has fully recovered. We will continue to monitor.

  2. monitoring

    Latencies have returned to normal and 429s have fully recovered. The issue was related to transient failures in the underlying infrastructure. Our engineers have applied mitigations, and we are continuing to monitor.

  3. investigating

    We are looking into an issue causing increased latency and 429s in us-east4.

Major

Elevated 500s in region gcp-us-west1

Started
Mon, Mar 16, 2026, 08:55:43 AM
Updated
Mon, Mar 16, 2026, 09:39:58 AM
Resolved
Mon, Mar 16, 2026, 09:39:58 AM
Duration
44m
  1. resolved

    The region has stabilized and returned to normal operation. We are continuing to monitor!

  2. monitoring

    A fix has been rolled out and 5xx errors in the region have reduced. We are continuing to monitor the situation!

  3. investigating

    We are investigating an increase 500's in the region!

Major

Partial outage in gcp-us-east4

Started
Thu, Oct 16, 2025, 09:12:29 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Oct 16, 2025, 10:52:20 AM
Duration
1h 39m
  1. resolved

    Full recovery as of 12:24 PM UTC

  2. monitoring

    We are seeing degraded performance again as latency to Google Cloud Storage. We are continuing to monitor and work with the GCP team.

  3. resolved

    Full recovery as of 10:44am UTC. We are investigating the root cause with the GCP team.

  4. monitoring

    Substantial improvement starting 10:35am UTC. We are continuing to investigate and work directly with GCP to resolve this fully.

  5. identified

    8:30am UTC - 9:25am UTC: Degraded performance on queries and writes in gcp-us-east4 from increased latency to Google Cloud Storage. 9:25am UTC: Escalated to partial outage both write and read paths due to climbing latency and timeouts from Google Cloud Storage. * Queries may see elevated rates of 429s, and 500s. * Writes may observe elevated rates of 408s, and 500s. Our team is working to restore full system performance, and we are working with GCP to restore service as soon as possible.

Major

Outage in aws us-east-1

Started
Mon, Oct 20, 2025, 07:58:57 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Oct 20, 2025, 11:18:06 PM
Duration
15h 19m
  1. resolved

    AWS has reported that the upstream incident is resolved.

  2. monitoring

    The turbopuffer service remains healthy in the aws-us-east-1 region. We continue to closely monitor the situation while AWS continues to report service disruption in the region.

  3. monitoring

    AWS is experiencing a major disruption in region us-east-1 impacting multiple services (<https://health.aws.amazon.com/health/status?eventID=arn:aws:health:us-east-1::event/MULTIPLE_SERVICES/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE/AWS_MULTIPLE_SERVICES_OPERATIONAL_ISSUE_BA540_514A652BE1A>). This caused turbopuffer to become unavailable in that region starting at 7:50 UTC. We've rolled out mitigations which resulted in a full recovery of turbopuffer at 8:20 UTC (despite the ongoing AWS outage). An existing mitigation to guard against EC2 instance launch issues has prevented the long-running EC2 degradation from further affecting turbopuffer availability. The AWS incident is ongoing. We will continue monitoring the cluster's health.

  4. monitoring

    We have deployed mitigations and are monitoring the situation.

  5. investigating

    [AWS us-east-1 is experiencing a major outage and many services are affected. ](https://health.aws.amazon.com/health/status) turbopuffer clients in aws-us-east-1 will see elevated error rate and latencies during this outage. We're looking at mitigating the impact.

Minor

Degraded performance in gcp-us-east4

Started
Mon, Oct 20, 2025, 09:55:09 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Oct 20, 2025, 05:22:02 PM
Duration
7h 26m
  1. resolved

    Service is stable. We are continuing to investigate the root cause with high priority.

  2. monitoring

    We have observed an increase in timeouts and a slow down in indexing starting at 9:38 UTC. The service has returned to normal at 10:00 UTC. We are monitoring and investigating the root cause.

  3. investigating

    we're currently observing degraded performance in gcp-us-east4 leading to elevated rates of 408s and 429s. Our team is currently investigating.

Major

Partial dashboard outage

Started
Mon, Oct 20, 2025, 07:42:21 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Tue, Oct 21, 2025, 02:31:00 AM
Duration
6h 48m
  1. resolved

    The upstream incident is resolved.

  2. monitoring

    Dashboard error rates have subsided. We continue to monitor the situation while we wait for confirmation from our authentication provider that the incident has resolved.

  3. identified

    Some users are reporting issues accessing the turbopuffer dashboard due to a partial outage with our authentication provider: <https://status.workos.com/incidents/rdz189wnjhyn> turbopuffer APIs are unaffected.

Minor

Elevated 500s in gcp-us-west1

Started
Mon, Nov 24, 2025, 01:24:36 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Nov 24, 2025, 01:27:41 PM
Duration
3m
  1. resolved

    We have taken steps to remediate the 500s, which have now subsided. We're investigating the root cause.

  2. identified

    We're investigating an increase in 500 responses in the region.

Minor

degraded performance in aws-us-east-1

Started
Thu, Dec 11, 2025, 05:52:34 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Dec 11, 2025, 07:31:53 PM
Duration
1h 39m
  1. resolved

    we have not seen any re-occurences of the issue.

  2. monitoring

    we've removed problematic query nodes from rotation and are monitoring for further issues.

  3. investigating

    some namespaces in aws-us-east-1 may experience degraded performance and request timeouts. our team is working to rectify the issue.

Minor

Degraded service (500 responses)

Started
Mon, Dec 15, 2025, 10:25:55 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Dec 15, 2025, 10:30:58 AM
Duration
5m
  1. resolved

    Service is stable.

  2. monitoring

    Some of our query nodes restarted repeatedly due to OOMs caused by a code path that was producing very large allocations. We have deployed a fix, and the service has recovered. We are monitoring.

Minor

Degraded performance in aws-us-west-2

Started
Mon, Dec 15, 2025, 05:14:51 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Mon, Dec 15, 2025, 06:03:54 PM
Duration
49m
  1. resolved

    The problematic queries have been isolated and traffic has been restored to normal.

  2. identified

    Some query nodes are currently experiencing excessive load, leading to 429s for certain namespaces. We have identified the problem and have a fix in preparation.

Critical

TLS certificate expiration in aws-us-east-1

Started
Tue, Dec 16, 2025, 05:24:23 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Sun, Dec 14, 2025, 09:28:00 PM
Duration
  1. resolved

    Service has been fully restored.

  2. monitoring

    The affected certificates were manually renewed. Connections to aws-us-east-1.turbopuffer.com should once again pass TLS validation.

  3. identified

    We've identified that a stuck ACME challenge prevented cert-manager from automatically renewing the TLS certificate for aws-us-east-1.turbopuffer.com.

  4. investigating

    We are investigating reports of connections to aws-us-east-1.turbopuffer.com failing TLS validation.

Minor

indexing backlog in aws-us-west-2

Started
Wed, Jan 7, 2026, 06:20:52 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 7, 2026, 03:00:00 PM
Duration
  1. resolved

    incident was resolved

  2. identified

    misconfiguration broke indexer auto-scaling, causing a backlog of jobs

Minor

Degraded performance in aws-us-west-2

Started
Wed, Jan 28, 2026, 03:19:02 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 04:12:41 AM
Duration
53m
  1. resolved

    Latencies have remained stable.

  2. monitoring

    The region has fully transitioned back to Intel hardware and API response latency has returned to normal. We're monitoring to ensure that latencies remain stable.

  3. identified

    We've identified a pathological performance issue when running on ARM hardware. We've shifted the region to run entirely on Intel hardware.

  4. investigating

    We're investigating reports of degraded performance in aws-us-west-2.

Minor

aws-us-west-2: Increased 429's on a subset of nodes

Started
Wed, Jan 28, 2026, 10:10:31 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 10:15:00 AM
Duration
4m
  1. resolved

    The issue has been resolved. We deployed a fix and confirmed that affected nodes have returned to normal operation.

  2. monitoring

    We identified an issue causing high CPU utilization on a subset of query nodes in us-west-2. Some customers experienced 429 errors and increased query latency during this time. We have identified the root cause and are deploying a fix. We are actively monitoring the situation and implementing additional safeguards to prevent this from recurring.

Minor

Elevated 500s in region gcp-us-east4 due to transient unavailability of turbopuffer query nodes

Started
Wed, Jan 28, 2026, 01:50:31 PM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Wed, Jan 28, 2026, 04:41:29 PM
Duration
2h 50m
  1. resolved

    Service has returned to normal. The cause has been identified and is being addressed.

  2. monitoring

    We've identified the problem and taken action to stop customer impact.

  3. investigating

    Customers may notice a small number of their requests receiving a 500 response due to transient unavailability of turbopuffer query nodes. We have identified a probable cause and are moving to confirming and addressing it.

Critical

Outage in aws-us-west-2

Started
Thu, Jan 29, 2026, 03:30:15 AM
Updated
Thu, Mar 12, 2026, 06:21:09 PM
Resolved
Thu, Jan 29, 2026, 04:47:29 AM
Duration
1h 17m
  1. resolved

    Service has been fully restored. We'll be sharing a full RCA with affected customers as soon as possible.

  2. monitoring

    We've identified the problem and rolled out a fix. We're closely monitoring the situation.

  3. investigating

    We are investigating reports of an outage in aws-us-west-2.

Watch Turbopuffer
Email alerts on every status change — outages, degradations, new incidents, and resolutions.

Watching all 1 providers. Customize on the alerts page.

Details

Aliases
turbo puffer
Indicator
none
Path
/turbopuffer

Related in Search & RAG