Degraded Performance on Lightning Cloud Workloads
- resolved
This incident has been resolved.
- investigating
Lightning Cloud workloads are experiencing degraded performance on the platform. We are investigating. In the meantime, please use GCP or AWS.
status.lightning.ai
Studios, jobs & model serving
All Systems Operational
This incident has been resolved.
Lightning Cloud workloads are experiencing degraded performance on the platform. We are investigating. In the meantime, please use GCP or AWS.
This incident has been resolved.
Lightning Cloud workloads are experiencing degraded performance on the platform. We are investigating. In the meantime, please use GCP or AWS.
This issue is now resolved
Lightning Cloud workloads are not currently able to start on the platform. We are investigating. In the meantime, please use GCP or AWS
This incident has been resolved.
The issue has been identified and a fix is being implemented.
We are experiencing request failures and issues loading the website. Running compute (Jobs, Studios) are unaffected.
This incident has been resolved.
The third-party outage is recovering and we are continuing to monitor the situation.
We are seeing degraded performance and intermittent errors on Studio environment restoration due to a third-party outage. Some Studios are experiencing slow or failed environment restoration.
This incident has been resolved.
The issue has been identified and a fix is being implemented.
We are experiencing issues with website loading and some http requests to Studios and Deployments. Running workloads and connections over SSH are not impacted.
This incident is now resolved
The incident is ongoing while we work with AWS towards resolution. Currently, new AWS workloads cannot be started on Lightning Cloud. Already running workloads and customers using bring your own cloud are unaffected. In the meantime, please consider using GCP or Nebius for new studios and jobs.
We have identified the cause and are working on resolution.
We are investigated degraded experience starting AWS machines on Lightning Cloud. Please use GCP or Nebius instead.
This incident has been resolved.
Nebius workloads are currently experiencing degraded performance on the platform.
This incident has been resolved.
Nebius workloads are currently experiencing degraded performance on the platform.
This incident has been resolved.
Nebius workloads are currently experiencing degraded performance on the platform.
We experienced an outage impacting access to lightning.ai between 2:00pm and 2:06pm Pacific. The situation is now resolved.
Resolved
The issue has been identified and we have applied a mitigation and are monitoring the situation.
We are investigating an issue causing intermittent outages of our website. Running machines and SSH access are not impacted.
This incident has been resolved.
Environment cloning and downloading plugin templates are experiencing failures
This incident has been resolved.
AWS has resolved this issue
Due to an ongoing issue with AWS some Studios and Jobs are experiencing issues including slow saving and starting.
Between 14:00 and 14:30 UTC, some customer workloads experienced a temporary interruption due to an unexpected platform issue. During this time, deployments may have paused or restarted. All systems are now fully operational, and no further impact is expected. We’re continuing to monitor closely and reviewing the cause to prevent similar incidents in the future.
Login was temporarily not functioning in some regions, active workloads not impacted Started: 10:51 AM EST Resolved: 11:02 AM EST
This incident has been resolved.
We are currently experiencing a major outage due to an outage on our cloud provider (GCP). All existing workloads are currently impacted and may be unresponsive or degraded. Additionally, new workloads cannot be created at this time. Our team is actively monitoring the situation and working with GCP to resolve the issue as quickly as possible. We will continue to provide updates as we learn more.
This incident has been resolved.
Issue resolved.
Users are seeing the aggregated credit activity of multiple users outside of their orgs/teamspaces. This is purely a UI error and a fix is coming shortly. No other systems impacted.
This incident has been resolved.
Currently metrics for stopped jobs and deployments are temporarily down while we investigate the issue
The issue has been identified and a fix is being implemented.
We are currently experiencing a platform outage and investigating the issue.
This incident has been resolved.
There is an outage with Lightning Studios. We are actively investigating this incident.
This incident has been resolved.
The current github degraded performance is effecting jobs plugins
This incident has been resolved.
This incident has been resolved.
We are currently identifying the root cause for this issue and working on getting a fix out as soon as possible