New B200 spot capacity is live in US East from $1.69 per GPU-hour. See availability
Incident history

January 2025.

Every incident and maintenance window that started this month, newest first, with what we posted at the time. Times are UTC. The record starts at launch and nothing is removed from it.

Incidents
3282 minutes of degraded or reduced service
Major outages
1Each one has a post-incident review
Maintenance windows
1Announced at least seven days ahead
Record since
14 Sep 2024Platform launch; every event since is listed

Console unavailable

Resolved
Major outage26 Jan 2025, 17:55 UTC · lasted 2 h 7 minConsole

Resolved — This incident has been resolved.

Storage cluster firmware update — US East

Completed
Scheduled maintenance22 Jan 2025, 07:00 UTC · lasted 3 h 8 minPersistent storage — US East (Virginia)

Completed — The maintenance is complete and every component is back to normal operation.

Volume attach failures in EU Central (Frankfurt)

Resolved
Partial outage18 Jan 2025, 05:23 UTC · lasted 45 minPersistent storage — EU Central (Frankfurt), Instance launch and lifecycle — EU Central (Frankfurt)

Resolved — This incident has been resolved. Launches that failed at the attach step were not billed.

Uptime counts major outages in full and partial outages at 60% of their duration; degraded performance and scheduled maintenance are not counted as downtime. Feeds: Atom · JSON.