UTHPC incident

GPU nodes on our Rocket cluster are currently unavailable

Major Resolved View vendor source →

UTHPC experienced a major incident on October 1, 2026, lasting —. The incident has been resolved; the full update timeline is below.

Started
Oct 01, 2026, 12:09 PM UTC
Resolved
Oct 01, 2026, 12:09 PM UTC
Duration
—
Detected by Pingoru
Oct 01, 2026, 12:09 PM UTC

Update timeline

  1. resolved Oct 01, 2026, 12:09 PM UTC

    Type: Incident Duration: 1 hour and 14 minutes Affected Components: ondemand.hpc.ut.ee, rocket.hpc.ut.ee, galaxy.hpc.ut.ee Oct 1, 12:09:48 GMT+0 - Investigating - GPU nodes on our Rocket cluster are currently unavailable due to a recently disclosed vulnerability. All ongoing jobs on them will be killed as the nodes will undergo emergency maintenance. We apologize for any inconvenience and will inform you when the nodes are available for submission. Oct 1, 12:44:43 GMT+0 - Identified - GPU nodes on our Rocket cluster are currently unavailable due to a recently disclosed vulnerability. All ongoing jobs on these nodes will be stopped as the nodes undergo emergency maintenance. Once GPU node availability is restored, the affected jobs will be returned to the job queue. We apologize for any inconvenience and will inform you when the nodes are available again. Oct 1, 13:24:15 GMT+0 - Resolved - The emergency maintenance has been completed, and the GPU nodes on the Rocket cluster are available again. Affected jobs have been requeued and will run again as resources become available. Thank you for your patience and understanding.