<?xml version="1.0" encoding="UTF-8"?>
<feed xml:lang="en-US" xmlns="http://www.w3.org/2005/Atom">
  <id>tag:status.hpc.ut.ee,2005:/history</id>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee"/>
  <link rel="self" type="application/atom+xml" href="https://status.hpc.ut.ee/history.atom"/>
  <title>UTHPC Status - Incident history</title>
  <updated>2026-07-11T06:00:00.000+00:00</updated>
  <author>
    <name>UTHPC</name>
  </author>
  
<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmrg2w1j408d60kmkyumjshqn</id>
  <published>2026-07-11T06:00:00.000+00:00</published>
  <updated>2026-07-11T06:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmrg2w1j408d60kmkyumjshqn"/>
  <title>University of Tartu cloud service emergency maintenance - mitigating a Linux kernel CVE</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 3 hours and 5 minutes</p>
    
    <p><small>Jul <var data-var='date'> 11</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Investigating</strong> -
  Due to a recently disclosed critical security vulnerability affecting Linux kernel-based virtualization platforms, we are implementing preventive security measures.

As part of these maintenance activities, UT Cloud virtual machines (VMs) will be restarted. The VMs will be brought back online in a rolling manner over the next couple of hours. Besides virtual machines, some websites are also affected. In case of questions, please contact us at support@hpc.ut.ee.

We apologize for any inconvenience and appreciate your understanding as we work to ensure the security and stability of our platform..</p>
<p><small>Jul <var data-var='date'> 11</var>, <var data-var='time'>09:05:09</var> GMT+0</small><br /><strong>Resolved</strong> -
  The maintenance is finished. The UT Cloud service is available, and affected virtual machines and websites are restored. .</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmrewic2606rp0zmvcb52dk4l</id>
  <published>2026-07-10T12:17:42.798+00:00</published>
  <updated>2026-07-10T12:17:42.806+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmrewic2606rp0zmvcb52dk4l"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>12:17:42</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>12:19:42</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmretvle102cc0rt94msvdpy3</id>
  <published>2026-07-10T11:04:02.568+00:00</published>
  <updated>2026-07-10T12:06:02.811+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmretvle102cc0rt94msvdpy3"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 2 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>12:06:02</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>
<p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>11:04:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmrdun1q402x20zs3q6j4jc55</id>
  <published>2026-07-09T18:37:36.855+00:00</published>
  <updated>2026-07-10T13:02:11.787+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmrdun1q402x20zs3q6j4jc55"/>
  <title>HPC cluster Rocket is temporarily unavailable due to ongoing security fixes (CVE-2026-46242 and CVE-2026-43499)</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 18 hours and 25 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand, rocket.hpc.ut.ee, Galaxy</p>
    <p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>13:02:11</var> GMT+0</small><br /><strong>Resolved</strong> -
  **HPC services are restored.** Access to the UT Rocket cluster and Galaxy, Open OnDemand, and their services RStudio and Jupyter have been fully restored.

The previously identified Linux kernel vulnerabilities have been addressed, and users can once again access the Rocket login nodes as usual and submit new jobs to the compute nodes..</p>
<p><small>Jul <var data-var='date'> 9</var>, <var data-var='time'>18:37:36</var> GMT+0</small><br /><strong>Investigating</strong> -
  **Due to recently identified Linux kernel vulnerabilities (**[**CVE-2026-46242**](https://access.redhat.com/security/cve/cve-2026-46242) **and** [**CVE-2026-43499**](https://access.redhat.com/security/cve/cve-2026-43499)**), we have temporarily disabled access to the UT HPC Rocket cluster as a precautionary measure to mitigate potential security risks.**

Currently, access to login nodes is unavailable. Also, new job submissions are temporarily disabled.  

Jobs that were already running before the access restrictions were applied continue to run and are not affected.  

Our system administrators are working to restore full service as quickly as possible. 

Thank you for your patience!.</p>
<p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>06:14:21</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have applied the security patches to mitigate the vulnerability. The login nodes are now accessible; however, job submission remains disabled while we complete additional validation and testing. We require more time to ensure the cluster is fully stable before restoring normal operation. 

We appreciate your patience. .</p>
<p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>07:04:18</var> GMT+0</small><br /><strong>Identified</strong> -
  Due to the same security vulnerability, all SAPU machines are currently unavailable while security patches are being applied. We expect service to be restored later tonight. Thank you for your patience..</p>
<p><small>Jul <var data-var='date'> 10</var>, <var data-var='time'>08:42:15</var> GMT+0</small><br /><strong>Identified</strong> -
  We are closing login nodes access again. The Rocket cluster is continuously not available. We expect the cluster to remain unavailable for most of the day. We apologize for the inconvenience and will provide updates as they become available..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmreyz56k07kn0rq6gbjx2v55</id>
  <published>2026-07-09T12:00:00.000+00:00</published>
  <updated>2026-07-09T12:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmreyz56k07kn0rq6gbjx2v55"/>
  <title>Mainenance on all SAPU machines due to Linux kernel security vulnerabilities - service inaccessible</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 3 days, 5 hours and 53 minutes</p>
    
    <p><small>Jul <var data-var='date'> 9</var>, <var data-var='time'>12:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Due to the Linux kernel security vulnerabilities that also affected the HPC Rocket cluster (incident on 9 July 2026), all SAPU virtual machines are temporarily unavailable while security updates are being applied.

We will provide updates as progress is made. Thank you for understanding..</p>
<p><small>Jul <var data-var='date'> 12</var>, <var data-var='time'>17:52:37</var> GMT+0</small><br /><strong>Resolved</strong> -
  All SAPU workstations are now patched and available. We apologize for the inconvenience and appreciate your understanding. If you encounter any issues with your workstation, be sure to contact support@hpc.ut.ee.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmqwpjt0p05sm2opc6f6bhbzx</id>
  <published>2026-06-27T18:43:02.952+00:00</published>
  <updated>2026-06-27T18:43:02.959+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmqwpjt0p05sm2opc6f6bhbzx"/>
  <title>support.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 4 minutes</p>
    <p><strong>Affected Components:</strong> support.hpc.ut.ee</p>
    <p><small>Jun <var data-var='date'> 27</var>, <var data-var='time'>18:43:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  support.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Jun <var data-var='date'> 27</var>, <var data-var='time'>18:47:02</var> GMT+0</small><br /><strong>Resolved</strong> -
  support.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmpwhyl5x00u4qim26p4s81bv</id>
  <published>2026-06-26T06:00:00.000+00:00</published>
  <updated>2026-06-26T06:00:01.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmpwhyl5x00u4qim26p4s81bv"/>
  <title>University of Tartu cloud service maintenance from 26 June to 3 July 2026.</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 7 days and 30 minutes</p>
    <p><strong>Affected Components:</strong> minu.etais.ee</p>
    <p><small>Jun <var data-var='date'> 26</var>, <var data-var='time'>06:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>
<p><small>Jul <var data-var='date'> 3</var>, <var data-var='time'>06:29:59</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully..</p>
<p><small>Jun <var data-var='date'> 26</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance of the University of Tartu HPC Center&#039;s cloud service will take place from 26 June to 3 July 2026\. 

Error-prone operations will be performed outside business hours to minimize impact on users. In the [minu.etais.ee](http://minu.etais.ee) self-service environment, there may be short-term interruptions in the UT HPC (Public) and UT HPC (Campus) cloud access or in virtual machines&#039; network connection. 

Regular service updates are essential to maintaining a secure and resilient infrastructure while ensuring the long-term reliability of the services. 

If you have any questions or encounter any issues, please contact the technical support team at support@hpc.ut.ee..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmq0p0w5r01h1nwa6rtwe8y0t</id>
  <published>2026-06-05T08:59:42.926+00:00</published>
  <updated>2026-06-05T08:59:42.936+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmq0p0w5r01h1nwa6rtwe8y0t"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 18 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Jun <var data-var='date'> 5</var>, <var data-var='time'>08:59:42</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Jun <var data-var='date'> 5</var>, <var data-var='time'>09:17:42</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmpwn89ok00j7qg37v3zzjgtr</id>
  <published>2026-06-02T12:58:22.791+00:00</published>
  <updated>2026-06-02T12:58:22.791+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmpwn89ok00j7qg37v3zzjgtr"/>
  <title>Rocket cluster file system instability</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 day, 2 hours and 2 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy, Open OnDemand, rocket.hpc.ut.ee</p>
    <p><small>Jun <var data-var='date'> 2</var>, <var data-var='time'>12:58:22</var> GMT+0</small><br /><strong>Investigating</strong> -
  We have identified an issue with the rocket cluster file systems that causes transient errors on nodes. Less nodes may be available while we work on fixing the underlying issue..</p>
<p><small>Jun <var data-var='date'> 3</var>, <var data-var='time'>15:00:15</var> GMT+0</small><br /><strong>Resolved</strong> -
  This underlying issues with the filesystem have been identified and resolved. All systems back operational..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmpwc9mhw0114p7m541wpskio</id>
  <published>2026-06-02T07:51:30.330+00:00</published>
  <updated>2026-06-02T09:50:18.390+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmpwc9mhw0114p7m541wpskio"/>
  <title>Issues with MyAccessID UT identity verification</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 59 minutes</p>
    <p><strong>Affected Components:</strong> kubernetes.hpc.ut.ee, puhuri.metacenter.no, puhuri-portal.neic.no, minu.etais.ee, account.lumi.cscs.ch, Open OnDemand, my.lumi-supercomputer.eu, lumi.deic.dk</p>
    <p><small>Jun <var data-var='date'> 2</var>, <var data-var='time'>09:50:18</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>
<p><small>Jun <var data-var='date'> 2</var>, <var data-var='time'>07:51:30</var> GMT+0</small><br /><strong>Investigating</strong> -
  The University of Tartu authentication system is currently experiencing technical issues. As a result, login using **MyAccessID UT identity verification** is temporarily unavailable. The University of Tartu Information Technology Office (ITO) has been informed and is working on resolving the issue. We apologize for the inconvenience and recommend trying again later..</p>
<p><small>Jun <var data-var='date'> 2</var>, <var data-var='time'>08:24:40</var> GMT+0</small><br /><strong>Investigating</strong> -
  This issue is also impacting the MyAccessID authentication process for Kubernetes..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmp6zkdev02pfpflrvvxpe40f</id>
  <published>2026-05-15T14:01:42.630+00:00</published>
  <updated>2026-05-15T14:01:42.642+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmp6zkdev02pfpflrvvxpe40f"/>
  <title>docs.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 12 minutes</p>
    <p><strong>Affected Components:</strong> docs.hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 15</var>, <var data-var='time'>14:01:42</var> GMT+0</small><br /><strong>Investigating</strong> -
  docs.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>May <var data-var='date'> 15</var>, <var data-var='time'>14:13:42</var> GMT+0</small><br /><strong>Resolved</strong> -
  docs.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmp6zjy8o00egqnlyjnbm7d9e</id>
  <published>2026-05-15T14:01:22.968+00:00</published>
  <updated>2026-05-15T14:01:22.984+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmp6zjy8o00egqnlyjnbm7d9e"/>
  <title>hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 14 minutes</p>
    <p><strong>Affected Components:</strong> hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 15</var>, <var data-var='time'>14:01:22</var> GMT+0</small><br /><strong>Investigating</strong> -
  hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>May <var data-var='date'> 15</var>, <var data-var='time'>14:15:22</var> GMT+0</small><br /><strong>Resolved</strong> -
  hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmp69k5pq0aajqjjinz3ehwiy</id>
  <published>2026-05-15T01:53:42.638+00:00</published>
  <updated>2026-05-15T01:53:42.649+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmp69k5pq0aajqjjinz3ehwiy"/>
  <title>docs.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 14 minutes</p>
    <p><strong>Affected Components:</strong> docs.hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 15</var>, <var data-var='time'>01:53:42</var> GMT+0</small><br /><strong>Investigating</strong> -
  docs.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>May <var data-var='date'> 15</var>, <var data-var='time'>02:07:42</var> GMT+0</small><br /><strong>Resolved</strong> -
  docs.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmp69jq9m0ex7qrnhjacxe7p8</id>
  <published>2026-05-15T01:53:22.618+00:00</published>
  <updated>2026-05-15T01:53:22.626+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmp69jq9m0ex7qrnhjacxe7p8"/>
  <title>hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 14 minutes</p>
    <p><strong>Affected Components:</strong> hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 15</var>, <var data-var='time'>01:53:22</var> GMT+0</small><br /><strong>Investigating</strong> -
  hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>May <var data-var='date'> 15</var>, <var data-var='time'>02:07:22</var> GMT+0</small><br /><strong>Resolved</strong> -
  hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmp69jb530a8sqjjtmipsohlv</id>
  <published>2026-05-15T01:53:03.014+00:00</published>
  <updated>2026-05-15T01:53:03.026+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmp69jb530a8sqjjtmipsohlv"/>
  <title>support.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 14 minutes</p>
    <p><strong>Affected Components:</strong> support.hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 15</var>, <var data-var='time'>01:53:03</var> GMT+0</small><br /><strong>Investigating</strong> -
  support.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>May <var data-var='date'> 15</var>, <var data-var='time'>02:07:02</var> GMT+0</small><br /><strong>Resolved</strong> -
  support.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmou0i1rk0057921zwef5b894</id>
  <published>2026-05-11T04:00:00.000+00:00</published>
  <updated>2026-05-11T04:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmou0i1rk0057921zwef5b894"/>
  <title>Maintenance on May 11, 2026 (InfiniBand network upgrade)</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 16 hours and 47 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand, Galaxy, rocket.hpc.ut.ee</p>
    <p><small>May <var data-var='date'> 11</var>, <var data-var='time'>04:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  The University of Tartu HPC **cluster Rocket and related computing services will be unavailable on May 11 due to the cluster’s InfiniBand network upgrade.** The purpose of the upgrade is to quadruple the data transfer throughput between two server rooms.

  
During the maintenance period, it will not be possible to run computing jobs on the following systems:

* UT HPC cluster **Rocket**
* **Open OnDemand** services ([ondemand.hpc.ut.ee](http://ondemand.hpc.ut.ee))
* **Galaxy (**[galaxy.hpc.ut.ee](http://galaxy.hpc.ut.ee)

Access to home directories, project folders, the SFTP service, and network drives (nfs, samba) will remain available.  

**What to keep in mind:**

* On May 9 at 07:00, the SLURM queue will be restricted; jobs running longer than 2 days will not start.
* On May 11 at 07:00, the SLURM queue will be closed for all computations.
* On May 12 at 07:00, the SLURM queue will reopen, and all services will be fully restored.

For additional questions, please contact our technical support at [support@hpc.ut.ee](mailto:support@hpc.ut.ee).  .</p>
<p><small>May <var data-var='date'> 9</var>, <var data-var='time'>08:16:14</var> GMT+0</small><br /><strong>Identified</strong> -
  Reminder: Starting today, May 9th at 07:00, the SLURM queue is restricted; jobs running longer than 2 days will not start..</p>
<p><small>May <var data-var='date'> 11</var>, <var data-var='time'>04:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>
<p><small>May <var data-var='date'> 11</var>, <var data-var='time'>20:46:39</var> GMT+0</small><br /><strong>Completed</strong> -
  University of Tartu HPC cluster’s InfiniBand network upgrade is finished. The cluster Rocket and related computing services are fully restored..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmol4d2b405n0zafan566ua73</id>
  <published>2026-04-30T06:45:03.229+00:00</published>
  <updated>2026-04-30T09:43:20.891+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmol4d2b405n0zafan566ua73"/>
  <title>Possible temporary service interruptions due to Critical Linux Kernel Vulnerability (CVE-2026-31431)</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 7 hours and 4 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, kubernetes.hpc.ut.ee, Open OnDemand, puhuri.metacenter.no, puhuri-portal.neic.no, minu.etais.ee, account.lumi.cscs.ch, my.lumi-supercomputer.eu, docs.hpc.ut.ee, lumi.deic.dk, hpc.ut.ee, Galaxy, support.hpc.ut.ee, registry.hpc.ut.ee, rocket.hpc.ut.ee</p>
    <p><small>Apr <var data-var='date'> 30</var>, <var data-var='time'>09:43:20</var> GMT+0</small><br /><strong>Identified</strong> -
  We have created a documentation for CVE-2026-31431 mitigation: &lt;https://docs.hpc.ut.ee/public/cve-2026-31431/&gt;  
This is primarily useful for UT Cloud virtual machine managers. We&#039;ll keep updating the document with the best approaches as we learn more. .</p>
<p><small>Apr <var data-var='date'> 30</var>, <var data-var='time'>13:48:46</var> GMT+0</small><br /><strong>Resolved</strong> -
  We have applied mitigation measures for the Critical Linux Kernel Vulnerability (CVE-2026-31431) across all affected services. The incident is resolved. 

However, the virtual machine managers are still required to apply the patches following the guides published here: &lt;https://docs.hpc.ut.ee/public/cve-2026-31431/&gt;

You are welcome to contact support for additional information: [support@hpc.ut.ee](mailto:support@hpc.ut.ee)

Best, 

UT HPC Center.</p>
<p><small>Apr <var data-var='date'> 30</var>, <var data-var='time'>06:45:03</var> GMT+0</small><br /><strong>Identified</strong> -
  Due to the recently disclosed “[Copy Fail](https://copy.fail/)” (CVE-2026-31431) Linux kernel vulnerability, our system administrators are actively applying mitigation measures and updates across all Tartu University HPC Center systems today.

As a result, you may experience temporary interruptions or reduced service availability while this work is in progress. All UT HPC services are affected. At this time, the work is expected to be completed today; however, we will provide further updates if the situation extends beyond today.continuing to work on a fix for this incident.

**Users running their own Linux virtual machines are advised to apply the recommended patches on their systems by following the instructions provided in the “Mitigation” section of the** [**CVE-2026-31431 advisory**](https://copy.fail/)**.**

Thank you for your understanding..</p>
<p><small>Apr <var data-var='date'> 30</var>, <var data-var='time'>08:41:41</var> GMT+0</small><br /><strong>Identified</strong> -
  We are providing an update on the mitigation for the Critical Linux Kernel Vulnerability (CVE-2026-31431).

We have learned that the mitigation described in CVE-2026-31431 is not effective on all Linux-based instances. Specifically, machines running RHEL or SUSE operating systems are currently not supported. As the respective OS providers have not yet released the required patches, we are recommending the following steps: 

```
python3 -c &quot;
import socket, sys
try:
 &amp;nbsp; &amp;nbsp;s = socket.socket(socket.AF_ALG, socket.SOCK_SEQPACKET, 0)
 &amp;nbsp; &amp;nbsp;s.bind((&#039;aead&#039;, &#039;authencesn(hmac(sha256),cbc(aes))&#039;))
 &amp;nbsp; &amp;nbsp;print(&#039;VULNERABLE - &amp;nbsp;continue with next steps&#039;)
 &amp;nbsp; &amp;nbsp;sys.exit(1)
except OSError as e:
 &amp;nbsp; &amp;nbsp;print(&#039;Not vulnerable:&#039;, e)
 &amp;nbsp; &amp;nbsp;sys.exit(0)
&quot;
```

  
2\. Add a kernel parameter of `initcall_blacklist=algif_aead_init` to `/etc/default/grub`:

```
sed -i &quot;s|^\(GRUB_CMDLINE_LINUX=\&quot;.*\)\&quot;\s*$|\1 initcall_blacklist=algif_aead_init\&quot;|&quot; /etc/default/grub
```

  
3\. Check the result:

```
grep GRUB_CMDLINE_LINUX /etc/default/grub

# The line needs to end with initcall_blacklist=algif_aead_init&quot;, for example
GRUB_CMDLINE_LINUX=&quot;... initcall_blacklist=algif_aead_init&quot;

Fix manually if necessary
```

  
4\. Update GRUB configuration:

```
grub2-mkconfig -o /boot/grub2/grub.cfg
```

  
5\. Reboot the machine:

```
reboot -h now
```

  
6\. Check again with the first script

For instance, running Debian or Ubuntu, mitigation can be applied by installing the latest available kernel version and rebooting the machine. .</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmnkgv9t03upsckqttx3ag5u</id>
  <published>2026-03-12T14:32:02.896+00:00</published>
  <updated>2026-03-12T14:32:02.903+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmnkgv9t03upsckqttx3ag5u"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 34 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Mar <var data-var='date'> 12</var>, <var data-var='time'>14:32:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 12</var>, <var data-var='time'>15:06:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmm6bbaa0c371b7r3rvcr6l1</id>
  <published>2026-03-11T15:08:02.914+00:00</published>
  <updated>2026-03-11T15:08:02.923+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmm6bbaa0c371b7r3rvcr6l1"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 18 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>15:08:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>15:26:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmm4bapf0aw9stprng50ekx7</id>
  <published>2026-03-11T14:12:02.930+00:00</published>
  <updated>2026-03-11T14:58:03.691+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmm4bapf0aw9stprng50ekx7"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 46 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>14:58:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>14:12:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmm13mvg03q8q0td9ib8deoh</id>
  <published>2026-03-11T12:42:06.603+00:00</published>
  <updated>2026-03-11T12:42:06.624+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmm13mvg03q8q0td9ib8deoh"/>
  <title>Galaxy outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 4 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>12:42:06</var> GMT+0</small><br /><strong>Investigating</strong> -
  Galaxy cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 11</var>, <var data-var='time'>12:46:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  Galaxy is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmiaqvop03l9lsgvrmpmoegk</id>
  <published>2026-03-08T22:01:02.952+00:00</published>
  <updated>2026-03-08T22:01:02.962+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmiaqvop03l9lsgvrmpmoegk"/>
  <title>support.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 6 minutes</p>
    <p><strong>Affected Components:</strong> support.hpc.ut.ee</p>
    <p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>22:01:02</var> GMT+0</small><br /><strong>Investigating</strong> -
  support.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>22:07:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  support.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmiak0r303nz10xeb69aezne</id>
  <published>2026-03-08T21:55:42.926+00:00</published>
  <updated>2026-03-08T21:55:42.941+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmiak0r303nz10xeb69aezne"/>
  <title>docs.hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 10 minutes</p>
    <p><strong>Affected Components:</strong> docs.hpc.ut.ee</p>
    <p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>21:55:42</var> GMT+0</small><br /><strong>Investigating</strong> -
  docs.hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>22:05:43</var> GMT+0</small><br /><strong>Resolved</strong> -
  docs.hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmmiah0rf03ya9hcelecz34dj</id>
  <published>2026-03-08T21:53:22.970+00:00</published>
  <updated>2026-03-08T21:53:22.984+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmmiah0rf03ya9hcelecz34dj"/>
  <title>hpc.ut.ee outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 14 minutes</p>
    <p><strong>Affected Components:</strong> hpc.ut.ee</p>
    <p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>21:53:22</var> GMT+0</small><br /><strong>Investigating</strong> -
  hpc.ut.ee cannot be accessed at the moment. This incident was created by an automated monitoring service..</p>
<p><small>Mar <var data-var='date'> 8</var>, <var data-var='time'>22:07:23</var> GMT+0</small><br /><strong>Resolved</strong> -
  hpc.ut.ee is now operational! This update was created by an automated monitoring service..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmm37bpja00f3114tp93wasvz</id>
  <published>2026-02-26T08:28:43.370+00:00</published>
  <updated>2026-02-26T08:28:43.370+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmm37bpja00f3114tp93wasvz"/>
  <title>Inaccessible services</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 38 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand, , kubernetes.hpc.ut.ee, puhuri-portal.neic.no, puhuri.metacenter.no, account.lumi.cscs.ch, my.lumi-supercomputer.eu, lumi.deic.dk, , Galaxy, support.hpc.ut.ee, registry.hpc.ut.ee, 
Waldur portals → 
Services →</p>
    <p><small>Feb <var data-var='date'> 26</var>, <var data-var='time'>08:28:43</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident. .</p>
<p><small>Feb <var data-var='date'> 26</var>, <var data-var='time'>08:37:47</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We implemented a fix and are currently monitoring the result. .</p>
<p><small>Feb <var data-var='date'> 26</var>, <var data-var='time'>09:06:52</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cml9hpa070ha38fqdikjn32ri</id>
  <published>2026-02-05T13:26:06.393+00:00</published>
  <updated>2026-02-05T16:31:08.762+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cml9hpa070ha38fqdikjn32ri"/>
  <title>HPC Cluster Software Module Load Interrupted</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 3 hours and 5 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand, rocket.hpc.ut.ee</p>
    <p><small>Feb <var data-var='date'> 5</var>, <var data-var='time'>16:31:08</var> GMT+0</small><br /><strong>Resolved</strong> -
  The cluster software stack has been restored from a prior point-in-time copy. All operations should be normal.</p>
<p><small>Feb <var data-var='date'> 5</var>, <var data-var='time'>13:26:06</var> GMT+0</small><br /><strong>Investigating</strong> -
  In HPC Cluster some software modules loading may give an error. We are fixing the issue. Thank you for your patience. .</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Incident/cmkv1f44a05nj14mmp7p2gx3a</id>
  <published>2026-01-26T10:41:32.703+00:00</published>
  <updated>2026-01-26T10:41:32.703+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/incident/cmkv1f44a05nj14mmp7p2gx3a"/>
  <title>Network issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 39 minutes</p>
    <p><strong>Affected Components:</strong> RStudio, kubernetes.hpc.ut.ee, puhuri.metacenter.no, minu.etais.ee, my.lumi-supercomputer.eu, Open OnDemand, puhuri-portal.neic.no, account.lumi.cscs.ch, docs.hpc.ut.ee, lumi.deic.dk, hpc.ut.ee, Galaxy, support.hpc.ut.ee, registry.hpc.ut.ee, rocket.hpc.ut.ee</p>
    <p><small>Jan <var data-var='date'> 26</var>, <var data-var='time'>10:41:32</var> GMT+0</small><br /><strong>Identified</strong> -
  We have detected a network issue that has caused services to be unavailable..</p>
<p><small>Jan <var data-var='date'> 26</var>, <var data-var='time'>11:20:58</var> GMT+0</small><br /><strong>Resolved</strong> -
  The issue has been resolved, and all services are now available again..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmhsyrm1v0001qx30kutebh5n</id>
  <published>2026-01-12T05:00:00.000+00:00</published>
  <updated>2026-01-15T15:46:14.402+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmhsyrm1v0001qx30kutebh5n"/>
  <title>UTHPC cluster Rocket, its file system and cooling maintenance with service interruption (12.–18. January, 2026)</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 7 days</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand, Galaxy, rocket.hpc.ut.ee</p>
    <p><small>Jan <var data-var='date'> 15</var>, <var data-var='time'>15:46:14</var> GMT+0</small><br /><strong>Identified</strong> -
  Part of the maintenance is completed, which means that the following services are **restored and accessible:**

* file systems **/gpfs/space and /gpfs/helios;**
* **SFTP** service
* SAPU machines

Kubernetes containers, or websites that use **NFS or Samba network drives** (mounted file systems), will be restored tomorrow, **Friday, the 16th, by 11 am.** 

**HPC Cluster and its services are still in maintenance and inaccessible:** 

* UTHPC **cluster Rocket**;
* **OpenOndemand** services (RStudio and Jupyter included)
* **Galaxy** ([galaxy.hpc.ut.ee](http://galaxy.hpc.ut.ee))

The maintenance will be completed by the 19th of January. 

Thank you for your patience.

HPC team.</p>
<p><small>Jan <var data-var='date'> 16</var>, <var data-var='time'>12:31:33</var> GMT+0</small><br /><strong>Identified</strong> -
  Confirming that Kubernetes containers with NFS or samba network drives are fully operational as of 11 am today.   
  
HPC Cluster and its services (see the list below) will be available on 19th of January. 

* UTHPC **cluster Rocket**;
* **OpenOndemand** services (RStudio and Jupyter included)
* **Galaxy** ([galaxy.hpc.ut.ee](http://galaxy.hpc.ut.ee)).</p>
<p><small>Jan <var data-var='date'> 19</var>, <var data-var='time'>05:00:00</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully.</p>
<p><small>Jan <var data-var='date'> 12</var>, <var data-var='time'>05:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>
<p><small>Jan <var data-var='date'> 12</var>, <var data-var='time'>05:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Dear User,

  
**We inform you that the University of Tartu HPC cluster Rocket, its file systems, and the cluster cooling system will undergo scheduled maintenance (will not be available) during the period of January 12-18, 2026.**

**The purpose** of the maintenance is to:

* increase the cluster&#039;s CPU and GPU capacity;
* expand the availability of GPU resources for cloud services and Kubernetes containers;
* enhance the cluster room cooling system to support the increasing computing power.

**During the maintenance period, the following services will not be available:**

* UTHPC **cluster Rocket**;
* file systems **/gpfs/space and /gpfs/helios;**
* **OpenOndemand** services
* **Galaxy** ([galaxy.hpc.ut.ee](http://galaxy.hpc.ut.ee))

The maintenance will not affect the UTHPC cloud service (virtual machines), Kubernetes containers, or websites, **except** in cases where they **use NFS or Samba network drives** (mounted file systems). In the near future, we will contact the master users of instances connected to NFS or Samba file systems.

  
The cluster, along with the file systems, will be available again for computations starting from **January 19, 2026.**

  
If you have any questions or concerns about whether the January maintenance may affect your work, please consult us at [support@hpc.ut.ee](mailto:support@hpc.ut.ee)

Thank you for your understanding!

Best regards, 

UTHPC team.</p>
<p><small>Nov <var data-var='date'> 10</var>, <var data-var='time'>16:10:09</var> GMT+0</small><br /><strong>Identified</strong> -
  The SFTP service will also be unavailable during this maintenance period..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmgtc18x809sr9jikfe3ue3wb</id>
  <published>2025-10-23T06:00:00.000+00:00</published>
  <updated>2025-10-24T14:52:27.994+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmgtc18x809sr9jikfe3ue3wb"/>
  <title>Galaxy maintenance</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 1 day, 8 hours and 52 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Oct <var data-var='date'> 24</var>, <var data-var='time'>14:52:27</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully..</p>
<p><small>Oct <var data-var='date'> 23</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  **Seoses plaaniliste hooldustöödega on Galaxy teenuse toimimine 23\. kuni 24\. oktoobrini häiritud.** 

Palun vältige pikaajaliste analüüside käivitamist, mis kattuksid hooldusperioodiga, kuna need võivad katkeda või ebaõnnestuda.

\----------------------------------------------

**Due to scheduled maintenance, Galaxy service operation will be disrupted between October 23rd and 24th.**

Please avoid starting long-running analyses that would overlap with the maintenance window, as they may be interrupted or fail to complete..</p>
<p><small>Oct <var data-var='date'> 23</var>, <var data-var='time'>06:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmfdum9wp00ag134i8vzh4o00</id>
  <published>2025-09-10T18:00:00.000+00:00</published>
  <updated>2025-09-10T18:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmfdum9wp00ag134i8vzh4o00"/>
  <title>Open Ondemand Maintenance Wednesday, September 10th, between 21:00 - 22:00</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 1 hour</p>
    <p><strong>Affected Components:</strong> RStudio, Open OnDemand</p>
    <p><small>Sep <var data-var='date'> 10</var>, <var data-var='time'>18:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  **Date:** Wednesday, September 10, 2025  
**Time:** 21:00 - 22:00

**Impact:** Users may experience temporary unavailability of Open OnDemand services (RStudio, Jupyter).

We apologize for any inconvenience and appreciate your patience as we work to improve our services.

For questions or concerns, please contact the IT support team at support@hpc.ut.ee..</p>
<p><small>Sep <var data-var='date'> 10</var>, <var data-var='time'>18:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>
<p><small>Sep <var data-var='date'> 10</var>, <var data-var='time'>19:00:00</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmdzv7wg4000yxpdo2jo1q23y</id>
  <published>2025-08-22T06:00:00.000+00:00</published>
  <updated>2025-08-22T06:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmdzv7wg4000yxpdo2jo1q23y"/>
  <title>Galaxy update 24.0 -&gt; 25.0</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 2 days, 14 hours and 58 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Aug <var data-var='date'> 22</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Ajavahemikus 22 kuni 24 august on meil kavas UT HPC riistvaral asuva kohaliku Galaxy tarkvara versiooni uuendus, et saaksite osa lisandunud platvormi kasutajamugavuse funktsioonidest. Kuna sisuliselt on tegemist terve süsteemi värskendusega, siis mõjutab see nii sisselogimist kui ka sel perioodil aktiivselt jooksvaid töid.

  
Tööde käivitamise paneme lukku juba eelneval päeval (21\. augustil), kuid tulemusi sirvida, töövooge koostada või mugandada jms. toiminguid saate teha kuni reede 22\. augusti hommikuni. Uuendatud serverisse saate taas sisse logida ja töid käivitada 24\. augustil.

Vabandame võimalike ebamugavuste pärast. Kui teil on kõnealusel nädala plaanis jooksutada ressursimahukaid ja pikki analüüse ning kahtlete, kas ja kuidas neid täies mahus serveri aktiivsesse ajaaknasse mahutada, siis võtke meiega julgesti ühendust kirjutades aadressil [support@hpc.ut.ee](mailto:support@hpc.ut.ee) või otse vestluskanalite kaudu.

  
\----------------------------------------------  

Between August 22nd and 24th, we plan to upgrade the version of the local Galaxy server running on UT HPC hardware to enable new user convenience features of the platform. Since it is essentially an system update, it will affect both login and actively running jobs during this period.  

We will lock the launch of jobs the day before (August 21st), but you can perform tasks such as browsing results, creating or customizing workflows, etc. until the morning of Friday, August 22st.  

We apologise for any inconvenience this may cause. If you have resource-intensive and long-running analyses planned for this week and are unsure whether and how to fit them into the server&#039;s active time window, please feel free to contact us via [support@hpc.ut.ee](mailto:support@hpc.ut.ee) or direct messages..</p>
<p><small>Aug <var data-var='date'> 22</var>, <var data-var='time'>06:00:01</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress.</p>
<p><small>Aug <var data-var='date'> 24</var>, <var data-var='time'>20:57:55</var> GMT+0</small><br /><strong>Completed</strong> -
  Galaxy uuendus versioonile 25.0 on edukalt lõppenud. Uues versioonis on täiendusi saanud andmekogude (_collection_) ja töövoogude (_workflow_) haldamine, lisaks on muudatusi kasutajaliideses.

Uuenduse täieliku ülevaate jaoks tutvuge palun [ametliku Galaxy teadaandega](https://docs.galaxyproject.org/en/master/releases/25.0%5Fannounce%5Fuser.html).

Kui uuenduse järgselt tekib Teil Galaxy kasutamisel probleeme, siis võtke meiega ühendust aadressil [support@hpc.ut.ee](mailto:support@hpc.ut.ee).

Kõik, kel on huvi tutvuda võimalustega, kuidas Galaxy töid nutikalt automatiseerida, võivad samuti endast märku anda. Leiame tutvustuse jaoks ühiselt sobiva aja. 

\----------------------------------------------

Our Galaxy service has been successfully upgraded to version 25.0, which brings some user interface changes and new features to collection and workflow management.

For a complete overview of new features, please refer to the [official Galaxy Release announcement](https://docs.galaxyproject.org/en/master/releases/25.0%5Fannounce%5Fuser.html).

If you encounter any issues following this update, please contact us at [support@hpc.ut.ee](mailto:support@hpc.ut.ee).

Anyone interested in learning about smart automation possibilities for Galaxy workflows is also welcome to reach out to us, and we can schedule a suitable time for a consultation..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmcw077kd000x39bgxmsu6ecx</id>
  <published>2025-07-15T06:00:00.000+00:00</published>
  <updated>2025-07-15T06:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmcw077kd000x39bgxmsu6ecx"/>
  <title>HPC Cluster upgrade</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 2 hours and 42 minutes</p>
    <p><strong>Affected Components:</strong> rocket.hpc.ut.ee</p>
    <p><small>Jul <var data-var='date'> 15</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  HPC cluster Rocket updates are scheduled for July 2025 that will improve the cluster&#039;s performance and capabilities.

**1) Login Node Updates**

We&#039;ll be performing system updates on both login nodes this month:

* **Login1**: July 15th
* **Login2**: July 22nd

To minimize disruption, we&#039;ll close new SSH connections one week before each update, allowing existing connections to naturally expire. One of the login nodes will remain available at all times, so you won&#039;t experience any service downtime.

**2) Slurm Update**

On **July 22nd, starting at 15:00**, we&#039;ll be upgrading Slurm from version 23.02 to 23.11\. Your running jobs won&#039;t be affected, and you&#039;ll be able to submit new jobs during the update. However, commands like sacct, sacctmgr, and related tools will be unavailable during the update. The process should take about two hours but may run longer. We

After July 22nd, the compute nodes will be updated in a rolling fashion. This means some nodes will be temporarily drained until all updates are complete, which may result in longer queue times depending on cluster usage..</p>
<p><small>Jul <var data-var='date'> 22</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress..</p>
<p><small>Jul <var data-var='date'> 16</var>, <var data-var='time'>13:11:18</var> GMT+0</small><br /><strong>Identified</strong> -
  We will be directing SSH to login1 today. The login2 internal route will still stay open until the 22nd, when we will be performing maintenance and rebooting the machine..</p>
<p><small>Jul <var data-var='date'> 15</var>, <var data-var='time'>08:41:38</var> GMT+0</small><br /><strong>Identified</strong> -
  Maintenance is now in progress..</p>
<p><small>Jul <var data-var='date'> 15</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cmcfygwew0047ahpirvmpmsvl</id>
  <published>2025-06-28T07:00:00.000+00:00</published>
  <updated>2025-06-28T07:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cmcfygwew0047ahpirvmpmsvl"/>
  <title>OpenStack cloud network replacement</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 1 day, 23 hours and 16 minutes</p>
    <p><strong>Affected Components:</strong> minu.etais.ee</p>
    <p><small>Jun <var data-var='date'> 28</var>, <var data-var='time'>07:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Our cloud engineers are replacing the virtual networks connecting cloud VMs to improve the service in the future. This might cause some downtime for the VMs..</p>
<p><small>Jun <var data-var='date'> 30</var>, <var data-var='time'>07:23:35</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully..</p>
<p><small>Jun <var data-var='date'> 28</var>, <var data-var='time'>08:07:54</var> GMT+0</small><br /><strong>Identified</strong> -
  The network update has started..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cm6tf8iih001wkut7f6h6tvlf</id>
  <published>2025-02-08T07:00:00.000+00:00</published>
  <updated>2025-02-10T12:13:43.060+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cm6tf8iih001wkut7f6h6tvlf"/>
  <title>Upcoming Estonia desynchronizing from the Russian electricity grid effect on UT HPC Center</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 2 days, 5 hours and 14 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy, RStudio, rocket.hpc.ut.ee, Open OnDemand</p>
    <p><small>Feb <var data-var='date'> 10</var>, <var data-var='time'>12:13:43</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully..</p>
<p><small>Feb <var data-var='date'> 8</var>, <var data-var='time'>07:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  On the 8th of February, Estonia and other Baltic countries will disconnect from the Russian electricity system and join the continental European frequency area on the 9th of February. We have been informed that Estonian residents will probably not notice the frequency band change. However, during this switchover, Estonia’s electricity supply will be more vulnerable than usual, and technical failures can never be completely ruled out.

If necessary, the University of Tartu HPC Center is prepared for potential electrical blackouts by activating our generator and UPS. Depending on the duration of the outage, we may need to shut down some non-critical services to optimize energy consumption. The HPC cluster&#039;s compute nodes would be the first units to shut down. We will update you about the changes in our service here on the status page. 

Stay safe,

UT HPC Center.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/cm4l66vcf00kfrmiff2daqda7</id>
  <published>2024-12-12T11:00:00.000+00:00</published>
  <updated>2024-12-12T11:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/cm4l66vcf00kfrmiff2daqda7"/>
  <title>Galaxy update</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 8 days, 23 hours and 59 minutes</p>
    <p><strong>Affected Components:</strong> Galaxy</p>
    <p><small>Dec <var data-var='date'> 12</var>, <var data-var='time'>11:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  UT Galaxy server upgrade and migration to a newer hardware resource has begun..</p>
<p><small>Dec <var data-var='date'> 12</var>, <var data-var='time'>13:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  UT Galaxy server upgrade and migration to a newer hardware resource has begun..</p>
<p><small>Jan <var data-var='date'> 2</var>, <var data-var='time'>09:13:58</var> GMT+0</small><br /><strong>Identified</strong> -
  Due to the holidays, the Galaxy update is behind schedule. The data has been moved to our newest filesystem, and access to the updated Galaxy can be granted if necessary.  

If you need immediate access to your Galaxy data or would like to test tools on the newer Galaxy, please write to [support@hpc.ut.ee](mailto:support@hpc.ut.ee). The official opening will take place in the coming weeks.  

Happy New Year!.</p>
<p><small>Feb <var data-var='date'> 18</var>, <var data-var='time'>12:58:45</var> GMT+0</small><br /><strong>Completed</strong> -
  Galaxy is now accessible to all UT HPC cluster (Rocket) users. If you want to use Galaxy but haven&#039;t used our HPC cluster before, please send an access request to [**support@hpc.ut.ee**](mailto:support@hpc.ut.ee).

Thank you all for your patience, and a special thanks to our pilot users!.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/clyr1npt0198716lgof6sa8kx0h</id>
  <published>2024-08-01T08:51:00.000+00:00</published>
  <updated>2024-09-19T08:30:27.441+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/clyr1npt0198716lgof6sa8kx0h"/>
  <title>UTHPC cluster maintenance in August, 2024</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 17 days, 23 hours and 39 minutes</p>
    <p><strong>Affected Components:</strong> rocket.hpc.ut.ee</p>
    <p><small>Sep <var data-var='date'> 19</var>, <var data-var='time'>08:30:27</var> GMT+0</small><br /><strong>Completed</strong> -
  Cluster update done along with home directory migration :).</p>
<p><small>Aug <var data-var='date'> 1</var>, <var data-var='time'>08:51:00</var> GMT+0</small><br /><strong>Identified</strong> -
  **Maintenance timetable**

\- The operation system (OS) switch will start on 01.08.2024 at 17.00.

\- The home directory update will be in the period 19.08-1.09.2024.

**Summary**

There will be two major updates to the UTHPC cluster that will affect users, including You:

1\. OS update

1.1 The OS will be updated to a new version

1.2 This also means an update to most of the software modules

2\. Home directory update

2.1 Change of the home directory path

**Please find the detailed overview of the update and mitigation steps in** [**HPC Docs.**](https://docs.hpc.ut.ee/public/HPC%5Fupdate%5Fsummer24/)

Please let us know if you have any questions or need help adjusting your scripts after the maintenance. Contact: support@hpc.ut.ee.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/clxvlzlsw406795e4n9xc1wotm3</id>
  <published>2024-06-27T06:18:35.810+00:00</published>
  <updated>2024-06-27T06:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/clxvlzlsw406795e4n9xc1wotm3"/>
  <title>Kubernetes storage system upgrade</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 18 hours and 36 minutes</p>
    <p><strong>Affected Components:</strong> hpc.ut.ee, registry.hpc.ut.ee, docs.hpc.ut.ee, puhuri-portal.neic.no, puhuri.metacenter.no, account.lumi.cscs.ch, lumi.deic.dk, my.lumi-supercomputer.eu, minu.etais.ee, kubernetes.hpc.ut.ee</p>
    <p><small>Jun <var data-var='date'> 27</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Kubernetes storage system will be upgraded. If everything goes correctly, there should be no impact to workloads, but due to historic reasons, caution is required..</p>
<p><small>Jun <var data-var='date'> 27</var>, <var data-var='time'>06:18:35</var> GMT+0</small><br /><strong>Completed</strong> -
  Storage system upgrade completed, everything went according to plan. Workloads were not affected, no new changes to users introduced..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/clsa5fqmd1087200h8omuf9lm9e0</id>
  <published>2024-02-06T21:00:00.000+00:00</published>
  <updated>2024-02-06T23:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/clsa5fqmd1087200h8omuf9lm9e0"/>
  <title>Kubernetes upgrade</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 3 hours and 30 minutes</p>
    <p><strong>Affected Components:</strong> hpc.ut.ee, registry.hpc.ut.ee, docs.hpc.ut.ee, puhuri-portal.neic.no, puhuri.metacenter.no, account.lumi.cscs.ch, lumi.deic.dk, my.lumi-supercomputer.eu, minu.etais.ee</p>
    <p><small>Feb <var data-var='date'> 6</var>, <var data-var='time'>23:00:00</var> GMT+0</small><br /><strong>Completed</strong> -
  Maintenance has completed successfully.</p>
<p><small>Feb <var data-var='date'> 6</var>, <var data-var='time'>21:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  The Kubernetes cluster hosting impacted services currently has issues due to two outstanding bugs, one in the Linux kernel and one in Longhorn. We will upgrade both of these today evening on all nodes, to prevent future issues. The plan is to not impact workload availability, but as this is a big upgrade, no promises can be made..</p>
<p><small>Feb <var data-var='date'> 6</var>, <var data-var='time'>19:30:16</var> GMT+0</small><br /><strong>Identified</strong> -
  Will need to start a bit earlier..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.hpc.ut.ee,2005:Maintenance/clpgicuqt1179248egn7ppbp1786</id>
  <published>2023-11-27T06:00:00.000+00:00</published>
  <updated>2023-11-27T06:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.hpc.ut.ee/maintenance/clpgicuqt1179248egn7ppbp1786"/>
  <title>Helpdesk upgrade</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Maintenance</p>
    <p><strong>Duration:</strong> 24 minutes</p>
    <p><strong>Affected Components:</strong> support.hpc.ut.ee</p>
    <p><small>Nov <var data-var='date'> 27</var>, <var data-var='time'>06:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  In progress..</p>
<p><small>Nov <var data-var='date'> 27</var>, <var data-var='time'>06:23:49</var> GMT+0</small><br /><strong>Completed</strong> -
  Upgrade has completed successfully..</p>

        ]]>
  </content>
</entry>

</feed>