FAS Research Computing - FASRC monthly maintenance will take place on September 14th, 2026. – Détails de la maintenance

Holyoke Tier 2 NFS connaît une panne partielle

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

SEPT14

FASRC monthly maintenance will take place on September 14th, 2026.

Terminéseptembre 14, 2026 à 13:00 – 17:00UTC

Affecte

Cannon Cluster

En maintenance depuis 1:00 PM à 5:00 PM

SLURM Scheduler - Cannon

En maintenance depuis 1:00 PM à 5:00 PM

Cannon Compute Cluster (Holyoke)

En maintenance depuis 1:00 PM à 5:00 PM

Boston Compute Nodes

En maintenance depuis 1:00 PM à 5:00 PM

GPU nodes (Holyoke)

En maintenance depuis 1:00 PM à 5:00 PM

seas_compute

En maintenance depuis 1:00 PM à 5:00 PM

Mises à jour
  • Terminé
    septembre 14, 2026 à 17:00UTC
    Terminé
    septembre 14, 2026 à 17:00UTC
    La maintenance s'est terminée avec succès
  • Mettre à jour
    septembre 14, 2026 à 13:08UTC
    Mettre à jour
    septembre 14, 2026 à 13:08UTC

    The Holyoke/MGHPCC row 7c power work has been CANCELLED by the facility and will be re-scheduled.

    The following work WILL NOT take pleace today:

    • Cancelled: MGHPCC will be upgrading power on Pod 7c Even Side on September 14th 7am-5pm. This necessitates idling half the nodes on that side of the pod. A blocking reservation has been put in place to accomplish this. No jobs will be canceled but users will notice degraded scheduling throughput due to half the nodes being closed in the following partitions: 

      arguelles_delgado, blackhole, conroy, davies, desai, doshi-velez, dsouza, eddy, edwards, geophysics, giribet, gpu_test, hernquist, huce_cascade, huttenhower, imasc, jacobsen2, janson_cascade, janson, ke, lukin, murphy, nguyen, ni_lab, olveczky, ortegahernandez, pehlevan, seas_compute, shared, shakhnovich, tambe, unrestricted, vishwanath, whipple, xlin, yin, zon

  • En cours
    septembre 14, 2026 à 13:00UTC
    En cours
    septembre 14, 2026 à 13:00UTC
    La maintenance est en cours
  • Planifiée
    août 31, 2026 à 18:56UTC
    Planifiée
    août 31, 2026 à 18:56UTC

    Our maintenance tasks should be completed between 9am-1pm.
    Some power work on row 7c will run 9-5 but will not affect running jobs or new jobs (see below).

    NOTICES:

    MAINTENANCE TASKS

    Cannon cluster will be paused during this maintenance?: YES
    FASSE cluster will be paused during this maintenance?: YES

    • Slurm Upgrade to 26.05.4

      • Audience: Cluster

      • Impact: The cluster will be paused during this maintenance

    • New: Enable Termination of User Processes on Logout

      • Audience: Cluster

      • Impact: Going forward all user processes will be terminated upon login session exit on the login nodes, excluding things running in screen and tmux. Users should leverage the cluster for non-interactive processes. This is to clean up after AI agents which tend to create orphaned processes which drag down login node performance.

    • New: watch command cadence limit

      • Audience: Cluster

      • Impact:The watch command will have a minimum cadence of 60s when used on commands talking to the slurm scheduler (i.e. squeue, showq, sinfo, scontrol, sdiag, lsload). Users desiring faster polling should leverage the sacct command that talks to the slurm database. In general users should not poll the scheduler more than once every minute, ideally once every 5-10 minutes. Polling more often slows the scheduler. FASRC reserves the right to ban users who who tax the scheduler with queries. For more on cluster customs and responsibilities see: https://docs.rc.fas.harvard.edu/kb/responsibilities/

    • Holyoke/MGHPCC row 7c power work - 9am-5pm (CANCELLED)

      • Audience: Cluster

      • Impact: MGHPCC will be upgrading power on Pod 7c Even Side on September 14th 7am-5pm. This necessitates idling half the nodes on that side of the pod. A blocking reservation has been put in place to accomplish this. No jobs will be canceled but users will notice degraded scheduling throughput due to half the nodes being closed in the following partitions: 

        arguelles_delgado, blackhole, conroy, davies, desai, doshi-velez, dsouza, eddy, edwards, geophysics, giribet, gpu_test, hernquist, huce_cascade, huttenhower, imasc, jacobsen2, janson_cascade, janson, ke, lukin, murphy, nguyen, ni_lab, olveczky, ortegahernandez, pehlevan, seas_compute, shared, shakhnovich, tambe, unrestricted, vishwanath, whipple, xlin, yin, zon

    • Login node reboots 

      • Audience: All login nodes

      • Impact: Login nodes will be unavailable until after maintenance

    • OOD/Open OnDemand down/reboots

      • Audience: All OOD users

      • Impact: OOD will be unavailable until after maintenance

    • Gurobi license key update

      • Audience: Anyone who uses Gurobi software on the clusters.

      • Impact: Jobs running Gurobi may fail. FASRC recommends waiting until maintenance is over to run Gurobi jobs.

    • Netscratch 90-day retention cleanup

      • Audience; All netscratch users

      • Impact: Files older than 90 days will be removed per our scratch policy. Please note that this cleanup can happen at any time, not just during maintenance.

    Thank you,
    FAS Research Computing
    https://docs.rc.fas.harvard.edu/
    https://www.rc.fas.harvard.edu/