FAS Research Computing - Monthly Maintenance and MGHPCC Power Work - Dec. 8, 2025 6am-6pm – تفاصيل الصيانة

النظام تحت الصيانة

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

Monthly Maintenance and MGHPCC Power Work - Dec. 8, 2025 6am-6pm

مكتمل
المقرر ل ديسمبر 08, 2025 في 11:00 – 23:00UTC

يؤثر

Cannon Cluster

صيانة من 11:00 AM ألى 11:00 PM

SLURM Scheduler - Cannon

صيانة من 11:00 AM ألى 11:00 PM

Cannon Compute Cluster (Holyoke)

صيانة من 11:00 AM ألى 11:00 PM

Boston Compute Nodes

صيانة من 11:00 AM ألى 11:00 PM

GPU nodes (Holyoke)

صيانة من 11:00 AM ألى 11:00 PM

seas_compute

صيانة من 11:00 AM ألى 11:00 PM

التحديثات
  • مكتمل
    ديسمبر 08, 2025 في 23:00UTC
    مكتمل
    ديسمبر 08, 2025 في 23:00UTC
    Maintenance has completed successfully
  • قيد التقدم
    ديسمبر 08, 2025 في 11:00UTC
    قيد التقدم
    ديسمبر 08, 2025 في 11:00UTC
    Maintenance is now in progress
  • مخطط
    ديسمبر 08, 2025 في 11:00UTC
    مخطط
    ديسمبر 08, 2025 في 11:00UTC

    Monthly maintenance will take place on December 8th. Our maintenance tasks should be completed between 9am-1pm. However: 

    Additionally, MGHPCC will be performing power upgrades on the odd side of Row 8A where much of our computer resides. This is the final upgrade for this row. Current estimate for this work is a 12 hour window 6am-6pm.

    A list of the affected partitions is provided at the bottom of this notice. The nodes in those partitions will be drained prior to the work and will be powered down. Once the work is completed, those nodes will be returned to service. 

    Notices:

    • New FASSE partition fasse_gpu_h200. This partitions has 2 H200 nodes and a 3day limit. It is available now.

    • 11/26 - 11/28 are university holidays (Thanksgiving). No on-site support, FASRC staff will return on 12/1.

    • Training: Upcoming training from FASRC and other sources can be found on our Training Calendar. at https://www.rc.fas.harvard.edu/upcoming-training/

    • Status Page: You can subscribe to our status to receive notifications of maintenance, incidents, and their resolution at https://status.rc.fas.harvard.edu/ (click Get Updates for options).

    • We'd love to hear success stories about your or your lab's use of FASRC. Submit your story here.

    MAINTENANCE TASKS

    Cannon cluster will be paused during this maintenance?: PARTIAL OUTAGE/YES
    FASSE cluster will be paused during this maintenance?: PARTIAL OUTAGE/YES

    • Power work on Row 8A odd

      • Audience: Users of the partitions listed below

      • Impact: These nodes and partitions will be fully or partially down all day

    • OneFS (Isilon) upgrade

      • Audience: All Isilon (Tier 1) shares

      • Impact: Some VMs will be impacted including Cannon OOD, CBScentral, MCZapps/MCZbase, Portal, and Rclic1 (license server)

    • Slurm upgrade to 25.05.5

      • Audience: All cluster users

      • Impact: Jobs will be paused during maintenance

    • Login node reboots

      • Audience: All login node users

      • Impact: Login nodes will reboot during the maintenance window

    Impacted Cannon Partitions (Full or Partial Outage):

    • arguelles_delgado_gpu_a100

    • arguelles_delgado_gpu_mixed

    • bigmem_intermediate

    • blackhole_gpu

    • eddy

    • gershman

    • gpu_requeue

    • hejazi

    • hernquist_ice

    • hoekstra

    • huce_ice

    • iaifi_gpu

    • iaifi_gpu_priority

    • iaifi_gpu_requeue

    • itc_gpu

    • jshapiro

    • kempner

    • kempner_dev

    • kempner_priority

    • kempner_h100

    • kempner_h100_priority

    • kempner_h100_priority2

    • kempner_h100_priority3

    • kempner_interactive

    • kempner_requeue

    • kovac

    • kozinsky

    • kozinsky_gpu

    • kozinsky_priority

    • kozinsky_requeue

    • murphy_ice

    • ortegahernandez_ice

    • rivas

    • seas_compute

    • seas_gpu

    • serial_requeue

    • siag_combo

    • siag_gpu

    • sur

    • zhuang