FAS Research Computing - Monthly Maintenance and MGHPCC Power Work - Dec. 8, 2025 6am-6pm – Detalji popravke

Sistem se popravlja

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

Monthly Maintenance and MGHPCC Power Work - Dec. 8, 2025 6am-6pm

Završeno
Zakazano za December 08, 2025 u 11:00 AM – 11:00 PMUTC

Utiče na

Cannon Cluster

Popravka u toku undefined 11:00 AM do 11:00 PM

SLURM Scheduler - Cannon

Popravka u toku undefined 11:00 AM do 11:00 PM

Cannon Compute Cluster (Holyoke)

Popravka u toku undefined 11:00 AM do 11:00 PM

Boston Compute Nodes

Popravka u toku undefined 11:00 AM do 11:00 PM

GPU nodes (Holyoke)

Popravka u toku undefined 11:00 AM do 11:00 PM

seas_compute

Popravka u toku undefined 11:00 AM do 11:00 PM

Ажурирања
  • Završeno
    December 08, 2025 u 11:00 PMUTC
    Završeno
    December 08, 2025 u 11:00 PMUTC
    Maintenance has completed successfully
  • U toku
    December 08, 2025 u 11:00 AMUTC
    U toku
    December 08, 2025 u 11:00 AMUTC
    Maintenance is now in progress
  • Planirano
    December 08, 2025 u 11:00 AMUTC
    Planirano
    December 08, 2025 u 11:00 AMUTC

    Monthly maintenance will take place on December 8th. Our maintenance tasks should be completed between 9am-1pm. However: 

    Additionally, MGHPCC will be performing power upgrades on the odd side of Row 8A where much of our computer resides. This is the final upgrade for this row. Current estimate for this work is a 12 hour window 6am-6pm.

    A list of the affected partitions is provided at the bottom of this notice. The nodes in those partitions will be drained prior to the work and will be powered down. Once the work is completed, those nodes will be returned to service. 

    Notices:

    • New FASSE partition fasse_gpu_h200. This partitions has 2 H200 nodes and a 3day limit. It is available now.

    • 11/26 - 11/28 are university holidays (Thanksgiving). No on-site support, FASRC staff will return on 12/1.

    • Training: Upcoming training from FASRC and other sources can be found on our Training Calendar. at https://www.rc.fas.harvard.edu/upcoming-training/

    • Status Page: You can subscribe to our status to receive notifications of maintenance, incidents, and their resolution at https://status.rc.fas.harvard.edu/ (click Get Updates for options).

    • We'd love to hear success stories about your or your lab's use of FASRC. Submit your story here.

    MAINTENANCE TASKS

    Cannon cluster will be paused during this maintenance?: PARTIAL OUTAGE/YES
    FASSE cluster will be paused during this maintenance?: PARTIAL OUTAGE/YES

    • Power work on Row 8A odd

      • Audience: Users of the partitions listed below

      • Impact: These nodes and partitions will be fully or partially down all day

    • OneFS (Isilon) upgrade

      • Audience: All Isilon (Tier 1) shares

      • Impact: Some VMs will be impacted including Cannon OOD, CBScentral, MCZapps/MCZbase, Portal, and Rclic1 (license server)

    • Slurm upgrade to 25.05.5

      • Audience: All cluster users

      • Impact: Jobs will be paused during maintenance

    • Login node reboots

      • Audience: All login node users

      • Impact: Login nodes will reboot during the maintenance window

    Impacted Cannon Partitions (Full or Partial Outage):

    • arguelles_delgado_gpu_a100

    • arguelles_delgado_gpu_mixed

    • bigmem_intermediate

    • blackhole_gpu

    • eddy

    • gershman

    • gpu_requeue

    • hejazi

    • hernquist_ice

    • hoekstra

    • huce_ice

    • iaifi_gpu

    • iaifi_gpu_priority

    • iaifi_gpu_requeue

    • itc_gpu

    • jshapiro

    • kempner

    • kempner_dev

    • kempner_priority

    • kempner_h100

    • kempner_h100_priority

    • kempner_h100_priority2

    • kempner_h100_priority3

    • kempner_interactive

    • kempner_requeue

    • kovac

    • kozinsky

    • kozinsky_gpu

    • kozinsky_priority

    • kozinsky_requeue

    • murphy_ice

    • ortegahernandez_ice

    • rivas

    • seas_compute

    • seas_gpu

    • serial_requeue

    • siag_combo

    • siag_gpu

    • sur

    • zhuang