FAS Research Computing - Istoricul notificărilor

Holyoke Tier 2 NFS: întrerupere parțială

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

Operațional

SLURM Scheduler - Cannon - Operațional

Cannon Compute Cluster (Holyoke) - Operațional

Boston Compute Nodes - Operațional

GPU nodes (Holyoke) - Operațional

seas_compute - Operațional

Operațional

SLURM Scheduler - FASSE - Operațional

FASSE Compute Cluster (Holyoke) - Operațional

Operațional

Kempner Cluster CPU - Operațional

Kempner Cluster GPU - Operațional

Operațional

FASSE login nodes - Operațional

Operațional

Cannon Open OnDemand - Operațional

FASSE Open OnDemand - Operațional

Întrerupere parțială

Netscratch (Global Scratch) - Operațional

Home Directory Storage - Boston - Operațional

Tape - (Tier 3) - Operațional

Holylabs - Operațional

Isilon Storage Holyoke (Tier 1) - Operațional

Holystore01 (Tier 0) - Operațional

HolyLFS04 (Tier 0) - Operațional

HolyLFS05 (Tier 0) - Operațional

HolyLFS06 (Tier 0) - Operațional

Holyoke Tier 2 NFS - Întrerupere parțială

Holyoke Specialty Storage - Operațional

holECS - Operațional

Isilon Storage Boston (Tier 1) - Operațional

BosLFS02 (Tier 0) - Operațional

Boston Tier 2 NFS - Operațional

CEPH Storage Boston (Tier 2) - Operațional

Boston Specialty Storage - Operațional

bosECS - Operațional

Samba Cluster - Operațional

Globus Data Transfer - Operațional

Istoricul notificărilor

Vezi starea actuală

ian 2024

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută1 oră 37 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută47 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută5 ore 24 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

dec 2023

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută47 ore 1 minut
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută6 ore 20 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută51 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

noi 2023

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută7 ore 16 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută2 ore 33 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Caeph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Partial Cannon outage
RezolvatÎntrerupere parțială4 ore 20 minute
  • Rezolvat
    UTC
    Rezolvat

    Cooling and power have been restored to the affected racks. The compute nodes have been resumed in Slurm and are now accepting jobs again.

    This incident has been resolved.

  • Identificat
    UTC
    Identificat

    We have identified the partitions that are impacted due to loss of cooling in holy7c[02-12] compute nodes. Some of these partitions are fully down, others are partially down.

    blackhole
    blackholepriority davies desai hucecascade
    hucecascadepriority
    huttenhower
    janson
    jansoncascade joonholee lukin seascompute
    shared
    tambe
    test
    vishwanath
    whipple

    Please submit to other partitions in order to run jobs.

    The spart command will show you all partitions you have access to, and our Running Jobs page provides a list of publicly available partitions for all cluster users. Please see our docs page for other helpful Slurm commands.

    The Holyoke MGHPCC data center is working to restore cooling, and FASRC staff are onsite to assist. No ETA.

  • În curs de investigare
    UTC
    În curs de investigare

    Compute nodes in holy7c[02-12] have experienced a power loss and are currently down. GPUs are not impacted at this time.

    The 'shared' partition is significantly impacted. Other public partitions and lab-owned partitions may be down or running at reduced capacity. Jobs are still being accepted/running, but may need to wait longer in the queue due to fewer resources being available.

    We are in contact with the Holyoke MGHPCC data center to investigate further. Updates to come. No ETA at this time.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută6 ore 55 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Caeph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Ceph instability - Affects Boston VMs (Virtual Machines) and Tier2 Ceph shares
RezolvatPerformanță scăzută5 ore 52 minute
  • Rezolvat
    UTC
    Rezolvat

    The Ceph instability has been resolved. Caeph Tier2 shares, VDI, and VMs should be back to their normal state.

    If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu

  • Identificat
    UTC
    Identificat

    The infrastructure behind Tier2 Ceph shares and VMs is unstable.
    This also affects VDI/OOD which relies on virtual machines.

    /net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

    Thanks for your patience.

Anterior

noi 2023 către ian 2024

Următorul