FAS Research Computing - Histórico de avisos

Experimentando Desempenho Parcialmente Degradado

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

Desempenho degradado

SLURM Scheduler - Cannon - Desempenho degradado

Cannon Compute Cluster (Holyoke) - Desempenho degradado

Boston Compute Nodes - Desempenho degradado

GPU nodes (Holyoke) - Desempenho degradado

seas_compute - Desempenho degradado

Operacional

SLURM Scheduler - FASSE - Operacional

FASSE Compute Cluster (Holyoke) - Operacional

Operacional

Kempner Cluster CPU - Operacional

Kempner Cluster GPU - Operacional

Operacional

FASSE login nodes - Operacional

Operacional

Cannon Open OnDemand - Operacional

FASSE Open OnDemand - Operacional

Operacional

Netscratch (Global Scratch) - Operacional

Home Directory Storage - Boston - Operacional

Tape - (Tier 3) - Operacional

Holylabs - Operacional

Isilon Storage Holyoke (Tier 1) - Operacional

Holystore01 (Tier 0) - Operacional

HolyLFS04 (Tier 0) - Operacional

HolyLFS05 (Tier 0) - Operacional

HolyLFS06 (Tier 0) - Operacional

Holyoke Tier 2 NFS (new) - Operacional

Holyoke Specialty Storage - Operacional

holECS - Operacional

Isilon Storage Boston (Tier 1) - Operacional

BosLFS02 (Tier 0) - Operacional

Boston Tier 2 NFS (new) - Operacional

CEPH Storage Boston (Tier 2) - Operacional

Boston Specialty Storage - Operacional

bosECS - Operacional

Samba Cluster - Operacional

Globus Data Transfer - Operacional

Histórico de avisos

set 2024

Virtual Machine hypervisor down - Affects FASSE login/OOD
  • Resolvido
    Resolvido
    Resolving. The hypervisor and all but one VM, which has separate issue, are operational.
  • Atualizar
    Atualizar

    FASSE Open OnDemand and FASSE login services should be operational now.

  • Monitorização
    Monitorização

    FASSE OOD is back up

    FASSE login nodes are still down

  • Identificado
    Identificado

    One of the hypervisors managing virtual machines is down. We are working to bring it back up. This does affect FASSE login and FASSE OOD nodes as well as may degrade OpenAuth (two-factor).

    Affected hosts are:
    HOST -- STATUS

    dataverse-backup UNKNOWN

    demo2-l3-fs UNKNOWN

    enos-vote-l3-fs UNKNOWN

    fasselogin01 UNKNOWN

    fasselogin02 UNKNOWN

    frontier-squid02 UNKNOWN

    frontier-squid03 UNKNOWN

    frontier-squid04 UNKNOWN

    goel-adm24-l3-fs UNKNOWN

    goel-blind-l3-fs UNKNOWN

    goel-l3-fs UNKNOWN

    h-dev-fasseooda-01 UNKNOWN

    h-dev-fasseooda-lb01 UNKNOWN

    h-dev-fasseoodb-lb11 UNKNOWN

    h-fasseooda-01 UNKNOWN

    h-fasseooda-lb02 UNKNOWN

    h-fasseoodb-lb11 UNKNOWN

    h-fasseoodb-lb12 UNKNOWN

    h-fasseoodc-lb21 UNKNOWN

    h-fasseoodc-lb22 UNKNOWN

    h-qa-fasseooda-01 UNKNOWN

    h-qa-fasseooda-lb02 UNKNOWN

    holy-es-master01 UNKNOWN

    holy-es-master02 UNKNOWN

    holy-es-master03 UNKNOWN

    holynagios UNKNOWN

    kreindlerl3-fs UNKNOWN

    martin-su-l3-fs UNKNOWN

    mcconnell-l3-fs UNKNOWN

    openauth02 jtriley UNKNOWN

    shleifer-dsl3-fs UNKNOWN

    stock-solar-l3-fs UNKNOWN

    stopsack-l3-fs UNKNOWN

    xcat UNKNOWN

ago 2024

Starfish upgrade
  • Concluído
    agosto 27, 2024 em 12:00
    Concluído
    agosto 27, 2024 em 12:00

    Starfish is back up

  • Atualizar
    agosto 26, 2024 em 14:35
    Atualizar
    agosto 26, 2024 em 14:35

    Starfish maintenance is still ongoing, no ETA at this time.

  • Em curso
    agosto 24, 2024 em 00:00
    Em curso
    agosto 24, 2024 em 00:00
    Maintenance is now in progress
  • Ainda não começou
    agosto 24, 2024 em 00:00
    Ainda não começou
    agosto 24, 2024 em 00:00

    The Starfish Zones Dashboard will be undergoing a few upgrades and maintenance this weekend from Friday, August 23rd at 8AM until Monday, August 26th at 8AM. The dashboard will not be accessible during this time. Further details will be provided, if needed. Please email rchelp@rc.fas.harvard.edu if you have any questions or concerns.

jul 2024

Authentication issues - Related to global Crowdstrike incident
  • Resolvido
    Resolvido

    All Crowdstrike-related resources are back up and operational.

  • Atualizar
    Atualizar
    For FASRC resources affected by the Crowdstrike issue, most are back in full services. A few remaining issues involving the following may not be resolved until Monda: - waywiser2 - proteomics2 - tmsdb3 - lic3
  • Atualizar
    Atualizar

    Please see HUIT Status (harvard.edu) for additional information on the global issue caused by Crowdstrike security which Harvard relies on. This is an ongoing issue university-wide.

    The systems that continue to be affected at FASRC are minimal, but some Windows-based systems managed by or connected to FASRC may still be affected.

  • Monitorização
    Monitorização

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

  • Identificado
    Identificado

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

  • Investigando
    Investigando

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

jul 2024 para set 2024

Seguinte