FAS Research Computing - 历史记录

HolyLFS06 (Tier 0) 目前性能下降

Status page for the Harvard FAS Research Computing cluster and other resources.

Cluster Utilization (VPN and FASRC login required): Cannon | FASSE


Please scroll down to see details on any Incidents or maintenance notices.
Monthly maintenance occurs on the first Monday of the month (except holidays).

GETTING HELP
Documentation: https://docs.rc.fas.harvard.edu | Account Portal https://portal.rc.fas.harvard.edu
Email: rchelp@rc.fas.harvard.edu | Support Hours


The colors shown in the bars below were chosen to increase visibility for color-blind visitors.
For higher contrast, switch to light mode at the bottom of this page if the background is dark and colors are muted.

运行正常

SLURM Scheduler - Cannon - 运行正常

Cannon Compute Cluster (Holyoke) - 运行正常

Boston Compute Nodes - 运行正常

GPU nodes (Holyoke) - 运行正常

seas_compute - 运行正常

运行正常

SLURM Scheduler - FASSE - 运行正常

FASSE Compute Cluster (Holyoke) - 运行正常

运行正常

Kempner Cluster CPU - 运行正常

Kempner Cluster GPU - 运行正常

运行正常

FASSE login nodes - 运行正常

运行正常

Cannon Open OnDemand - 运行正常

FASSE Open OnDemand - 运行正常

性能下降

Netscratch (Global Scratch) - 运行正常

Home Directory Storage - Boston - 运行正常

Tape - (Tier 3) - 运行正常

Holylabs - 运行正常

Isilon Storage Holyoke (Tier 1) - 运行正常

Holystore01 (Tier 0) - 运行正常

HolyLFS04 (Tier 0) - 运行正常

HolyLFS05 (Tier 0) - 运行正常

HolyLFS06 (Tier 0) - 性能下降

Holyoke Tier 2 NFS - 运行正常

Holyoke Specialty Storage - 运行正常

holECS - 运行正常

Isilon Storage Boston (Tier 1) - 运行正常

BosLFS02 (Tier 0) - 运行正常

Boston Tier 2 NFS - 运行正常

CEPH Storage Boston (Tier 2) - 运行正常

Boston Specialty Storage - 运行正常

bosECS - 运行正常

Samba Cluster - 运行正常

Globus Data Transfer - 运行正常

历史记录

9月 2024

Virtual Machine hypervisor down - Affects FASSE login/OOD
  • 已解决
    UTC
    已解决
    Resolving. The hypervisor and all but one VM, which has separate issue, are operational.
  • 更新
    UTC
    更新

    FASSE Open OnDemand and FASSE login services should be operational now.

  • 持续监控中
    UTC
    持续监控中

    FASSE OOD is back up

    FASSE login nodes are still down

  • 已确认问题
    UTC
    已确认问题

    One of the hypervisors managing virtual machines is down. We are working to bring it back up. This does affect FASSE login and FASSE OOD nodes as well as may degrade OpenAuth (two-factor).

    Affected hosts are:
    HOST -- STATUS

    dataverse-backup UNKNOWN

    demo2-l3-fs UNKNOWN

    enos-vote-l3-fs UNKNOWN

    fasselogin01 UNKNOWN

    fasselogin02 UNKNOWN

    frontier-squid02 UNKNOWN

    frontier-squid03 UNKNOWN

    frontier-squid04 UNKNOWN

    goel-adm24-l3-fs UNKNOWN

    goel-blind-l3-fs UNKNOWN

    goel-l3-fs UNKNOWN

    h-dev-fasseooda-01 UNKNOWN

    h-dev-fasseooda-lb01 UNKNOWN

    h-dev-fasseoodb-lb11 UNKNOWN

    h-fasseooda-01 UNKNOWN

    h-fasseooda-lb02 UNKNOWN

    h-fasseoodb-lb11 UNKNOWN

    h-fasseoodb-lb12 UNKNOWN

    h-fasseoodc-lb21 UNKNOWN

    h-fasseoodc-lb22 UNKNOWN

    h-qa-fasseooda-01 UNKNOWN

    h-qa-fasseooda-lb02 UNKNOWN

    holy-es-master01 UNKNOWN

    holy-es-master02 UNKNOWN

    holy-es-master03 UNKNOWN

    holynagios UNKNOWN

    kreindlerl3-fs UNKNOWN

    martin-su-l3-fs UNKNOWN

    mcconnell-l3-fs UNKNOWN

    openauth02 jtriley UNKNOWN

    shleifer-dsl3-fs UNKNOWN

    stock-solar-l3-fs UNKNOWN

    stopsack-l3-fs UNKNOWN

    xcat UNKNOWN

8月 2024

Starfish upgrade
  • 已完成
    八月 27, 2024 在 下午 12:00UTC
    已完成
    八月 27, 2024 在 下午 12:00UTC

    Starfish is back up

  • 更新
    八月 26, 2024 在 下午 2:35UTC
    更新
    八月 26, 2024 在 下午 2:35UTC

    Starfish maintenance is still ongoing, no ETA at this time.

  • 进行中
    八月 24, 2024 在 上午 12:00UTC
    进行中
    八月 24, 2024 在 上午 12:00UTC
    Maintenance is now in progress
  • 已计划
    八月 24, 2024 在 上午 12:00UTC
    已计划
    八月 24, 2024 在 上午 12:00UTC

    The Starfish Zones Dashboard will be undergoing a few upgrades and maintenance this weekend from Friday, August 23rd at 8AM until Monday, August 26th at 8AM. The dashboard will not be accessible during this time. Further details will be provided, if needed. Please email rchelp@rc.fas.harvard.edu if you have any questions or concerns.

7月 2024

Authentication issues - Related to global Crowdstrike incident
  • 已解决
    UTC
    已解决

    All Crowdstrike-related resources are back up and operational.

  • 更新
    UTC
    更新
    For FASRC resources affected by the Crowdstrike issue, most are back in full services. A few remaining issues involving the following may not be resolved until Monda: - waywiser2 - proteomics2 - tmsdb3 - lic3
  • 更新
    UTC
    更新

    Please see HUIT Status (harvard.edu) for additional information on the global issue caused by Crowdstrike security which Harvard relies on. This is an ongoing issue university-wide.

    The systems that continue to be affected at FASRC are minimal, but some Windows-based systems managed by or connected to FASRC may still be affected.

  • 持续监控中
    UTC
    持续监控中

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

  • 已确认问题
    UTC
    已确认问题

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

  • 调查中
    UTC
    调查中

    Authentication is back up and running. Windows machines are still in a bad state and will need remedial work to get them back in service.

7月 2024 9月 2024

下一页