அறிவிப்பு வரலாறு

செயல்திறன் குறைந்துள்ளது

இயங்குகிறது

மார். 2024

தீர்க்கப்பட்டது
மார்ச் 21, 2024 இல் பிற்பகல் 12:00
தீர்க்கப்பட்டது
மார்ச் 21, 2024 இல் பிற்பகல் 12:00
There was a brief disruption in the connectivity from MGHPCC to the outside world/Internet. During this period outside attempts to connect to services hosted in Holyoke would time out. Internal data center network remained normal and no jobs were affected.

தீர்க்கப்பட்டது
மார்ச் 19, 2024 இல் பிற்பகல் 1:35
தீர்க்கப்பட்டது
மார்ச் 19, 2024 இல் பிற்பகல் 1:35
After making some hardware and network changes, we monitored holylabs yesterday and overnight and have determined that it is stable again.
We will continue to monitor but currently believe this issue has been resolved.
விசாரிக்கப்படுகிறது
மார்ச் 18, 2024 இல் பிற்பகல் 1:51
விசாரிக்கப்படுகிறது
மார்ச் 18, 2024 இல் பிற்பகல் 1:51
Holylabs was again unstable over the weekend. We are actively working this issue and will update as soon as possible.
புதுப்பிப்பு
மார்ச் 15, 2024 இல் பிற்பகல் 1:31
புதுப்பிப்பு
மார்ச் 15, 2024 இல் பிற்பகல் 1:31
Holylabs locked up again overnight. We are investigating.
கண்காணிக்கப்படுகிறது
மார்ச் 14, 2024 இல் பிற்பகல் 1:38
கண்காணிக்கப்படுகிறது
மார்ச் 14, 2024 இல் பிற்பகல் 1:38
Holylabs became stuck overnight and was restarted. It is back up and we are monitoring.

தீர்க்கப்பட்டது
மார்ச் 18, 2024 இல் பிற்பகல் 3:21
தீர்க்கப்பட்டது
மார்ச் 18, 2024 இல் பிற்பகல் 3:21
holyscratch01 is currently stable. However, it continues to be under heavy utilization as a matter of course due to job loads.
There is a plan underway to replace holyscratch01, but no ETA at this time as we are still qualifying and planning purchasing. We will notify the community when an ETA is known. Until then please be aware that load on scratch may continue along this trend. Thanks for your understanding.
கண்காணிக்கப்படுகிறது
மார்ச் 13, 2024 இல் பிற்பகல் 2:15
கண்காணிக்கப்படுகிறது
மார்ச் 13, 2024 இல் பிற்பகல் 2:15
Performance has improved, but we are still monitoring some high loads across object storage units.
விசாரிக்கப்படுகிறது
மார்ச் 12, 2024 இல் பிற்பகல் 7:54
விசாரிக்கப்படுகிறது
மார்ச் 12, 2024 இல் பிற்பகல் 7:54
holyscratch01 performance degraded
We are currently investigating this incident.

தீர்க்கப்பட்டது
மார்ச் 04, 2024 இல் பிற்பகல் 5:55
தீர்க்கப்பட்டது
மார்ச் 04, 2024 இல் பிற்பகல் 5:55
VMs are returning to operation. If you receive an error (on Portal, Coldfront, etc) please wait a few minutes and try again.
Ticket system is online and starting to receive delayed emails.
விசாரிக்கப்படுகிறது
மார்ச் 04, 2024 இல் பிற்பகல் 4:41
விசாரிக்கப்படுகிறது
மார்ச் 04, 2024 இல் பிற்பகல் 4:41
The RT ticket system is currently offline due to a VM issue.
Any emails sent to the ticket system will eventually be delivered once it recovers, but until then expect delayed response.

தீர்க்கப்பட்டது
மார்ச் 04, 2024 இல் பிற்பகல் 9:44
தீர்க்கப்பட்டது
மார்ச் 04, 2024 இல் பிற்பகல் 9:44
This incident has been resolved.
கண்காணிக்கப்படுகிறது
மார்ச் 04, 2024 இல் பிற்பகல் 4:21
கண்காணிக்கப்படுகிறது
மார்ச் 04, 2024 இல் பிற்பகல் 4:21
Informational Notice
The Slurm upgrade to 23.11.4 was completed successfully during maintenance. However a complication with the automation of Slurm's cryptographic keys occurred during the upgrade which caused nodes to lose the ability to talk to the Slurm master. The Slurm master therefore viewed those nodes as down and requeued their jobs.
All jobs on Cannon and FASSE were requeued.
This is deeply regrettable but the chain of events which caused this could not be foreseen.
To check the status of your jobs, see the common Slurm commands at:
https://docs.rc.fas.harvard.edu/kb/convenient-slurm-commands/#Information_on_jobs
FAS Research Computing
https://docs.rc.fas.harvard.edu/
rchelp@rc.fas.harvard.edu

பிப். 2024

தீர்க்கப்பட்டது
பிப்ரவரி 28, 2024 இல் பிற்பகல் 7:59
தீர்க்கப்பட்டது
பிப்ரவரி 28, 2024 இல் பிற்பகல் 7:59
The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu
அடையாளம் காணப்பட்டது
பிப்ரவரி 28, 2024 இல் பிற்பகல் 6:24
அடையாளம் காணப்பட்டது
பிப்ரவரி 28, 2024 இல் பிற்பகல் 6:24
The infrastructure behind Tier2 Ceph shares and VMs is unstable.
This also affects VDI/OOD which relies on virtual machines.

/net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

Thanks for your patience.

தீர்க்கப்பட்டது
பிப்ரவரி 27, 2024 இல் பிற்பகல் 6:31
தீர்க்கப்பட்டது
பிப்ரவரி 27, 2024 இல் பிற்பகல் 6:31
This incident has been resolved.
விசாரிக்கப்படுகிறது
பிப்ரவரி 26, 2024 இல் பிற்பகல் 2:00
விசாரிக்கப்படுகிறது
பிப்ரவரி 26, 2024 இல் பிற்பகல் 2:00
holyscratch01 is experiencing high load. Performance may be delayed.
OOD is also affected as a result of this.
We are currently investigating this incident.

தீர்க்கப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 8:56
தீர்க்கப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 8:56
Clusters are back in production status. We will continue to monitor for any aberrant behavior, but this incident has been resolved.
அடையாளம் காணப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 8:07
அடையாளம் காணப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 8:07
I spoke too soon, partial recovery, I'll update here when we are sure everything is back in production, apologies.
கண்காணிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:48
கண்காணிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:48
All cannon nodes are back in service and slurm is resuming jobs, fasse is coming back up as well, we will monitor the situation, but anticipate a return to full normal operations shortly.
புதுப்பிப்பு
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:35
புதுப்பிப்பு
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:35
The nodes are coming back into normal service, we anticipate this to be fairly quick
அடையாளம் காணப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:20
அடையாளம் காணப்பட்டது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:20
slurm is operational, jobs are idled and should resume as normal
விசாரிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:18
விசாரிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 7:18
We are currently investigating this incident. We will update here as we can

தீர்க்கப்பட்டது
பிப்ரவரி 23, 2024 இல் பிற்பகல் 3:10
தீர்க்கப்பட்டது
பிப்ரவரி 23, 2024 இல் பிற்பகல் 3:10
This incident has been resolved.
விசாரிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 6:55
விசாரிக்கப்படுகிறது
பிப்ரவரி 08, 2024 இல் பிற்பகல் 6:55
Boslfs02 (/n/boslfs02) is currently experiencing instability.
We are working on the issue and will update this incident as the situation progresses.

தீர்க்கப்பட்டது
பிப்ரவரி 05, 2024 இல் பிற்பகல் 4:37
தீர்க்கப்பட்டது
பிப்ரவரி 05, 2024 இல் பிற்பகல் 4:37
This incident has been resolved.
அடையாளம் காணப்பட்டது
பிப்ரவரி 05, 2024 இல் பிற்பகல் 4:33
அடையாளம் காணப்பட்டது
பிப்ரவரி 05, 2024 இல் பிற்பகல் 4:33
We are working on a fix for this incident.

ஜன. 2024

தீர்க்கப்பட்டது
ஜனவரி 31, 2024 இல் பிற்பகல் 8:53
தீர்க்கப்பட்டது
ஜனவரி 31, 2024 இல் பிற்பகல் 8:53
fasselogin machines are back in service
அடையாளம் காணப்பட்டது
ஜனவரி 31, 2024 இல் பிற்பகல் 7:26
அடையாளம் காணப்பட்டது
ஜனவரி 31, 2024 இல் பிற்பகல் 7:26
We have identified the underlying issue and are working on fixing the FASSE login nodes. In the meantime, please use FASSE VDI. Updates to come.
விசாரிக்கப்படுகிறது
ஜனவரி 31, 2024 இல் பிற்பகல் 4:00
விசாரிக்கப்படுகிறது
ஜனவரி 31, 2024 இல் பிற்பகல் 4:00
fasselogin01/02 are currently down. We are investigating this incident.

தீர்க்கப்பட்டது
ஜனவரி 18, 2024 இல் முற்பகல் 2:36
தீர்க்கப்பட்டது
ஜனவரி 18, 2024 இல் முற்பகல் 2:36
The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu
அடையாளம் காணப்பட்டது
ஜனவரி 18, 2024 இல் முற்பகல் 12:58
அடையாளம் காணப்பட்டது
ஜனவரி 18, 2024 இல் முற்பகல் 12:58
The infrastructure behind Tier2 Ceph shares and VMs is unstable.
This also affects VDI/OOD which relies on virtual machines.

/net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

Thanks for your patience.

தீர்க்கப்பட்டது
ஜனவரி 17, 2024 இல் பிற்பகல் 11:31
தீர்க்கப்பட்டது
ஜனவரி 17, 2024 இல் பிற்பகல் 11:31
The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu
அடையாளம் காணப்பட்டது
ஜனவரி 17, 2024 இல் பிற்பகல் 10:44
அடையாளம் காணப்பட்டது
ஜனவரி 17, 2024 இல் பிற்பகல் 10:44
The infrastructure behind Tier2 Ceph shares and VMs is unstable.
This also affects VDI/OOD which relies on virtual machines.

/net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

Thanks for your patience.

தீர்க்கப்பட்டது
ஜனவரி 12, 2024 இல் பிற்பகல் 5:34
தீர்க்கப்பட்டது
ஜனவரி 12, 2024 இல் பிற்பகல் 5:34
Coldfront is accessible again. Thanks for your patience so that would could investigate further. We've added new logging to have a better view into this should it happen again.
விசாரிக்கப்படுகிறது
ஜனவரி 12, 2024 இல் பிற்பகல் 3:22
விசாரிக்கப்படுகிறது
ஜனவரி 12, 2024 இல் பிற்பகல் 3:22
Coldfront is currently up but inaccessible. Login attempts will result in a timeout/proxy error.
We would like time to gather forensics to sort out why this is happening before restarting the service.

If you need to access Coldfront, please try again later today.

தீர்க்கப்பட்டது
ஜனவரி 11, 2024 இல் பிற்பகல் 1:22
தீர்க்கப்பட்டது
ஜனவரி 11, 2024 இல் பிற்பகல் 1:22
The Ceph instability has been resolved. Ceph Tier2 shares, VDI, and VMs should be back to their normal state.

If your VM, /net/fs-[labname] share, or VDI session is still impacted, please contact rchelp@rc.fas.harvard.edu
அடையாளம் காணப்பட்டது
ஜனவரி 11, 2024 இல் முற்பகல் 7:57
அடையாளம் காணப்பட்டது
ஜனவரி 11, 2024 இல் முற்பகல் 7:57
The infrastructure behind Tier2 Ceph shares and VMs is unstable.
This also affects VDI/OOD which relies on virtual machines.

/net/fs-[labname] shares, new OOD/VDI sessions, and VMs are affected and may will be inaccessible until this is resolved.

Thanks for your patience.

ஜன. 2024 வரை மார். 2024

FAS Research Computing - அறிவிப்பு வரலாறு

அறிவிப்பு வரலாறு

மார். 2024

பிப். 2024

ஜன. 2024