All Systems Operational

About This Site

Welcome to Dartmouth College's home for real-time incident and maintenance communications.

Academic Services Operational
Administrative Services Operational
Authentication Operational
Email Operational
M365 email Operational
Gmail Operational
Infrastructure Operational
Load Balancer Operational
Containerization Operational
RDS Services Operational
Active Directory Operational
Computing Infrastructure Operational
Network Operational
Research Computing Operational
90 days ago
99.48 % uptime
Today
General Research Services Operational
AI Operational
90 days ago
99.48 % uptime
Today
Granite Operational
HPC Operational
Rapport Operational
Storage Operational
Security Operational
Telephone Operational
Web Operational
Operational
Degraded Performance
Partial Outage
Major Outage
Maintenance
Major outage
Partial outage
No downtime recorded on this day.
No data exists for this day.
had a major outage.
had a partial outage.

Scheduled Maintenance

Network Maintenance for Dartmouth Chat Sep 2, 2026 06:00-07:00 EDT

Network Services will be performing maintenance on the network connection used to access Dartmouth Chat.

The maintenance will be taking place from 6:00am to 7:00am on Tuesday, September 2nd.

During this time Dartmouth Chat will be unavailable.

Posted on Aug 21, 2026 - 10:30 EDT

Discovery Cluster: Slurm Upgrade to v.26.05.01 Sep 3, 2026 08:00-11:00 EDT

Research Computing and Data will be upgrading the HPC workload environment to Slurm version 26.05.01.

This update will deliver enhanced job‑scheduling performance, improved stability , and minor patch fixes across the cluster.

Impact: New jobs will not be started or queued after 8 a.m. on September 3rd.

Posted on Aug 19, 2026 - 15:55 EDT
Wi-Fi Devices ?
Fetching
Aug 28, 2026

No incidents reported today.

Aug 27, 2026
Completed - The scheduled maintenance has been completed.
Aug 27, 22:32 EDT
Verifying - Verification is currently underway for the maintenance items.
Aug 27, 22:00 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 27, 21:30 EDT
Scheduled - ITC Network Services will perform network maintenance on the primary network firewall. Network traffic will be migrated to an alternate firewall before the maintenance and migrated back after the maintenance is complete. No user outage is expected, other than some network sessions needing to be reestablished after each traffic move.
Aug 27, 12:12 EDT
Completed - The scheduled maintenance has been completed.
Aug 27, 00:00 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 26, 21:00 EDT
Scheduled - We will be undergoing scheduled maintenance during this time.
Aug 25, 10:37 EDT
Aug 26, 2026
Resolved - The hardware has been replaced and the affected nodes are online.
Aug 26, 15:07 EDT
Identified - Due to a switch outage several of our HPC nodes are in a down state. We have ordered the replacement part and estimate the nodes to be restore after lunch tomorrow.
Aug 25, 15:56 EDT
Completed - Weve completed our update.
Aug 26, 15:06 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 26, 10:00 EDT
Scheduled - Research Computing and Data will be installing routine firmware and software updates on the Dell PowerScale system which provides network file services to the campus. This includes DartFS, ThayerFS, and others.

These reboots are staggered so that the service never goes down, but any clients connected to a node that is in the process of rebooting obviously need to reconnect to a different node.

NFS clients running Linux do this automatically and seamlessly.

Unfortunately, the SMB protocol does not allow for the same kind of seamless transition to a different node. The actual SMB client behavior depends on the operating system (Mac, Windows, or Linux) and the application using the storage.

In general, Windows systems and applications will recognize that the connection was interrupted and automatically reconnect to a different node after just a few seconds.

Mac and Linux SMB clients will keep trying the same node for longer. After a timeout period, they declare a network error. Some applications will try to restart the connection from scratch and that will work. Most commonly, Mac users can simply reconnect from their saved list of favorite servers.

Please note that if you were actively using any of these file servers during the entire maintenance window you would experience at least one such disconnection and probably several as we cycle through the nodes to reboot them all.

We will post an update here when the work is complete.

Aug 25, 09:36 EDT
Resolved - The failed hardware has been replaced and affected BMR datacenter network connections are back online.
Aug 26, 14:25 EDT
Identified - We are working with the equipment vendor to replace the failed hardware.
Aug 25, 12:03 EDT
Investigating - ITC is investigating an equipment failure affecting some network connections in the BMR datacenter.
Aug 25, 11:01 EDT
Completed - Closing status - Maintenance complete
Aug 26, 11:22 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 26, 08:00 EDT
Scheduled - Updates to email security for Dartmouth email will be undergone to reduce spoofed and phishing mail. No action is required from users, and no impact to email delivery is expected.
Aug 21, 07:40 EDT
Aug 25, 2026
Completed - Clsoing status - Maintenance window is complete
Aug 25, 17:58 EDT
In progress - Hello all ,
Due to the critical technical issue, Planon cloud team is performing emergency maintenance for the next hour to address and fix this issue. Efforts are made to minimize disruption to your services , however a few minutes of downtime can be expected.

Aug 25, 11:03 EDT
Resolved - We have resolved the issue though we are continuing to monitoring our systems.
Aug 25, 08:54 EDT
Update - Discovery is now stable. We are waiting for jobs to finish on about 30 more nodes so we can reboot those and get them ready for new jobs.  Until those 30 are reintegrated, you may see increased wait times in the queue for resources to be free.
Aug 21, 13:58 EDT
Monitoring - We experienced a large cluster issue while updating the Slurm database host to a newer version.
This resulted in some jobs becoming stuck in the COMPLETING (CG) state and caused broader job instability.

We have since upgraded the cluster to the same Slurm version as the database and are rebooting affected
nodes to clear jobs stuck in the CG state. The cluster is recovering, and stability is improving.

We appreciate your patience as we continue working to fully resolve the issue and return the cluster to a steady state.

Please check the status and results of your recent jobs. If you experienced job failures or lost compute time as a result
of this outage, please reach out to research.computing@dartmouth.edu.

Where helpful, we can temporarily increase your
available resources to help account for lost compute time.

Aug 21, 10:21 EDT
Update - Weve applied an update to the scheduler. All new jobs will begin as normal however as running jobs finish, we will need to reboot compute nodes.
Aug 21, 09:00 EDT
Update - HPC systems remain operational but degraded due to an ongoing issue. Troubleshooting is paused for tonight and will resume tomorrow morning. We appreciate your patience.
Aug 20, 19:12 EDT
Identified - The scheduler on Discovery(Slurm) is behaving erratically and we have a ticket open with the vendor. We will update when we know more.
Aug 20, 15:33 EDT
Completed - The scheduled maintenance has been completed.
Aug 25, 01:00 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 24, 22:00 EDT
Scheduled - Niagara will be intermittently unavailable as we perform maintenance. We appreciate your patience.
Aug 24, 15:52 EDT
Aug 24, 2026
Aug 23, 2026

No incidents reported.

Aug 22, 2026
Resolved - Everything is should be resolved. We will continue to monitor to ensure network performance.
Aug 22, 09:53 EDT
Investigating - We are investigating network issues detected in Sudikoff Hall. We will update as we learn more.
Aug 22, 08:51 EDT
Aug 21, 2026
Resolved - Data load completed.
Aug 21, 11:18 EDT
Investigating - The STUDENT to Warehouse data feed failed last night. Information, Technology, and Consulting (ITC) is looking into the issue and will follow up as soon as further information is available.

We apologize for any inconvenience.

Aug 21, 08:29 EDT
Resolved - Load completed.
Aug 21, 11:18 EDT
Identified - The issue has been identified, corrected and the load was restarted.
Aug 21, 08:45 EDT
Investigating - The STUDENT to Warehouse data feed failed last night. Information, Technology, and Consulting (ITC) is looking into the issue and will follow up as soon as further information is available.

We apologize for any inconvenience.

Aug 21, 08:27 EDT
Aug 20, 2026
Resolved - Maintenance has been completed and Dartmouth Chat is now at Dartmouth Chat v0.10.2-dc
Aug 20, 18:43 EDT
Monitoring - An issue has been identified with the Dartmouth Chat database upgrade for the August release and we are monitoring the situation.
Aug 20, 10:33 EDT
Investigating - Dartmouth Chat Upgrade is still offline for maintenance, an update will be posted once it is finished.
Aug 20, 07:35 EDT
Completed - The scheduled maintenance has been completed.
Aug 20, 07:30 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 20, 06:30 EDT
Scheduled - Research Computing and Data will be releasing an update to Dartmouth Chat that will include the following :

- Update to Open WebUI v0.10.2
- Gemini 3.7 Flash

Further details will be published on our AI Blog:

https://ai-tools.dartmouth.edu/blog

The upgrade will be take place from 6:30 to 7:30AM on Thursday, August 20th.

During this time Dartmouth Chat will be intermittently unavailable.

Aug 17, 10:59 EDT
Completed - The scheduled maintenance has been completed.
Aug 20, 07:00 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 20, 02:00 EDT
Scheduled - Monthly OS patch updates will be applied to all production servers and databases between 2:00 AM and 7:00 AM.

The patch makeup window is Friday morning 2:00 AM and 7:00 AM.

It should be expected that all systems and services will be unavailable during this time.

For more information, please reference the following Team Dynamix article.
https://services.dartmouth.edu/TDClient/1806/Portal/KB/ArticleDet?ID=105996

Aug 17, 08:11 EDT
Aug 19, 2026
Resolved - This incident has been resolved.
Aug 19, 09:32 EDT
Investigating - We are currently investigating an issue with the Niagara supervisor.
Aug 19, 08:53 EDT
Aug 18, 2026
Completed - Planon L132 upgrade is complete
Aug 18, 20:56 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 18, 18:00 EDT
Scheduled - Planon is ready to be upgraded to L132 .
End users can expect a downtime approximately lasting for 4 hours , starting 6 PM - 10 PM .

Aug 18, 15:48 EDT
Completed - The scheduled maintenance has been completed.
Aug 18, 15:46 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 18, 07:00 EDT
Update - We will be undergoing scheduled maintenance during this time.
Aug 17, 17:05 EDT
Scheduled - Dartmouth ITC is rolling out security updates to several custom web applications.
Aug 17, 17:03 EDT
Resolved - ITC has resolved the incident. If you are still experiencing issues with authentication then please submit a ticket through the Services Portal.
Aug 18, 11:48 EDT
Identified - ITC is investigating issues that are preventing some applications from loading. We will provide another update as more information becomes available.
Aug 18, 09:47 EDT
Monitoring - A fix has been applied and all systems are performing as expected. We're continuing to monitor for any recurring problems.
Aug 18, 08:37 EDT
Update - We are continuing to work on a fix for this issue.
Aug 18, 07:49 EDT
Identified - ITC has identified the issue affecting authentication and is working to restore service. We will continue to provide updates here.
Aug 18, 07:07 EDT
Investigating - ITC is investigating degraded performance with authentication. We will provide another update as more information becomes available.
Aug 18, 06:30 EDT
Completed - The scheduled maintenance has been completed.
Aug 18, 06:30 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 18, 05:00 EDT
Scheduled - The HAProxy Enterprise Edition (HAPEE) load balancer data plane will be upgraded from version 2.8 to version 3.2. This is a rolling upgrade, and load balanced services will remain online during the upgrade window.
Aug 17, 10:04 EDT
Completed - The scheduled maintenance has been completed.
Aug 18, 05:01 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 17, 17:00 EDT
Update - We will be undergoing scheduled maintenance during this time.
Aug 17, 16:20 EDT
Scheduled - We will be undergoing scheduled maintenance for Dartmouth Claude during this time to move to a new configuration, which will allow for more features. We anticipate that Dartmouth Claude will not be available again until tomorrow morning about 8 am.
Aug 17, 16:18 EDT
Aug 17, 2026
Resolved - GitHub has resolved the incident. If you are still experiencing issues with github.com then please submit a ticket through the Services Portal.
Aug 17, 22:25 EDT
Investigating - GitHub is investigating degraded performance. We will follow up as additional information is available from the vendor.
Aug 17, 11:33 EDT
Aug 16, 2026

No incidents reported.

Aug 15, 2026
Resolved - IRA data is available for reporting
Aug 15, 18:18 EDT
Monitoring - The ETL load got stuck partway through one of the processing steps. We restarted it this morning, and it completed successfully at approximately 10:41 AM.
The data is currently being copied over to the reporting database. We expect all data to be available for reporting within the next hour.
Apologies for the inconvenience

Aug 15, 12:19 EDT
Aug 14, 2026
Completed - ITC Network Services has completed the network maintenance on the primary network firewall. All traffic is now flowing through the primary firewall again.
Aug 14, 22:28 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 14, 21:00 EDT
Scheduled - ITC Network Services will perform network maintenance on the primary network firewall. Network traffic will be migrated to an alternate firewall before the maintenance and migrated back after the maintenance is complete. No user outage is expected, other than some network sessions needing to be reestablished after each traffic move.
Aug 14, 16:09 EDT
Completed - Consolidated Oracle prod db/Copper performance issues have been resolved. Please submit a ticket to the student team if you notice any issues.
Aug 14, 16:44 EDT
In progress - Oracle consolidated prod db/Copper is currently experiencing degraded performance and we are looking into it.
Aug 14, 15:57 EDT
Completed - this is now completed
Aug 14, 16:32 EDT
In progress - we need to reboot the system now as it is no longer responding. this will happen at 4:10pm now.
Aug 14, 16:04 EDT
Scheduled - the database will be restarted so performance enhancements can take affect.
Aug 14, 13:16 EDT
Completed - ITC has completed the maintenance on Oracle consolidated prod db/Copper. If you experience any issues please submit a ticket at our Services Portal.
Aug 14, 06:41 EDT
In progress - Scheduled maintenance is currently in progress. We will provide updates as necessary.
Aug 14, 05:45 EDT
Scheduled - ITC will be performing scheduled maintenance on Oracle consolidated PROD DB/COPPER Services related to the database will be unavailable during this event. We appreciate your patience. If you have questions about this event then please contact the student team.
Aug 13, 14:59 EDT