All systems are operational.
FeM AdminDB and network services (DHCP, DNS...)
We are still monitoring the situation as parts of the system continue to behave strange.
The storage updates have been deployed and most of the systems seem to have stabilized. We are monitoring the situation bevor returning back to normal operations.
We were able to pinpoint a potential issue. After the core systems maintenance last thursday no schema version upgrade to the underlying storage pool was applied, which caused I/O problems, high load and drive fragmentation. These are potentially valid causes for the current outages.
After applying the storage upgrade on parts of the core systems, partial network recovery was observed.
We are still monitoring the situation and will keep you updated. We will continue the upgrades, if the systems stabilize further.
The root cause of this issue is still not found.
We are still experiencing regular partial network outages.
We are currently experiencing an unexpected outage with our DHCP service. As a result, the automatic assignment of IPv4 addresses is failing across the network.