Hello ManageEngine Support Team,
We are experiencing a critical issue with OpManager Failover after rebuilding the secondary server.
Our environment:
Previously, these two servers were configured as a working Failover pair.
During the recent upgrade, the secondary OpManager server stopped starting. We therefore completely uninstalled OpManager and Applications Manager from the secondary server and performed a clean installation.
After the clean installation, we followed the Failover configuration procedure:
.bat script to import/synchronize the required files to the secondary server.After enabling Failover, the system entered an unstable state.
The servers continuously attempt to change their roles:
The role switching appears to continue indefinitely, and the Failover pair does not reach a stable Primary/Standby state.
Could you please help us determine the cause of this behaviour and restore a stable Failover configuration?
Please also specify exactly which diagnostic data we should provide to speed up the investigation, including:
Please let us know whether we should temporarily stop one of the OpManager servers before collecting the logs or performing any further actions.
At the moment, the monitoring system is frequently unavailable because of the continuous Active/Standby role switching, so we would appreciate your assistance as soon as possible.