Issue
- In a clustered environment if we restart the nodes at the same time for a deployment, schedulers and information can be lost.
Environment
- Liferay DXP 7.3
Resolution
- The issue described above is caused by not starting clustered nodes sequentially which is considered to be a bad practice.
- With the non-sequential restart we introduce a race condition to our system. If one of the nodes with a previous 'slave' status gets elected as a 'master' that node will not have the information we initialized before, therefore it will be lost.
-
In general for DXP it is a requirement to start the nodes sequentially. This means that you need to wait until a node fully starts up, before starting the second one.
Additional information
- If you would like to gather some additional insight in the Schedulers topic you can visit the related developer blog post: Evolution of the Liferay Scheduler API