Etcd pod in CrashLoopBackOff after rebuilding master node in RHOCP 4
Issue
-
After rebuilding or replacing a bare-metal control-plane node, the newly added node shows a
Readystatus inoc get nodes, but the correspondingetcdstatic pod remains inCrashLoopBackOffor fails to join the cluster. -
The
BareMetalHost(BMH) object exists but has noCONSUMERassigned, and there is no correspondingMachineobject for the new node:NAME STATUS STATE CONSUMER ONLINE ERROR AGE <rebuilt_node_name> discovered unmanaged true 19h -
The
openshift-etcd-operatorpod logs continuously output the following warning:Ignoring node (<rebuilt_node_name>) for scale-up: no Machine found referencing this node's internal IP (<node_ip>)
Environment
- Red Hat OpenShift Container Platform (RHOCP)
- 4
Subscriber exclusive content
A Red Hat subscription provides unlimited access to our knowledgebase, tools, and much more.