Etcd pod in CrashLoopBackOff after rebuilding master node in RHOCP 4

Solution Verified - Updated -

Issue

  • After rebuilding or replacing a bare-metal control-plane node, the newly added node shows a Ready status in oc get nodes, but the corresponding etcd static pod remains in CrashLoopBackOff or fails to join the cluster.

  • The BareMetalHost (BMH) object exists but has no CONSUMER assigned, and there is no corresponding Machine object for the new node:

    NAME                 STATUS       STATE       CONSUMER   ONLINE   ERROR   AGE
    <rebuilt_node_name>  discovered   unmanaged               true            19h
    
  • The openshift-etcd-operator pod logs continuously output the following warning:

    Ignoring node (<rebuilt_node_name>) for scale-up: no Machine found referencing this node's internal IP (<node_ip>)
    

Environment

  • Red Hat OpenShift Container Platform (RHOCP)
    • 4

Subscriber exclusive content

A Red Hat subscription provides unlimited access to our knowledgebase, tools, and much more.

Current Customers and Partners

Log in for full access

Log In

New to Red Hat?

Learn more about Red Hat subscriptions

Using a Red Hat product through a public cloud?

How to access this content