Emergency Stop For Rancher System Upgrade Controller
Use this when a Rancher system-upgrade-controller worker Plan is actively cordoning or draining nodes and a normal live-object patch does not stick because GitOps restores it. Confirm The Controller Is Active kubectl get plan -n system-upgrade kubectl get jobs,pods -n system-upgrade -o wide kubectl get events -n system-upgrade --sort-by=.lastTimestamp Look for jobs like: apply-agent-plan-on-worker-1-... Repeated jobs can mean the upgrade job is timing out while trying to stop rke2-agent or rke2-server. In that case the controller may start over and attempt the shutdown again, keeping the selected worker in the disruption path. ...