Kubernetes Troubleshooting

Comprehensive Kubernetes and OpenShift cluster health analysis and troubleshooting. Use this skill when: (1) Proactive cluster health assessment and security analysis (2) Analyzing pod/container logs for errors or issues (3) Interpreting cluster events (kubectl get events) (4) Debugging pod failures: CrashLoopBackOff, ImagePullBackOff, OOMKilled (5) Diagnosing networking issues: DNS, Service connectivity, Ingress/Route problems (6) Investigating storage issues: PVC pending, mount failures (7) Analyzing node problems: NotReady, resource pressure, taints (8) Troubleshooting OCP-specific issues: SCCs, Routes, Operators, Builds (9) Performance analysis and resource optimization (10) Security vulnerability assessment and RBAC validation (11) ARO (Azure Red Hat OpenShift) cluster troubleshooting (12) ROSA (Red Hat OpenShift on AWS) cluster troubleshooting

kcns008 Updated

File contents

kcns008/cluster-skills/tree/main/skills/kubernetes-troubleshooting commit 18f4b0f8ac

Frequently asked questions

npx skillmds@latest add kcns008-cluster-skills/kubernetes-troubleshooting