Kubernetes guides
Understand cluster failures and restore workloads safely.
Diagnose Kubernetes liveness probe failures by checking events, restart history, probe configuration, application stalls, dependencies, and resource pressure.
Diagnose MountVolume.SetUp failed events by checking volume references, Secrets, ConfigMaps, PVCs, CSI drivers, node access, and mount permissions.
Diagnose CreateContainerConfigError by checking pod events, missing Secrets and ConfigMaps, invalid keys, service accounts, volumes, and security context.
Diagnose Kubernetes pod eviction by identifying memory, disk, inode, PID, taint, and node-pressure causes before replacing workloads.
Diagnose Kubernetes readiness probe failures by checking events, probe configuration, application listeners, dependencies, timing, and resource pressure.
Trace Kubernetes DNS failures through pod configuration, service names, CoreDNS, endpoints, NetworkPolicies, node resolvers, and upstream DNS.
Trace Kubernetes Service connectivity from DNS and ClusterIP through ports, EndpointSlices, pods, readiness, NetworkPolicies, and kube-proxy.
Diagnose kubectl Forbidden errors by checking identity, context, RBAC permissions, namespace scope, service accounts, and admission policy.
Determine why Kubernetes containers are OOMKilled by comparing memory limits, actual usage, node pressure, application behavior, and restart history.
Find why a Kubernetes container repeatedly crashes by checking pod events, previous logs, exit codes, probes, configuration, and resource limits.
Resolve Kubernetes image pull failures by checking image names, tags, registry authentication, node connectivity, and architecture compatibility.
Determine why Kubernetes cannot schedule or start a Pending pod by checking scheduler events, resources, constraints, volumes, and quotas.