Vendor586 problems· 4 reviewed

Kubernetes

586 incident problems in Kubernetes environments.

All problems (586)

CICD-1484A Jenkins Kubernetes agent pod terminates and the workspace stays lockedOne Jenkins controller begins leaving locked workspaces only on ephemeral Kubernetes agents after a sidecar image update.CI/CDIntermediate10 minProK8S-1572A Karpenter node class updates and one nodegroup still launches with the old profileOne Karpenter nodegroup launches with the old instance profile after a rename.KubernetesIntermediate10 minProK8S-1528A Karpenter provisioner launches nodes and pods remain pendingPods stay pending on one scheduling path after a label schema migration.KubernetesIntermediate10 minProK8S-1518A Karpenter provisioner launches nodes and pods still stay pendingPods remain pending only on one scheduling path after a Karpenter label schema migration.KubernetesIntermediate10 minProSECURITY-1457A Secrets Store CSI mount rotates correctly and the Java app still trusts the old chainTLS still fails in one Java app after certificate rotation through CSI secrets until the pod restarts.SecurityIntermediate10 minProK8S-1595A StatefulSet returns and one PVC still lands wrongOne PVC still lands in the wrong place after topology label cleanup.KubernetesAdvanced10 minProK8S-1605A volume expansion succeeds and one pod still sees ENOSPCOne pod still sees ENOSPC after successful volume expansion.KubernetesAdvanced10 minProK8S-1458An ExternalDNS rollout updates public records and one cluster path still flapsAn ExternalDNS rollout updates public records and one cluster path still flaps focuses on Service Discovery and asks the reader to isolate the key signal in AWS. Record flapping often comes from ownership-marker collisions rather than from...KubernetesIntermediate10 minProK8S-1508An HPA sees demand and never scalesAutoscaling stops after adapter query-step tuning even though Prometheus shows rising demand.KubernetesIntermediate10 minProK8S-1489An HPA sees rising load and refuses to scaleAutoscaling stops only after a deployment rename and metrics adapter refresh that otherwise looks healthy.KubernetesIntermediate10 minProK8S-1498An HPA sees rising metrics and never scalesAutoscaling stops only after query step tuning on the metrics adapter and Prometheus backend.KubernetesIntermediate10 minProK8S-1460An ingress snippet policy blocks only one classAn ingress snippet policy blocks only one class focuses on edge-routing and asks the reader to isolate the key signal in NGINX. Admission allowlists often key off controller metadata that seems unrelated to the snippet itself.KubernetesIntermediate10 minProK8S-1487A cert-manager DNS01 challenge updates Route53 and still times outACME issuance fails intermittently only in hosted zones managed by both cert-manager and ExternalDNS.KubernetesIntermediate11 minProK8S-1575A CSI expansion succeeds and one pod still sees the old sizeOne pod still sees the old filesystem size after successful PVC expansion.KubernetesAdvanced11 minProK8S-1505A StatefulSet restore finishes and one workload still hangsA restored stateful workload can not mount on replacement nodes after a node pool refresh.KubernetesAdvanced11 minProK8S-1529A Velero restore completes and fresh pods still blockNew pods are blocked only in restored namespaces after policy changes.KubernetesAdvanced11 minProK8S-1539A Velero restore completes and fresh pods still blockNew pods are blocked only in restored namespaces after policy changes.KubernetesAdvanced11 minProK8S-1519A Velero restore succeeds and quotas still block new podsFresh pods are blocked in a restored namespace even though the quota object looks correct.KubernetesAdvanced11 minProK8S-1585A volume expansion completes and one pod still sees ENOSPCOne pod still sees ENOSPC after a successful volume expansion.KubernetesAdvanced11 minProK8S-1451An HPA stops scaling from custom metricsCustom metric scaling fails only for one workload after an adapter config reload.KubernetesIntermediate11 minPro