kubernetes

Featured

Kubernetes operations: debugging, security, RBAC, and infrastructure tooling.

DevOps & Infrastructure 425 stars 46 forks Updated yesterday MIT

Install

View on GitHub

Quality Score: 96/100

Stars 20%
88
Recency 20%
100
Frontmatter 20%
70
Documentation 15%
100
Issue Health 10%
80
License 10%
100
Description 5%
100

Skill Content

# Kubernetes Skill Three domains: **debugging** (pod triage, networking, resources), **security** (RBAC, pod hardening, network isolation, supply chain), and **cobaltcore** (KVM exporter, hypervisor metrics). Select by request signal, then follow the phases below. Always specify `-n <namespace>` in every kubectl command. Use read-only commands to gather evidence before proposing changes. --- ## Domain Selection | Signal | Domain | |--------|--------| | CrashLoopBackOff, OOMKilled, ImagePullBackOff, Pending | Debugging | | Service unreachable, DNS failure, port-forward | Debugging (network) | | CPU throttling, memory limit, disk pressure | Debugging (resources) | | RBAC, permissions, roles, ServiceAccount | Security (access) | | Pod hardening, container security, PodSecurity | Security (pods) | | NetworkPolicy, default-deny, namespace isolation | Security (network) | | Image signing, secrets, admission control | Security (supply chain) | | KVM exporter, cobaltcore, hypervisor metrics | Cobaltcore | --- ## Phase 1: TRIAGE ### Debugging Triage Flow Follow this sequence for every pod or workload issue. Do not skip steps -- many failures are only visible in events and describe output, not in logs. ```bash kubectl get pods -n <namespace> -o wide kubectl describe pod <pod-name> -n <namespace> kubectl logs <pod-name> -n <namespace> -c <container-name> kubectl logs <pod-name> -n <namespace> -c <container-name> --previous kubectl get events -n <namespace> --sort-by='.lastTime...

Details

Author
notque
Repository
notque/vexjoy-agent
Created
6 months ago
Last Updated
yesterday
Language
Python
License
MIT

Integrates with

Similar Skills

Semantically similar based on skill content — not just same category

DevOps & Infrastructure Listed

kubernetes-troubleshooting

Diagnose unhealthy Kubernetes workloads from symptom to root cause - CrashLoopBackOff, Pending, ImagePullBackOff, OOMKilled, Init:Error, readiness probe failures, pods stuck Terminating, and rollouts that never complete. Use whenever a pod, deployment, job or statefulset is not running as expected and the cause is not yet known, before changing any manifest.

0 Updated 2 weeks ago
riteshsonawane1372
DevOps & Infrastructure Listed

kubernetes-troubleshooting

Systematic triage of failing Kubernetes workloads using kubectl. Use when a pod is Pending, CrashLoopBackOff, ImagePullBackOff, OOMKilled, Error, or stuck Terminating; when a PersistentVolumeClaim will not bind; when a Service returns no endpoints or connections are refused; when an app is reachable inside the cluster but not from outside via Ingress or the Gateway API; when a Deployment's replicas never appear because a ResourceQuota, LimitRange, Pod Security, or admission webhook rejected them; when DNS resolution fails inside the cluster; or when a node is NotReady or reporting disk, memory, or PID pressure. Use when a Secret or config is missing or a live change keeps reverting (operator-synced secrets, GitOps reconciliation). Use when someone asks why a workload is not running, not reachable, or not scheduling. Use it too for proactive health checks — when someone asks "is the cluster OK?" or whether something is healthy even though nothing is obviously failing, since a green-looking cluster can still hi

0 Updated 2 months ago
ngaxavi
DevOps & Infrastructure Listed

kubernetes-operations

Debugs Kubernetes pods and controllers — FailedCreate, ImagePullBackOff, init-container failures, probe flapping, missing service endpoints, GKE NEG readiness. Use when a pod is not Running, a Deployment/StatefulSet shows FailedCreate, image pulls fail, or services lack endpoints.

1 Updated 2 months ago
Goodsmileduck