Kubernetes for AI agents: when code is cheap, proof matters
How MCP, Cluster API and Argo CD can give a small team the infrastructure to verify AI-generated code, run parallel QA and retire environments automatically.
Sourced guidanceThe Kubedex library
Learn what Kubernetes tools do and how to choose between them. Browse comparisons, step-by-step guides and clearly dated experiences from the original Kubedex site.
108 articles
Original context. Clear evidence.How MCP, Cluster API and Argo CD can give a small team the infrastructure to verify AI-generated code, run parallel QA and retire environments automatically.
Sourced guidanceUnderstand MinIO and compare self-hosted S3 storage options for application uploads, backups and model files, then plan a safe migration.
Sourced guidanceLearn how Argo CD keeps deployments in sync with Git, then manage ApplicationSets, investigate failed changes and recover safely.
Sourced guidanceUnderstand OpenTelemetry Collector, Prometheus and Grafana Alloy, then choose how Kubernetes metrics, logs and traces reach your monitoring tools.
Sourced guidanceUnderstand what vLLM, KServe and llm-d do, when to combine them, and how to measure an AI service before adding more infrastructure.
Sourced guidanceUnderstand device plugins, Dynamic Resource Allocation and Kueue, then find why a GPU application is waiting or cannot use its device.
Sourced guidanceCompare Trivy and Grype for finding known vulnerabilities in container images, and turn scan results into clear release decisions.
Sourced guidanceLearn how image signatures identify a publisher and how Kyverno can reject unapproved images before they run in Kubernetes.
Sourced guidanceUnderstand Bitnami charts and images, find maintained replacements and move applications without losing configuration or data.
Sourced guidanceMove a PostgreSQL database from Bitnami packaging to CloudNativePG, including roles, extensions, backups and the point when applications switch writes.
Sourced guidanceCompare Karpenter, Cluster Autoscaler and EKS Auto Mode for adding Kubernetes machines when applications need them and removing unused capacity.
Sourced guidanceReplace the retired community ingress-nginx controller with a maintained gateway while preserving application routes, HTTPS and access checks.
Sourced guidanceCompare Argo CD and Flux for deploying applications from Git: dashboards, Helm releases, access controls and recovery after a bad change.
Sourced guidancePrepare for the practical CKA exam with current rule checks and a lab-based study plan, alongside the original 2019 tips.
Historical recordRead an early EKS production experience and understand which lessons need checking against current Amazon EKS behavior.
Historical recordFind out what 90% CPU utilization means in Kubernetes and choose a target that leaves enough room for traffic spikes and new Pods to start.
Sourced guidanceLearn what the K3s Helm Controller does and when to use it instead of running Helm commands or adopting a wider GitOps workflow.
Sourced guidanceLearn when to use HPA, KEDA and a node autoscaler to add application capacity, handle demand spikes and remove idle machines.
Sourced guidanceCompare Alpine, distroless and full Linux base images by application compatibility, security updates and how you will debug a running container.
Sourced guidanceLearn what Helm charts, values and releases are, then plan your first Kubernetes application installation and update.
Sourced guidanceA historical account of contracting in London; its day rates, tax assumptions and business setup advice are not current guidance.
Historical recordChoose a practical learning objective and use the historical DevOps and SRE course directory as a starting point for further study.
Historical recordLearn how cloud automation tools create repeatable infrastructure, with a study plan for identity, Terraform and cleanup.
Historical recordBuild Linux skills for running applications: processes, files, permissions, networks, packages and logs.
Historical recordLearn the programming skills useful for infrastructure work, including Git, Python, tests and small maintainable tools.
Historical recordLearn the foundations of site reliability engineering through one service, its users, failure signals and recovery procedures.
Historical recordDescribe the engineering work you owned, the constraints you faced and the outcomes you can substantiate; distinguish current resume advice from historical LinkedIn features.
Historical recordUnderstand the roles of Django, Gunicorn, a reverse proxy and Kubernetes before deploying a Python web application.
Sourced guidanceFind a maintained replacement for an old Helm chart without mistaking an archived package for a discontinued application.
Sourced guidanceCompare EKS Pod Identity and IRSA for letting applications access AWS services without storing long-lived access keys.
Sourced guidanceUnderstand why vulnerability scanners disagree and read the original scanner follow-up without treating its old counts as current results.
Historical recordLearn what Kubernetes does, then deploy one small application and practice inspecting, changing and recovering it in a local cluster.
Historical recordRead the historical AKS versus GKE provisioning test, understand what was measured and plan a fair current comparison.
Historical recordCompare Amazon EKS, Azure AKS and Google GKE: what each managed Kubernetes service provides and which fits your applications and cloud services.
Sourced guidancePlan a Helm 3 to Helm 4 upgrade, check plugins and deployment automation, and reproduce the recorded small upgrade-and-rollback test.
Sourced guidanceCompare direct Helm commands with Argo CD and Flux for deploying and updating applications from charts.
Historical recordLearn what an OCI Helm registry is and how to distribute charts alongside container images without confusing their versions or credentials.
Sourced guidanceImprove software delivery by fixing one recurring delay or failure, then check whether the change helps the people doing the work.
Sourced guidanceCompare Envoy Gateway, Cilium, Traefik and managed gateways for routing web traffic into Kubernetes applications.
Sourced guidanceRead the preserved December 2018 AKS test and distinguish its historical failures from a current Azure Kubernetes assessment.
Historical recordCompare Istio and Linkerd for encrypting traffic between services, controlling access and investigating requests across applications.
Sourced guidanceArchived Eaglecliff Recruitment advert for a Kubernetes-focused DevOps contract in Central London. Historical listing; current availability is unverified.
Historical recordHistorical Kubedex job board. Recovered advertisements are dated records; job applications and employer submissions are no longer accepted.
Historical recordFind out why Karpenter is keeping a lightly used node, including Pod disruption budgets, resource requests and placement rules.
Sourced guidanceUnderstand the historical Helm metrics exporter, why its Tiller dependency no longer applies and what to monitor in current deployments.
Historical recordA historical August 2021 update on restoring Kubedex, repairing comparison tables and plans for the Kubernetes chart directory.
Historical recordA September 2018 site update covering the original Helm chart catalog, Linkerd 2.0 release and Tiller metrics work.
Historical recordA historical October 2018 Kubedex update on scanner experiments, autoscaling, serverless research and the growing Helm chart directory.
Historical recordHistorical October 2018 site update covering cloud comparisons, contracting, base images and newly indexed Helm charts.
Historical recordAn October 2018 site update covering observability, Project Dolos cloud tests, Helm chart changes and early traffic statistics.
Historical recordA historical November 2018 site update covering DevOps resumes, cloud and on-premises comparisons, local Kubernetes and early Kubedex traffic.
Historical recordDecember 2018 Kubedex update covering EKS, Kubernetes operators, container runtimes, networking, Helm additions and historical site statistics.
Historical recordKubedex’s December 2018 site update: new chart listings, audience observations and predictions for 2019, preserved as historical context.
Historical recordKubedex’s September 2018 launch update: early catalogue growth, traffic and Tiller-exporter plans, preserved as site history.
Historical recordUnderstand Kubernetes configuration backups, persistent-volume backups and database recovery before reading the historical tool comparison.
Historical recordCompare containerd and CRI-O, understand what a container runtime does, and separate Kubernetes node choices from Docker development tools.
Historical recordUnderstand Kubernetes cost allocation, idle capacity and cloud bills before comparing tools in the historical research workbook.
Historical recordChoose a Kubernetes course for the skills you need and pair it with a small application you can deploy and troubleshoot.
Historical recordCompare Cilium and Calico for connecting Kubernetes Pods, enforcing network rules and investigating traffic problems.
Sourced guidanceCompare Talos, Bottlerocket and general-purpose Linux for the machines running Kubernetes, including updates and debugging.
Historical recordBuild a repeatable way for developers to create and deploy one kind of service, using templates, Git and clear support instructions.
Sourced guidanceLearn how Kubernetes Secrets supply application credentials and choose where passwords and tokens should be created, stored and replaced.
Sourced guidanceUnderstand External Secrets, Sealed Secrets, SOPS and Vault, then choose how applications receive passwords, tokens and certificates.
Sourced guidanceUnderstand what Trivy, Grype, Kyverno, Falco and secret-management tools protect, then choose the checks your Kubernetes applications need.
Sourced guidanceFollow a practical route from Linux and networking to containers, Kubernetes and reliable application delivery.
Historical recordLearn how small, frequent changes and feedback help a team improve software without committing to a large untested plan.
Sourced guidanceUnderstand how Ansible inventories and playbooks automate repeated server configuration, then practice on one disposable machine.
Sourced guidanceLearn what the AWS CLI does and how profiles, accounts and regions determine where a command runs.
Sourced guidancePlan an AWS learning account, understand Free and Paid account plans and keep track of the resources your exercises create.
Sourced guidanceUnderstand DNS, addresses, ports and TLS, then trace why a request cannot reach a service.
Sourced guidanceLearn how a controlled failure experiment checks whether an application recovers as expected.
Sourced guidanceUnderstand continuous integration and delivery, from checking a code change to deploying a known version and recovering a failed release.
Sourced guidanceBuild a small CI pipeline that checks each code change, reports useful failures and produces an identifiable application build.
Sourced guidanceAdd practical security checks to software delivery, including who can change source, build an image and approve deployment.
Sourced guidanceLearn how disks, partitions, filesystems and mount points relate before adding storage to a Linux machine.
Sourced guidanceUnderstand timeouts, partial failures and retries when an application depends on more than one service.
Sourced guidanceLearn how Docker images, containers and volumes work, then build and run a small application locally.
Sourced guidanceUnderstand Amazon EC2 virtual machines, their disks and network access, and what to remove when a lab is finished.
Sourced guidanceLearn how Git records changes, how branches support collaboration and how to inspect a change before sharing it.
Sourced guidanceLearn how Go can turn a small automation task into a command-line program with clear inputs, errors and tests.
Sourced guidanceUnderstand AWS IAM users, roles and permissions, then design access for one narrow task.
Sourced guidanceInstall Ubuntu in a disposable virtual machine and learn the basic disk, account and network choices.
Sourced guidanceLearn what Kubernetes does and how Pods, Deployments and Services work together to run an application.
Sourced guidanceLearn how metrics, logs and traces help explain an application failure, then instrument one small request path.
Sourced guidanceUnderstand how Linux package managers install software, resolve dependencies and receive updates from repositories.
Sourced guidanceUse Python to automate a small read-only task, handle invalid input and keep dependencies separate from the system installation.
Sourced guidanceIdentify what a small cloud service needs to protect, who can access it and how to respond if access or data is lost.
Sourced guidanceUnderstand Vagrant, Vagrantfiles, providers and boxes before building a repeatable local virtual-machine lab.
Sourced guidanceLearn how Terraform configuration, providers, plans and state work together to create and change infrastructure.
Sourced guidanceChoose software tests that catch real mistakes and understand the difference between an isolated test and a check against a real dependency.
Sourced guidanceInvestigate an application failure by recording the symptom, narrowing its cause and testing one explanation at a time.
Sourced guidanceUnderstand Linux accounts and SSH key pairs, then practice a restricted login to a disposable machine.
Sourced guidanceLearn what virtual machines provide and how guest storage, networking and snapshots differ from the host computer.
Sourced guidanceUnderstand an AWS VPC, subnets, routes and security rules before connecting cloud applications to each other or the internet.
Sourced guidanceCompare kind, minikube and Docker Desktop for learning Kubernetes or testing applications on your own computer.
Historical recordLearn what metrics, logs, traces and profiles tell you about an application, and how Prometheus, OpenTelemetry and storage backends fit together.
Sourced guidanceUnderstand Kubernetes operators, how they differ from Helm charts, and when automated upgrades, backups and recovery justify using one.
Sourced guidanceUnderstand Project Dolos, the historical automation that tested cluster creation on Google GKE and Azure AKS.
Historical recordUnderstand Prometheus Operator and kube-prometheus-stack, then upgrade monitoring without losing metrics, dashboards or alerts.
Sourced guidanceUnderstand OKD, the former PKS product and Google Distributed Cloud before choosing Kubernetes for your own data center.
Sourced guidanceExplain OpenShift, Rancher and the historical PKS product, then compare what they provide for running Kubernetes in your own infrastructure.
Historical recordUnderstand Redis and Valkey, decide whether they fit your cache or persistent data needs, and plan a Kubernetes migration around real compatibility.
Sourced guidanceUnderstand serverless on Kubernetes and choose between request-driven serving with Knative and event-driven worker scaling with KEDA.
Historical recordRead the original 2019 Helm chart shortlist and learn how to find maintained packages for the same Kubernetes tasks today.
Historical recordFind why an AWS identity cannot access Kubernetes, and distinguish the IAM Authenticator from EKS access entries and Pod credentials.
Historical recordUnderstand Terraform and Azure AKS, then plan a cluster that can survive the failures your application must tolerate.
Sourced guidanceUnderstand what Terraform and Helmfile each manage and how to combine infrastructure creation with Helm application deployment.
Historical recordA historical May 2019 site update about last-updated labels, community comparison spreadsheets and the Helm chart catalogue.
Historical recordTry a broader term or reset the category and evidence filters.