Spotlight
Samir Savla
This case study shows how a telemetry gateway handles millions of events per second by rejecting overload at the edge and using two levels of autoscaling on Kubernetes.
Devarshi Shimpi
This tutorial shows you how to secure Kubernetes with five Kyverno policies for resource limits, image tags, labels, restricted privileges, and signed container images through YAML examples.
Nameeta Limje
This case study shows why a Kubernetes liveness probe must not depend on downstream services after one caused payment pods to restart repeatedly, and how separate readiness and liveness checks avoid the same failure.
Prasad Ekke
This article explains how Go 1.25 makes GOMAXPROCS respect container CPU limits, why GOMEMLIMIT still needs manual setup, and how you can prevent CPU throttling and OOM kills in Kubernetes.
Tools and utilities
HelmSharp is a managed .NET library for rendering Helm charts and managing Kubernetes releases in-process, including packaging, repositories, upgrades, rollbacks, and storage, without the Helm executable.
KubeIntellect lets you troubleshoot Kubernetes in plain English by checking kubectl, Prometheus, and Loki, while requiring human approval before it changes the cluster.
obleth is a multi-tenant AI gateway that fairly shares GPU capacity, routes OpenAI-compatible requests to the right model, starts Slurm jobs, and tracks cost and energy use.
Kagent is a Kubernetes-native framework for building and running AI agents as custom resources, with multiple model providers, reusable MCP tools, a UI, CLI, and tracing.
WaaS creates browser-accessible Linux and Windows desktops as Kubernetes resources, with GitOps workflows, secure remote access, quotas, and OIDC and RBAC controls.
Events starting soon
October 14, 2026
Location: Sydney, AU and virtual
This is a free event.
October 14, 2026
Location: New York, NY, USA
This event requires an entrance fee
October 14, 2026
This is a virtual event
This is a free event.
October 14, 2026
Location: Trondheim, NO
This event requires an entrance fee
October 14, 2026
Location: Dubai, AE
This event requires an entrance fee
October 15, 2026
This is a virtual event
This is a free event.
Non-deterministic agents pose specific challenges for platform teams in observability, state management, governance, and trust.
Mauricio (Salaboy) Salatino explains why agentic applications behave like distributed multi-agent systems. His test assigned agents to take an order, cook the pizza, deliver it, and charge the customer. One order crossed 15 containers and produced 200 traces.
In this interview:
Learn from production
Hakan Kaya
This case study shows how an AKS-based GitOps platform gave on-premises clusters workload identity by publishing static OIDC discovery and JWKS files, avoiding API server changes after provisioning.
Elad Cohen
This case study shows how WSC Sports rebuilt their entire deployment model from helm upgrade pipelines to full GitOps using ArgoCD ApplicationSets, covering:
Nic Cope
This case study explains how Modelplane was built entirely as a Crossplane configuration, with no custom controllers, to turn GPUs spread across clouds into one fleet behind a stable inference API.
Aswin A
This case study shows how a team ran one Kafka cluster stretched across three separate Kubernetes clusters with Strimzi, Submariner and Cilium, and kept it alive when a whole cluster went down.
Matching jobs
DevOps Engineer with Clara
Salary: $122.4K to $440K a year
Location: remote from
Tech stack: Kubernetes, AWS, Java, Javascript, Python, Ruby, Typescript, Redis, PostgreSQL, MySQL
DevOps Engineer with Qualysoft
Salary: $93.6K to $258.5K a year
Location: based in the office in Bucharest, RO
Tech stack: Kubernetes, AWS, Docker, Python, Terraform, Gitlab, Ansible, Grafana, Prometheus, Loki
DevOps Engineer with Quindar
Salary: $69.55K to $330K a year
Location: based in the office in Denver, CO, USA
Tech stack: Kubernetes, Shell
DevOps Engineer with SAP IT Business Systeme
Salary: $93.6K to $258.5K a year
Location: based in the office (and remote from home) in Timisoara, RO
Tech stack: Kubernetes
DevOps Engineer with Swarm Aero
Salary: $69.55K to $330K a year
Location: remote from
Tech stack: Kubernetes, Go, Python, Terraform
Build something
Tayeb gasmi
This tutorial shows how to move a daily Spring Boot batch job out of the web app and run it as a Kubernetes CronJob so memory spikes do not take down the API.
PuneetPahuja
This tutorial teaches how to secure Kubernetes AI platforms with RBAC, NetworkPolicy, secrets, resource limits, Kyverno, OPA Gatekeeper, and automatic policy checks.
Sowmithdurusoju
This tutorial shows how to serve and scale large language models on Kubernetes with KServe, use KEDA for demand-based scaling, and run inference with vLLM.
Sabbir Ahmed
This tutorial teaches how to preserve the real client IP behind an Istio ambient gateway with externalTrafficPolicy: Local and safely read X-Forwarded-For in Go.
Call for Papers closing soon
1
days
This is a virtual event
Online conference organized by Anaconda.
The conference starts on the 22 October 2026.
2
days
Location: San Francisco, CA, USA
In-person conference organized by SOFTWARE 3.0.
The conference starts on the 12 November 2026.
2
days
KubeCon + CloudNativeCon Europe 2027
Location: Barcelona, ES
In-person conference organized by CNCF.
The conference starts on the 18 March 2027.
3
days
Location: Utrecht, NL
In-person conference organized by S&S Media.
The conference starts on the 11 March 2027.
6
days
Global Summit on Data Science and Cloud Computing
Location: Rome, IT
In-person conference organized by Noveltics Group.
The conference starts on the 20 October 2026.
6
days
GENAIX: Global Generative & Agentic AI Summit
Location: Singapore, SG
In-person conference organized by The People Events.
The conference starts on the 16 April 2027.
9
days
Cloud Native AI + Inference Day Europe
Location: Barcelona, ES
In-person conference organized by CNCF.
The conference starts on the 15 March 2027.
More articles
Mahesh Devendran
This article explains how to keep public EKS load balancer protected with AWS WAF, Shield Advanced, and Route 53 through automatic checks and cleanup.
Shobhit Paliwal
This article explains why common GPU utilization numbers can hide idle compute and how allocation, active time, and useful model work reveal different kinds of GPU waste.
Bart Rijnders
This case study shows how Albert Heijn built a centralized LGTM-stack observability platform across 1800 engineers, replacing ELK, Nagios, Dynatrace and Azure Monitor setups.
Heba Elayoty
This article explains how Kubernetes 1.37 gives controllers shared APIs and a Go library for workload-aware scheduling, including gang scheduling and grouped workloads.