All posts by Kim McMahon

  • Reset

Found 2662 posts


Platform engineering maturity: From toolchain to self-service
Platform engineering maturity: From toolchain to self-service
Most platform engineering conversations tend to split into two rooms pretty quickly. The first room is full of teams who don’t have a platform yet. Scattered scripts, tribal knowledge, and every team is doing the same...
September 1, 2026 | Atulpriya Sharma | CNCF Ambassador and Platform Engineering TCG Organizer

OpenTelemetry has graduated… now what?
Project Maintainer Post OpenTelemetry has graduated… now what?
In case you missed it: OpenTelemetry (OTel) has officially achieved CNCF graduated status! It now stands proudly alongside amazing open source projects such as Kubernetes and Prometheus, to name just a few. It’s been a long...
August 31, 2026 | Adriana Villela (OTel Community Manager | OTel End User SIG Maintainer | CNCF Ambassador) and Reese Lee (OTel Community Manager | OTel End User SIG Maintainer)

Observability in Kubernetes: From metrics to meaning
Member Post Observability in Kubernetes: From metrics to meaning
Kubernetes made infrastructure more programmable, scalable, and resilient. It also made production systems harder to reason about. Workloads move, replicas churn, dependencies multiply, and a single user request can cross ingress, services, queues, storage, and background...
August 31, 2026 | Neel Shah, Stackgen

Scale before the spike: Predictive autoscaling for GPU workloads on Kubernetes
Scale before the spike: Predictive autoscaling for GPU workloads on Kubernetes
The 3 AM Call We got paged one Tuesday morning. A critical production service had crashed under traffic—not gradually degraded, but crashed. Hundreds of pending pods. Users were seeing 15–20% error rates. The incident postmortem was...
August 28, 2026 | Ramkumar Nagaraj (Golden Kubestronaut, Adobe) and Bingi Narasimha Karthik (Golden Kubestronaut, Adobe)

Your Kubernetes platform is ready for containers. Is it ready for AI?
Member Post Your Kubernetes platform is ready for containers. Is it ready for AI?
Kubernetes has given platform teams a consistent way to deploy, scale, and operate containerized applications. Now, many of those same teams are being asked to support AI. The transition is already underway. According to the CNCF...
August 28, 2026 | Kasia Hilborne, Vultr

Building an AI factory on Kubernetes
Ambassador Post Building an AI factory on Kubernetes
An AI factory is not just a model or a cluster. It is a pool of GPUs that many teams draw from at once: one team fine-tuning, another serving inference, a third running evaluations, all on...
August 27, 2026 | Hrittik Roy | CNCF Ambassador and Platform Advocate at vCluster

Governance guidance for CNCF projects: Choosing the right structure for your project’s size and stage
Community Post Governance guidance for CNCF projects: Choosing the right structure for your project’s size and stage
Clear patterns have emerged from governance reviews across 72 CNCF projects, distinguishing between what the CNCF requires at each maturity level versus what the data recommends for long-term project health. This post captures those patterns as...
August 26, 2026 | CNCF Technical Oversight Committee

The lazy developer’s guide to observing your own code
Ambassador Post The lazy developer’s guide to observing your own code
It’s no secret that developers are increasingly being asked to shift left. It seems there’s always something new to shift left on. And now developers are being asked to shift left on observability. This means that...
August 25, 2026 | Adriana Villela | OTel Community Manager & CNCF Ambassador) and Diana Todea | OTel Docs Approver & CNCF Ambassador

Stop trying to learn all of Kubernetes at once
Member Post Stop trying to learn all of Kubernetes at once
As a recovering VMware architect, it took me a little while to grasp Kubernetes. And I noticed I’m not alone in this.. From developers on our own team who need to get fluent in Kubernetes fast...
August 25, 2026 | Joep Piscaer, Portainer.io

Automating root cause analysis at scale: Multi-signal correlation for cloud native incident response
Member Post Automating root cause analysis at scale: Multi-signal correlation for cloud native incident response
The problem: Humans shouldn’t be correlation engines At Atlassian’s scale, hundreds of interconnected microservices distributed across multiple regions mean a production incident generates an overwhelming volume of telemetry. The problem is that finding the causal factor...
August 24, 2026 | Santosh Balaranganathan, Michael Yoo, James Moessis, James Kieltyka, Jason Lee, Lavender Neesham - Atlassian

How to turn slow queries into actionable reliability metrics with OpenTelemetry
Member Post How to turn slow queries into actionable reliability metrics with OpenTelemetry
Slow SQL queries degrade user experience, cause cascading failures, and turn simple operations into production incidents. The traditional fix? Collect more telemetry. But more telemetry means more things to look at, not necessarily more understanding. Instead...
August 21, 2026 | Severin Neumann, Causely

Announcing H1 2027 KCDs
Staff Post Announcing H1 2027 KCDs
Get ready to connect, learn, and innovate right in your backyard. Kubernetes Community Days (KCDs) are officially kicking off for H1! Supported by the Cloud Native Computing Foundation (CNCF), these community-organized events bring open source adopters...
August 20, 2026 | Helena Spease | Community & Outreach, CNCF

German ciphers, telegrams, and cloud native data sovereignty
Member Post German ciphers, telegrams, and cloud native data sovereignty
A lesson from 1917 In January 1917, Germany sent a secret telegram. It went to Mexico. The offer: join the war against the United States, and you can have Texas, Arizona and New Mexico back. The...
August 20, 2026 | James Hirst and Budhaditya Bhattacharya, Tyk

Kyverno is a platform primitive, not a security tool
Ambassador Post Kyverno is a platform primitive, not a security tool
Where does Kyverno live in your organization? I don’t mean which cluster! On which team’s slide deck does it show up? Whose budget line?  For most companies I’ve talked to, the answer is security. Kyverno is...
August 19, 2026 | Koray Oksay | CNCF Ambassador

Cloud Native platform sovereignty through multi-plane architecture
Ambassador Post Cloud Native platform sovereignty through multi-plane architecture
When people talk about cloud sovereignty, the conversation often starts with regions: where a workload runs and where its data is stored. But choosing a region is only part of the story. The architecture of the...
August 18, 2026 | Chamod Perera | CNCF Ambassador and Suvin Kodituwakku, Senior Software Engineer at WSO2

Welcome Falkey the Falco and Ky the Kyverno Pyrenees
Welcome Falkey the Falco and Ky the Kyverno Pyrenees
If you have yet to meet Phippy, she’s a friendly PHP app exploring the cloud native world with her pals. Over the last decade, Phippy’s circle has grown to include eighteen friends, with the newest members...
August 17, 2026 | Audra Montenegro | Community & Outreach, CNCF

Eleven minutes, zero humans: Building a self-healing Kubernetes upgrade pipeline on Kairos
Kubestronaut Post Eleven minutes, zero humans: Building a self-healing Kubernetes upgrade pipeline on Kairos
Once upon a time, upgrading a Kubernetes control plane meant staying awake for it. SSH into every node. Run the upgrade by hand. Watch etcd health the whole time, hoping quorum holds through every reboot. This...
August 14, 2026 | Olivier Calzi | CNCF Golden Kubestronaut

Lightweight Dragonfly Deployment: P2P Distribution Without the Database Stack
Project Maintainer Post Lightweight Dragonfly Deployment: P2P Distribution Without the Database Stack
Dragonfly speeds up file and container image distribution using peer-to-peer (P2P) technology, but a standard installation deploys several components and dependencies. Beyond the Scheduler, Seed Client, and Client that move data, a traditional setup requires a...
August 13, 2026 | Wenbo Qi (Gaius), Dragonfly Maintainer

LLMOps and platform engineering: Who should own the AI pipeline?
Member Post LLMOps and platform engineering: Who should own the AI pipeline?
A few years ago, getting a model into production meant a data scientist, a DevOps engineer, and a narrow set of tools: train it, test it, ship it, watch the dashboards. Large language models broke that...
August 13, 2026 | Daniel Bryant, Syntasso

Good apps aren’t born, they’re guided: Building observable policy as code
Community Post Good apps aren’t born, they’re guided: Building observable policy as code
As parents in tech, we’ve learned that neither children nor applications thrive without clear boundaries. There are no “good” or “bad” kids, just as there are no inherently “good” or “bad” applications, only behaviors shaped by...
August 12, 2026 | Diana Todea (Head of Developer Relations Engineering, VictoriaMetrics) & Cortney Nickerson (Community at Kyverno)