Komodor Blog

All articles
Page 1
Welcome to Komodor's blog, your go-to resource for insights on all things Kubernetes. Stay tuned for expert advice, in-depth tutorials, and the latest industry trends to help you throughout your K8s journey.

Komodor Appoints Ziv Harfenist as Chief Financial Officer 

2 min read

Komodor, the autonomous AI SRE platform for cloud-native infrastructure and operations, today announced the appointment of Ziv Harfenist as Chief Financial Officer (CFO) and the promotion of Yogev Goldis to Chief People Officer (CPO).

AI SRE in Practice: Diagnosing Configuration Drift in Deployment Failures

5 min read

Part 3 of our AI SRE in Practice Series. In this part we cover how an AI SRE helps diagnose configuration drift in deployment failures.

AI SRE in Practice: Resolving GPU Hardware Failures in Seconds

4 min read

Part 2 of the AI SRE in Practice Series. In this post we discuss: Resolving GPU Hardware Failures in Seconds

When is it ok or not ok to trust AI SRE with your production reliability?

3 min read

This series demonstrates what AI SRE trained on real workloads actually looks like in practice. We're going to walk through real troubleshooting scenarios that our customers encounter daily, showing the before and after of AI-powered investigations.

From Promise to Practice: What Real AI SRE Can Actually Do When Production Breaks

4 min read

This series demonstrates what AI SRE trained on real workloads actually looks like in practice. We're going to walk through real troubleshooting scenarios that our customers encounter daily, showing the before and after of AI-powered investigations.

7 Kubernetes Predictions for 2026 – AI Will Push SRE to its Limit

2 min read

SRE teams are about to feel even more pressure. GPU-heavy computing is breaking the assumptions today's clusters were built on, while enterprises are beginning to trust autonomous operations and cost pressure is pushing consolidation across the cloud-infrastructure stack. Based on these forces, here are my 2026 Kubernetes predictions as well as some best practice recommendations to help platform teams prepare for what reliable operations will mean next year. 

Kubernetes v1.35: The Release That Tackles the Industry’s $100 Billion Waste Problem

7 min read

There's a bigger story here that every platform team needs to understand: K8s is finally acknowledging that cluster utilization is fundamentally broken.

komodor-at-kubecon

KubeCon Atlanta 2025 & the AI-Native Shift

6 min read

If you missed the event or couldn't attend every session, here are the talks that captured some interesting (IMO) technical shifts happening in the Kubernetes ecosystem.

building-trust-ai-kubernetes

Building Trust in AI-Powered Kubernetes Ops: Why “Good Enough” Is a Production Killer

3 min read

We aren't building a chatbot to suggest recipes. We are building systems that, armed with kubectl permissions, have the potential to take down production with a single, wrong command. This demands we elevate our standards far beyond "good enough."