Writing on Developer Experience, Cloud Native, AI and more.
All of my long-form thoughts on developer experience, cloud native, AI and more.
GPU utilization for LLM inference often sits as low as 5%, yet most teams respond by buying more hardware instead of tuning what they already have. Learn why the real bottleneck is configuration, not capacity, and how to unlock 2x throughput from existing GPUs.
Break the FinOps/SRE/Developer friction by transforming optimization into an automated, invisible platform capability. Achieve continuous cost and performance balance through GitOps.
A first look at survey data from 133 contributors across nearly 100 CNCF projects, revealing how AI tools like Claude Code and GitHub Copilot are reshaping open-source workflows and why governance policies haven't kept pace.
FinOps wants savings. SREs want reliability. Developers just want to ship. See how Akamas Insights aligns teams around one shared view of Kubernetes efficiency.
Design an edge observability pipeline using OpenTelemetry and Fluent Bit to capture, sample, and reliably process telemetry under strict bandwidth limits.
Cloud native optimization 2026 explores trends, challenges, and gaps in Kubernetes, FinOps, and platform engineering.
Stop treating cloud efficiency as a quarterly cleanup. Learn how to transform optimization into a core platform capability by bridging the gap between SRE reliability and FinOps goals.
Why 60% of Java workloads on Kubernetes are wasting money and the 4 lessons from Microsoft & Akamas to fix it.
Most Java apps on Kubernetes still use bad defaults. Learn how to fix them to cut waste, boost performance, and avoid hidden latency issues.
Kubernetes 1.35 makes resources mutable. Here is how In-Place Resizing works under the hood and where it still requires caution.
