How we reduced cloud monitoring costs by over 50% on GKE by architecting a lightweight, highly available Prometheus + Thanos cluster instead of managed Prometheus.
How we evaluated Longhorn for RWX shared storage in Kubernetes, diagnosed stability bottlenecks in dynamic clusters, and migrated to a hardened off-cluster NFS solution.
How we implemented request-body-aware dynamic routing in Ingress Nginx using custom Lua plugins to simplify complex microservice migrations without downtime.
How we designed and automated on-demand ephemeral staging environments in a single Kubernetes cluster to enable reliable pre-release integration testing across microservices.