
Most Kubernetes reliability discussions assume stable cloud networking, elastic infrastructure, and always-available control planes. But the real operational challenge starts when Kubernetes has to run closer to the edge: in constrained, bandwidth-limited, intermittently connected, or partially air-gapped environments where standard cloud assumptions no longer hold. This talk shares practical SRE lessons from operating managed Kubernetes environments across edge and hybrid infrastructure using technologies such as GKE Enterprise, EKS Anywhere, Bottlerocket OS, GitOps workflows, observability tooling, and production-grade operational runbooks. I will cover what changes when clusters are no longer “just in the cloud”: upgrade planning, image and artifact distribution, node OS lifecycle, observability under constrained bandwidth, incident response, storage behavior, and the tradeoffs between automation and safe human control. The session is intended to be candid and technical, focused on lessons learned rather than theory. Attendees will walk away with a practical mental model for designing and operating Kubernetes platforms in environments where connectivity is imperfect, upgrades require choreography, observability must be selective, and reliability depends as much on operational discipline as it does on tooling.
Edward Rodriguez is the founder of FUSSMOBILE, a Kubernetes and cloud managed services provider focused on production-grade platform operations, SRE practices, and hybrid infrastructure. He and his team work with enterprise Kubernetes environments across cloud and edge deployments, including EKS, GKE Enterprise, Rancher, EKS Anywhere, Container based OS, GitOps, observability, and operational automation. His work focuses on helping organizations run reliable Kubernetes platforms in complex real-world environments where availability, upgrade safety, and operational discipline matter.