Lead Platform Engineer – Cloud & Developer Platforms, Kubernetes
Listed on 2026-08-02
-
IT/Tech
SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations
Design, build, and operate Guardian's enterprise cloud application platforms and automation ecosystems. Own critical components of the Amazon EKS platform lifecycle. Support automation, system design, and operational tooling for application teams. Balance deep Kubernetes platform ownership with versatile engineering skills across scripting, Linux systems, automation, Helm charts, CI/CD, and serverless workflows. Focus on operational reliability, security-by-default patterns, and disciplined platform lifecycle management.
Mentor engineers and contribute to raising overall platform engineering, automation, and operational maturity.
- 7+ years of experience in platform engineering, cloud engineering, Dev Ops, SRE, or closely related roles.
- Strong hands-on experience with Kubernetes and Amazon EKS, including real-world production operations and upgrades.
- Broad engineering depth beyond Kubernetes, including:
- Linux systems troubleshooting
- Container runtimes and Docker
- Automation and scripting for operational workflows
- Hands-on experience with AWS Lambda, Step Functions, and serverless execution models.
- Strong proficiency in Infrastructure as Code using Terraform or equivalent tools.
- Experience building and operating CI/CD automation using Jenkins and/or Git Hub Actions.
- Strong scripting and automation skills using Python and Shell.
- Solid understanding of how application architecture decisions impact reliability, security, and operations.
- Experience designing and operating platforms consumed by multiple application teams at enterprise scale.
- Proven ability to troubleshoot complex, cross-layer issues spanning:
- Kubernetes platforms
- Linux and containers
- CI/CD and automation tooling
- Security and observability systems.
Demonstrates extensive experience in Platform Engineering and Cloud Engineering, with a strong focus on Kubernetes and Amazon EKS management, automation, and operational reliability. Proficient in Infrastructure as Code and CI/CD practices, ensuring secure and efficient application lifecycle management.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).