Cloud Architect
Listed on 2026-09-04
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Requisition : 2068
Standard Weekly
Hours:
40.00
Location:ULA - Denver
Relocation:Yes
- Relocation may be available
Travel Requirements:10%
At ULA, success comes through the efforts of a strong, united team.
Thanks for your interest in United Launch Alliance, the world's most experienced and reliable space launch company! Successfully launching more than 155 consecutive missions with 100% mission success doesn't happen by accident. It's a testament to the commitment and dedication of our team of rocket scientists and support employees combined with the systems and processes we use to pull them together.
As a ULA employee, you'll have the opportunity to grow in your career while working in a team-oriented culture that combines technology, innovation, ingenuity and a commitment to the extraordinary. Whether you are in college just launching your career, or, have experience and want to come work with the best rocket team in the world, our unshakable unity yields stronger solutions and better results as we carry out our mission to save lives, explore the universe, and connect the world.
Our team is excited to meet you!
At ULA, the Cloud Architect 5 serves as the senior technical authority for Kubernetes platform engineering, multienclave cloud operations, and mission critical workload integration. This role defines enterprise architecture standards, drives reliability and security across clusters, and mentors Cloud Ops engineers while partnering closely with mission program teams.
Key responsibilities include:- Govern and maintain Kubernetes clusters supporting engineering workloads, CI/CD environments, and Cloud Ops automation.
- Design and manage cluster topology including nodes, node groups, autoscaling rules, taints/tole rations, and workload affinity.
- Own cross cluster networking and service mesh architecture, including Ingress, load balancers, and CNI strategy.
- Define and enforce Kubernetes security posture: RBAC models, OPA/Gatekeeper policies, pod security standards, and secrets management approaches.
- Architect and manage storage strategies involving CSI drivers, EBS/EFS provisioning, persistent volumes, and backup/restore workflows.
- Drive observability practices using Prometheus, Grafana, and distributed tracing systems.
- Lead incident response and reliability engineering for EKS clusters, addressing node failures, scheduler issues, cluster upgrades, and workload migration.
- Coach Cloud Ops team members to strengthen cluster operations, readiness, automation, and troubleshooting capabilities.
- Integrate Kubernetes platforms with mission program workflows, tool chains, and specialized workloads.
Bachelor
Required Years of ExperienceMinimum of 8 years of related work experience
Basic Qualifications- Bachelor's degree in a STEM (Science, Technology, Engineering, Mathematics) field from an accredited college or university
- 8+ years of cloud engineering or infrastructure experience, including handson production Kubernetes administration
- Deep expertise with EKS or equivalent managed Kubernetes platforms
- Strong proficiency in cloud networking, service mesh concepts, and core Kubernetes networking models
- Demonstrated experience implementing security controls such as RBAC, OPA/Gatekeeper, pod security standards, and secrets management systems
- Handson experience with CSI drivers, persistent storage architectures, and backup/restore design
- Proficiency with observability stacks including Prometheus, Grafana, and distributed tracing
- Proven ability to lead incident response, diagnose complex cluster issues, and ensure reliability in mission critical environments
- Strong communication skills and ability to coach and mentor junior and midlevel engineers
- Experience operating Kubernetes in multienclave or regulated/governed environments
- Deep knowledge of AWS cloud architecture including VPC design, IAM security, EKS internals, and cloud native networking
- Familiarity with Git Ops workflows, ArgoCD, Flux, or similar declarative deployment systems
- Strong background in automation using Terraform, Helm, Ansible, or Python
- Experience integrating Kubernetes systems with complex engineering…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).