Site Reliability Engineer
Listed on 2026-09-04
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Spacecraft represent the most pressing unmet need across the entire aerospace industry. As more launch vehicles come online and the cost to orbit decreases, more companies launching payloads to space continue to emerge.
For the first time in history, this influx of payload companies combined with reduced launch costs has resulted in a massive increase in need for commercial spacecraft platforms, known as satellite buses. These buses hold the payloads of our customers and are flown on launch vehicles.
Apex manufactures these satellite buses at scale using a combination of software, vertical integration, and hardware that is designed for manufacturing. Our spacecraft enable the future of society: ranging from earth observation to communications and more.
We’d love for you to join us on our mission of providing humankind access to the galaxy beyond our planet.
About the RoleWe are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site network infrastructure that connects Apex facilities and mission operations across both commercial and secured environments. This is a hands-on engineering role for someone who is equally comfortable architecting a highly available Kubernetes platform, and defining the security posture that keeps it all compliant.
You will be a primary technical owner
, setting the standards that the rest of the team builds on.
Architect and build ground and site network infrastructure spanning commercial and secured or classified environments through the platform layer.
Design, deploy, and scale highly available Kubernetes clusters that support workloads across multiple security and classification levels.
Own the security posture of ground infrastructure, defining and enforcing controls, hardening, and compliance for sensitive government and commercial programs.
Build and maintain infrastructure as code using Terraform, Terragrunt, and similar tooling so that environments are repeatable, reviewable, and fast to stand up.
Establish observability across the ground network, including monitoring, metrics, logging, and alerting, so issues are caught before they reach a mission.
Design and run CI/CD pipelines and automation that streamline deployments and reduce manual, error-prone work.
Act as a primary site reliability engineer for ground systems, driving uptime, incident response, and reliability standards on programs where downtime is not an option.
Partner with security, mission operations, and program teams to translate requirements into infrastructure that is both compliant and operable.
Set architectural standards and mentor other engineers, raising the bar for how ground infrastructure is built at Apex.
Applicants must be U.S. persons as defined by U.S. export-control law.
5+ years of experience in site reliability, infrastructure, or network engineering, with a meaningful portion in aerospace, defense, satellite, or another mission-critical domain.
Hands-on experience architecting and operating ground network or large-scale production network infrastructure.
Deep expertise with Kubernetes and containers, including building and scaling high availability clusters in production.
Strong networking fundamentals across routing, switching, segmentation, and secure network design.
Proven experience with infrastructure as code with CI/CD, automation, and observability tooling such as Prometheus, Grafana, or similar.
Working knowledge of security and compliance for regulated or classified environments, and the judgment to design a defensible security posture.
Proficiency scripting and building tooling in Python, Go, or a comparable language.
Active Top Secret or Top Secret/SCI clearance
Prior experience supporting classified or government space programs or standing up infrastructure across multiple classification levels.
Familiarity with Git Ops workflows and tools such as ArgoCD, and with modern observability stacks including Mimir.
Apex believes in creating a work environment that you look forward to embracing every day. Our employees love working…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).