Senior Cloud Ops Engineer
Listed on 2026-09-12
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer
Senior Cloud Ops Engineer
Location:
Denver, CO (Hybrid), San Antonio, TX (Hybrid), Brooklyn, NY or Remote (US Based).
As the Senior Cloud Ops Engineer, you are responsible for leading the production reliability of our cloud-hosted applications. You will act as the technical owner for specific cloud systems, transitioning them from reactive support to a proactive model that prioritizes automation, system optimization, and stability. This position is a powerhouse focused on supporting application delivery and ensuring the cloud infrastructure is a "paved road" for our development teams.
We recognize that talent comes in many forms. If you’re excited about this mission but your path hasn’t perfectly mirrored our requirements list, apply anyway. We value skills, grit, and potential.
Work Model:
We prioritize candidates in the Denver, CO, San Antonio, TX, and Brooklyn, NY area, but are open to remote talent.
- Locals: In person for sprint planning and quarterly planning meetings.
- Remote: Quarterly travel for team meetings.
- End-to-End System Ownership: Own the health and performance of assigned production systems, defining Service Level Objectives (SLOs) and building the monitoring/alerting necessary to meet them.
- Architect CI/CD & Automation: Design and maintain automated delivery pipelines (Infrastructure as Code) to ensure repeatable, secure feature deployments.
- Network Infrastructure Design: Manage scalable cloud network topologies, including transit gateways, VPCs, subnets, security groups, and route tables, ensuring low-latency and secure communication between distributed services.
- Eliminate Operational Toil: Proactively identify repetitive manual tasks and engineer automated solutions to prevent recurrence.
- Tier 3 Support & RCA: Act as the second escalation point for complex incidents, leading thorough root cause analysis (RCA) and implementing systemic, long-term fixes.
- Cross-Functional Collaboration: Partner with development teams to translate software requirements into robust operational designs, effectively communicating design decisions and resulting system trade-offs.
Required Qualifications
- Experience: 5+ years in cloud operations, systems administration, or Dev Ops, with a heavy focus on managing production workloads in mission‑critical environments.
- Technical Mastery: Advanced proficiency in AWS (EC2, VPC, IAM), Cloud Formation or Terraform, and CI/CD tools such as Git Hub Actions.
- Cloud Networking: Deep understanding of VPC design, subnetting, Route
53, and VPN/Direct Connect configurations; experience troubleshooting complex connectivity issues across multi-account or hybrid on‑prem/cloud environments. - Software Development: Proficient in one or more software languages such as Node.js, Python, or Go.
- Observability: Strong experience with observability stacks (e.g., Splunk, Data Dog, or Cloud Watch) including logging, tracing, and metrics tuning.
- Travel: You will be required to travel quarterly to Product Increment planning.
- U.S. Citizenship: Must be a U.S. Citizen and able to obtain a DoD NIPR network account and Common Access Card (CAC).
- Security Clearance: Must have, or be able to obtain, a Secret Clearance.
- Based in the Denver, CO, San Antonio, TX, or Brooklyn, NY area.
- Experience working within DoD, Air Force, or Federal contractor environments.
- AWS Certified Sys Ops Administrator or equivalent technical certification.
- Safety & Innovation: You embed security and reliability practices into daily work to drive continuous improvement and mitigate risk.
- People & Communication: You invite vigorous debate and offer "kindly blunt" feedback, always maintaining empathy and assuming noble intent.
- Integrity & Ethics: You build trust by…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).