SrStaffDevOps Engineer
Listed on 2026-07-24
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, AWS
Overview
Vantor is forging the new frontier of spatial intelligence, helping decision makers and operators navigate what’s happening now and shape what’s coming next. Vantor is a place for problem solvers, changemakers, and go-getters—where people work together to help our customers see the world differently. This position is hybrid with three days a week on-site in Westminster, CO. To be eligible, you must be a U.S. citizen and able to obtain a U.S. Government security clearance.
Export Control/ITAR rules may apply.
Vantor’s Platform Dev Ops team is growing. We are seeking a Senior Staff Dev Ops Engineer with Site Reliability Engineering experience to support the build, deployment, reliability, and operations of the Vantor Hub software suite. The team designs, secures, and operates a wide range of custom and third-party software solutions deployed in AWS and governs dozens of AWS accounts across commercial and government environments.
Responsibilities- Lead reliability engineering efforts for the Vantor Hub platform and related infrastructure services.
- Design, implement, and maintain scalable CI/CD pipelines for software and infrastructure delivery.
- Build, operate, and improve cloud infrastructure automation using Terraform, Cloud Formation, Kubernetes, Docker, and AWS-native services.
- Troubleshoot complex infrastructure, deployment, networking, performance, reliability, and production issues.
- Improve service availability, scalability, maintainability, reliability, security, and overall operational readiness.
- Design and operate highly available, resilient, observable, and secure infrastructure across commercial and government AWS environments.
- Partner with engineering teams to improve deployment safety, rollback capabilities, observability, production readiness, and operational support.
- Participate in a team on-call rotation (approximately one week every twelve weeks) for production reliability and incident response.
- Leverage AI development tools to accelerate software design, implementation, testing, and documentation while maintaining reliability and security.
- Contribute to shared engineering standards for responsible AI-assisted development, including validation practices, documentation expectations, and review patterns.
- Use AI-assisted tooling to accelerate infrastructure automation, documentation, runbooks, test scaffolding, incident analysis, and repeatable operational workflows; validate outputs through peer review, testing, and security practices.
- Bachelor’s degree in Software Engineering, Computer Science, or related field, or equivalent experience.
- 8+ years of experience in Dev Ops, Platform Engineering, Site Reliability Engineering, or cloud operations with production ownership.
- Proven experience owning, operating, and improving production systems in cloud-based environments.
- Strong experience with Site Reliability Engineering practices (SLOs/SLIs, error budgets, incident response, post-incident reviews, reliability metrics, automation).
- Strong proficiency with AWS, including multi-account production workloads.
- Strong proficiency with Docker and Kubernetes; experience with infrastructure-as-code tools (Terraform or Cloud Formation).
- Proficiency in scripting languages (Python, Bash, or Power Shell).
- Solid understanding of networking concepts and protocols.
- Strong communication and collaboration skills.
- U.S. citizenship and willingness to obtain a U.S. Government security clearance.
- Experience with high-availability systems, database replication, backup/restore, disaster recovery, and business-continuity planning.
- Experience with zero-downtime deployment strategies (blue/green, canary, rolling deployments, feature flags, automated rollback).
- Experience with observability and incident management platforms (Prometheus, Grafana, Cloud Watch, ELK, Pager Duty, etc.).
- Experience with secure cloud operations, least-privilege IAM, secrets management, vulnerability remediation, and audit-ready infrastructure.
- Experience supporting government, regulated, or compliance-driven environments.
- Practical experience using AI-assisted development or automation tools with a focus on validation,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).