Senior Engineer - Cloud Operations
Listed on 2026-07-14
-
IT/Tech
AWS, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Overview
Hypori is seeking a Senior Cloud Ops Engineer to architect, build, and evolve the secure, scalable cloud infrastructure that powers our SaaS platform. This senior technical leadership role combines deep hands‑on cloud operations expertise with architectural judgment to define how the platform scales, secures, and recovers — across both AWS and AWS Gov Cloud environments.
You will independently own ambiguous, high‑impact infrastructure problems, guide technical direction for the team, and act as a senior escalation point when production systems are on the line. This builder‑designer hybrid role is equally comfortable writing Terraform and designing infrastructure modernization solutions.
You will drive the strategy and execution of Infrastructure as Code, observability, CI/CD, and operational frameworks, while mentoring engineers and raising the technical bar across the organization. This role sits at the intersection of Dev Ops, SRE, and cloud architecture, with a strong emphasis on engineering rigor, operational maturity, and continuous improvement in a mission‑critical, regulated environment.
Responsibilities- Own the architecture and operation of secure, scalable, highly available AWS and AWS Gov Cloud infrastructure supporting the Hypori SaaS platform end to end.
- Set the technical direction for Infrastructure as Code (Terraform, Open Tofu, Cloud Formation), establishing org‑wide standards, reusable modules, and governance patterns other engineers build on.
- Drive engineering tasks for platform reliability, scalability, and capacity; helping define SLA/SLO targets.
- Lead the end‑to‑end observability strategy — monitoring, logging, alerting, telemetry — across Datadog, Splunk, Prometheus, Grafana, and/or ELK.
- Serve as incident commander for high‑severity production events; drive root cause analysis and lead systemic, multi‑phase remediation — not just fixes, but prevention.
- Design and implement CI/CD platforms and deployment automation (Git Hub Actions, Jenkins, AWS Code Pipeline), championing progressive delivery patterns as the org standard.
- Lead the operation and evolution of containerized workloads on Kubernetes and Docker, making the orchestration and scaling trade‑off calls for the platform.
- Proactively identify architectural scalability and performance bottlenecks before they become incidents, and design the systems that eliminate them.
- Partner directly with Security and Compliance leadership to implement secure‑by‑design infrastructure supporting FedRAMP, DoD IL5, and SOC 2 Type 2 requirements.
- Participate in owning disaster recovery and business continuity — design, test, and continuously validate RTO/RPO targets.
- Identify and develop solutions for cloud cost optimization as a discipline, not a cleanup task — identifying inefficiencies at the design level and implementing measurable, sustained cost controls.
- Mentor engineers, lead design reviews, and shape the technical standards the broader team is measured against.
- Lead cross‑team resolution of systemic operational issues, delivering road‑mapped improvements with measurable, reported impact.
- Represent Hypori in technical conversations with external cloud and infrastructure vendors, holding them to performance, reliability, and security commitments.
- Participate in a 24/7 on‑call rotation.
- Bachelor’s in Computer Science, Engineering, or equivalent hands‑on experience in cloud infrastructure, SRE, or systems engineering.
- 8+ years in Cloud Ops, SRE, Dev Ops, or Infrastructure Engineering, with a demonstrated track record of progression into senior/architectural ownership.
- Deep AWS expertise (EC2, VPC, IAM, S3, EBS, EFS, Route 53, Cloud Watch); AWS Gov Cloud experience strongly preferred.
- Proven experience designing multi‑account AWS environments and landing zone frameworks (e.g., AWS LZA), including isolation, governance, and guardrail design.
- Expert‑level Infrastructure as Code (Terraform and/or Cloud Formation), with a track record of building reusable modules and org‑wide standards, not just consuming them.
- Production‑grade Kubernetes and Docker experience, with the judgment to make orchestration and architectural trade‑off…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).