Cloud Platform Engineer
Listed on 2026-07-14
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, AWS
About Us:
Havoc is a leader in all-domain collaborative autonomy. Its software-defined hardware approach powers military and commercial-grade autonomous systems across sea, air, and land to sense, decide, and act together in complex and contested environments. Havoc connects assets, enabling them to share information, adapt in real time, and continue operating even when communications are disrupted or denied. Havoc optimizes mission performance and minimizes human risk.
Havoc was founded in 2024 and headquartered in Providence, Rhode Island. Learn more at Havoc:
All-Domain Collaborative Autonomy.
We are seeking a Staff Cloud Platform Engineer to drive the design and operation of the cloud platform that powers our autonomous systems.
In this role, you will be a hands‑on staff‑level individual contributor, owning the design of secure, reliable, and cost‑effective cloud infrastructure while helping set the technical direction for how we provision, deploy, and operate services.
You will work closely with backend, SRE, data, autonomy, and security teams to deliver platform capabilities that support mission‑critical workloads in real‑world environments.
This role is ideal for someone who combines deep cloud and infrastructure expertise with strong technical judgment, thrives in fast‑paced, high‑ownership environments, and is energized by complex and ambiguous platform challenges.
Key Responsibilities Technical OwnershipDrive execution and technical excellence across cloud platform projects
Set infrastructure standards, conduct design and code reviews, and guide architectural decisions
Shape the cloud platform roadmap in partnership with product and engineering leadership
Foster a culture of ownership, quality, automation, and continuous improvement
Design, build, and operate scalable AWS infrastructure using Infrastructure as Code
Build and maintain Kubernetes‑based platforms, including EKS, and container workflows for application teams
Build self‑service tooling and paved paths that enable teams to deploy and operate services safely
Ensure the platform reliably supports both application and data workloads
Design platform capabilities with reliability, scalability, security, and cost efficiency in mind
Optimize cloud systems for performance, scalability, resilience, and cost efficiency
Identify and resolve infrastructure bottlenecks, scaling limits, and reliability risks
Establish reliability practices, including SLIs, SLOs, error budgets, and incident response
Improve observability, maintainability, and operational readiness across cloud platform services
Partner with backend, SRE, data, autonomy, and security teams to deliver platform capabilities across the stack
Support end‑to‑end system design from edge devices to cloud infrastructure
Coordinate across teams to build and maintain release processes and CI/CD pipelines
Help align platform capabilities with product, engineering, security, and mission needs
Enforce secure cloud practices, including IAM least privilege, secrets management, and network segmentation
Maintain high standards for automation, testing, observability, documentation, and infrastructure reliability
Support technical reviews, design discussions, and production readiness efforts
Build systems with security, scalability, and long‑term maintainability in mind
10+ years of experience in cloud platform, infrastructure, Dev Ops, or related engineering roles
Deep expertise with AWS core services, including compute, networking, storage, IAM, and managed data services
Strong experience with Infrastructure as Code
Hands‑on experience operating Kubernetes, including EKS, and containerized environments
Strong experience building CI/CD pipelines and deployment automation
Proficiency in Go, Python, or similar languages for platform tooling and automation
Solid understanding of cloud networking, security, and distributed systems fundamentals
Ability to drive projects from concept through production
Strong technical judgment and ability to operate independently in…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).