Director, Engineering – Release Engineering, DevOps & SRE
Listed on 2026-08-09
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Infrastructure, Systems Engineer
We’re looking for a hands‑on engineering leader to build and own the Release Engineering, Sec Dev Ops , and Site Reliability Engineering (SRE) functions for Infinia. This is a foundational role: you’ll define how our software is built, secured, released, and kept running at enterprise scale.
If you thrive at the intersection of infrastructure automation, operational excellence, and team building — and you want your work to directly shape the reliability and velocity of a market‑leading storage platform — this is your role.
What You’ll DoRelease Engineering
Own the end-to-end release pipeline — build systems, artifact management, versioning, and release gating
Drive automation of build, test, and packaging workflows to maximize developer velocity
Define release cadence, branching strategies, and code freeze processes with engineering leadership
Architect and operate scalable CI/CD pipelines and developer tool chains for on‑prem and cloud
Embed security into the SDLC — SAST, DAST, dependency scanning, secrets management
Champion automation‑first Dev Ops practices from commit to production delivery
Define SLOs, SLIs, and error budgets; own incident response and post‑mortem culture
Drive observability — logging, metrics, distributed tracing — across the platform
Influence reliability and operability early, at design and code‑review stages
Lead capacity planning and infrastructure scaling decisions
Build, mentor, and grow teams across all three disciplines in a global, distributed org
Communicate roadmap and risks to senior leadership; partner cross‑functionally with product, QA, and security
12+ years in Dev Ops, Release Engineering, or SRE — with 5+ years managing engineering teams
Deep experience with CI/CD platforms (Git Hub Actions, Jenkins, Git Lab CI, Tekton, or similar)
Infrastructure-as-code fluency:
Terraform, Ansible, Pulumi, or equivalentContainer orchestration expertise:
Kubernetes and Docker in productionPractical Dev Sec Ops background — you’ve shipped security tooling, not just talked about it
SRE chops: SLO/SLI design, on‑call frameworks, observability stacks (Prometheus, Grafana, Open Telemetry)
Strong communicator — you can translate complex technical tradeoffs for exec and non‑technical audiences
Experience with storage, distributed systems, or infrastructure software
Familiarity with HPC, AI/ML infrastructure, or enterprise data platforms
Background scaling Dev Ops in high‑growth, globally distributed environments
Exposure to SOC 2, ISO 27001, or FedRAMP compliance requirements
Work on infrastructure that underpins some of the world’s most demanding AI and research workloads
Greenfield opportunity to shape Release Engineering, Dev Ops, and SRE from the ground up for Infinia
Collaborative, engineering‑first culture that values autonomy, technical depth, and continuous learning
Competitive compensation, remote‑first flexibility, and a team that genuinely enjoys solving hard problems
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).