Director, Engineering – Release Engineering, DevOps & SRE
Listed on 2026-07-31
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
We are on our way to being the first company to power 1 MILLION GPUs and want world-class talent to join our amazing team!
The world is moving faster than ever, and yet, it will never move this slowly again. We are at the forefront of an incredible technological revolution but, at its core, it is fueled by incredible people. People like you.
1,000
Global Employees
11k
Happy Customers
16
International Offices
We are the world’s leading data intelligence platform that reliably accelerates massive datasets for actionable real-time insights. Join our team to help the best and brightest minds tackle the world’s biggest challenges in business, science, medicine, academia and government.
Do What Can’t be DoneFor the past 20 years, our team has kept us at the forefront of storage technology and has provided the foundation for enabling researchers to push the limits of “what can be done.”
These innovations take research and discovery to the next level, enabling them to discover cures to disease, observe global warming patterns, model innovative automotive and aerospace designs, discover new sources of energy, make communities safer, and accelerate business results across a wide variety of industries.
At DDN, we understand our customers’ diverse needs. Whether you’re a data scientist, IT professional, executive, or researcher, our solutions empower you with cutting-edge technology and unparalleled support.
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
Tuition Reimbursement
Employee Referral Program
Excellent Medical, Dental and Vision Benefits
Paid Leave Programs
Anniversary and Recognition Awards
Director, Engineering – Release Engineering, Dev Ops & SRE LocationEmployment Type
Full time
Location TypeHybrid
We're looking for a hands-on engineering leader to build and own the Release Engineering, Sec Dev Ops , and Site Reliability Engineering (SRE) functions for Infinia. This is a foundational role: you'll define how our software is built, secured, released, and kept running at enterprise scale.
If you thrive at the intersection of infrastructure automation, operational excellence, and team building — and you want your work to directly shape the reliability and velocity of a market‑leading storage platform — this is your role.
What You'll Do Release EngineeringOwn the end-to-end release pipeline — build systems, artifact management, versioning, and release gating
Drive automation of build, test, and packaging workflows to maximize developer velocity
Define release cadence, branching strategies, and code freeze processes with engineering leadership
Dev Ops and Sec Dev OpsArchitect and operate scalable CI/CD pipelines and developer tool chains for on-prem and cloud
Embed security into the SDLC — SAST, DAST, dependency scanning, secrets management
Champion automation‑first Dev Ops practices from commit to production delivery
Site Reliability EngineeringDefine SLOs, SLIs, and error budgets; own incident response and post‑mortem culture
Drive observability — logging, metrics, distributed tracing — across the platform
Influence reliability and operability early, at design and code‑review stages
Lead capacity planning and infrastructure scaling decisions
LeadershipBuild, mentor, and grow teams across all three disciplines in a global, distributed org
Communicate roadmap and risks to senior leadership; partner cross‑functionally with product, QA, and security
What You Bring12+ years in Dev Ops, Release Engineering, or SRE — with 5+ years managing engineering teams
Deep experience with CI/CD platforms (Git Hub Actions, Jenkins, Git Lab CI, Tekton, or similar)
Infrastructure‑as‑code fluency:
Terraform, Ansible, Pulumi, or equivalent
Container orchestration expertise:
Kubernetes and Docker in production
Practical Dev Sec Ops background — you've shipped security tooling, not just talked about it
SRE chops: SLO/SLI design, on‑call frameworks, observability stacks (Prometheus, Grafana, Open Telemetry)
Strong communicator — you can translate complex technical tradeoffs for exec and non‑technical audiences
Bonus PointsExperience with storage, distributed systems, or…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).