Lead DevOps Engineer
Listed on 2026-06-11
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Project Manager
Overview
My client is looking for a Lead Dev Ops Engineer to define and drive platform engineering practices that enable secure, scalable, and automated environments across data, AI, and application workloads. This is a hands‑on leadership role, responsible for setting standards across CI/CD, Infrastructure as Code, Kubernetes/Open Shift, observability, and security, while ensuring engineering teams can deliver efficiently within a governed and resilient platform.
You will act as a key bridge between Infrastructure, Security, Data Ops, and Platform Engineering teams, helping to establish automation, reliability, and compliance as core principles across the organisation.
- Define and enforce Dev Ops standards across CI/CD, IaC, and container orchestration.
- Design, build, and operate Kubernetes/Open Shift clusters in hybrid and on-prem environments.
- Own cluster lifecycle management, including upgrades, scaling, networking, storage integration, and resilience.
- Champion Dev Ops best practices across a hybrid architecture (Azure + on‑prem).
- Standardise and optimise CI/CD frameworks (Azure Dev Ops, Git Hub Actions).
- Drive adoption of Infrastructure as Code (Terraform, Ansible) across all environments.
- Enable automated, consistent, and scalable deployments across platforms.
- Establish observability frameworks (metrics, logs, traces, dashboards, alerting).
- Define SLOs/SLAs and ensure platform reliability and performance.
- Oversee incident management, runbooks, and automated remediation strategies.
- Embed security‑first practices including RBAC, secrets management, encryption, and vulnerability scanning.
- Ensure compliance with PHI/PII regulations through secure platform design.
- Support auditability and traceability across all deployments.
- Contribute to platform governance and design authority forums.
- Mentor and develop Dev Ops engineers, raising platform maturity and best practices.
- Collaborate with Data Ops, AI/ML, and Engineering teams to support scalable delivery.
- Partner with Infrastructure and Security teams to implement shared services and guardrails.
- 10+ years’ experience in Dev Ops / Platform Engineering, with 5+ years in a leadership role.
- Strong hands‑on expertise in Kubernetes and Open Shift (cluster operations and lifecycle management).
- Deep experience with CI/CD tools (Azure Dev Ops, Git Hub Actions).
- Advanced knowledge of Terraform and Ansible (Infrastructure as Code).
- Strong Linux systems experience in hybrid and on‑prem environments.
- Experience implementing observability tools (Prometheus, Grafana, ELK, Open Telemetry).
- Solid understanding of security frameworks (RBAC, secrets management, compliance).
- Proven track record delivering enterprise‑scale, cloud‑native platforms.
- Experience working with data and AI platforms (Databricks, Kafka, Airflow, MLflow).
- Exposure to hybrid infrastructure (Azure + on‑prem storage such as MinIO or VAST).
- Experience contributing to architecture governance or design councils.
Please apply to be contacted with further information. Salt is acting as an Employment Agency in relation to this vacancy.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).