Senior DevOps Engineer
Listed on 2026-09-24
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Infrastructure
Position Objective:
Bridge development and operations so that enterprise applications and IT infrastructure are highly available, secure, scalable, and delivered through reliable automation of IT infrastructure and enterprise applications. The role sits at the intersection of engineering velocity and operational excellence, improving how systems are built, deployed, monitored, and maintained. Intended for a practitioner who can translate platform and infrastructure needs into resilient, repeatable operational outcomes.
Connect software delivery and operational reliability in a way that strengthens uptime, security, scalability, and automated delivery across enterprise applications and infrastructure. The scope spans infrastructure operations, deployment automation, reliability enablement, and the continuous improvement of environments and delivery processes that support development teams and business-critical systems. This role exists because the organization needs a senior operator-builder who can reduce manual operational overhead, improve consistency across environments, and create dependable delivery mechanisms that support growth without compromising control or stability.
In team context, this person will work across development and operations boundaries, aligning engineering practices with operational needs and helping ensure that infrastructure and application delivery are both efficient and resilient.
- Own outcomes that improve the availability of enterprise applications and supporting infrastructure by strengthening operational reliability, reducing preventable downtime, and ensuring that production environments remain stable under normal and peak conditions.
- Drive secure and automated delivery outcomes by improving the mechanisms through which infrastructure and enterprise applications are built, tested, released, and changed, reducing manual intervention and increasing repeatability.
- Improve scalability across platforms and environments by evolving infrastructure patterns and operational practices so that systems can support growth in usage, complexity, and business demand without disproportionate increases in operational effort.
- Strengthen the connection between development and operations teams by establishing workflows, handoffs, and technical guardrails that enable faster delivery while preserving system integrity, operational visibility, and accountability.
- Increase the security posture of infrastructure and delivery processes by embedding operational controls, secure configuration practices, and disciplined change management into day-to-day engineering and platform workflows.
- Reduce operational risk by identifying fragile processes, single points of failure, and environment inconsistencies, then driving remediation efforts that improve resilience and recovery readiness.
- Improve observability and operational decision-making by ensuring that systems produce the telemetry, alerts, and service-level signals needed to detect issues early, respond effectively, and prioritize reliability investments.
- Enable development teams to deliver more efficiently by providing dependable infrastructure foundations, standardized deployment pathways, and operational support models that reduce friction across the software lifecycle.
- Drive continuous improvement in infrastructure and application operations by using incidents, recurring issues, and delivery bottlenecks as inputs for durable process, tooling, and architecture improvements.
- Contribute senior-level operational leadership by shaping best practices, influencing engineering standards, and helping teams make pragmatic trade-offs between speed, stability, security, and scale.
- Lead the deployment, health monitoring, and lifecycle management of Container Orchestration platforms (Open Shift/Kubernetes), ensuring alignment with the organization’s strategic objectives.
- Manage a portfolio of mixed OS environments (Linux and Windows), handling server troubleshooting, kernel tuning, and automated provisioning via Infrastructure as Code (Terraform, Ansible).
- Oversee public/hybrid cloud infrastructure management (Open Shift, Azure, AWS), optimizing cloud resources and ensuring high availability across zones.
- To develop and execute project plans for CI/CD pipeline optimizations, reducing build times via caching, parallel execution, and multi-stage builds.
- Implement Git Ops methodologies and execute zero-downtime releases utilizing Blue/Green or Canary…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).