Operations Engineer; Senior; DevOps & Cloud
Listed on 2026-09-14
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Infrastructure, Systems Administrator
Introduction
An exciting opportunity is available for an experienced Senior Operations Engineer to join a highly technical, internationally integrated IT environment.
The successful candidate will play a key role in ensuring the reliability, availability, security and operational stability of enterprise applications and infrastructure across cloud and on-premises environments.
This position requires a strong combination of Dev Ops engineering, infrastructure automation, cloud operations, CI/CD, Infrastructure as Code, observability, incident management and IT Service Management
.
- Lead configuration management and infrastructure automation initiatives.
- Design, implement and maintain Infrastructure as Code (IaC) and automation solutions.
- Build, maintain and optimise CI/CD pipelines
. - Support and improve cloud and on-premises production environments.
- Drive effective change, release and transition management
. - Ensure configuration, release and change governance standards are maintained.
- Act as a senior escalation point for complex production incidents.
- Lead troubleshooting,
Root Cause Analysis (RCA) and post-incident improvement activities. - Implement and enhance monitoring, logging and observability capabilities.
- Collaborate with infrastructure and development teams to ensure solutions are designed for operational reliability and supportability.
- Develop and maintain runbooks, operational procedures and technical documentation
. - Monitor operational KPIs, service quality, availability and reliability.
- Support IT Service Continuity
, resilience and disaster-recovery requirements. - Facilitate technical onboarding, knowledge transfer and training.
- Mentor and support junior Operations/Dev Ops engineers.
- Drive continuous improvement and automation to reduce manual operational effort.
- Ensure infrastructure and applications comply with security, lifecycle and governance requirements.
- Maintain effective communication with technical and business stakeholders.
- Support SLA/OLA and availability requirements.
- Collaborate with geographically distributed and international Dev Ops teams.
ssential Technical Skills
Applicants should have strong practical experience in most of the following:
- Configuration Management: Ansible, Puppet, Chef and/or Salt
- Scripting & Automation: Python, Bash and/or Power Shell
- CI/CD: Jenkins, Git Lab CI, Git Hub Actions or equivalent
- Infrastructure as Code: Terraform, AWS Cloud Formation or equivalent
- Cloud Platforms: AWS, Azure or equivalent
- Monitoring & Observability: Prometheus, Grafana, ELK/EFK or comparable enterprise monitoring solutions
- Linux and/or Windows infrastructure administration
- Production troubleshooting and Root Cause Analysis
- Change, Release and Transition Management
- ITSM/ITIL processes and operational governance
- Git/version-control environments
Advantageous Skills
Experience in the following will be beneficial:
- Docker and Kubernetes
- Dev Ops and Site Reliability Engineering (SRE) practices
- SLIs, SLOs and error budgets
- Middleware and enterprise platform technologies
- Infrastructure security, hardening and lifecycle management
- Data-centre infrastructure, networking, storage and servers
- Highly regulated enterprise environments, particularly financial services or automotive
- Deployment and testing automation
- Data pipelines and/or ML deployment environments
- Communities of Practice or Centres of Excellence
- Mentoring/coaching junior technical professionals
- German language capability
Qualifications & Experience
- Relevant IT degree, diploma or equivalent qualification
- Approximately 6–10 years' broad IT experience
- At least 3–5 years' experience in IT Operations, Dev Ops, SRE, Cloud Operations or a closely related environment
- Proven experience supporting complex production environments
- Demonstrable experience with transition, change and release management
- ITIL Foundation and Service Transition experience/qualification or equivalent is advantageous/highly preferred
- Strong troubleshooting and analytical capability
- Excellent stakeholder communication skills
- Demonstrated ability to operate effectively within multidisciplinary technical teams
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).