Lead Engineer – Infrastructure & Cloud Engineering
Listed on 2026-08-24
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer, SRE/Site Reliability, IT Infrastructure
About the Client
Our client is a high-growth technology powerhouse operating at the forefront of AI, cloud and digital infrastructure. With an ambitious vision for the future of technology, the organization is bringing together world-class engineering talent, advanced infrastructure and next-generation platforms to solve some of the most complex technology challenges at scale.
This is an opportunity to join an environment where deep engineering, ambitious technology and genuine innovationcome together, with the chance to influence the architecture of critical platforms and help shape what comes next.
The OpportunityWe are looking for an experienced Lead Engineer – Infrastructure & Cloud Engineering to provide technical leadership across private cloud, virtualization, container and observability platforms.
The role combines deep hands-on engineering expertise with technical leadership, helping shape platform strategy, engineering standards, operational excellence and the adoption of emerging technologies.
You will work closely with architecture, security, platform engineering, SRE and operations teams to build, operate and continuously improve highly available infrastructure platforms supporting demanding technology environments.
This is a hands-on technical leadership role for someone who enjoys solving complex infrastructure challenges, leading engineering initiatives and helping teams adopt modern approaches to cloud, automation and platform engineering.
Key Responsibilities- Lead the technical design, implementation and lifecycle management of large-scale private cloud, virtualization and container platforms
, including Open Stack and Open Shift or comparable technologies. - Define and drive engineering standards, design principles, operational best practices and technical governance across infrastructure and cloud platforms.
- Provide technical leadership and mentorship to engineering teams, supporting technical decision-making, complex problem-solving, knowledge sharing and continuous development.
- Lead the strategy, design and adoption of observability platforms
, covering metrics, logs, traces, dashboards, alerting, service health and SLO/SLI capabilities. - Drive the adoption of AI-assisted operations and intelligent automation to improve operational efficiency, incident management, root cause analysis, platform reliability and service automation.
- Lead the evaluation, integration and adoption of new technologies, products and architectural approaches in collaboration with Architecture, Product Engineering, SRE, Security and Operations teams.
- Provide technical oversight for complex platform upgrades, migrations, production changes and infrastructure transformation initiatives.
- Act as a senior technical escalation point for critical production incidents, leading complex troubleshooting, incident response, root cause analysis and service recovery.
- Drive capacity planning, scalability, performance optimisation, resilience engineering and operational readiness across infrastructure platforms.
- Ensure observability, automation and operational tooling are effectively integrated with service management and incident management processes.
- Work closely with security, compliance and risk teams to ensure infrastructure platforms meet appropriate cybersecurity, governance and regulatory requirements.
- Define and champion automation strategies using Infrastructure as Code, Git Ops, CI/CD and platform engineering practices.
- Lead the development and maintenance of technical standards, architecture documentation, operational procedures, design guides and knowledge resources.
- Collaborate with leadership teams on infrastructure strategy, roadmap planning, technology evaluation and long-term platform evolution.
- Bachelor's or Master's degree in Computer Science, Engineering, Software Engineering or a related technology discipline, or equivalent practical experience.
- 8+ years of experience designing, implementing, operating, troubleshooting and leading large-scale cloud, infrastructure or platform engineering environments.
- Strong hands-on experience with private cloud and virtualization platforms
, with proven…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).