Site Reliability Engineer – Datadog/Azure
Job in
Birmingham, West Midlands, B1, England, UK
Listed on 2026-10-08
Listing for:
MYO Talent
Part Time
position Listed on 2026-10-08
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Engineer, Azure
Job Description & How to Apply Below
Site Reliability & Observability Engineer / Platform Engineer / Datadog – Synthetic Monitoring, APM, RUM, Log Management, SLO's, Alerting / Azure / Azure Dev Ops / Cloudflare / 6-month contract / Hybrid – West Midlands / Remote / £450 – 600 per day Inside IR35.
One of our leading clients is seeking a Lead Site Reliability & Observability Engineer to build and operate a world-class monitoring, synthetic testing, and reliability platform.
Location – West Midlands / Remote – 5 days per week with 1-2 days per week onsite
Duration – 6 months +
Day rate – £450 – 600 per day Inside IR35
This role will lead the implementation of Datadog across Azure and Cloudflare, creating a comprehensive early warning system that continuously validates APIs, integrations, and customer user journeys in production.
Key Responsibilities:
· ?
Own and evolve the Datadog observability platform.
· Design and maintain synthetic monitoring for critical API and UI workflows.
· Build continuous production validation covering business-critical customer journeys.
· Integrate monitoring, testing, dashboards, and alerting into Azure Dev Ops and Git Hub pipelines.
· Develop monitoring-as-code and testing-as-code practices using Terraform.
· Create actionable dashboards, SLOs, SLIs, alerts, and anomaly detection.
· Integrate Datadog with Azure, Cloudflare, and modern SaaS architectures.
· Drive reliability, performance, and root-cause analysis across production systems.
Required Experience:
· Strong hands-on Datadog expertise, including:
o Synthetic Monitoring
o APM
o RUM
o Log Management
o SLOs and Alerting
· Experience operating large-scale global SaaS platforms.
· Deep Azure experience.
· Experience integrating Cloudflare services.
· ?
Strong CI/CD experience with Azure Dev Ops and Git Hub.
· ?
Expertise in API, integration, and browser-based testing.
· Infrastructure as Code experience using Terraform.
· Experience with distributed systems, microservices, and cloud-native architectures.
Desirable:
· Datadog certifications.
· Azure certifications.
· ?
Cloudflare administration experience.
· Background in Site Reliability Engineering (SRE) or Platform Engineering leadership roles.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×