Leave us your email address and we'll send you all the new jobs according to your preferences.
Site Reliability & Observability Engineer - Datadog / Azure
Posted 6 hours 21 minutes ago by MYO Talent
Site Reliability & Observability Engineer / Datadog - Synthetic Monitoring, APM, RUM, Log Management, SLO's, Alerting / Azure / Azure DevOps / Cloudflare / 6-month contract / Hybrid - West Midlands / Remote / £450 - 600 per day Inside IR35.
One of our leading clients is seeking a Lead Site Reliability & Observability Engineer to build and operate a world-class monitoring, synthetic testing, and reliability platform.
Location - West Midlands / Remote - 5 days per week with 1-2 days per week onsite
Duration - 6 months +
Day rate - £450 - 600 per day Inside IR35
This role will lead the implementation of Datadog across Azure and Cloudflare, creating a comprehensive early warning system that continuously validates APIs, integrations, and customer user journeys in production.
Key Responsibilities:
- Own and evolve the Datadog observability platform.
- Design and maintain synthetic monitoring for critical API and UI workflows.
- Build continuous production validation covering business-critical customer journeys.
- Integrate monitoring, testing, dashboards, and alerting into Azure DevOps and GitHub pipelines.
- Develop monitoring-as-code and testing-as-code practices using Terraform.
- Create actionable dashboards, SLOs, SLIs, alerts, and anomaly detection.
- Integrate Datadog with Azure, Cloudflare, and modern SaaS architectures.
- Drive reliability, performance, and root-cause analysis across production systems.
Required Experience:
- Strong hands-on Datadog expertise, including:
- Synthetic Monitoring
- APM
- RUM
- Log Management
- SLOs and Alerting
- Experience operating large-scale global SaaS platforms.
- Deep Azure experience.
- Experience integrating Cloudflare services.
- Strong CI/CD experience with Azure DevOps and GitHub.
- Expertise in API, integration, and browser-based testing.
- Infrastructure as Code experience using Terraform.
- Experience with distributed systems, microservices, and cloud-native architectures.
Desirable:
- Datadog certifications.
- Azure certifications.
- Cloudflare administration experience.
- Background in Site Reliability Engineering (SRE) or Platform Engineering leadership roles.
MYO Talent
Related Jobs
Bank - Consultant - Haematology
- Essex, United Kingdom
.NET Developer
- Cambridgeshire, Cambridge, United Kingdom, CB1 0
Senior Quantexa Data Engineer - London / Hybrid
- £60,000 - £80,000 Annual
- London, United Kingdom
Bank - Consultant - Paediatrics
- Essex, United Kingdom
.NET Developer
- Kent, Maidstone, United Kingdom, ME141