Back to search
MYO Talent Reedcouk · Posted yesterday

Site Reliability & Observability Engineer – Datadog / Azure

West Midlands ( ) GBP 450-600 Contract Remote

Contract Reedcouk
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Site Reliability & Observability Engineer / Datadog – Synthetic Monitoring, APM, RUM, Log Management, SLO’s, Alerting / Azure / Azure DevOps / Cloudflare / 6-month contract / Hybrid – West Midlands / Remote / £450 – 600 per day Inside IR35.


One of our leading clients is seeking a Lead Site Reliability & Observability Engineer to build and operate a world-class monitoring, synthetic testing, and reliability platform.


Location – West Midlands / Remote – 5 days per week with 1-2 days per week onsite

Duration – 6 months +

Day rate – £450 – 600 per day Inside IR35


This role will lead the implementation of Datadog across Azure and Cloudflare, creating a comprehensive early warning system that continuously validates APIs, integrations, and customer user journeys in production.


Key Responsibilities:

  • Own and evolve the Datadog observability platform.
  • Design and maintain synthetic monitoring for critical API and UI workflows.
  • Build continuous production validation covering business-critical customer journeys.
  • Integrate monitoring, testing, dashboards, and alerting into Azure DevOps and GitHub pipelines.
  • Develop monitoring-as-code and testing-as-code practices using Terraform.
  • Create actionable dashboards, SLOs, SLIs, alerts, and anomaly detection.
  • Integrate Datadog with Azure, Cloudflare, and modern SaaS architectures.
  • Drive reliability, performance, and root-cause analysis across production systems.


Required Experience:

  • Strong hands-on Datadog expertise, including:
  • Synthetic Monitoring
  • APM
  • RUM
  • Log Management
  • SLOs and Alerting
  • Experience operating large-scale global SaaS platforms.
  • Deep Azure experience.
  • Experience integrating Cloudflare services.
  • Strong CI/CD experience with Azure DevOps and GitHub.
  • Expertise in API, integration, and browser-based testing.
  • Infrastructure as Code experience using Terraform.
  • Experience with distributed systems, microservices, and cloud-native architectures.


Desirable:

  • Datadog certifications.
  • Azure certifications.
  • Cloudflare administration experience.
  • Background in Site Reliability Engineering (SRE) or Platform Engineering leadership roles.


Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search