Lead Site Reliability & Observability Engineer
Amtis Professional Ltd · Acocks Green, Birmingham ·
- Seniority
- Lead
- Category
- SRE
Salary: £? - ? per year
Requirements:- Strong hands-on Datadog experience across Synthetic Monitoring, APM, RUM, Log Management, SLOs, and alerting
- Deep Azure experience and experience integrating Cloudflare services
- Experience operating large-scale production environments
- Strong CI/CD experience with Azure DevOps and GitHub
- Expertise in API, integration, and browser-based testing
- Terraform experience and a strong understanding of distributed systems and microservices
- A background in SRE or Platform Engineering leadership
- Datadog or Azure certifications and Cloudflare administration experience are advantageous
- Lead the implementation of Datadog across Azure and Cloudflare
- Build synthetic monitoring and automated production validation for critical APIs, integrations, and customer journeys
- Own and evolve our Datadog observability platform
- Design synthetic monitoring for critical API and browser workflows
- Integrate monitoring, testing, and release validation into Azure DevOps and GitHub pipelines
- Develop monitoring-as-code and testing-as-code using Terraform
- Create actionable dashboards, SLOs, SLIs, alerts, and anomaly detection
- Drive reliability improvements, performance investigations, and root-cause analysis
- API
- Azure
- CI/CD
- Datadog
- DevOps
- GitHub
- Terraform
- microservices
- Cloud
More:
Were recruiting a Lead Site Reliability & Observability Engineer for a contract opportunity based in Birmingham, initially for six months. The rate is up to £575 per day, inside IR35, with hybrid working that includes two days on-site per week. The role focuses on improving observability and production validation so issues can be identified within minutes of a software release.
last updated 41 week of 2026