Site Reliability Engineer
Indexed description
What You'll Be Doing
- Design and implement monitoring and alerting systems for critical services
- Automate operational tasks to improve efficiency and reduce manual effort
- Collaborate with development teams to enhance system reliability and performance
- Manage incident response and post-incident reviews
- Analyse system metrics to identify trends and areas for improvement
- Contribute to capacity planning and scalability strategies
- Experience in site reliability engineering or DevOps roles
- Strong scripting and automation skills (e.g., Python, Powershell, Azure Automation)
- Strong experience working in public cloud environments such as Microsoft Azure (preferred), AWS, GCP etc.
- Knowledge of monitoring tools and observability practices
- Understanding of cloud infrastructure and containerisation
- Excellent problem-solving and analytical abilities
- Commitment to continuous improvement and operational excellence
- 26 days annual leave plus bank holidays (increasing with length of service)
- Flexible working options
- Private health care
- Competitive pension scheme – Anglian Water double-matches your contributions up to 6%
- Life assurance at eight times your salary
- Annual bonus
- Personal medical assessments
- Virtual GP service
- Cancer screening
- Financial wellbeing support and salary finance benefits
- Lifestyle Savings including discounts on retail, travel, and utilities
- Employee Assistance Programme
- Volunteer days
- Environmental and wellbeing initiatives
Closing Date: 9th August 2026
#loveeverydrop!
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search