Principal SRE / Hybrid / Tempe
Indexed description
This is a high-impact principal engineering role where you’ll help shape the reliability strategy for a large cloud platform. The team is looking for someone who can dive deep technically while also influencing architecture decisions and mentoring other engineers. Engineers here work in modern environments built around Terraform infrastructure automation, Kubernetes and ECS container platforms, and advanced monitoring and observability tools. It’s an opportunity to step into a leadership-level engineering role where you can guide best practices, build resilient systems, and help scale cloud infrastructure supporting millions of users.
Required Skills & Experience
- 5+ years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles
- Strong experience supporting large-scale AWS environments
- Expertise with infrastructure as code tools such as Terraform or CloudFormation
- Experience designing or supporting containerized platforms such as Kubernetes or ECS
- Strong Linux systems knowledge
- Scripting experience using Bash
- Python scripting experience
- Experience supporting multi-cloud environments such as AWS, Azure, or GCP
- Experience implementing monitoring and observability platforms at scale
- Experience driving reliability or automation initiatives across engineering teams
- Exposure to modern DevOps or platform engineering practices
- 45% Cloud Infrastructure (AWS, Azure, GCP)
- 25% Infrastructure as Code (Terraform / CloudFormation)
- 15% Containers (Kubernetes / ECS)
- 15% Observability & Monitoring (Datadog, Splunk, Logging Platforms)
- 70% Hands On
- 15% Architecture & Technical Leadership
- 15% Team Collaboration & Mentorship
- Bonus OR Commission eligible
- Medical, Dental, and Vision Insurance
- Vacation Time
- Stock Options
Posted By: Isabella Sweet
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search