Site Reliability Engineer
Indexed description
Overview
We're looking for a Site Reliability Engineer who's passionate about infrastructure, hungry to learn, and committed to delivering great results. If you're a team player with a positive mindset who wants to make a real impact, we'd love to hear from you.
What You'll Do & How You'll Make Your Mark.
- Manage distributed infrastructure across multiple datacenters using open-source technologies
- Ensure product SLAs through proactive monitoring, capacity planning, and issue resolution, including participation in a 24/7 on-call rotation
- Evaluate and implement Platform-as-a-Service solutions to improve efficiency across SRE teams
- Use data and metrics to guide decisions, with a strong focus on security and best practices
- Build robust automation and scripting to reduce manual work
- Automate infrastructure provisioning using Terraform, Puppet, or Ansible
- Bachelor’s or Master’s degree in Computer Science, Information Technology, or a related field
- 1-3 years of relevant experience
- Hands-on experience with at least one public cloud platform (OCI, AWS, GCP, or Azure)
- Solid understanding of Linux fundamentals and networking concepts
- Proficiency in Terraform, Puppet, and Ansible
- Experience with web servers such as Nginx, Apache, or Tomcat
- Strong scripting skills in at least one language (e.g., Python, Go)
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search