Sr. Site Reliability Engineer
Indexed description
At Visa, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.
Join Visa and do work that matters – to you, to your community, and to the world. Progress starts with you.
Summary
Job Description
The DevOps squad is dedicated to implementing and optimizing the Pismo CI/CD Platform, automating infrastructure provisioning, and improving operational tasks using best practices in CI/CD and IaC. The squad stays updated with the latest technology trends, conducts POCs for new technologies, and provides strategic insights to drive continuous improvement and innovation.
What You'll Do
- Implement, optimize, and maintain reliable and scalable CI/CD pipelines.
- Develop and maintain Infrastructure as Code scripts to automate infrastructure provisioning and management.
- Apply Infrastructure as Code best practices to ensure consistency, repeatability, security, and compliance across environments.
- Identify opportunities to automate routine operational tasks, improving efficiency and reducing manual effort.
- Implement automation solutions to streamline deployment, scaling, maintenance, and operational processes.
- Make well-informed SRE and engineering decisions, considering technical debt, system design, stability, reliability, monitoring, observability, and business requirements.
- Monitor and maintain the resilience of the CI/CD platform, avoiding unplanned or uncommunicated changes.
- Troubleshoot complex platform and pipeline issues, collaborating with senior and staff engineers when broader technical alignment is required.
- Write clear and effective post-mortem documentation for internal and external stakeholders.
- Apply and contribute to the continuous improvement of coding standards, engineering practices, and non-functional requirements.
- Participate in code and Pull Request reviews, providing constructive feedback focused on CI/CD, automation, reliability, and code quality.
- Mentor junior and mid-level engineers, sharing technical knowledge and supporting their professional development.
- Act as a technical reference within the squad for CI/CD, Infrastructure as Code, automation, and platform reliability topics.
- Stay up to date with emerging technologies and industry practices, sharing relevant insights with the squad.
- Support and execute Proofs of Concept to evaluate and introduce new CI/CD and automation technologies.
- Contribute high-quality technical solutions that support platform resilience, continuous improvement, and the squad’s strategic goals.
Qualifications
For this role, you must be based in Brazil
Language Skills
- Proficiency in English at B2 level or above (Upper-Intermediate)
- 5 or more years of relevant work experience with a Bachelor's or Associate’s Degree or at least 2 years of work experience with an Advanced degree (e.g. Masters, MBA, JD, MD)
- Experience implementing, optimizing, and maintaining CI/CD pipelines in production environments.
- Experience with Infrastructure as Code and infrastructure automation practices
- Experience working in DevOps, Site Reliability Engineering, Platform Engineering, Cloud Engineering, or a related technical area.
- Strong experience implementing, optimizing, and maintaining CI/CD pipelines in production environments.
- Experience with Infrastructure as Code and infrastructure automation practices.
- Ability to develop and maintain automation scripts for infrastructure provisioning, deployment, scaling, and operational activities.
- Understanding of SRE principles, including system reliability, availability, resilience, monitoring, observability, and incident management.
- Experience supporting reliable and scalable platforms and troubleshooting complex issues in distributed environments.
- Knowledge of software engineering standards, coding practices, version control, and Pull Request review processes.
- Ability to evaluate technical trade-offs involving system design, technical debt, stability, reliability, maintainability, and business requirements.
- Experience identifying opportunities to automate manual operational processes and improve engineering efficiency.
- Knowledge of monitoring, logging, metrics, alerting, and observability practices.
- Experience creating technical documentation and post-mortem reports for technical and non-technical stakeholders.
- Ability to collaborate with engineers across different experience levels and provide constructive technical feedback.
- Strong analytical and problem-solving skills, with the ability to work on well-scoped and moderately ambiguous technical challenges.
- Ability to contribute to technical discussions and escalate broader or cross-squad decisions to senior and staff engineers when appropriate.
- Experience maintaining CI/CD platforms or internal developer platforms used by multiple engineering teams.
- Experience working with critical or mission-critical production systems.
- Experience with cloud infrastructure, container orchestration, and modern deployment practices.
- Familiarity with GitOps, Infrastructure as Code, and automated infrastructure management.
- Experience conducting Proofs of Concept and evaluating new technologies for implementation in production environments.
- Experience mentoring junior and mid-level engineers.
- Experience with incident response, root cause analysis, and post-mortem practices.
- Experience working in global or distributed engineering teams.
- Relevant cloud, DevOps, Kubernetes, or Infrastructure as Code certifications.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search