SRE (Site Reliability Engineer)
Indexed description
Required Skills *
Linux administration, cloud platforms (AWS/Azure/GCP), Kubernetes & Docker, monitoring and observability, scripting (Python/Bash), CI/CD, networking, incident management, automation, troubleshooting, and reliability engineering practices (SLA/SLO/SLI).
Required Experience / Knowledge Areas *
- Strong experience in Linux/Unix administration, troubleshooting, and system performance tuning.
- Hands-on expertise with cloud platforms such as AWS, Azure, or GCP.
- Proficiency in Kubernetes, Docker, automation, and CI/CD pipelines.
- Experience with monitoring, observability, incident management, and root cause analysis.
- Preferred experience in banking, financial services, fintech, or enterprise application
- Strong understanding of networking, security, and infrastructure management.
- Experience in incident management, root cause analysis (RCA), and problem resolution.
- Knowledge of high availability, disaster recovery, SLA/SLO/SLI, and performance optimization.
- Experience supporting banking, financial services, fintech, or enterprise application environments is preferred.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search