Site Reliability Engineer
Indexed description
W2 ONLY | NO C2C | NO CORP-TO-CORP
We have 4 locations: Columbia SC, Knoxville TN, Lafayette LA, Burmingham, AL
Job Description:
We are seeking an experienced Site Reliability Engineer (SRE) to ensure the reliability, scalability, and performance of cloud-based applications and infrastructure.
Key Responsibilities:
- Manage and maintain Kubernetes environments and containerized applications.
- Design, deploy, and support infrastructure on AWS.
- Develop automation, monitoring, and operational tools using Python.
- Implement observability and monitoring using Grafana.
- Troubleshoot production issues, improve system performance, and ensure high availability.
- Collaborate with development and DevOps teams to improve deployment, automation, and reliability processes.
Required Skills:
- Kubernetes
- AWS
- Python
- Grafana
- Cloud infrastructure & monitoring
- Troubleshooting and production support
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search