DevOps Engineer
Indexed description
This engineer will partner with software teams to improve release velocity without compromising availability or security. The work directly impacts deployment frequency, recovery time, system performance, and the day-to-day operating efficiency of the engineering organization.
Key Responsibilities
- Build and maintain AWS infrastructure using Terraform, including VPCs, IAM, ECS or EKS, RDS, S3, CloudFront, and managed messaging services
- Operate Kubernetes-based production environments, improving cluster reliability, resource utilization, deployment safety, and horizontal scaling
- Develop and optimize CI/CD pipelines with GitHub Actions, GitLab CI, or Jenkins for automated testing, container builds, progressive delivery, and rollbacks
- Implement observability using Prometheus, Grafana, CloudWatch, OpenTelemetry, or equivalent tools to monitor service health, latency, capacity, and error rates
- Automate operational workflows with Python, Go, or Bash, reducing manual intervention in provisioning, incident response, and routine maintenance
- Lead or participate in incident response, including troubleshooting production failures, coordinating remediation, and documenting blameless post-incident reviews
- Partner with engineering and security teams to establish platform standards for secrets management, access control, vulnerability remediation, disaster recovery, and compliance
- 3–8 years of experience in DevOps, site reliability engineering, platform engineering, or a closely related infrastructure role
- Strong hands-on experience with AWS and production infrastructure spanning networking, IAM, compute, storage, databases, and security controls
- Proficiency with Terraform or an equivalent infrastructure-as-code tool, including reusable modules, state management, and code review practices
- Production experience operating Kubernetes and Docker, including deployments, services, ingress, autoscaling, troubleshooting, and resource management
- Demonstrated ability to build CI/CD pipelines and release automation using GitHub Actions, GitLab CI, Jenkins, Argo CD, or comparable tools
- Working knowledge of Linux systems, networking fundamentals, observability practices, and scripting with Python, Go, or Bash
- Bonus: Experience with service meshes, GitOps, Helm, OpenTelemetry, Kafka, multi-region architecture, SOC 2 controls, or formal computer science and engineering education
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search