DevOps Engineer
Indexed description
This role partners with software engineers and security teams to make releases safer, reduce operational toil, and improve system reliability. The team needs an engineer who can troubleshoot complex production issues, define practical infrastructure standards, and strengthen the platform through automation and measurable service-level objectives.
Key Responsibilities
- Build and maintain AWS infrastructure using Terraform, including networking, IAM, compute, storage, and managed database services
- Operate Kubernetes workloads across development, staging, and production environments; manage deployments, scaling, ingress, secrets, and resource policies
- Automate CI/CD pipelines with GitHub Actions, Argo CD, or equivalent tools, including testing gates, environment promotion, and rollback procedures
- Implement observability with Prometheus, Grafana, CloudWatch, and centralized logging to track service health, latency, capacity, and error rates
- Define and improve reliability practices, including SLOs, alerting standards, incident response procedures, and post-incident corrective actions
- Partner with application teams to improve containerization, release workflows, configuration management, and production readiness
- Troubleshoot infrastructure and distributed-system failures, document root causes, and deliver durable fixes rather than manual workarounds
- 3–8 years of experience in DevOps, site reliability engineering, platform engineering, or infrastructure engineering
- Hands-on experience managing AWS production environments and writing maintainable Terraform modules
- Strong Kubernetes and Docker experience, including debugging workloads, networking issues, deployments, and resource constraints
- Proficiency with Linux, Bash or Python, Git, and CI/CD systems such as GitHub Actions, GitLab CI, Jenkins, or Argo CD
- Working knowledge of observability practices and tools, including metrics, logs, tracing, alert design, and incident response
- Bachelor’s degree in computer science, engineering, information systems, or a related technical field; equivalent practical experience is also considered
- Bonus: Experience with Helm, Argo CD, service meshes, PostgreSQL or Redis operations, compliance controls, and building internal developer platforms
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search