AWS Cloud DevOps / SRE Engineer
Indexed description
Role Summary:
AWS Cloud DevOps/SRE Engineer responsible for CI/CD automation, Terraform-based Infrastructure as Code, Kubernetes platform operations, observability, GitOps delivery, release reliability, and DevSecOps controls for cloud platforms. The role requires strong AWS experience, hands-on Terraform and Kubernetes expertise, experience with CloudWatch or similar observability tools, and solid working knowledge of Bamboo, Tekton, and ArgoCD.
Experience: 7 - 10 years
Shift : Day
Work Model : Hybrid
Core Skills/ Required Skills
● Strong AWS experience across infrastructure, networking, IAM, monitoring, automation, and cloud operations.
● DevOps and SRE expertise covering reliability engineering, incident response, production readiness, automation, SLIs, SLOs, and runbook management.
● Hands-on Terraform experience for Infrastructure as Code, including reusable modules, environment provisioning, state management, and standardized deployments.
● Strong Kubernetes and EKS experience with container orchestration, deployment operations, scaling, troubleshooting, and platform reliability.
● Experience with CloudWatch or similar observability tools for logs, metrics, dashboards, alarms, and proactive issue detection.
● Strong working knowledge of Bamboo for enterprise CI/CD orchestration, build automation, and release workflows.
● Strong working knowledge of Tekton for Kubernetes-native pipeline automation and reusable pipeline tasks.
● Strong working knowledge of ArgoCD for GitOps deployments, synchronization, drift detection, and environment promotion.
● Hands-on scripting experience with Python, Bash, PowerShell, or AWS CLI for validation and operational automation.
● Experience supporting reliable CI/CD, GitOps, and release workflows across AWS and Kubernetes environments.
● Strong troubleshooting and automation skills for infrastructure, deployments, monitoring, and production operations.
● Nice to have: Docker, Helm, Prometheus, Grafana, Jenkins, GitHub Actions, CloudFront, and DevSecOps tooling.
Roles & Responsibilities:
● Design, build, and maintain CI/CD pipelines for application and infrastructure deployments on AWS.
● Develop reusable Terraform Infrastructure as Code modules and deployment standards for repeatable, secure, and governed environments.
● Operate Kubernetes and EKS platforms while supporting workload reliability, scaling, deployments, and production troubleshooting.
● Implement Bamboo, Tekton, and ArgoCD workflows for automated builds, deployments, rollbacks, and GitOps delivery.
● Configure CloudWatch and similar observability tools for metrics, logs, dashboards, alerts, and reduced mean time to resolution.
● Integrate DevSecOps controls, including least-privilege IAM, secrets management, encryption, security scanning, and audit evidence.
● Create and maintain runbooks, deployment guides, rollback plans, platform documentation, and operational procedures.
● Improve release reliability by validating automated deployment, rollback, and environment-promotion workflows.
● Troubleshoot CI/CD, infrastructure, Kubernetes, monitoring, and production deployment issues.
Qualifications:
- 7–10 years of experience in AWS cloud engineering, DevOps, SRE, platform engineering, release engineering, or production operations.
- Hands-on AWS experience with IAM, CloudWatch, Amazon EKS, VPC, Load Balancers, security services, CloudFront, and AWS CLI.
- Expert-level Terraform experience, including reusable modules, workspaces and environments, state management, and controlled delivery patterns.
- Strong Kubernetes administration and troubleshooting experience, with Amazon EKS experience preferred.
- Proven experience with Bamboo, Tekton, ArgoCD, or similar enterprise CI/CD and GitOps platforms.
- Preferred certifications include AWS Certified DevOps Engineer – Professional, Certified Kubernetes Administrator, HashiCorp Certified Terraform Associate, AWS Certified Solutions Architect – Associate, CKS, or CKAD.
- Senior AWS DevOps/SRE expertise focused on Terraform, Kubernetes, observability, Bamboo, Tekton, and ArgoCD.
- Strong experience supporting reliable cloud platforms, production operations, automated deployments, and GitOps delivery workflows.
- Ability to troubleshoot AWS, Terraform, Kubernetes, CI/CD, observability, and production environment issues.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search