Sr. Cloud Engineer
Indexed description
Your responsibilities will be:
- Build and maintain a globally distributed cloud infrastructure platform using Infrastructure as Code (e.g. Terraform, Ansible, among others).
- Collaborate with Cloud Platform Engineering and Business Unit applications team to engineer infrastructure solutions to support their cloud-based applications
- Design, deploy, and operate Kubernetes-based platforms and workloads (EKS/GKE), including GitOps workflows (e.g., Argo CD).
- Perform system and database maintenance activities, including application deployment, backup and restoration procedures, patching, upgrades, etc.
- Support and enhance CI/CD pipelines and infrastructure automation.
- Maintain and support cloud environments, including incident response, troubleshooting, and performance tuning.
- Document infrastructure designs, standard operating procedures, and support runbooks.
- Create and maintain clear, accurate documentation tailored to the target audience (e.g., engineers, support staff).
- Collaborate with development, security, and operations teams to ensure systems are secure and compliant.
- Participate in on-call rotations and support operational readiness.
- Participate in retrospectives to identify opportunities to improve platform scalability and supportability.
Expectations for Project and Initiative Work
- Contribute to implementation efforts for cloud infrastructure projects.
- Take ownership of discrete technical tasks and follow through to completion.
- Scope and track work progress using tools like Jira.
- Communicate status, blockers, and risks clearly to project leads.
- Participate in design discussions and reviews with guidance from senior team members.
- Follow defined patterns and standards to ensure consistency and maintainability.
The ideal candidate has
- 6+ years of experience working with public cloud platforms (AWS and GCP preferred) in a professional setting.
- Deep, hands-on experience deploying and managing Kubernetes (EKS, GKE) + GitOps (Argo CD) or other orchestration platform (e.g., Hashicorp Nomad / Consul)
- Deep, hands-on experience using Terraform, Ansible, CI/CD tools (e.g., Jenkins, GitHub Actions), and scripting languages (Python) to deploy, maintain, and support cloud infrastructure.
- Strong experience supporting containerized applications and Linux-based systems Deep experience building and maintaining databases, including MongoDB and PostgreSQL
- Strong understanding of networking, IAM, cloud observability, and common cloud services such as AWS VPC, EC2, S3, SNS, SQS, and Lambda.
- Experience managing secrets and sensitive configuration using tools such as HashiCorp Vault.
- Experience designing, deploying, and operating fault-tolerant systems with on-call responsibilities.
- Strong working knowledge of Git and common branching and collaboration workflows.
- Growth mindset and interest in continuous learning Strong collaboration and communication skills
You’ll really stand out with:
- Deep experience with building complex cloud infrastructure using Terraform Strong focus on security best practices (cloud IAM, least privilege, encryption)
- Experience with multi-cloud environments (AWS + GCP) Deep experience building and managing CI/CD pipelines (Jenkins, GitHub Actions)
- Deep experience with cloud observability tooling (Datadog, Prometheus, Grafana, Loki, ELK, Cloudwatch)
- Experience with data streaming technologies and middleware such as Kafka, MQ, etc.
- A track record of leading projects end-to-end, from design through delivery and operations.
- Experience designing for scalability, resilience, and operability in complex cloud environments
“Our specialized recruiting professionals apply their expertise and utilize our proprietary AI to find you great job matches faster.”
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search