DevOps Engineer
Indexed description
We work with marquee customers including Tata Steel, JSW, ArcelorMittal, Vedanta, Godrej & Boyce, Grasim, Holcim, and Jindal Steel, and are scaling globally across India, the Middle East, Europe, and North America. As we move into our next phase of growth, we are building the operating backbone that will take Ripik from a high-velocity scale-up to a category-defining global industrial AI company.
The Role
We are looking for a DevOps Engineer to own the availability, reliability, and security of Ripik's production infrastructure. You will manage our AWS environment end-to-end, build and maintain CI/CD pipelines that reduce manual deployment effort, keep a close eye on infrastructure performance and cost, and be the first responder when production incidents occur — diagnosing root cause and driving fixes. This role suits someone equally comfortable deep in an AWS console, writing a deployment script, and running a structured RCA after an incident.
Key Responsibilities
- Maintain the availability and reliability of production infrastructure across our AWS environment.
- Own deployment and CI/CD pipelines — reduce manual deployment work and keep releases fast and safe.
- Monitor infrastructure and cost — identify performance bottlenecks and cost-optimisation opportunities.
- Respond to and perform root-cause analysis (RCA) of production incidents.
- Set up backup and recovery procedures, and maintain disaster-recovery readiness.
- Identify and implement cybersecurity measures across the infrastructure.
- Maintain documentation of infrastructure and architecture diagrams, keeping them current as systems evolve.
- B.Tech in Computer Science or a related discipline, with 3+ years of hands-on DevOps experience.
- Strong working experience with AWS — EC2, S3, ECR, RDS, ECS, IAM, CloudWatch, and VPC.
- Strong working experience with CI/CD — GitHub Actions, secrets management, rollback and release strategies.
- Strong working experience with databases and queues (SQL / NoSQL) — MongoDB, Postgres, Redis, Kafka, SQS.
- Experience and familiarity with monitoring and observability tools — Beats, Logstash, CloudWatch — and RCA of production incidents.
- Familiarity with Linux, Python, Bash, Windows, PowerShell scripting, cron, systemd, NSSM, AWS CLI, CUDA, and NVIDIA tooling.
- Experience with Terraform for AWS infrastructure-as-code.
- Experience with Kubernetes.
- Experience with Prometheus + Grafana.
- FinOps / cloud cost analysis experience.
- GPU infrastructure experience.
- Experience managing on-prem + cloud hybrid infrastructure.
- Basic understanding of ML / computer vision workloads.
- Ability to shape the future of manufacturing by leveraging best-in-class AI and software; we are a unique organization with niche skill set that you would also develop while working with us
- World class work culture, coaching and development
- Mentoring from highly experienced leadership from world class companies (refer to Ripik.AI website for details)
- International exposure
- https://www.youtube.com/watch?v=YRuz3XxPfcE
- https://www.ripik.ai
- https://www.linkedin.com/company/ripik-ai/
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search