Senior DevOps Engineer
Indexed description
We are looking for a high-impact Senior DevOps Engineer to join our team. You won't just be "keeping clusters alive" in a vacuum; you will own the reliability, security, and evolution of our production infrastructure. We are looking for someone who brings a distributed-systems architect's vision, both within and beyond the cloud.
Why join AZmed?You will work within the whole R&D team (developers and data scientists) and doctors, daily. We are experiencing rapid growth and working with over 2,000 sites in 50 countries, making it an exciting time to join our team and contribute to AZmed's mission.
You will help us run and scale AZmed’s products, Rayvolve and Rayscan, an AI-based software for abnormality detection in standard X-rays and Lung CT scans. Our systems analyze medical images in production, with strong requirements around availability, traceability, and security of health data, powering the first French deep learning software in radiology. Let’s grow together by strengthening AZmed’s infrastructure and shaping the future of healthcare.
ResponsibilitiesArchitect our cloud infrastructure: Contribute to the design and evolution of our AWS architectures: reliability, scalability, fault tolerance, cost efficiency. Act as a driving force on architecture decisions and their trade-offs.
Operate our Kubernetes clusters: Maintain, secure, and evolve our Kubernetes clusters (managed EKS and self-hosted k3s) on AWS. Manage lifecycles: upgrades, patching, autoscaling, capacity planning.
Upgrade our observability stack: Maintain and evolve our observability stack. Operate the collection of metrics, traces, and logs, and guarantee its reliability.
Guarantee resilience: Ensure high availability, backup/restore, and disaster recovery across a hybrid cloud and on-premise environment, where downtime directly impacts radiologists.
Make alerting actionable: Define meaningful SLOs, and build dashboards and alerting rules that teams actually trust.
Collaborate for Change: Work alongside product, AI, and engineering peers to support instrumentation and observability adoption, and make reliability everyone’s business.
Flexible work : Possibility of working remotely two days per week 🌻
Annual off-sites : A company holiday each year in amazing locations 🪅
Atmosphere : Young team (29 y.o average) 🐣
Localisation : Great office in Paris 2nd arrondissement (Grands Boulevards) 🏣
5+ years of professional experience as a DevOps, SRE, or Platform Engineer.
Architecture and distributed systems: designing secure, reliable, scalable, and secure systems.
AWS: designing and operating cloud architectures (VPC, IAM, EKS, ECR, S3, networking, security…).
Kubernetes in production, particularly on managed services (EKS)
Observability experience: Prometheus / Alertmanager / Grafana and the OpenTelemetry ecosystem.
Comfort with scripting and automation (Bash, Python, Git).
Rigor and a strong sense of reliability; a "you build it, you run it" mindset.
Architectural vision and the ability to drive proposals.
Autonomy, a teaching mindset, and clear communication in both French and English.
SRE culture
Awareness of highly regulated environments and sensitive health data
Experience with GitOps (ArgoCD) and Infrastructure as Code (Terraform)
Experience running on-premise / edge infrastructure
Experience in securing a production infrastructure
Kubernetes (EKS, k3s)
Python 3.12+
AWS
MongoDB
Postgresql
GitOps with ArgoCD
Terraform / Helm
GitHub CI/CD
Screening interview with our Head of Engineering (30 min, remote)
In-depth technical interview with our engineering team (2 hours, at the office)
Meet the team and founders (2h30 hours, at the office)
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search