Infrastructure Engineer
Indexed description
Job Summary
We are seeking a Cloud Infrastructure / Platform Engineer to design, build, and operate highly scalable cloud infrastructure and platform services. This role is ideal for a hands-on engineer who combines strong cloud infrastructure expertise with recent production software development experience.
The ideal candidate will have deep experience with AWS, Terraform, Linux, Docker, Kubernetes/EKS/ECS, cloud networking, distributed systems, and CI/CD, along with the ability to develop and maintain production-quality applications and platform services using Node.js/TypeScript and/or C#.
This is not a traditional DevOps or infrastructure administration role. The engineer will be expected to write and maintain production code recently, while also owning infrastructure, platform reliability, scalability, and operational excellence.
Key Responsibilities
- Design, build, and maintain scalable AWS cloud infrastructure and platform services.
- Develop and maintain production software using Node.js and TypeScript and/or C#.
- Build reusable and maintainable infrastructure using Terraform and Infrastructure-as-Code.
- Design and manage multi-environment cloud infrastructure across development, testing, and production.
- Deploy, operate, and troubleshoot containerized workloads using Docker and Kubernetes, EKS, and/or ECS.
- Work extensively in Linux-based environments such as Ubuntu, Amazon Linux, RHEL, or Alpine.
- Design and troubleshoot cloud networking components including VPCs, DNS, load balancing, routing, security groups, and service connectivity.
- Build and support distributed systems capable of handling high-volume data ingestion and real-time telemetry.
- Implement event-driven and streaming architectures using technologies such as Kafka, Kinesis, or other messaging/pub-sub platforms.
- Develop and maintain automated CI/CD pipelines with testing, quality controls, deployment gates, and repeatable environment promotion.
- Monitor and troubleshoot production systems using Datadog, CloudWatch, Splunk, or comparable observability platforms.
- Participate in incident response, root-cause analysis, performance optimization, and reliability improvements.
- Design solutions that support backward compatibility and long-lived connected devices/machines.
- Develop technical documentation covering architecture, operational procedures, incidents, technical decisions, and system tradeoffs.
- Collaborate with software engineers, architects, and platform teams through design reviews, code reviews, and architecture discussions.
- Mentor engineers on cloud-native architecture, infrastructure patterns, software development, and production operations.
Location: Oak Brook, IL or Sioux Falls, SD
Work Arrangement: Hybrid – 2–3 days per week onsite
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search