Platform Engineer
Indexed description
Must Have Technical/Functional Skills:
• Strong expertise in Python for platform automation, API development, and engineering tooling.
• Deep hands-on experience with Kubernetes, including cluster administration, Helm, networking, RBAC, storage, security, and troubleshooting.
• Proven expertise in GitOps practices using Argo CD and/or FluxCD for automated, declarative application and infrastructure deployments.
• Extensive experience in DevOps practices, including CI/CD pipeline design, Infrastructure as Code (Terraform), Docker, and release automation.
• Strong knowledge of Developer Experience (DevEx) tooling, including Internal Developer Platforms (IDPs), Backstage, self-service workflows, and developer automation.
• Experience with major cloud platforms (AWS, Azure, or GCP) and cloud-native architectures.
• Proficiency in Observability tools such as Prometheus, Grafana, OpenTelemetry, ELK/OpenSearch, and distributed tracing solutions.
• Strong understanding of DevSecOps, Kubernetes security, secrets management, Policy-as-Code (OPA/Kyverno), and secure software delivery practices.
• Experience building scalable, resilient, and highly available platform services with a focus on automation, reliability, and operational excellence.
• Excellent troubleshooting, scripting, and cross-functional collaboration skills to enable high-performing engineering teams.
Roles & Responsibilities
• Design, develop, and maintain scalable Internal Developer Platforms (IDPs) and platform services.
• Develop automation tools and platform services using Python.
• Build and maintain Kubernetes-based infrastructure for containerized workloads.
• Implement GitOps practices for infrastructure and application deployments using tools such as Argo CD or FluxCD.
• Design and optimize CI/CD pipelines for secure, reliable, and automated software delivery.
• Build and enhance Developer Experience (DevEx) capabilities by creating self-service portals, templates, reusable automation, and platform APIs.
• Automate infrastructure provisioning using Infrastructure as Code (Terraform, Pulumi, or similar).
• Develop platform observability using monitoring, logging, and tracing solutions.
• Collaborate with development, security, and operations teams to improve platform reliability and deployment velocity.
• Implement security best practices, policy-as-code, and compliance automation.
• Troubleshoot production platform issues and continuously improve system reliability and performance.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search