Platform Engineer
Indexed description
Job Description
Your Role:
We are seeking a highly skilled and motivated Platform Engineer to join our Infrastructure Customer Engineering and Support team. This is a special Engineering task force dedicated to design, deploy and troubleshoot OpenStack/ OpenShift environments. They are deeply involved in technologies like Nokia Container Services (NCS) and CloudBand Infrastructure Software (CBIS), private clouds based on Kubernetes and OpenStack, and collaborate closely with developers and product engineers to bridge the gap between infrastructure and software.
- Design, deploy, configure, manage, troubleshoot, and optimize containerized applications and infrastructure deployed on platforms like Kubernetes, OpenShift, and OpenStack.
- Develop, test, and maintain robust automation scripts using Python and Ansible to streamline daily operational tasks and improve overall service efficiency.
- Lead the investigation and resolution of complex, high-severity customer incidents.
- Utilize expertise to quickly identify root causes and implement effective, durable solutions.
- Prepare and conduct rigorous Root Cause Analysis (RCA) for critical incidents to identify systemic issues and prevent recurrence.
- Stay current with industry’s best practices and emerging technologies in cloud and containerization.
- Networking Foundations: Strong knowledge of core networking principles (TCP/IP, routing, load balancing, firewalls) in a cloud environment. A solid grasp of computer networking fundamentals, such as understanding of VLANs and IP routing, is a must.
- Containerization & Virtualization: Strong knowledge of Kubernetes orchestration, OpenStack platforms, and Docker/Containerization. Knowledge in areas like Podman, Kubernetes, Helm, and/or OpenStack, KVM/QEMU is a significant advantage.
- Scripting and Automation: Solid Python scripting skills for task automation and system management. Proficiency in scripting with Bash and Python, or the willingness to learn and adapt, as well as familiarity with Ansible is required.
- Root Cause Analysis (RCA): Expertise in preparation and implementation of RCAs.
- Escalation and Monitoring: Proven experience with EME (Escalation, Monitoring, and Emergency) management processes.
- Database Expertise: Understanding of relational databases such as MySQL and MariaDB, as well as experience with ETCD.
- Red Hat Certified Specialist in Cloud Infrastructure (EX210).
- Red Hat Certified Engineer (RHCE).
- Red Hat OpenStack (EX310).
- RedHat Certified Specialist in OpenShift Administration (EX280).
- RedHat Certified Specialist in OpenShift Automation and API Management (EX380).
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search