PayNet (Payments Network Malaysia)
Linkedin · Posted 9d ago
Principal Engineer – IT Operations | RPP Squad
Continue to application
Add your email once, then Caio opens the original posting.
Indexed description
Why PayNet / Why Now What You Will Actually Do
- Own 24/7 operations for mission‑critical real‑time payment systems (RPP) where availability, stability, and security are non‑negotiable
- Strengthen the ACI UP platform to improve scalability, performance, and reliability as operational demands increase
- Ensure day‑to‑day operations and change delivery consistently meet payment standards (ISO 20022 & ISO 8583)
- Reduce downtime and operational risk by leading incident response, complex troubleshooting, and faster recovery across core platforms
- Lead operations ownership for RPP and the ACI UP platform in a high‑availability environment
- Drive system resilience through strong engineering across Java, Linux, RDBMS, and automation
- Be the technical authority during incidents, ensuring rapid recovery and stable service
- Raise team capability through mentoring and continuous improvement aligned to business objectives
- Real‑time payment operations require deep technical judgment to keep services stable under pressure
- Operational excellence here directly reduces incident frequency, impact, and recovery time
- Compliance with ISO 20022 & ISO 8583 must be sustained through both BAU operations and implementation change
- Principal‑level leadership is needed to balance BAU, improvements, and high‑stakes troubleshooting in one operating environment
- Lead and oversee 24/7 operations for RPP, ensuring high availability, stability, and security
- Provide technical leadership for ACI UP, improving scalability, performance, and reliability
- Ensure operational and implementation compliance with ISO 20022 & ISO 8583 standards
- Manage and optimise backend systems across Java, Linux, and RDBMS, including automation via shell scripting
- Lead incident management and complex troubleshooting to minimise downtime and ensure rapid recovery
- Drive continuous improvement and team capability through mentoring and alignment to business objectives
- Leading a high‑severity production incident, coordinating deep troubleshooting and restoring service rapidly
- Identifying a recurring stability issue and implementing automation to prevent repeated operational failures
- Improving platform performance through system‑level analysis and tuning in a high‑availability environment
- Validating that operational changes and implementations remain compliant with ISO 20022 & ISO 8583 requirements
- Coaching engineers through BAU + project delivery pressures while maintaining a stable, secure operating posture
Required
- Extensive hands‑on capability in Java, Linux administration (RHEL/Oracle), and shell scripting for operations and automation
- Practical knowledge of payment standards (ISO 20022 & ISO 8583) and how they apply in payment operations
- Proven experience operating Linux systems in a high‑availability, mission‑critical environment, including performance analysis and optimization
- Demonstrated leadership across BAU and project operations, including complex troubleshooting under pressure
- Hands‑on experience with the ACI UP platform or similar payment processing systems
- RHCSA / RHCE certification in Linux administration
- Exposure to DevOps and observability tools such as Jenkins, Terraform, Grafana, Prometheus, ELK
- Familiarity with core infrastructure services such as SMTP, LDAP, DNS, Apache, Nginx, HAProxy, WebLogic, JBoss/WildFly
- Advanced performance tuning using tools such as vmstat, iostat, sar, lsof
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search
Want help applying to roles like this?
Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search