Principal / Distinguished Solution Architect
Indexed description
We are looking for a Principal / Distinguished Solution Architect to lead the architecture of large-scale customer AI infrastructure environments. This is a highly impactful, senior individual-contributor role for a visionary technologist who can operate seamlessly from high-level customer strategy down to deep technical design and production delivery.
What You Will Do
- Strategic Advisory: Serve as a trusted technical advisor to strategic customers, engaging directly with CTOs, infrastructure, platform, and engineering teams.
- End-to-End Design: Translate complex workload, scale, performance, security, availability, and operational requirements into elegant, high-performing architectures.
- Infrastructure at Scale: Design massive-scale GPU infrastructure spanning compute, InfiniBand / ROCE / RDMA fabrics, high-performance storage, Kubernetes / Slurm orchestration, security, observability, SRE, and datacenter connectivity.
- Deep System Integration: Drive the complex integration of hardware fabrics, orchestration layers, and tenant environments to ensure optimal performance of distributed AI training andI nference workloads.
- Technical Leadership: Lead architecture workshops, high-level/low-level designs (HLD/LLD), capacity sizing, BOM creation, design reviews, and production-readiness decisions.
- Product Influence: Collaborate intimately with the Nava Product & Engineering team to shape our platform capabilities based on real-world customer requirements.
- Ecosystem Partnership: Partner with leading GPU, server, networking, storage, security, and software vendors to build seamlessly integrated solutions.
- Lifecycle Ownership: Stay deeply engaged through the entire lifecycle—from architecture and deployment to validation and production readiness—resolving intricate cross-stack issues.
- Engineering Standards: Pioneer reusable reference architectures, design patterns, and engineering frameworks that scale across customers and regions.
- Architectural Excellence: Exceptional systems architecture depth spanning compute, networking, storage, distributed systems, and modern cloud infrastructure.
- Domain Expertise: Deep, hands-on experience with advanced GPU infrastructure, HPC, Kubernetes, Slurm, high-performance networking, or hyperscale platforms.
- Complex Problem Solving: A proven track record of solving cross-stack technical bottlenecks and making pragmatic design trade-offs across performance, reliability, scalability, security, and cost.
- Executive Communication: The ability to engage confidently with customer engineering teams, senior technology leaders, and internal/external product partners.
- Impact Driven: Strong leadership influence without formal authority, and a history of transforming whiteboard architectures into robust production outcomes.
Ideal candidates will have held roles such as Principal Engineer, Distinguished Engineer, Principal Architect, Field CTO, or senior Customer Engineering / Solution Architecture at hyperscalers, GPU cloud providers, AI infrastructure innovators, HPC organizations, or large-scale platform teams.
The Opportunity: This is a rare opportunity to define how large-scale AI infrastructure is
architected and delivered across APAC. You will remain deeply technical, stay close to
customers, and play a pivotal role alongside an elite engineering team to build the future of AI
infrastructure.
Skills: storage,customer,architecture,infrastructure,cloud,design,security,networking,customer engineering,teams
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search