AI/ML Software Engineer
Indexed description
Our constant innovation and ongoing success is down to our amazing teams of incredibly talented people, who collaborate and support each other to come up with truly groundbreaking ideas and solutions. Solutions that will have a huge impact on people's lives; making the world a better place, one processor at a time.
Are you ready?
To learn more about SiFive’s phenomenal success and to see why we have won the GSA’s prestigious Most Respected Private Company Award (for the fourth time!), check out our website and Glassdoor pages.
Job Description:
The Role:
Join the SiFive AI/ML Software Team to build the high-performance stack for Next-Gen AI. We are seeking engineers to optimize and deploy LLMs and Generative AI models on RISC-V architectures, spanning from compiler infrastructure to distributed runtime systems.
Responsibilities:
- Compiler Infrastructure: Develop and maintain MLIR/IREE/Triton compiler stacks; optimize end-to-end (e2e) LLM performance and contribute to relevant open-source communities.
- Runtime Systems: Design and implement single/multi-device scheduling and memory management layers; define hardware-software interfaces for high-throughput AI workloads.
- Model Infrastructure: Implement model sharding and distribution strategies; lead the integration between high-level frameworks (e.g., PyTorch) and hardware backends.
- Performance Optimization: Analyze and profile AI models to identify bottlenecks; develop high-performance kernels leveraging RISC-V Vector (RVV) and custom ISA extensions.
- Co-design: Collaborate with hardware architects to influence the design of future AI accelerators and microarchitectures.
- Education: Master’s or PhD in Computer Science, Applied Mathematics, or a related field.
- Programming: Strong proficiency in C++ and Python.
- Domain Expertise: Solid understanding of AI/ML models (LLMs, Diffusion, Transformers) and their deployment challenges.
- Technical Track (One or more of the following):
- Experience with compiler frameworks like MLIR, IREE, LLVM, or Triton.
- Background in system programming, memory management, or multi-device orchestration.
- Knowledge of distributed computing, model parallelism, or tensor sharding.
- Experience with AI/ML frameworks like PyTorch, ONNX Runtime, or TF/TFLite.
- Familiarity with distributed training/inference and collective communications.
- Active contributions to open-source AI/ML or compiler projects.
- Experience in low-level performance tuning or MLPerf benchmarking.
Additional Information:
This position requires a successful background and reference checks and satisfactory proof of your right to work in:
Taiwan
Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
SiFive is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search