Multimodal Data Engine Expert
Indexed description
1.1 Position Overview
We are seeking a Multimodal Data Engine Expert to lead the development of next-generation data infrastructure for the AI Agent era.
The successful candidate will drive the architectural evolution and core technology development of a next-generation multimodal intelligent data platform, with a focus on heterogeneous compute resource pooling, multimodal computing engines, vector storage systems, high-performance distributed caching, AI4DB, and Data Agent technologies.
This role will be responsible for building end-to-end capabilities for massive-scale heterogeneous and unstructured data, spanning resource scheduling, computation, retrieval, storage, and intelligent autonomous management. The goal is to establish core data infrastructure optimized for large language models and AI Agent ecosystems, while driving technological innovation and practical implementation at the intersection of Data and AI.
1.2 Key Responsibilities
- Multimodal Data Processing Platform Design and build a unified scheduling and management platform for heterogeneous computing resources, including CPUs, GPUs, and NPUs. Develop serverless resource pooling and management capabilities across multiple computing engines, continuously optimize heterogeneous resource scheduling performance, and improve overall resource utilization.
- Multimodal Computing Engine Development Develop and optimize computing engines for large-scale unstructured and multimodal data, enabling efficient processing and analysis of text, images, video, and other data types. Design next-generation multimodal query optimizers and hybrid execution engines, and support high-performance retrieval and processing of vector, textual, geospatial, and other data.
- Multimodal Vector Indexing and Management Systems Design and develop unified systems for multimodal vector retrieval, storage, metadata management, and access control. Integrate with open-source ecosystems and enable high-performance, intelligent storage optimization, data management, and index lifecycle management.
- Distributed Caching Services for Multimodal Data Build high-performance distributed caching services for multimodal data platforms. Provide high-speed near-compute caching for computing engines and efficient east-west data transfer across engines, improving the overall performance and efficiency of multimodal data processing pipelines.
- AI-Powered Agent Development Leverage emerging AI4DB (AI for Databases) and LLM Agent technologies to build next-generation intelligent data platforms with autonomous decision-making capabilities. Develop intelligent Agents and Skills that enable end-to-end automation and intelligence across workload development, data storage, data analysis, operations, and system maintenance.
- Strong proficiency in programming languages such as C, C++, Python, and Java.
- Strong R&D background and technical expertise in areas such as database systems, big data systems, distributed systems, and high-performance computing. Research or hands-on development experience with core system components, such as query optimizers, execution engines, storage engines, or distributed storage systems, is highly preferred.
- Strong understanding of the architecture, characteristics, and performance considerations of heterogeneous computing resources, including CPUs, GPUs, and NPUs. Hands-on experience in heterogeneous resource management and scheduling, with the ability to enable unified resource management and efficient resource utilization.
- Experience in cloud computing platform development and maintenance, as well as DevOps-related engineering practices.
- Experience with LLM fine-tuning and reinforcement learning, multimodal data processing technologies such as NLP, computer vision, and time-series analysis, and the integration of AI technologies with database or data systems is highly preferred.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search