Data Platform Architect
Indexed description
About Us:
A global technology company specializing in e-commerce, cloud computing, digital payments, logistics, and artificial intelligence, serving businesses and consumers across more than 200 countries and regions.
Job Responsibilities
- Design, develop, and maintain enterprise-level data warehouse models, ETL pipelines, and scheduling workflows to ensure data accuracy, consistency, and timeliness
- Analyze business requirements, identify data sources, and design end-to-end data processing workflows across the data warehouse layers, including ODS, DWD, DWS, and ADS
- Develop and optimize large-scale ETL/ELT processes using big data technologies such as Hive, Spark, Flink, HDFS, and Kafka for data extraction, cleansing, transformation, and loading
- Write efficient, maintainable SQL queries and scripts (e.g., Python and Shell) to improve data processing performance and execution efficiency
- Collaborate closely with Data Analysts, Product Managers, and BI Engineers to support reporting, dashboards, and analytical requirements
- Develop data visualizations using BI tools such as Tableau, Power BI, FineBI, or QuickBI to support business decision-making
- Participate in data governance initiatives, including metadata management, data quality monitoring, and data lineage analysis
.Job Requirements
- Bachelor's degree or above in Computer Science, Data Engineering, Data Science or a related courses
- Minimum 4 years of relevant working experience in data engineering or data warehouse development
- Strong command of both English and Chinese Mandarin languages with good communication and writing skills
- Hands-on experience with cloud-native technologies, including Kubernetes (K8s), Docker, Jaeger, SkyWalking, and Harbor
- Experience with Public Cloud (AWS, Azure, GCP) or Private Cloud (Alibaba Cloud) services such as ACK and EDAS is highly preferred
- Solid understanding of enterprise data platform or data warehouse architecture, with experience participating in the design or implementation of a data warehouse from scratch (0→1)
- Strong knowledge of data warehouse methodologies and dimensional modeling, including: Multi-layer architecture (Staging → ODS → DW/DWD → DWS → ADS, Star Schema, Snowflake Schema
- Hands-on experience with ETL orchestration tools such as Kettle (Pentaho Data Integration) and Apache Airflow
- Familiar with Alibaba Cloud Data Platform products such as MaxCompute (ODPS) and Data Transmission Service (DTS) is an advantage
- Strong SQL programming skills and scripting experience using Python and Shell
- Experience working with large-scale distributed data processing technologies is highly desirable
- Excellent analytical, problem-solving, and cross-functional collaboration skills
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search