Senior Data Engineer
Indexed description
The Opportunity
As a Senior Data Engineer, you will be responsible for building and maintaining the infrastructure that supports data collection, processing, and storage, working closely with data scientists, analysts, and other stakeholders to ensure that data systems are reliable, scalable, and secure. Your work will be crucial in enabling data-driven decision-making across the organization. This is a key technical role focused on developing and optimizing the company's data infrastructure which involves designing and implementing data pipelines, ensuring data quality, and collaborating with cross-functional teams to support various data initiatives.
Responsibilities:
As a Senior Data Engineer, you will be responsible for developing and maintaining data systems to support the company’s strategic goals. Your role will encompass a range of activities focused on data pipeline development, data quality, and cross-functional collaboration.
- Data Pipeline Architecture and Development
- Data Integration
- API and Data Services
- Data Storage
Understand data engines and structure to effectively design solutions for transactional, analytics, and search purposes.
- Data Quality and Governance
- Collaboration and Support
- Security and Compliance
- Documentation
Qualifications:
Skills and attributes for success
- Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field; or equivalent work experience.
- 7+ years of experience in a Data Engineering role.
- Programming languages like Python and SQL and managing huge scale data potentially Terabyte to Petabyte.
- Hands-on experience with big data technologies like Spark (Using PySpark / Scala) and Flink.
- Familiarity with machine learning frameworks such as TensorFlow, PyTorch, or similar.
- Strong understanding of data warehousing or Lake-house concepts, ETL processes, and data modeling.
- Experience with API development and integration with data services.
- Experience with cloud platforms like Azure.
- Knowledge in DevOps, CI/CD methods, and containerization technologies like Docker or Kubernetes.
- Experience with real-time / streaming data processing.
- Programming Languages: Python, SQL
- Query Engine: Trino
- Big Data Technologies: Spark, Flink
- Unstructured Data: Text, Image, Audio & Video
- Databases: Clickhouse, MySQL, PostgreSQL, MongoDB, Cassandra, HBase, Redis
- Cloud Platforms: Azure
- API Development: RESTful APIs, GraphQL, OpenAPI
- Data Services: Kafka, RabbitMQ
- Containers: Docker, Kubernetes
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search