GCP & Python Data Engineer (Hibrido, Porto)
Indexed description
Descrição da Função
Principais Responsabilidades: Design and maintain end-to-end ELT/ETL pipelines for high-volume, multi-source data. Build data models, warehouses and lakes on GCP using BigQuery, Cloud Storage, Dataflow, Dataproc and Pub/Sub. Develop Python transformations, orchestration, automation and data-quality checks. Integrate and operationalise AI prediction models within cloud data pipelines. Plan and execute UAT and QA activities, data validation and release-readiness controls. Implement monitoring, logging, alerting, governance, security, privacy and compliance controls. Maintain technical documentation, data dictionaries, runbooks and test plans. Requisitos Mínimos: Strong Python development, including design patterns, testing, packaging and performance considerations. Hands-on GCP experience with BigQuery, Cloud Storage, Dataflow / Beam, Dataproc, Pub/Sub and Cloud Composer / Airflow. SQL, data modelling, schema design and production data-pipeline experience. Understanding of AI prediction models, data preparation, feature stores and MLOps principles. Git, orchestration and CI/CD for data pipelines. French at C1 minimum, written and spoken, for active stakeholder engagement. Professional English and strong communication, documentation and autonomous delivery skills.
Localização
- Porto, Portugal
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search