SoftServe
Linkedin · Posted 7d ago
Senior Big Data Engineer (Python + GCP)
Continue to application
Add your email once, then Caio opens the original posting.
Indexed description
About The RoleIn this role, you will contribute to building scalable and reliable data solutions on Google Cloud Platform, leveraging modern Big Data technologies and cloud-native services. You will work with batch and streaming data processing systems, helping organizations transform, manage, and unlock the value of their data. As part of a collaborative engineering team, you will participate in the full project lifecycle, from discovery and solution design to implementation and production deployment.
Responsibilities
- Design, develop, and maintain scalable data pipelines for batch and streaming workloads
- Build and optimize data processing solutions using Python (must), SQL, Java, Apache Spark, and Databricks
- Create cloud-native data architectures leveraging GCP services including BigQuery, Dataflow, Cloud Composer, Pub/Sub, and Cloud Storage
- Implement data transformation, modeling, and analytics engineering practices using dbt and Dataform
- Collaborate with business stakeholders, architects, and engineering teams to translate data requirements into effective technical solutions
- Support development of modern data platforms and analytics solutions on GCP with GCP native services and/or Databricks
- Contribute to technical design discussions, architecture decisions and continuous improvement initiatives
- Ensure data quality, reliability and performance across data processing workflows
- Support end-to-end project delivery, including PoCs, MVPs, production deployments and platform enhancements
- 5+ years of professional experience in Big Data or Data Engineering
- Advanced expertise in Python and SQL (Java nice to have) for large-scale data processing and transformation
- Hands-on experience developing data solutions on Google Cloud Platform (GCP)
- Experience with Apache Spark and data processing frameworks such as Cloud Dataflow or Apache Beam
- Proven background building scalable solutions using Databricks and Lakehouse concepts
- Experience with orchestration tools such as Apache Airflow or Cloud Composer
- Knowledge of streaming technologies such as Apache Kafka or Google Cloud Pub/Sub
- Strong knowledge of BigQuery and modern cloud data architectures
- Experience with data transformation and modeling tools such as dbt and Dataform
- Strong analytical thinking, troubleshooting capabilities, and problem-solving skills
- Effective communication with both technical and non-technical stakeholders
- Upper-intermediate or higher level of English
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search
Want help applying to roles like this?
Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search