Senior Data Engineer
Indexed description
Senior Data Engineer – Databricks / AI Data Privacy
We’re looking for a Senior Data Engineer to help build an enterprise-grade, AI-driven data masking and synthetic data framework from the ground up using Databricks.
📍 Remote – Midwest, Chicago, New York & New Jersey only
Key Responsibilities
- Design and build scalable solutions using Databricks Lakehouse, Unity Catalog, Delta Lake & PySpark
- Develop data masking and classification capabilities using AI/ML, NLP/NER
- Build high-volume synthetic data generation solutions while preserving referential integrity
- Develop enterprise ETL/ELT pipelines and automation
- Support CI/CD, cloud deployments, governance, security, and data privacy
- Contribute to architecture, implementation, SIT, deployment, and knowledge transfer
Requirements
- Strong hands-on Databricks, Unity Catalog, Delta Lake & PySpark experience
- Strong Python and cloud data engineering background
- Experience with Azure and/or AWS
- Experience with data architecture, ETL/ELT, CI/CD, and data governance
- Strong understanding of data security, quality, and integrity
Preferred
- NLP / NER / AI-ML experience
- Synthetic data generation
- Data masking / de-identification / TDM
- FPE and secure key management
- MLflow / Databricks ML
- Financial services or regulatory experience
Ideal candidate: A hands-on Senior Data Engineer with strong Databricks and cloud experience who can contribute to both architecture and implementation of an enterprise data privacy platform.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search