Improving
Getonbrd · Posted today
Senior Data Engineer
Continue to application
Add your email once, then Caio opens the original posting.
Indexed description
- 5+ years of experience in data engineering with a strong focus on pipeline development.
- Hands-on expertise with Databricks, including Delta Lake, Databricks Workflows, and Unity Catalog.
- Proficiency in PySpark and Apache Spark for large-scale distributed data processing.
- Experience building pipelines with Spark Declarative Pipelines (Delta Live Tables).
- Working knowledge of dbt for SQL-based transformation development and testing.
- Experience processing healthcare data including claims, clinical, EMR/EHR, eligibility, or provider data.
- Proficiency in Python and SQL with a disciplined approach to code quality and testing.
- Experience with cloud infrastructure (Azure, AWS, or GCP) and CI/CD pipeline tooling.
- Understanding of data modeling methodologies: Data Vault 2.0, 3NF, and dimensional modeling (Kimball).
- Familiarity with healthcare data compliance considerations (HIPAA) and data standards (HL7, FHIR, ICD-10, CPT, NPI).
- Experience with data orchestration tools such as Apache Airflow or Azure Data Factory.
- Exposure to streaming data pipelines using Kafka or Databricks Structured Streaming.
- Databricks Certified Data Engineer Associate or Professional certification.
Projects
We are seeking a skilled Senior Data Engineer to design, build, and deliver data engineering solutions on our Databricks-based healthcare data platform. This is a hands-on individual contributor role, executing within established engineering standards to build and maintain production-grade data pipelines, contribute to our migration from SQL Server to Databricks, and uphold platform quality and reliability. The Senior Data Engineer works closely with the Principal Data Engineer and broader team to deliver data engineering workstreams across ingestion, transformation, and serving layers.- Design, build, and maintain production-grade data pipelines using PySpark and Spark Declarative Pipelines (Delta Live Tables) on Databricks.
- Develop modular, tested, and documented dbt transformation models as part of the team's analytics engineering layer.
- Implement and maintain multi-layer data architectures (Bronze/Silver/Gold) on Delta Lake following medallion architecture principles.
- Contribute to the migration of data pipelines, stored procedures, and SSIS packages from SQL Server to Databricks, ensuring logic fidelity and data integrity.
- Ensure pipeline reliability through robust error handling, data quality checks, and SLA-aligned alerting.
- Manage Databricks compute resources, cluster configurations, and job scheduling via Databricks Workflows.
- Enforce data security and access controls within pipeline and storage design via Unity Catalog.
- Participate in code reviews, contribute to engineering discussions, and help uphold team-wide coding standards.
- Write clean, maintainable Python and SQL code following version control and CI/CD best practices.
Benefits
- Major Medical Expense Insurance
- Life Insurance
- Dental and Vision Insurance
- Mental Health Support
- IMSS (Mexican Social Security)
- Seniority Bonus
- Savings Fund Program
- Career Development Plan
- Christmas Bonus (Aguinaldo)
- Vacation Bonus
- Corporate Retirement Plan
- Certifications and Training Programs
- Internal Events
- TotalPass Wellness Program
- Additional Paid Time Off
Additional Protection and Discounts
- Auto and Motorcycle Insurance
- Pet Insurance
- Personal Belongings Insurance
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search
Want help applying to roles like this?
Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search