Job Summary
Proven hands-on experience with Databricks for large-scale data engineering and processing.
Experience with healthcare data formats: X12 (834, 837), JSON, XML, flat files, and Excel.
Strong knowledge of MongoDB and relational databases (Oracle, SQL Server) for querying, reconciliation, and data validation.
Programming experience in Python and basic to intermediate Java.
Responsibilities
- Minimum 2+ years of hands-on Databricks experience for building and maintaining ETL/data pipelines (PySpark, SQL) and should have worked on Databricks in recent project.
- Process healthcare data formats (X12, JSON, XML, flat files, Excel) at scale.
- Implement data ingestion, transformation, validation, and optimization workflows.
- Work with:
- NoSQL databases (MongoDB)
- Relational databases (Oracle, SQL Server)
- Perform data validation, reconciliation, and quality checks.
- Collaborate with business stakeholders, data architects, and analysts.