Proven hands-on experience with Databricks for
large-scale data engineering and processing.
Experience with healthcare data formats: X12
(834, 837), JSON, XML, flat files, and Excel.
Strong knowledge of MongoDB and relational
databases (Oracle, SQL Server) for querying, reconciliation, and data
validation.
Programming experience in Python and basic to
intermediate Java.
🔹 Responsibilities
Minimum 2+ years of hands-on Databricks
experience for building and maintaining ETL/data pipelines (PySpark, SQL)
and should have worked on Databricks in recent project.
Process healthcare data formats (X12, JSON, XML,
flat files, Excel) at scale.
Implement data ingestion, transformation,
validation, and optimization workflows.
Work with:
NoSQL databases (MongoDB)
Relational databases (Oracle, SQL Server)
Perform data validation, reconciliation, and
quality checks.
Collaborate with business stakeholders, data
architects, and analysts.
Optimize performance, scalability, and
reliability of data pipelines.