Description
The Data Quality Engineer will design and operate a Databricks lakehouse data quality layer using Databricks Labs DQX. Responsibilities include profiling source and Bronze data, defining completeness, uniqueness, validity, range, referential, and custom SQL checks, quarantining invalid records, embedding checks in Databricks Jobs and Lakeflow pipelines, storing results in Delta tables under Unity Catalog, translating legacy SQL validation logic, setting quality-gate thresholds, and deploying rules through CI/CD. The role requires at least three years of data engineering or data quality experience, PySpark and Spark SQL expertise, DQX or comparable framework experience, Delta Lake and Unity Catalog knowledge, and familiarity with Git and CI/CD.
