Description
The role focuses on building and maintaining batch and streaming data pipelines, moving production data into archival or data-lake storage, improving dataset quality and observability, designing data-validations, creating reusable datasets, and optimizing queries and processing jobs. It requires senior-level data-pipeline experience, strong Python and SQL, production Python or Scala experience, and expertise with PySpark, Spark SQL, Spark Streaming, Parquet, and Iceberg. The position is onsite five days per week in central Paris and includes 100% health-care coverage.
