Description
Databricks is hiring an Ingestion Core Engineer to build distributed platform systems that ingest high-volume, petabyte-scale structured and unstructured data from cloud storage, databases, and file sources into Delta Lake. The role focuses on streaming ingestion, incremental processing, replication, throughput, latency, reliability, cost optimization, monitoring, observability, and AI/ML evaluation, while collaborating with teams supporting RAG and AI agents. Candidates should have at least five years of production coding experience in Java, Scala, Go, C++, or Python, along with experience in large-scale distributed and asynchronous systems, streaming, Spark, databases, data processing, or CDC.
