To enable all of this on Databricks, making data ingestion seamless is crucial. That’s the mission of the Ingestion Core Team: to make the ingestion of all data—structured and unstructured—simple, reliable, and efficient. Simplifying the complex is hard, and that’s where you come in. This role requires building distributed platform systems to incrementally ingest high-volume, petabyte-scale data from diverse sources—including cloud storage (SQS, ADLS, GCS), databases (Oracle, SQL Server, MySQL, Postgres), and file sources (Google Drive, SharePoint)—at high throughput and low cost. The data includes structured formats (JSON, Parquet, CSV) as well as unstructured data (text, images, docs, PPTs, and blobs), all of which land in Delta Lake with schema evolution and change data capture (CDC) capabilities.