Production experience with a distributed processing engine, a cloud data warehouse or lakehouse platform (e.g., Databricks/Spark, Snowflake, BigQuery, Redshift, Synapse/Fabric) and orchestration tooling (e.g., Airflow, Dagster, Databricks Workflows, Azure Data Factory), including batch and incremental ingestion from heterogeneous external sources (APIs, flat-file/SFTP feeds, third-party vendors, database replication/CDC)