Design and build complex ETL/ELT pipelines for structured and unstructured data sources using GCP native services like Cloud Storage, Dataflow, and Dataproc, Data Integration & ETL/ELT
•
Build and orchestrate automated data pipelines using Fivetran, Google DTS, and custom ETL/ELT frameworks.
•
Create robust ingestion patterns for batch and streaming data.
•
Troubleshoot data integration issues, latency problems, and pipeline failures.
•
Partner with application teams to integrate SaaS, on-prem, and cloud sources.
Implement BigQuery best practices, including data modeling, partitioning, clustering, and cost/performance tuning.
Orchestration & Automation:
Build reusable workflow frameworks using orchestration tools like Cloud Composer and Git-based CI/CD pipelines.
Data Quality & Governance:
Enforce strict data validation, observability, schema handling, and reconciliation for enterprise data.
Cross-Functional Collaboration:
Partner with data science, Data Analytics, AI, and business teams to translate requirements into technical solutions.
Security, Compliance & Governance
•
Implement data access controls leveraging IAM, service accounts, and VPC-SC.
•
Ensure compliance with organizational and industry data standards.
•
Design disaster recovery, backup, and high-availability strategies.
•
Experience with Data Governance Framework, Data Quality, Data Stewardship, Ontology and Data Lineage.
Collaboration & Leadership
•
Work closely with product owners, data engineers, analysts, and architects across the organization.
•
Serve as a technical advisor for large enterprise reporting and integration initiatives.
•
Mentor junior engineers and support cross-functional architecture reviews.