• 5+ years of hands-on data engineering experience, building and maintaining production-grade data platforms and pipelines.
• Strong programming skills in general-purpose language (such as Python or Scala) for data processing, and SQL for data analytics.
• Deep experience with distributed data processing frameworks, such as Apache Spark, including performance tuning and optimization.
• Proven experience building data solutions using services on AWS (EMR, Lambda, s3, etc).
• Strong understanding of data modeling and data warehousing concepts, including partitioning, schema design for large-scale datasets.
• Experience operating and supporting production pipelines, including monitoring, alerting, incident response, and improving reliability over time.
• Solid foundation in software engineering practices, including version control, CI/CD, testing strategies, and code review.
• Strong communication and collaboration skills, with the ability to work effectively with both technical and non-technical stakeholders.