All About You
• Hands-on experience in data engineering, data integration, and large-scale data processing.
• Strong proficiency in SQL and Python for data transformation, automation, and analytics workloads.
• Solid understanding of data modeling, database design, and performance optimization techniques.
• Experience designing and developing ETL/ELT solutions across modern data platforms.
• Knowledge of data quality frameworks, testing methodologies, and data validation practices.
• Familiarity with cloud-based or modern enterprise data platforms.
• Experience with source control, CI/CD pipelines, and software engineering best practices.
• Strong analytical and problem-solving skills with the ability to troubleshoot complex data issues.
• Excellent communication and collaboration skills with the ability to work effectively across technical and business teams.
• Self-motivated, detail-oriented, and committed to delivering high-quality, scalable solutions.
Preferred Qualifications
• Experience with PySpark and Apache Spark for distributed data processing.
• Experience working within the Hadoop ecosystem and large-scale data environments.
• Familiarity with GitLab, Jenkins, or similar DevOps and CI/CD platforms.
• Exposure to Power BI or other business intelligence and data visualization tools.
• Understanding of modern data architecture patterns supporting analytics, machine learning, and AI workloads.
Education
• Bachelor’s or Master’s Degree in a Computer Science, Information Technology, Engineering, Mathematics, Statistics, M.S./M.B.A. preferred