Job description

We are seeking an experienced **Data Engineer** with strong expertise in **Google Cloud Platform (GCP), PySpark, and Data Product development**. The ideal candidate should have hands-on experience designing and building scalable data pipelines, data lakes, and data products using GCP-native services. The candidate must be proficient in distributed data processing, cloud data engineering, and modern data architectures. Key Responsibilities - Design, develop, and maintain large-scale data pipelines using **PySpark**. - Build and deliver reusable **Data Products** for analytics, reporting, and business consumption. - Develop ETL/ELT processes on **Google Cloud Platform (GCP)**. - Work with structured, semi-structured, and streaming data sources. - Design and implement data lake and lakehouse solutions. - Optimize Spark workloads for performance and cost efficiency. - Collaborate with business, architecture, and engineering teams to define data product requirements. - Implement data quality, governance, lineage, and monitoring frameworks. - Support CI/CD automation and DevOps practices for data engineering solutions. - Contribute to Data Mesh and modern data platform initiatives.