PySpark Data Engineer
Tata Consultancy Services · Kolkata, West Bengal, India
Tata Consultancy Services · Kolkata, West Bengal, India
TCS is Hiring PySpark Data Engineer Location: **Face to Face Interview on 8th August at Tata Consultancy Services Limited, Gitanjali Park Campus, IT/ITES SEZ, Plot- IIF / 3, (Block K - Ground Floor) Action Area - II, New Town, Rajarhat, Kolkata - 700156, West Bengal** Role Summary We are seeking a **PySpark Data Engineer** with strong experience in building scalable data pipelines, big data processing, and cloud-based data platforms. The ideal candidate will have expertise in PySpark, Databricks, SQL, and modern data engineering practices to support enterprise analytics and data-driven initiatives. Key Responsibilities - Design, develop, and maintain ETL/ELT pipelines using PySpark and Databricks. - Process and transform large-scale structured and unstructured datasets. - Build scalable batch and real-time data processing solutions. - Optimize Spark jobs for performance, scalability, and cost efficiency. - Implement data quality, governance, and monitoring frameworks. - Integrate data from multiple enterprise and cloud sources. - Collaborate with business, analytics, and architecture teams to deliver data solutions. - Support CI/CD, deployment automation, and production troubleshooting.\\ Mandatory Skills - Strong hands-on experience in **PySpark, Python, and SQL** - Expertise with **Azure Databricks** or similar Spark platforms - Experience in designing and optimizing **data pipelines** - Knowledge of **Spark SQL, DataFrames, UDFs, performance tuning** - Experience with **Azure Data Factory (ADF)** and Azure Data Lake - Strong understanding of ETL/ELT concepts and data modeling