Databricks Developer
Persistent Systems · State of Mahārāshtra, India
Free to search · AI fit score against your CV · tailor your résumé in one click
Persistent Systems · State of Mahārāshtra, India
Job Description About Persistent We are an AI-led, platform-driven Digital Engineering and Enterprise Modernization partner, combining deep technical expertise and industry experience to help our clients anticipate what's next. Our offerings and proven solutions create a unique competitive advantage for our clients by giving them the power to see beyond and rise above. We work with many industry-leading organizations across the world, including 20 Fortune 50 companies and 4 of the 5 top banks in both the US and India, and numerous innovators across the healthcare ecosystem. Our disruptor's mindset, commitment to client success, and agility to thrive in the dynamic environment have enabled us to sustain our growth momentum. Persistent has been recognized across top industry platforms for innovation, leadership, and inclusion. We reported $1,654.4M FY26 revenue with 17.4% Y-o-Y growth. We have delivered 24 sequential quarters of growth with $436.0M in Q4 FY26 revenue, up 3.2% Q-o-Q and 16.2% Y-o-Y growth. Our 27,500+ global team members, located in 18 countries, have been instrumental in helping the market leaders transform their industries. About Position: We are looking for an experienced Databricks Developer / Lead with 7+ years of experience in Data Engineering, Analytics, and Cloud Data Platforms. The ideal candidate will possess deep expertise in Databricks, Apache Spark, distributed data processing, and cloud-native data architectures. This role involves designing, developing, and optimizing large-scale data pipelines while ensuring high performance, scalability, reliability, and operational excellence across enterprise data platforms. • Role: Databricks Developer • Location: Pune • Experience: Between 7 to 12 Years • Job Type: Full Time Employment What You'll Do: • Design, develop, and maintain scalable ETL and ELT pipelines using Databricks, PySpark, SQL, and Notebooks. • Build and manage data processing workflows across Bronze, Silver, and Gold layers using Delta Lake architecture. • Develop batch and real-time streaming data pipelines using Spark Structured Streaming. • Optimize Spark jobs, clusters, and workloads for performance, scalability, and cost efficiency. • Implement and maintain data quality checks, validation frameworks, and monitoring processes. • Collaborate with data architects, analysts, and business stakeholders to understand requirements and deliver robust data solutions. • Develop and maintain reusable data engineering frameworks and components. • Integrate Databricks solutions with cloud-native services on Azure, AWS, or GCP. • Manage Databricks cluster configurations, autoscaling, job orchestration, and workload optimization. • Implement CI/CD pipelines and code deployment processes using Git and DevOps best practices. • Troubleshoot and resolve issues related to data pipelines, cluster performance, storage, and platform environments. • Support production deployments and ensure platform reliability and operational stability. • Contribute to architecture discussions and recommend improvements to data platform design and implementation. • Mentor junior team members and share best practices in Databricks and Spark development. • Drive continuous improvement initiatives across the data engineering ecosystem. Expertise You'll Bring: • 7 to 12 years of experience in Data Engineering, Big Data, Analytics, or Data Platform development. • Hands-on experience delivering multiple enterprise-scale Databricks implementations and migration projects. • Strong expertise in Databricks workspace administration, job orchestration, and notebook development. • Extensive experience with Apache Spark, PySpark, Spark SQL, and distributed computing concepts. • Deep understanding of Spark internals, execution plans, partitioning, caching, and optimization techniques. • Strong experience implementing Delta Lake architecture and Medallion data design patterns. • Hands-on experience building batch and streaming data processing solutions. • Experience with cloud platforms such as Azure, AWS, or GCP, with strong expertise in at least one cloud ecosystem. • Strong knowledge of Databricks cluster management, autoscaling, and resource optimization. • Experience designing and implementing large-scale data lakes and modern data architectures. • Proficiency in SQL development, query optimization, and performance tuning. • Experience with Git, CI/CD pipelines, DevOps practices, and automated deployments. • Strong understanding of data governance, security, access controls, and compliance standards. • Experience working with structured, semi-structured, and unstructured datasets. • Excellent analytical, debugging, troubleshooting, and problem-solving skills. • Strong communication and stakeholder management abilities. • Experience leading teams, mentoring engineers, and driving technical excellence. Benefits: • Competitive salary and benefits package • Culture focused on talent development with quarterly growth opportunities and company-sponsored higher education and certifications • Opportunity to work with cutting-edge technologies • Employee engagement initiatives such as project parties, flexible work hours, and Long Service awards • Annual health check-ups • Insurance coverage: group term life, personal accident, and Mediclaim hospitalization for self, spouse, two children, and parents Values-Driven, People-Centric & Inclusive Work Environment: Persistent is dedicated to fostering diversity and inclusion in the workplace. We invite applications from all qualified individuals, including those with disabilities, and regardless of gender or gender preference. We welcome diverse candidates from all backgrounds. • We support hybrid work and flexible hours to fit diverse lifestyles. • Our office is accessibility-friendly, with ergonomic setups and assistive technologies to support employees with physical disabilities. • If y