PySpark Developer – ETL & SQL
On-siteHyderabad, Telangana, India
Job Summary
Develop and maintain robust data pipelines using PySpark to transform large datasets into meaningful insights. Design scalable ETL processes for business and analytics needs while writing optimized SQL queries for extraction, transformation, validation, and reporting. Ensure data quality and accuracy across sources, identify performance bottlenecks, and collaborate with cross-functional teams to deliver reliable data solutions. Participate in code reviews and contribute to engineering best practices. This role supports a data engineering team in Hyderabad with hybrid work options, offering comprehensive medical coverage and flexible hours.
Required Qualifications
- 4-9 years experience
- 4–8 years of experience in Data Engineering or related roles
- Strong hands-on experience with PySpark and large-scale data processing
- Solid expertise in designing and developing ETL pipelines
- Advanced SQL skills, including query optimization and performance tuning
- Good understanding of data warehousing and data modeling concepts
- Strong problem-solving skills and attention to detail
- Excellent communication and collaboration skills
Desired Qualifications
- Experience with AWS
- Exposure to Airflow or other workflow orchestration tools
- Familiarity with Hadoop ecosystem technologies
- Experience working with modern data lake or cloud data platforms
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.