Data Engineer
HybridBengaluru, Karnataka, India
Bengaluru, Karnataka, IndiaHybridFull TimeSmall
Full TimeSmall
Job Summary
Design and maintain scalable data pipelines using Python, PySpark, and SQL across AWS and Databricks environments. Build and orchestrate ETL workflows with Airflow, Kafka, and Hive while managing infrastructure via Terraform. Implement CI/CD processes using GitHub and ensure data quality through exposure to Parquet and Delta Lake formats. This role requires 6–8 years of experience and is based in Gurgaon, Pune, Bangalore, Chennai, Hyderabad, Bhopal, or Jaipur. Applicants must provide clean profiles with actual experience proofs and documentation.
Required Qualifications
- 6+ Years
- 7+ Years
- 8+ Years
- Programming languages -Python/PySpark + Java or Scala
- Database expertise -SQL, NoSQL
- Cloud data platforms - AWS
- Big data frameworks -Databricks, Spark, Kafka, Hive
- Workflow orchestration (Airflow, ETL, Data Pipelines)
- Infrastructure management skills (Terraform)
- Datawarehouse Exposure
- Fundamental understanding of Parquet, Delta Lake and other OTFs file formats
- Strong experience on building CI/CD processes, experience with GitHub
- Actual experience
- Proofs of docs
- On Site: (3 Days / week or 12 Days / Month Work from Office)
- Locations: Gurgaon, Pune, Bangalore, Chennai, Hyderabad, Bhopal, Jaipur
Desired Qualifications
- GCP
- dbt
- Gurgaon
- Pune
- Bangalore
- Chennai
- Hyderabad
- Bhopal
- Jaipur
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.