Clanx logo
ClanxPosted 1 month ago

Senior Data Engineer - Gurugram

$3,000,000–$4,000,000 year

HybridGurugram, Haryana, India

Full TimeSenior LevelBachelors Degree

Job Summary

Design, build, and maintain scalable batch ETL/ELT and CDC pipelines that ingest data from client systems into the data platform. Develop and optimize Spark/PySpark and SQL workloads for large-scale processing, ensuring pipelines are reliable, idempotent, and backfill-safe. Model and transform raw data into trusted analytical and ML-ready datasets while implementing data quality checks, lineage, monitoring, and alerting mechanisms. Collaborate with ML, product, and engineering teams to support feature pipelines and analytics platforms, driving best practices for scalability and maintainability across the data stack. This role supports Mechademy's enterprise AI solutions for industrial asset monitoring and predictive maintenance across oil & gas, power generation, and LNG sectors.

Required Qualifications

  • 3–7 years of hands-on experience building production-grade data pipelines
  • Strong expertise in SQL, including query optimization and complex transformations
  • Strong Python programming skills for data engineering workloads
  • Experience with Spark/PySpark or equivalent distributed processing frameworks
  • Strong understanding of data modeling, ETL/ELT, CDC, incremental processing, and partitioning
  • Experience with workflow orchestration tools such as Dagster, Airflow, or Prefect
  • Experience working with AWS or Azure cloud platforms
  • Strong focus on data quality, observability, validation, and monitoring
  • Experience building scalable datasets for analytics and machine learning use cases
  • Bachelor's degree in Computer Science, Engineering, Mathematics, or a related field
  • Gurugram - Hybrid (2–3 days on-site)

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce