Software Engineer
On-siteBengaluru, Karnataka, India
Job Summary
Develop deep understanding of business context to present feature recommendations in agile environments. Lead, design, and code database solutions enabling data-driven decision-making for multi-faceted ad serving operations. Design, develop, and manage ETL/ELT pipelines in Databricks using PySpark/SparkSQL, integrating various data sources to support business operations. Mentor junior engineers while staying abreast of developments in data governance, quality, and performance optimization. Work closely with global engineering resources to ensure enterprise data warehouse solutions evolve in lockstep with changing business models.
Required Qualifications
- Bachelor's Degree in Computer Science or equivalent degree
- 2.5 – 5 years of data engineering experience with expertise using Apache Spark and Databases (preferably Databricks) in marketing technologies and data management, and technical understanding in these areas
- Monitor and tune Databricks workloads to ensure high performance and scalability, adapting to business needs as required
- Good experience in Basic and Advanced SQL writing and tuning
- Experience with Python
- Good understanding of CI/CD practices with experience in Git for version control and integration for spark data projects
- Solid understanding of Disaster Recovery and Business Continuity solutions
- Experience with scheduling applications with complex interdependencies, preferably Airflow
- Good experience in working with geographically and culturally diverse teams
- Understanding of data management concepts in both traditional relational databases and big data lakehouse solutions such as Apache Hive, AWS Glue or Databricks
- Excellent written and verbal communication skills
- Ability to handle complex products
- Excellent communication and problem-solving skills, with the ability to manage multiple priorities
- Ability to diagnose and troubleshoot problems quickly
- Detail oriented, able to multi-task, prioritize and able to quickly change priorities
- Good time management
- Experience working as a Data Engineer with strong database fundamentals and ETL background
- Experience working in a Data warehouse environment and dealing with data volume in terabytes and above
- Experience working in relation data systems, preferably PostgreSQL and SparkSQL
- Excellent designing and coding skills
- Proficient with bug tracking and test management toolsets to support development processes such as CI/CD
- Develop a deep understanding of the business context under which your team operates and present feature recommendations in an agile working environment
- Lead, design and code solutions on and off database for ensuring application access to enable data-driven decision making for the company's multi-faceted ad serving operations
- Working closely with Engineering resources across the globe to ensure enterprise data warehouse solutions and assets are actionable, accessible and evolving in lockstep with the needs of the ever changing business model
- Design, develop, and manage ETL/ELT pipelines in Databricks using PySpark/SparkSQL, integrating various data sources to support business operations
- Lead in the areas of solution design, code development, quality assurance, data modelling, business intelligence
- Mentor Junior engineers in the team
- Stay abreast of developments in the data world in terms of governance, quality and performance optimization
- Able to have effective client meetings, understand deliverables, and drive successful outcomes
Desired Qualifications
- Good to have knowledge of cloud platforms (cloud security) and familiarity with Terraform or other infrastructure-as-code tools
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.