Senior Data Engineer (Databricks)
Remote
Job Summary
Design, build, and maintain ETL/ELT pipelines using Python, Databricks, and Spark to move data from multiple source systems into a data lake and warehouse. Create transformation rules, data models, and quality checks while implementing monitoring, logging, and orchestration practices. Maintain data catalogues, support clear data lineage, and prepare technical documentation using Git, Docker, and CI/CD tools. Collaborate with Product Analysts, Data Scientists, and ML Engineers to improve data availability and usability. Requires 7+ years of experience, strong Python and SQL skills, and knowledge of AWS cloud services. Based in the EU for a global pharmaceutical company in Prague, with a start date of July/August 2026.
Required Qualifications
- 7+ years of relevant data engineering experience
- Strong hands-on experience with Python
- Practical ETL/ELT experience on Databricks
- Strong knowledge of AWS cloud services
- Solid database experience and strong SQL skills
- 2–3 years of Spark/PySpark experience
- Experience with Git, Docker, and CI/CD tools
- Good understanding of SDLC documentation
- Ability to communicate clearly with engineering, analytics, and data science teams
- Experience in healthcare, pharma or similar
Desired Qualifications
- Experience with data governance, data catalogues, or data lineage tools
- Experience supporting machine learning or advanced analytics teams
- Experience in regulated or enterprise environments
- Knowledge of modern data lakehouse architecture
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.