Senior Data Engineer EMEA
On-siteLondon, England, United Kingdom
Job Summary
Drive the transition from code writing to AI orchestration by building agentic workflows for data engineering, CI/CD, and governance. Steward the AI-amplified Databricks framework while designing semantic models and ingestion pipelines for domains like custom MDM and geocoding. Own production troubleshooting and root cause analysis for pipeline failures, performance issues, and data quality problems by systematically working through code, logs, and documentation. Design cross-functional data products that bring global sources into a single coherent shape using advanced SQL, Python, and PySpark.
Required Qualifications
- Around 6 or more years of relevant experience
- A degree in a quantitatively rigorous field such as computer science, data science, econometrics, mathematics, or physics
- Systems thinking
- Production troubleshooting and root cause analysis
- Data modeling
- Advanced SQL, including window functions, query optimization, and MERGE/UPSERT operations
- Python and PySpark
- Deep experience with Databricks (Declarative Pipelines, DABS, Delta Lake, Spark optimization, job orchestration)
- Familiarity with the Azure ecosystem
- Working understanding of Unity Catalog
- Strong written communication
- Async updates and clear documentation
Desired Qualifications
- Real estate context is a plus but not required
- Advanced Spark optimization (broadcast joins, salting, partitioning strategies)
- Geospatial data processing (H3 indexes, spatial SQL, point in polygon at scale)
- Recursive CTEs and complex SQL patterns
- Infrastructure familiarity (Azure Portal, resource management, CLI)
- Strong Git workflows and code review habits
- Real estate domain experience
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.