Data Engineer Greece - Mid Level
On-siteLondon, England, United Kingdom
Job Summary
Design, build, and maintain large-scale ETL/ELT pipelines using Apache Spark to ingest and transform terabyte-scale multimodal datasets. Operate polyglot data infrastructure across relational, vector, and graph databases to support semantic search, RAG systems, and attribution modeling. Develop documented dbt models that ensure data quality and traceability for downstream AI and BI systems. Instrument pipelines with logging and metrics to diagnose production issues while contributing to CI/CD and infrastructure-as-code. Participate in code reviews and design discussions to ship reliable services in production.
Required Qualifications
- 2–4 years of professional data engineering experience, with work deployed to production
- Strong proficiency in at least one general-purpose language (e.g., Python, Scala, or Java) and comfort working across a codebase
- Solid fundamentals in designing and building data pipelines (ETL/ELT), with experience in Apache Spark for large-scale data processing
- Experience with cloud platforms (GCP/AWS), containers (Docker), and CI/CD
- Strong software engineering practices — Git, testing, code review, CI/CD
- Clear communication — you can explain technical choices and trade-offs to both technical and non-technical colleagues
Desired Qualifications
- Experience integrating ML/LLM systems into production applications (model serving, RAG, agents)
- Familiarity with infrastructure-as-code (Terraform), orchestration, and observability tooling
- Experience with event-driven or streaming architectures (e.g., Pub/Sub, Kafka)
- Exposure to security, IAM, and data governance in cloud environments
- Background in marketing technology, ad tech, or large-scale data products
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.