Senior Data Engineer Greece
RemoteGreece
Job Summary
Architect and build production data pipelines and platforms using Apache Spark to ingest, transform, and serve terabyte-scale multimodal datasets. Design and operate multi-model data stores including relational, vector, and graph databases for semantic search, RAG systems, and attribution modelling. Orchestrate analytics workflows with dbt to create version-controlled data models that power downstream AI and BI systems. Ensure platform reliability through robust monitoring, data quality checks, and lineage tracking under live traffic. Lead technical direction by writing design docs, making build-vs-buy decisions, and mentoring mid-level and junior engineers. Own non-functional quality metrics such as latency, throughput, scalability, and cost for systems in your domain.
Required Qualifications
- 5+ years of professional software engineering experience
- shipping and operating production systems
- experience with scaling
- experience with reliability
- experience with on-call
- experience with the gap between a working prototype and a dependable service
- Deep, demonstrable expertise designing and building distributed data pipelines with Apache Spark
- strong data modelling across relational, vector, and graph databases
- Strong proficiency in at least one general-purpose language (e.g., Python, Scala, or Java)
- the ability to work effectively across others
- Hands-on experience with cloud platforms (GCP/AWS)
- experience with containers (Docker)
- experience with CI/CD
- experience with infrastructure-as-code (Terraform)
- Strong software engineering habits — version control
- Strong software engineering habits — testing
- Strong software engineering habits — code review
- Strong software engineering habits — CI/CD
- Comfort with ambiguity
- Clear communication
- the ability to write a one-page design doc that is useful for both product managers and staff engineers
Desired Qualifications
- Experience building and operating ML/LLM-powered production systems (model serving, RAG, agents) at scale
- Experience with event-driven or streaming architectures (e.g., Pub/Sub, Kafka) and real-time systems
- Depth in security, IAM, networking, and data governance in cloud environments
- Background in marketing technology, ad tech, or large-scale data products
- Meaningful open-source contributions
- a track record of technical leadership
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.