Senior AI Data Engineer
On-siteKraków, Lesser Poland, Poland
Job Summary
Design, build, and operate enterprise-grade data and AI platforms as code using GitOps principles, GitHub, and Azure DevOps. Implement pull-request-driven change control, automated testing, and CI/CD pipelines while defining Infrastructure-as-Code for data and AI systems. Engineer scalable ETL/ELT pipelines supporting AI/ML model training, feature stores, and near-real-time inference workflows. Enable GenAI and RAG by curating data sources, managing embeddings, and maintaining vector databases. Implement proactive monitoring for data quality, pipeline performance, and distribution shifts, integrating security and compliance controls directly into pipelines. Build and operate Azure-based platforms including storage, compute, and orchestration, optimizing for cost, performance, and scale. Partner with Product, Engineering, and Security teams to translate requirements into durable architectures and produce clear documentation.
Required Qualifications
- Bachelor's Degree in Computer Science, Data Engineering, Engineering, or equivalent practical experience.
- 5 - 7 years of experience in data engineering, platform engineering, or infrastructure roles.
- Strong proficiency in Python and SQL, with working fluency in JSON, YAML, and shell scripting.
- Experience using Gitbased workflows, Infrastructure as Code, and CI/CD pipelines to build and operate data and AI platforms in production environments.
- Experience operating workloads in Azure and AWS.
- Has performed direct and operational applications of large language models (LLMs) and GenAI platforms (OpenAI, Anthropic Claude, Google Gemini) within enterprise controlled environments.
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.