Junior SRE
$150,000–$150,000 year
On-siteLondon, England, United Kingdom
Job Summary
Design and maintain data pipelines for machine learning training and inference, while developing workflow orchestration components for scheduling and resource allocation. Analyze system performance and scalability, then improve the reliability and observability of ML or HPC platforms. Collaborate with quantitative researchers to translate requirements into robust technical solutions, contributing production-quality code that supports research and trading workloads. You will learn existing systems, contribute independently to ongoing work, and progressively take ownership of well-defined platform areas with support from senior engineers. This role offers strong mentorship and exposure to real-world challenges in large-scale infrastructure.
Required Qualifications
- A degree (BSc or MSc) in Computer Science, Engineering, Applied Mathematics, or a related field, or equivalent practical experience.
- Strong foundations in computer science (e.g. algorithms, data structures, operating systems, distributed systems).
- Experience developing software in at least one language such as Python, C++, or Rust.
- Familiarity with writing, testing, and maintaining production‐quality code.
- Ability to reason clearly about technical problems and communicate effectively with teammates.
- Interest in infrastructure, performance, and large‐scale systems.
- Exposure to machine learning workflows, data engineering, or ML platforms.
- Experience working in Linux environments.
- Introductory knowledge of distributed systems, HPC, or cloud platforms for CPU and GPU workloads.
- Experience with profiling, benchmarking, or performance analysis.
Desired Qualifications
- Familiarity with CUDA or other GPU development frameworks.
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.