Responsibilities
- Accelerate researchers by taking ownership of complex parts of large-scale pipelines and delivering reliable internal tools
- Bridge research and product by exposing clean APIs, automating model pushes, and surfacing live metrics
- Write efficient, well-tested Python and systems code, enforcing code review, CI, and observability
- Design and optimize distributed services for Kubernetes or SLURM environments with thousands of GPU jobs
- Prototype utilities like CLI tools and dashboards, then evolve them into stable, shared libraries
Requirements
- Master’s degree in Computer Science or equivalent experience
- 4+ years building and operating large-scale or distributed systems
- Strong software design instincts including modular code, tests, CI/CD, and observability
- Fluency in Python plus one systems language such as C++, Rust, Go, or Java
- Hands-on experience with container orchestration and schedulers like Kubernetes, K8s, SLURM, or similar
- Comfortable profiling performance, optimizing I/O, and automating workflows
- Self-starter with low ego, collaborative, and high energy
Nice to Have
- Exposure to ML workloads or data-processing pipelines
- Experience with GPU clusters or CUDA
- Open-source contributions or widely used internal tools
Benefits
- Healthcare coverage
- Parental leave
- Retirement plans
- Relocation support
- Wellness programs
- Meal and transportation allowances
- Other location-specific perks
Work Arrangement
Hybrid — Paris, Warsaw, Zurich, London
Other
- Language: Not specified
- Travel: Not specified
- Hours: Not specified
- Shifts: Not specified
- Equipment: Not specified
- Clearance: Not specified
- Background checks: Not specified
- Contract duration: Not specified
- Probation: Not specified
- Relocation: Relocation support offered
- Training: Not specified