Backend Engineer – ML Infrastructure
Build and own distributed data pipelines, orchestration, and production services for AI training and evaluation. Ensure reliability and performance. Use Python, JAX, C++, Rust, Go, Docker. Mid-level role in Israel.
Responsibilities
- Design and implement distributed data pipelines for training and evaluation data at scale.
- Build orchestration and infrastructure layer for researchers and AI engineers.
- Ensure production systems are fast, observable, and reliable.
- Write clean, well-tested code under production load.
- Use coding agents and AI-assisted development daily.
- Own well-scoped components and collaborate with senior engineers.
- Debug across unfamiliar codebases.
Requirements
- BSc/MSc/PhD in Computer Science or related field.
- 3+ years professional software engineering experience.
- Strong coding ability with production systems, open-source, or personal projects.
- Mastery of Python/JAX and at least one of C++, Rust, or Go.
- Hands-on experience with distributed systems, data pipelines, or backend services at scale.
- Solid computer science fundamentals.
- Comfort with Docker, cloud services, CI/CD, observability.
- Fluency with coding agents.
- Excellent English communication.
Nice to have
- 5+ years production engineering experience.
- ML infrastructure experience (training pipelines, model serving, experiment tracking).
- GPU programming or performance optimization.
- Cloud infrastructure expertise (AWS, GCP).
- Open-source contributions to infrastructure or data tooling.