Backend Software Engineer
Join a growing engineering team building an orchestration platform for AI data centers. Develop backend services that bring compute, networking, and accelerators together, and own features from implementation through production.
Responsibilities
- Build and maintain core services and features for an orchestration platform spanning compute, networking, and accelerators.
- Translate customer requirements into APIs, workflows, and data models in collaboration with the team.
- Deliver features from prototyping and implementation through testing, deployment, and production support.
- Write reliable, maintainable code and contribute to shared libraries, automated tests, and CI/CD pipelines.
- Contribute to design discussions and make practical tradeoffs involving performance, reliability, and simplicity.
- Participate in code reviews, share knowledge, and improve engineering practices.
- Troubleshoot issues and improve service performance, observability, and operational readiness.
Requirements
- At least 5 years of professional software development experience, including building and maintaining production backend services.
- Strong backend development skills, preferably in Python, and experience building REST or GraphQL APIs.
- Practical understanding of distributed systems, including scalability, resilience, and failure handling.
- Experience with SQL or NoSQL databases, data modeling, and query performance.
- Experience writing automated tests and working with Git and CI/CD workflows.
- Familiarity with multithreading, locks, synchronization, and performance analysis and optimization.
- Ability to own features and collaborate and communicate effectively within a small team.
Nice to have
- Familiarity with TypeScript, React, or modern web application architecture.
- Experience with cloud platforms, containers, or Kubernetes.
- Experience with observability tools such as OpenTelemetry, Prometheus, Grafana, or Datadog.
- Familiarity with secure coding practices and application security risks.
- Experience with orchestration systems, provisioning pipelines, cluster management, or schedulers.
- Exposure to AI or GPU infrastructure, or high-performance networking.