SVP of Production Engineering & Autonomous Operations
Seeking an executive leader for production engineering and autonomous operations. Report directly to CEO, build AI-driven incident response agents, ensure platform reliability, and lead a small senior team. Hands-on coding required. Fully remote.
Responsibilities
- Own platform reliability and autonomous operations.
- Track availability, MTTR, and customer satisfaction weekly.
- Design, extend, and govern autonomous agents for incident response, auto-remediation, and RCA generation.
- Personally drive high-severity incidents and write root-cause analyses.
- Write and ship agent code.
- Serve as executive-facing operations leader for enterprise customers.
- Recruit and retain top-tier senior engineers.
Requirements
- 10+ years in production engineering, SRE, or SaaS operations at scale.
- 3+ years at SVP, VP, or Head-of level leading senior-only engineering org.
- Experience building, deploying, or leading agent-driven incident response and auto-remediation systems.
- Deep AWS operational expertise.
- Executive-caliber communication skills.
- Fluent or advanced English.
- US-morning time-zone overlap (~13:00–17:00 UTC).
- OFAC-clear country of residence.
Nice to have
- Recognized voice in autonomous operations (writing, talks, open-source).
- Multi-tenant B2B SaaS background.
- Familiarity with observability stacks (Grafana, Prometheus, Loki, Datadog, PagerDuty, OpsGenie).
- Side project or deep hobby demonstrating problem-solving.