DevOps Lead
Lead cloud architecture and DevOps operations for a company developing training and simulation systems. Own the DevOps roadmap, guide infrastructure delivery, and provide hands-on technical leadership across engineering.
Обязанности
- Own cloud architecture vision, design, and implementation with a focus on security, scalability, and reliability.
- Define and evolve the DevOps roadmap, tooling, and operating model across research and development.
- Partner with software architects and developers on deployment patterns, service architecture, and production infrastructure.
- Lead, mentor, and develop the DevOps team while setting technical standards and conducting design reviews.
- Build and maintain CI/CD pipelines, release processes, and delivery automation.
- Implement Infrastructure as Code and configuration management for reproducible, auditable environments.
- Establish cloud security practices covering IAM, secrets management, network segmentation, secure baselines, and compliance.
- Design monitoring, logging, alerting, and tracing practices; define SLOs and SLAs.
- Improve incident response readiness, cloud cost efficiency, performance, capacity planning, and scaling.
- Provide hands-on technical execution for early-stage architecture and key infrastructure components.
Требования
- At least 7 years of experience in DevOps, SRE, platform engineering, or infrastructure roles.
- At least 2 years of team leadership or technical leadership of a DevOps function.
- Strong experience designing and operating AWS, Azure, or GCP infrastructure.
- Experience with secure, scalable architectures, IAM, encryption, secrets management, and network security.
- Hands-on experience with CI/CD and modern release practices.
- Strong Infrastructure as Code experience with Terraform, CloudFormation, Pulumi, or similar tools.
- Experience with automation scripting using Python, Bash, or similar languages.
- Production experience with Docker, Kubernetes, and Helm, including scaling.
- Experience with observability platforms and incident management processes.
- Solid Linux, networking, and complex production troubleshooting knowledge.
- Strong communication and cross-functional collaboration skills.