DevOps Team Lead
Lead and develop DevOps and site reliability engineering practices for high-scale cloud services or SaaS products. The role combines hands-on cloud infrastructure work with technical leadership, automation, and production operations.
Обязанности
- Lead and mentor DevOps and site reliability engineers.
- Design, implement, and operate cloud infrastructure across AWS, GCP, or Azure.
- Develop and maintain infrastructure-as-code and automation workflows.
- Operate Kubernetes and container-based production environments.
- Design and improve CI/CD pipelines and release processes.
- Monitor production systems and manage logging and reliability operations at scale.
- Apply AI tools to engineering automation, troubleshooting, and productivity.
- Support cloud networking, IAM, load balancing, and managed services.
- Operate managed databases and production data infrastructure.
- Collaborate with engineering teams on cloud-native and distributed-system delivery.
Требования
- At least 6 years of DevOps or site reliability engineering experience, including at least 2 years in a leadership role.
- Production experience with cloud infrastructure and cloud-native architectures, microservices, and distributed systems.
- Hands-on experience with Docker, Kubernetes, or other container orchestrators.
- Strong infrastructure-as-code experience with Terraform, Ansible, CloudFormation, Pulumi, or similar tools.
- Strong experience designing and operating CI/CD pipelines and release processes.
- Hands-on use of AI tools to improve engineering productivity, automation, troubleshooting, and daily workflows.
- Experience with monitoring and logging solutions in large-scale production environments.
- Strong networking fundamentals, including load balancing and network protocols.
- Ability to lead and mentor engineers in a fast-paced environment.
- Strong scripting or software development skills, including the ability to read, write, and review production code.
- Experience operating PostgreSQL or other managed databases in production.
- Strong interpersonal, collaboration, and teamwork skills.
Будет плюсом
- Experience with MongoDB and PostgreSQL or other databases.
- Experience with Node.js or other backend technologies.
- Experience with Pulumi, Terraform, Ansible, CloudFormation, or comparable infrastructure tools.
- Experience with AWS, GCP, or Azure cloud environments.