Senior IT/Lab Manager
Lead the planning, deployment, and day-to-day operation of physical engineering labs and IT infrastructure supporting research, quality assurance, validation, and other engineering teams.
Обязанности
- Own operations, planning, and the roadmap for engineering labs and IT infrastructure, including servers, storage, networking, and related services.
- Lead and mentor the IT/Lab team while establishing standards, guidelines, ownership, and continuous improvement practices.
- Partner with engineering teams to design, provision, and maintain secure, reliable, high-performance environments.
- Manage data-center and lab operations, including rack layouts, cabling, power, cooling, hardware lifecycles, and resource availability.
- Lead procurement and vendor relationships for hardware, software, and infrastructure services.
- Automate system provisioning, configuration, and operational processes using shell, Perl, and Ansible.
- Design and maintain monitoring, logging, and alerting for server, network, and storage systems.
- Investigate and resolve complex infrastructure issues and support rapid incident response.
Требования
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
- At least 10 years of experience in IT or systems administration, including extensive Linux/Unix experience.
- At least 3 years in an IT, lab, or infrastructure management or team-lead role.
- Extensive Linux/Unix administration experience, including installation, configuration, troubleshooting, and performance tuning.
- Experience supporting engineering organizations such as R&D, quality engineering, and verification teams.
- Experience managing data-center and lab environments, including server, network, and storage equipment.
- Experience with infrastructure procurement and vendor management.
- Proficiency in shell or Perl scripting and Ansible for provisioning, configuration, and operations.
- Hands-on experience with infrastructure monitoring and alerting.
- Strong debugging skills across operating systems, networking, storage, virtualization, and application layers.
Будет плюсом
- Experience with Kubernetes in on-premises or hybrid environments.
- Hands-on experience with Slurm, HPC clusters, or large-scale compute environments.
- Background in HPC, large-scale Linux clusters, or performance-sensitive engineering environments.