← Back to jobs
N

Infrastructure Engineering Manager

NVIDIA AI·Israel·en
Not specifiedFull-timeEngineering ManagementComputer Hardware

Lead and grow an infrastructure engineering team supporting firmware, driver, hardware, software, and verification organizations. Own technical direction, delivery, customer support, and operational reliability for Linux-based engineering infrastructure.

Responsibilities

  • Lead and develop an infrastructure engineering team responsible for bare-metal provisioning, virtual machine infrastructure, server fleet automation, CI/CD infrastructure, debug support, and high-performance networking environments.
  • Set the technical roadmap, priorities, execution plans, and delivery commitments across infrastructure initiatives.
  • Own VM inventory and lifecycle management, including image readiness, OS compatibility, package baselines, kernel configuration, provisioning, and production availability.
  • Build reliable infrastructure for provisioning, testing, validation, and debugging workflows used by engineering and verification teams.
  • Lead customer support, technical triage, root-cause analysis, bottleneck removal, and workflow optimization for internal engineering stakeholders.
  • Guide system debugging and recovery involving server bring-up, driver and firmware interactions, boot failures, networking, lab instability, automation failures, and environment recovery.
  • Provide technical leadership for Linux automation platforms covering server lifecycle management, OS installation, driver setup, inventory, resource allocation, observability, and production readiness.
  • Partner with firmware, driver, hardware, software, cloud, and verification teams to define requirements and improve infrastructure reliability.

Requirements

  • Bachelor’s degree in Computer Engineering, Computer Science, Electrical Engineering, or a related technical field, or equivalent experience.
  • At least 8 years of experience in Linux systems administration, infrastructure automation, DevOps, systems software, firmware infrastructure, lab infrastructure, or a related engineering field.
  • At least 3 years leading or managing engineering teams, technical projects, or cross-functional infrastructure initiatives.
  • Strong Linux expertise, including systemd, package management, kernel parameters, GRUB, sysctl, NFS, networking, boot flows, and service management.
  • Hands-on experience designing, implementing, and debugging Python and scripting automation, CI/CD workflows, and software development practices.
  • Experience managing multiple Linux distributions, OS images, provisioning flows, package dependencies, compatibility, and environment consistency.
  • Ability to support internal customers through triage, root-cause analysis, workflow optimization, incident handling, and cross-functional communication.
  • Experience coaching, mentoring, hiring, managing performance, prioritizing work, and building inclusive engineering teams.

Nice to have

  • Experience leading infrastructure teams supporting firmware R&D, hardware bring-up, driver development, lab automation, cloud provisioning, or large-scale engineering environments.
  • Knowledge of RDMA, InfiniBand, Ethernet, OFED, SR-IOV, PCI passthrough, VFIO/IOMMU, or related Linux networking technologies.
  • Experience with Ansible, infrastructure as code, Jenkins, Kubernetes, Docker, KVM, QEMU, libvirt, Vagrant, or multi-architecture environments.
  • Familiarity with NVIDIA or Mellanox hardware, firmware tools, hardware diagnostics, BIOS/BMC automation, and server recovery workflows.
  • Experience with Redfish, iDRAC, iLO, IPMI, automated remediation, and fleet incident reduction.

Benefits

  • Competitive salary
  • Generous benefits package
  • Inclusive and equal-opportunity work environment

Relevance

More opportunities

Similar jobs

The newest open roles in Engineering Management.

Questions, answered

Frequently asked questions

Infrastructure Engineering Manager – Linux and DevOps | CVZilla