DevOps Data Infrastructure Team Lead
Lead a team operating and improving production data infrastructure across cloud and on-premises environments. Combine hands-on engineering with technical leadership to maintain reliable, scalable streaming platforms and services for internal teams.
Responsabilidades
- Lead and mentor data infrastructure engineers, supporting technical growth and knowledge sharing
- Own the reliability, stability, and operational health of production data infrastructure
- Operate and maintain Kafka, Kafka Connect, Elasticsearch, Logstash, and Kibana platforms
- Design, deploy, and monitor scalable data flows across cloud and on-premises environments
- Drive infrastructure-as-code, CI/CD, and deployment automation
- Oversee internal tools and microservices that provide self-service access to data platform capabilities
- Collaborate with engineering, data, analytics, and platform teams on reliable data integration and delivery
- Define monitoring, alerting, and observability practices using tools such as Prometheus and Grafana
- Lead incident response, root-cause analysis, post-mortems, and preventive improvements
- Plan platform upgrades, capacity, and security improvements
- Manage operational priorities, technical roadmaps, and infrastructure improvements
Requisitos
- At least 5 years in platform engineering, SRE, DevOps, or production streaming infrastructure
- At least 3 years leading engineering teams
- Strong Linux and Windows administration, troubleshooting, shell scripting, and performance-tuning skills
- Experience with CI/CD, Git, and infrastructure-as-code tools such as Terraform or Ansible
- Production Kubernetes and Docker experience, including storage, networking, RBAC, and resource management
- Proficiency in Python or a similar language for infrastructure tools, automation, and microservices
- Hands-on production experience with Apache Kafka, Kafka Connect, and Schema Registry
- Working knowledge of Elasticsearch, Logstash, and Kibana
- Experience with Prometheus, Grafana, monitoring, and observability
- Understanding of distributed systems, event streaming, DevOps, and infrastructure automation
- Familiarity with AWS, GCP, or Azure and hybrid cloud/on-premises environments
- Understanding of production operations, incident management, infrastructure security, access control, and secrets management
- Track record improving large-scale data platforms and streaming pipelines
- Experience leading incident response and collaborating across engineering, data, analytics, and platform teams
Se valora
- Experience with Microsoft SQL Server or other RDBMS platforms
- Familiarity with Apache Flink or similar stream-processing frameworks
- Experience with multi-tenant platforms, RBAC, and platform governance
- Experience building self-service infrastructure platforms
Beneficios
- Hybrid work model
- Free parking and electric vehicle charging
- Health insurance, with family options and extensions
- Birthday gift and a day off during the birthday month
- Referral bonus or gift card
- HitechZone membership
- Gifts for holidays and life events
- Ten Bis