Are you a HPC Infrastructure Reliability Engineer seeking a new interesting challenge? If your answer is yes, it’s your lucky day so keep reading, it can be just what you’re looking for. WHAT WILL YOU DO?
Manage and optimize high-performance physical infrastructure (servers, GPUs, and advanced networking) Ensure availability, performance, and reliability of HPC and AI environments Drive infrastructure automation (IaC) and enable zero-touch provisioning Oversee the full hardware lifecycle (capacity planning, deployment, and decommissioning) Work with tools such as HPE OneView, Lenovo XClarity, and ServiceNow CMDB Collaborate with R&D, science, and engineering teams to design optimal infrastructure solutions Optimize resource utilization (CPU/GPU) and improve overall infrastructure efficiency WHAT ARE WE LOOKING FOR? ~5–7+ years of experience in Data Center Engineering, Bare Metal, or HPC Infrastructure ~ Strong expertise in enterprise hardware (HPE, Lenovo) and high-performance systems ~ Hands‑on experience with GPUs (NVIDIA) and AI/HPC environments ~ Solid knowledge of high-speed networking (e.g., InfiniBand, high‑throughput Ethernet) ~ Proven experience in Infrastructure as Code (IaC) and automation (Python, Bash, Ansible, or similar) ~ Experience with infrastructure management tools such as HPE OneView and/or Lenovo XClarity ~ Good understanding of the hardware lifecycle (capacity planning, deployment, decommissioning) ~ Strong communication skills in English, with the ability to collaborate with technical and business stakeholders WHERE AND WHEN?
Workplace: Madrid (Hybrid model) Work Schedule: Business Hours WHAT CAN WE OFFER YOU? Permanent contract – We offer indefinite contracts from the first day.
Pay and benefits – Competitive salary and a versatile compensation plan adapted to your needs (Ticket restaurant plan, Childcare T
¿Te interesa este puesto?