Principal HPC Network Engineer (remote in the EU) (Barcelona)

Principal HPC Network Engineer (remote in the EU) (Barcelona)

09 ago
|
Mirantis
|
Barcelona

09 ago

Mirantis

Barcelona

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

¿Es usted el solicitante adecuado para esta ocasión? Descúbralo leyendo el resumen del puesto a continuación.

Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen. Learn more at Description

We are seeking a highly skilled Senior HPC Networking Engineer to design, deploy, manage, and troubleshoot high-performance networking environments. The ideal candidate will have deep expertise in InfiniBand technologies, strong general networking knowledge, and hands‑on experience with Fortinet solutions. You will play a critical role in ensuring the performance, reliability, and scalability of HPC infrastructure.

Key Responsibilities

-

- Design, deploy, and maintain high-performance network infrastructures for HPC environments, with a strong focus on InfiniBand fabrics.





- Troubleshoot complex network issues across InfiniBand and Ethernet environments, ensuring minimal downtime and optimal performance.

- Manage and optimize InfiniBand components, including switches, HCAs, subnet managers, and fabric configurations.

- Perform performance tuning, monitoring, and capacity planning for HPC networking systems.

- Implement and maintain network security using Fortinet solutions (FortiGate, FortiManager, FortiAnalyzer).

- Diagnose and resolve issues related to routing, switching, latency, and throughput across hybrid network environments.

- Collaborate with compute, storage, and platform teams to support HPC workloads and cluster operations.

- Develop and maintain documentation for network architecture, configurations, and operational procedures.

- Participate in on‑call rotations and provide escalation support for critical incidents.

- Lead or contribute to network upgrades, migrations, and new deployments.

Qualifications
Required

-

- 5+ years of experience in network engineering, with a focus on HPC or data center environments.

- Strong hands‑on experience with InfiniBand technologies (e.g., Mellanox/NVIDIA).

- Solid understanding of networking fundamentals: TCP/IP, routing protocols (BGP, OSPF), VLANs, QoS, and network design.

- Proven experience deploying and troubleshooting Fortinet solutions (FortiGate, FortiManager, VPNs, firewall policies).





- Experience with network performance analysis and troubleshooting tools.

- Familiarity with Linux systems and scripting for automation (e.g., Bash, Python).

- Strong analytical and problem‑solving skills.

Preferred

-

- Experience with large‑scale HPC clusters or AI/ML infrastructure.

- Knowledge of RDMA, MPI, and low‑latency networking concepts.

- Certifications such as FCSS/FCNSP (Fortinet), CCNP/CCIE, or equivalent.

- Experience with automation and Infrastructure as Code tools (e.g., Ansible, Terraform).

Soft Skills

-

- Strong communication and collaboration skills.

- Ability to work independently and handle complex technical challenges.

- Detail‑oriented with a proactive approach to problem‑solving.

Additional Information
What We Offer

-

- Operate some of the most advanced AI infrastructure environments in production today.

- Work with the latest NVIDIA GPU technologies, Kubernetes platforms, and high‑performance networking environments.

- Help define operational standards and reliability practices for next‑generation AI infrastructure services.

- Influence the adoption of AI‑powered operational capabilities through k0rdent AI.

- Work alongside highly skilled engineers solving complex infrastructure and platform challenges at scale.

- Join a growing organisation investing heavily in AI infrastructure, platform services, and operational innovation.

We are a Leader for Container Management in G2 (#2 after AWS)!

We are a Leader for Container Management in G2 (#2 after AWS)! xqbhyrx

#J-18808-Ljbffr
Hay opciones de teletrabajo/trabajo desde casa disponibles para este puesto.

📌 Principal HPC Network Engineer (remote in the EU) (Barcelona)
🏢 Mirantis
📍 Barcelona

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: principal hpc network engineer (remote in the eu) (barcelona) / barcelona

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: principal hpc network engineer (remote in the eu) (barcelona) / barcelona