Principal HPC Network Engineer (remote in the EU)

Mirantis

Barcelona

Presencial

EUR 90.000 - 120.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Mirantis seeks a Senior HPC Networking Engineer to design, deploy, and manage high-performance networks for AI and data-intensive workloads. The role focuses on InfiniBand fabrics, Fortinet security, and seamless integration with Kubernetes-based infrastructure.

You will diagnose and optimize hybrid network environments, collaborating with compute and storage teams to ensure scalable, reliable HPC operations and secure data flows across on-prem and cloud-like environments.

Formación

  • 5+ years of experience in network engineering, with a focus on HPC or data center environments.
  • Strong hands-on experience with InfiniBand technologies (e.g., Mellanox/NVIDIA).
  • Solid understanding of TCP/IP, routing (BGP, OSPF), VLANs, QoS, and network design.
  • Proven experience deploying and troubleshooting Fortinet solutions (FortiGate, FortiManager, VPNs).
  • Experience with network performance analysis and troubleshooting tools.
  • Familiarity with Linux systems and scripting for automation (Bash, Python).
  • Strong analytical and problem-solving skills.

Responsabilidades

  • Design, deploy, and maintain high-performance network infrastructures for HPC environments, focusing on InfiniBand fabrics.
  • Troubleshoot complex network issues across InfiniBand and Ethernet environments to minimize downtime.
  • Manage InfiniBand components: switches, HCAs, subnet managers, fabric configs.
  • Perform performance tuning, monitoring, and capacity planning for HPC networking systems.
  • Implement and maintain network security using Fortinet solutions (FortiGate, FortiManager, FortiAnalyzer).
  • Collaborate with compute/storage/platform teams to support HPC workloads and cluster operations.
  • Develop and maintain documentation for network architecture and procedures.
  • Participate in on-call rotations and provide escalation support for incidents.
  • Lead or contribute to network upgrades, migrations, and new deployments.

Conocimientos

InfiniBand
Fortinet
Networking fundamentals
Linux scripting
Troubleshooting

Herramientas

FortiGate
FortiManager
FortiAnalyzer
Bash
Python
Mellanox InfiniBand

Descripción del empleo

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen. Learn more at www.mirantis.com.

Job Description

We are seeking a highly skilled Senior HPC Networking Engineer to design, deploy, manage, and troubleshoot high-performance networking environments. The ideal candidate will have deep expertise in InfiniBand technologies, strong general networking knowledge, and hands‑on experience with Fortinet solutions. You will play a critical role in ensuring the performance, reliability, and scalability of HPC infrastructure.

Key Responsibilities
  • Design, deploy, and maintain high-performance network infrastructures for HPC environments, with a strong focus on InfiniBand fabrics.
  • Troubleshoot complex network issues across InfiniBand and Ethernet environments, ensuring minimal downtime and optimal performance.
  • Manage and optimize InfiniBand components, including switches, HCAs, subnet managers, and fabric configurations.
  • Perform performance tuning, monitoring, and capacity planning for HPC networking systems.
  • Implement and maintain network security using Fortinet solutions (FortiGate, FortiManager, FortiAnalyzer).
  • Diagnose and resolve issues related to routing, switching, latency, and throughput across hybrid network environments.
  • Collaborate with compute, storage, and platform teams to support HPC workloads and cluster operations.
  • Develop and maintain documentation for network architecture, configurations, and operational procedures.
  • Participate in on‑call rotations and provide escalation support for critical incidents.
  • Lead or contribute to network upgrades, migrations, and new deployments.
Qualifications
Required
  • 5+ years of experience in network engineering, with a focus on HPC or data center environments.
  • Strong hands‑on experience with InfiniBand technologies (e.g., Mellanox/NVIDIA).
  • Solid understanding of networking fundamentals: TCP/IP, routing protocols (BGP, OSPF), VLANs, QoS, and network design.
  • Proven experience deploying and troubleshooting Fortinet solutions (FortiGate, FortiManager, VPNs, firewall policies).
  • Experience with network performance analysis and troubleshooting tools.
  • Familiarity with Linux systems and scripting for automation (e.g., Bash, Python).
  • Strong analytical and problem‑solving skills.
Preferred
  • Experience with large‑scale HPC clusters or AI/ML infrastructure.
  • Knowledge of RDMA, MPI, and low‑latency networking concepts.
  • Certifications such as FCSS/FCNSP (Fortinet), CCNP/CCIE, or equivalent.
  • Experience with automation and Infrastructure as Code tools (e.g., Ansible, Terraform).
Soft Skills
  • Strong communication and collaboration skills.
  • Ability to work independently and handle complex technical challenges.
  • Detail‑oriented with a proactive approach to problem‑solving.
Additional Information
What We Offer
  • Operate some of the most advanced AI infrastructure environments in production today.
  • Work with the latest NVIDIA GPU technologies, Kubernetes platforms, and high‑performance networking environments.
  • Help define operational standards and reliability practices for next‑generation AI infrastructure services.
  • Influence the adoption of AI‑powered operational capabilities through k0rdent AI.
  • Work alongside highly skilled engineers solving complex infrastructure and platform challenges at scale.
  • Join a growing organisation investing heavily in AI infrastructure, platform services, and operational innovation.

We are a Leader for Container Management in G2 (#2 after AWS)!

We are a Leader for Container Management in G2 (#2 after AWS)!

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

HPC Network Engineer
HPC Network Engineer

Mirantis • Barcelona

Presencial
EUR 60.000 - 90.000
Senior HPC Networking Architect - InfiniBand & Fortinet
Senior HPC Networking Architect - InfiniBand & Fortinet

Mirantis • Barcelona

Presencial
EUR 90.000 - 120.000
Remote HPC Network Engineer: InfiniBand & AI Infra
Remote HPC Network Engineer: InfiniBand & AI Infra

Mirantis • Barcelona

Presencial
EUR 60.000 - 90.000
HPC Networking Engineer — InfiniBand & AI Infra
HPC Networking Engineer — InfiniBand & AI Infra

Mirantis • Barcelona

Presencial
EUR 45.000 - 55.000
Work with advanced AI infrastructure environments
Exposure to NVIDIA GPU technologies
Involvement in defining next-gen AI operations
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Mirantis • Barcelona

Presencial
EUR 90.000 - 120.000
Competitive compensation package
Professional development and training
Conferences and working groups
+1
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Mirantis • Bellprat

Presencial
EUR 90.000 - 120.000
Competitive compensation package
Professional development and training
Conference attendance and tech talks
+1
Senior Network Engineer
Senior Network Engineer

Hamilton Barnes ? • España

Presencial
EUR 80.000 - 110.000
AI Infrastructure Solutions Engineer
AI Infrastructure Solutions Engineer

DDN • Madrid

Presencial
EUR 70.000 - 110.000
HPC Administrator
HPC Administrator

Atos SE • Madrid

Presencial
EUR 42.000 - 65.000
Network Engineer
Network Engineer

Barcelona Technical Center • l'Hospitalet de Llobregat

Presencial
EUR 60.000 - 90.000