Senior HPC & AI Infrastructure Engineer

Hermes Corporate

Venezia

In loco

EUR 70.000 - 100.000

Tempo pieno

3 giorni fa
Candidati tra i primi
Generatore di candidature

Una candidatura completa in un minuto — curriculum e lettera di presentazione personalizzati, pronti da inviare.

Supera i filtri ATS

Descrizione del lavoro

Hermes Corporate is seeking an experienced AI & HPC Infrastructure Engineer to operate, optimise, and evolve high-performance computing infrastructure supporting engineering and scientific workloads.

This hands-on role requires strong Linux and HPC expertise, ownership of production platforms, and the ability to troubleshoot complex systems, drive improvements across GPU compute, storage, networking, containers, and platform performance.

Competenze

  • Substantial professional experience in Linux systems engineering / administration at an advanced level.
  • Proven experience operating production HPC environments, including cluster administration and large-scale parallel workloads.
  • Hands-on experience administering GPU compute infrastructure, particularly NVIDIA-based environments and the associated software stack.
  • Practical experience with containerisation in HPC environments, including GPU device access and multi-node workloads.
  • Strong troubleshooting and performance-tuning skills across complex compute infrastructure.
  • Demonstrated ability to take end-to-end ownership of production infrastructure and act as a senior technical escalation point.
  • Strong analytical and problem-solving skills with a production-focused mindset.
  • Professional working proficiency in English.

Mansioni

  • Administer and maintain GPU compute infrastructure, including system configuration, NVIDIA drivers and software stack, firmware, and hardware health monitoring.
  • Operate, monitor, and optimise HPC clusters to ensure reliability, performance, and efficient use of compute resources.
  • Define and refine resource allocation policies across different workload types to ensure effective and equitable use of available capacity.
  • Manage containerised environments for platform users, including user provisioning, environment maintenance, GPU access, and standards for container usage.
  • Administer shared and high-performance storage and support the operation and troubleshooting of high-speed interconnects.
  • Lead performance engineering activities, including profiling, benchmarking, bottleneck analysis, and platform optimisation.
  • Support the integration and optimisation of engineering and scientific applications on GPU-accelerated and HPC platforms.
  • Provide incident response, troubleshooting, root-cause analysis, and change management, including planning and execution of maintenance activities.
  • Maintain platform monitoring, alerting, and operational reporting.
  • Develop and maintain technical documentation, operational procedures, and platform standards.
  • Contribute to security, architecture governance, and compliance activities.

Conoscenze

Linux systems engineering
HPC environments
GPU compute infrastructure
Containerisation
Troubleshooting
Production ownership
English proficiency

Strumenti

NVIDIA drivers
Container runtimes

Descrizione del lavoro

Hermes Corporate is seeking an experienced AI & HPC Infrastructure Engineer to operate, optimise, and evolve high-performance computing infrastructure supporting engineering and scientific workloads.

This hands-on role requires strong Linux and HPC expertise, ownership of production platforms, and the ability to troubleshoot complex systems, drive improvements across GPU compute, storage, networking, containers, and platform performance.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Senior GPU HPC & AI Infrastructure Engineer
Senior GPU HPC & AI Infrastructure Engineer

Hermes Corporate • Italia

In loco
EUR 70.000 - 120.000
Senior AI & HPC Infrastructure Engineer
Senior AI & HPC Infrastructure Engineer

Hermes Corporate • Lodi

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Lodi

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Cuneo

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Lazio

In loco
EUR 60.000 - 90.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Venezia

In loco
EUR 70.000 - 100.000
AI/ML Engineer: Deploy & Optimize LLMs at Scale
AI/ML Engineer: Deploy & Optimize LLMs at Scale

Hermes Corporate • Italia

In loco
EUR 70.000 - 100.000
AI/ML Engineer — Design, Fine-Tune & Deploy Models
AI/ML Engineer — Design, Fine-Tune & Deploy Models

Hermes Corporate • Pistoia

In loco
EUR 65.000 - 90.000
AI/ML Engineer: Build Scalable Real-World AI Solutions
AI/ML Engineer: Build Scalable Real-World AI Solutions

Hermes Corporate • Plasencia

In loco
EUR 50.000 - 80.000
HPC & AI Infrastructure Engineer
HPC & AI Infrastructure Engineer

The Italian Institute of Artificial Intelligence (AI4I) • Torino

Ibrido
EUR 40.000 - 60.000
Office in Torino technology hub
Competitive compensation
Access to advanced computing resources
+1