Senior HPC & AI Infrastructure Engineer

Hermes Corporate

Cuneo

In loco

EUR 70.000 - 110.000

Tempo pieno

2 giorni fa
Candidati tra i primi
Generatore di candidature

Non inviare un curriculum generico — genera un curriculum e una lettera di presentazione personalizzati per questo specifico impiego.

Supera i filtri ATS

Descrizione del lavoro

Hermes Corporate is seeking an experienced AI & HPC Infrastructure Engineer to operate, optimise, and evolve GPU-accelerated computing platforms that support engineering and scientific workloads.

The role demands hands-on Linux and HPC proficiency, ownership of production platforms, and strong problem-solving skills to drive performance and reliability across GPU compute, storage, networking, containers, and platform performance.

Competenze

  • Advanced Linux systems engineering experience.
  • Production HPC environment operation experience.
  • Hands-on GPU compute infrastructure administration (NVIDIA stack).

Mansioni

  • Administer and maintain GPU compute infrastructure, including system configuration, NVIDIA drivers and software stack, firmware, and hardware health monitoring.
  • Operate, monitor, and optimise HPC clusters to ensure reliability, performance, and efficient use of compute resources.
  • Define and refine resource allocation policies across workload types for effective and equitable capacity use.
  • Manage containerised environments for platform users, including user provisioning, environment maintenance, GPU access, and container usage standards.
  • Administer shared and high-performance storage and support troubleshooting of high-speed interconnects.
  • Lead performance engineering activities, including profiling, benchmarking, bottleneck analysis, and platform optimisation.
  • Support integration and optimisation of engineering and scientific applications on GPU-accelerated and HPC platforms.
  • Provide incident response, troubleshooting, root-cause analysis, and change management, including maintenance planning and execution.
  • Maintain platform monitoring, alerting, and operational reporting.
  • Develop and maintain technical documentation, procedures, and platform standards.
  • Contribute to security, architecture governance, and compliance activities.

Conoscenze

Linux systems engineering
Production HPC ownership
Troubleshooting and performance tuning
Analytical and problem-solving
English proficiency

Strumenti

NVIDIA GPU software stack
Containerisation (GPU)
GPU device access
Cluster administration tools

Descrizione del lavoro

Hermes Corporate is seeking an experienced AI & HPC Infrastructure Engineer to operate, optimise, and evolve GPU-accelerated computing platforms that support engineering and scientific workloads.

The role demands hands-on Linux and HPC proficiency, ownership of production platforms, and strong problem-solving skills to drive performance and reliability across GPU compute, storage, networking, containers, and platform performance.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Senior HPC & AI Infrastructure Engineer
Senior HPC & AI Infrastructure Engineer

Hermes Corporate • Venezia

In loco
EUR 70.000 - 100.000
Senior AI & HPC Infrastructure Engineer
Senior AI & HPC Infrastructure Engineer

Hermes Corporate • Lodi

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Cuneo

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Lodi

In loco
EUR 70.000 - 110.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Lazio

In loco
EUR 60.000 - 90.000
Hpc Engineer
Hpc Engineer

Hermes Corporate • Venezia

In loco
EUR 70.000 - 100.000
AI/ML Engineer: Deploy & Optimize LLMs at Scale
AI/ML Engineer: Deploy & Optimize LLMs at Scale

Hermes Corporate • Italia

In loco
EUR 70.000 - 100.000
AI/ML Engineer — Design, Fine-Tune & Deploy Models
AI/ML Engineer — Design, Fine-Tune & Deploy Models

Hermes Corporate • Pistoia

In loco
EUR 65.000 - 90.000
AI/ML Engineer: Build Scalable Real-World AI Solutions
AI/ML Engineer: Build Scalable Real-World AI Solutions

Hermes Corporate • Plasencia

In loco
EUR 50.000 - 80.000
HPC & AI Infrastructure Engineer
HPC & AI Infrastructure Engineer

The Italian Institute of Artificial Intelligence (AI4I) • Torino

Ibrido
EUR 40.000 - 60.000
Office in Torino technology hub
Competitive compensation
Access to advanced computing resources
+1