Senior GPU HPC & AI Infrastructure Engineer

Hermes Corporate

Italia

In loco

EUR 70.000 - 120.000

Tempo pieno

3 giorni fa
Candidati tra i primi

Ricevi più risposte dai datori di lavoro

Invia un CV specifico per questa offerta in pochi minuti.

Descrizione del lavoro

Hermes Corporate seeks an experienced AI & HPC Infrastructure Engineer to operate, optimise and evolve high-performance computing platforms that power advanced engineering and scientific workloads.

This hands-on role requires strong Linux and HPC expertise, ownership of production platforms, and the ability to troubleshoot complex systems across GPU compute, storage, networking, containers, and platform performance.

Competenze

  • Proven experience in Linux systems engineering at an advanced level.
  • Proficient in operating production HPC environments and large-scale parallel workloads.
  • Hands-on administration of GPU compute infrastructure and NVIDIA software stacks.
  • Experience with containerisation in HPC environments and multi-node workloads.
  • Strong debugging, profiling, and performance-tuning skills in complex compute infra.
  • Ability to own production infrastructure and serve as senior escalation point.

Mansioni

  • Administer GPU compute infrastructure including drivers, software stack, firmware, and hardware health.
  • Operate and optimise HPC clusters for reliability and performance.
  • Refine resource allocation policies across workloads for efficient capacity use.
  • Manage containerised environments for platform users, including provisioning and GPU access.
  • Maintain shared high-speed storage and support interconnects operation and troubleshooting.
  • Lead performance engineering activities: profiling, benchmarking, bottleneck analysis, platform optimisation.
  • Support integration and optimisation of engineering and scientific applications on GPU-accelerated platforms.
  • Provide incident response, root-cause analysis, and change management for maintenance planning.
  • Maintain platform monitoring, alerting, and operational reporting.
  • Develop and maintain technical docs, procedures, and platform standards.
  • Contribute to security, architecture governance, and compliance activities.

Conoscenze

Linux systems engineering
HPC cluster administration
NVIDIA GPU compute stack
Containerisation in HPC
Performance troubleshooting
End-to-end ownership

Strumenti

NVIDIA drivers
CUDA toolkit
Container tooling (Docker/Singularity)
Storage & networking tech

Descrizione del lavoro

Hermes Corporate seeks an experienced AI & HPC Infrastructure Engineer to operate, optimise and evolve high-performance computing platforms that power advanced engineering and scientific workloads.

This hands-on role requires strong Linux and HPC expertise, ownership of production platforms, and the ability to troubleshoot complex systems across GPU compute, storage, networking, containers, and platform performance.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

HPC ENGINEER
HPC ENGINEER

Hermes Corporate • Italia

In loco
EUR 70.000 - 120.000
AI/ML Engineer: Deploy & Optimize LLMs at Scale
AI/ML Engineer: Deploy & Optimize LLMs at Scale

Hermes Corporate • Italia

In loco
EUR 70.000 - 100.000
Linux Systems Administrator - HPC Simulations
Linux Systems Administrator - HPC Simulations

Allegro MicroSystems, LLC • Milano

Ibrido
EUR 55.000 - 85.000
Hybrid work model
Staff AI Storage Platform Architect
Staff AI Storage Platform Architect

Radian Arc • Lombardia

In loco
EUR 120.000 - 190.000
Infrastructure Engineer
Infrastructure Engineer

European Tech Recruit • Millan

In loco
EUR 50.000 - 70.000
Artificial Intelligence Specialist
Artificial Intelligence Specialist

Hermes Corporate • Italia

In loco
EUR 70.000 - 110.000
HPC & AI Infrastructure Engineer (Hybrid, Rome)
HPC & AI Infrastructure Engineer (Hybrid, Rome)

ELT Group • Roma

Ibrido
EUR 36.000 - 44.000
Premio legato ai risultati di business
Welfare
Assicurazione sanitaria
+2
Senior Linux HPC Systems Engineer (Hybrid)
Senior Linux HPC Systems Engineer (Hybrid)

Do IT Now • Emilia-Romagna

Ibrido
EUR 60.000 - 90.000
Birthday off
Professional development
Remote work possibility
+1
Senior HPC/AI Platform Delivery Lead (Hybrid)
Senior HPC/AI Platform Delivery Lead (Hybrid)

AI4I Foundation • Piemonte

Ibrido
EUR 70.000 - 90.000
Staff Network Engineer for AI Fabric, Data Center & Edge
Staff Network Engineer for AI Fabric, Data Center & Edge

Triwill Group • Italia

Remoto
EUR 120.000 - 180.000