Senior Storage Infrastructure Engineer for AI/ML & HPC

Nvidia

New South Wales

On-site

AUD 180,000 - 240,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NVIDIA is seeking a Production Storage Engineer to design, deploy, and optimize large-scale storage systems for GPU cloud services. You will focus on high availability, low latency, and data integrity across distributed storage environments, supporting AI/ML workflows and enterprise-grade storage deployments.

You will automate operations, monitor performance with the Elastic Stack and Prometheus, and lead capacity planning while maintaining strict security and compliance.

Qualifications

  • BS degree or equivalent in Computer Science, Storage Systems, or related field with 8+ years practical experience.
  • Experience with distributed and high‑performance storage solutions including clustered/parallel file systems and enterprise storage.
  • Solid understanding of block, file, object storage and their performance characteristics.
  • Experience with storage networking protocols such as NFS, SMB, iSCSI, S3, Fibre Channel, RDMA, NVMe over Fabrics.
  • Expertise in algorithms, data structures, software design, and automating maintenance of large-scale Linux‑based storage systems.
  • Experience in one or more: C/C++, Java, Python, Go, NodeJS, Bash for storage automation and tuning.
  • Hands-on with Ansible, Chef, Puppet, Terraform for automating deployments; observability with InfluxDB, Prometheus, Grafana, Elastic Stack.
  • Excellent written and oral communication; strong teamwork and commitment to daily delivery.

Responsibilities

  • Design, build, and support large-scale storage clusters ensuring scalability, high availability, and data integrity.
  • Develop and maintain storage monitoring, logging, and alerting systems for proactive issue resolution.
  • Collaborate on AI/ML workloads to optimize storage for low latency and high throughput.
  • Enhance storage service lifecycles from design to deployment and ongoing optimization.
  • Supervise production storage infrastructure, monitor latency and system health using predictive analytics and automation.
  • Improve storage efficiency via compression, deduplication, tiering, and intelligent workload placement.
  • Scale storage with AI/ML automation and secure data access with encryption and auditing.
  • Participate in on-call rotations for storage and production system reliability.

Skills

Distributed storage
High-performance storage
Automation
Linux
C/C++
Python
Go
Networking
Data management
Block/File/Object storage

Education

BS degree or equivalent in Computer Science/Storage Systems

Tools

Ansible
Chef
Puppet
Terraform
InfluxDB
Prometheus
Grafana
Elastic stack

Job description

NVIDIA is seeking a Production Storage Engineer to design, deploy, and optimize large-scale storage systems for GPU cloud services. You will focus on high availability, low latency, and data integrity across distributed storage environments, supporting AI/ML workflows and enterprise-grade storage deployments.

You will automate operations, monitor performance with the Elastic Stack and Prometheus, and lead capacity planning while maintaining strict security and compliance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Storage Systems Architect for AI/ML & HPC
Senior Storage Systems Architect for AI/ML & HPC

NVIDIA • Sydney

On-site
AUD 180,000 - 240,000
Senior Storage Production Engineer - DGX Cloud
Senior Storage Production Engineer - DGX Cloud

NVIDIA • Sydney

On-site
AUD 180,000 - 240,000
Senior High-Performance Storage Architect for AI Infrastructure
Senior High-Performance Storage Architect for AI Infrastructure

NVIDIA • Sydney

On-site
AUD 251,000 - 363,000
Senior High-Performance Storage Architect - NVIS
Senior High-Performance Storage Architect - NVIS

NVIDIA • Sydney

On-site
AUD 251,000 - 363,000
Senior Solution Architect, AI Compute Engineer - NVIS
Senior Solution Architect, AI Compute Engineer - NVIS

NVIDIA • Sydney

On-site
AUD 120,000 - 160,000
AI Domain Architect (AI Storage)
AI Domain Architect (AI Storage)

World Wide Technology • City of Melbourne

On-site
AUD 180,000 - 240,000
AI Storage Domain Architect – High-Performance NVMe & GPU AI
AI Storage Domain Architect – High-Performance NVMe & GPU AI

World Wide Technology • City of Melbourne

On-site
AUD 180,000 - 240,000
Senior Storage Architect for Exabyte AI Infrastructure
Senior Storage Architect for Exabyte AI Infrastructure

Sharon AI, Inc • Australia

Hybrid
AUD 160,000 - 270,000
Hybrid working
Birthday leave
Employee Assistance Program (EAP)
+1
Senior HPC Network Engineer: GPU Scale & Fabric Design
Senior HPC Network Engineer: GPU Scale & Fabric Design

Pathway Search • Sydney

On-site
AUD 120,000 - 190,000
Hybrid Senior Infrastructure Engineer - HPC & Storage
Hybrid Senior Infrastructure Engineer - HPC & Storage

Experis • City of Melbourne

Hybrid
AUD 140,000 - 200,000
Weekly pay
Hybrid work arrangements
Cutting-edge technology access