Senior Storage Engineer

Hydra Host, Inc.

Miami (FL)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hydra Host, Inc. is seeking a Senior Storage Engineer to design and build their first production-grade storage platform for AI and HPC workloads. The role involves leading technical decisions, implementing storage solutions, and optimizing performance for GPU clusters.

The ideal candidate has over 8 years of hands-on experience in high-performance storage systems, including block and object storage solutions, along with a solid background in Linux systems engineering and parallel file systems.

Qualifications

  • 8+ years of experience in designing high-performance storage systems for compute clusters.
  • Proven expertise in building storage infrastructure from scratch.
  • Strong background in parallel file systems and GPU workloads.

Responsibilities

  • Define and implement Hydra Host's first production storage platform.
  • Lead decisions on storage stack design and performance tuning.
  • Integrate and optimize parallel file systems for AI training workloads.

Skills

High-performance storage systems
Block storage (NVMe, SAN, Ceph)
Object storage (S3, MinIO)
Parallel file systems (WekaIO, BeeGFS, Lustre)
Linux systems engineering
Automation and scripting
Understanding of AI/ML data pipelines

Job description

# Senior Storage Engineer at Hydra HostJob Title: Storage Engineer**About Hydra Host**Hydra Host is a Founders Fund-backed NVIDIA cloud partner building the infrastructure platform that powers AI at scale. We connect AI Factories - high-performance GPU data centers - with the teams that depend on them: research labs training foundation models, enterprises running production inference, and developer platforms demanding scalable compute capacity. Hydra Host is building the next-generation bare-metal GPU infrastructure network and marketplace under its Brokkr platform. The company enables independent data centers to monetize GPU capacity while providing enterprises with scalable, high-performance access to NVIDIA-based compute (e.g., H100, H200, B200, L40S, RTX 4090). As we expand our infrastructure capabilities, Hydra Host is now seeking a Storage Engineer to lead the architecture, development, and deployment of our next-generation AI/HPC storage platform.**The role:**As a Storage Engineer, you will be responsible for designing and building Hydra Host’s first production-grade storage platform from the ground up, supporting the company’s rapidly expanding network of bare-metal GPU clusters.You’ll own the architecture, technology selection, implementation, and evolution of this platform, defining how Hydra Host manages data for large-scale, distributed AI workloads across global data centers.This is a senior, hands-on role for an engineer who has built storage systems for GPU clusters before, with deep expertise in both block and object storage and a strong understanding of parallel file systems, performance optimization, and large-scale orchestration.**Key Responsibilities**· Define, architect, and implement Hydra Host’s first production storage platform tailored for bare-metal GPU clusters and AI/HPC workloads.· Lead all technical decisions around storage stack design, from hardware infrastructure to parallel file system orchestration and performance tuning.· Select, build, and maintain storage solutions spanning both block (NVMe, SAN, Ceph, etc.) and object storage (S3-compatible, custom, or Ceph Object Gateway) layers.· Design for high-throughput, low-latency access, supporting large datasets, rapid checkpointing, and parallel access for distributed AI training workloads.· Integrate and optimize parallel file systems such as Lustre, BeeGFS, Spectrum Scale, WekaIO, or CephFS, ensuring maximum performance and fault tolerance.· Ensure compatibility across Hydra’s diverse GPU/OEM ecosystem, accounting for unique firmware, BMC/Redfish APIs, and hardware configurations.· Develop automation, observability, and management tooling for storage, focusing on reliability, scalability, and efficiency.· Act as a builder and architect: deeply hands-on in deployment, troubleshooting, and optimization, while guiding long-term storage roadmap.· Collaborate cross-functionally with GPU, HPC, and platform engineering teams to integrate storage with compute and network layers.· Interface with customers and product leadership to define feature priorities, performance benchmarks, and future enhancements.**Must-Have Qualifications**· 8+ years of progressive, hands-on experience designing and implementing high-performance storage systems for compute clusters in HPC, AI, or bare-metal cloud environments.· Proven track record building storage infrastructure from scratch, not just operating existing systems.· Deep expertise in block storage (NVMe, SAN, Ceph, distributed block systems) and object storage (S3, MinIO, Ceph Object Gateway, etc.).· Strong background in parallel file systems (WekaIO, BeeGFS, Lustre, Spectrum Scale, or similar) supporting GPU or AI cluster workloads.· Solid foundation in Linux systems engineering, automation, and scripting for distributed environments.· Familiarity with BMC, Redfish APIs, and OEM server firmware for bare-metal management.· Deep understanding of AI/ML data pipelines: model checkpointing, data locality, and multi-tiered storage optimization.· Excellent problem-solving, debugging, and communication skills, able to translate technical decisions into clear architectural direction.**Preferred Qualifications**· Experience building storage solutions for large-scale GPU or HPC infrastructure.· History of technical leadership or mentorship, growing teams or owning a product roadmap.· Experience evaluating and managing vendor relationships and negotiating storage hardware/software contracts.· Contributions to open-source HPC or storage projects (Ceph, Lustre, BeeGFS, etc.).· Familiarity with confidential computing, secure data handling, or high-availability architectures.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Storage Engineer
Senior Storage Engineer

Kindredventures • Miami (FL)

On-site
USD 120,000 - 150,000
Senior Storage Engineer
Senior Storage Engineer

Hydra Host • Miami (FL)

On-site
USD 120,000 - 160,000
Senior AI/HPC Storage Architect
Senior AI/HPC Storage Architect

Hydra Host, Inc. • Miami (FL)

On-site
USD 120,000 - 150,000
Senior AI/HPC Storage Architect
Senior AI/HPC Storage Architect

Hydra Host • Miami (FL)

On-site
USD 120,000 - 150,000
Lead AI/HPC Storage Architect for Production Platform
Lead AI/HPC Storage Architect for Production Platform

Hydra Host • Miami (FL)

On-site
USD 120,000 - 160,000
Operations and Support Lead
Operations and Support Lead

Hydra Host, Inc. • Miami (FL)

On-site
USD 120,000 - 180,000
Storage Engineer - Hosting
Storage Engineer - Hosting

Hamilton Barnes Associates Limited • United States

On-site
USD 170,000 - 230,000
Stock options
Bonus 10%
Partner Manager – AI Infrastructure
Partner Manager – AI Infrastructure

Hydra Host, Inc. • United States

On-site
USD 100,000 - 130,000
Senior HPC Storage Engineer
Senior HPC Storage Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
HPC Storage Engineer
HPC Storage Engineer

GTN Technical Staffing • Dallas (TX)

Hybrid
USD 120,000 - 170,000
Competitive base salary
100% company-paid benefits