Low-Level Senior Software Engineer, Xet Storage - US Remote

Hugging Face

New York (NY)

On-site

USD 190,000 - 260,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity in company
Remote options
Health, dental and vision benefits
Flexible paid time off

Job summary

Hugging Face is seeking a senior engineer for the Xet Storage team to help scale and operate the storage backend powering the Hugging Face platform. You will work on xet-core in Rust and contribute to hf-xet, the Python library that underpins the Hub client and open-source ecosystem.

We value low-level performance, ownership end-to-end, and the ability to thrive in a fast-moving, remote environment. This role offers opportunities to influence infrastructure used by researchers and builders

Qualifications

  • 8+ years building and scaling distributed systems, storage, or networking infrastructure.
  • Proficiency in a low-level systems language, with Rust strongly preferred; Python/Typescript/Go/C++ used across the stack.
  • Ability to work independently in a high-trust, low-process environment and take ownership end to end.
  • Comfort operating in a fast-moving, async, fully remote environment.

Responsibilities

  • Contribute to xet-core Rust project powering hf-xet and the Hub client.
  • Design, build, and operate features in the Xet Storage backend within the Infrastructure organization.

Skills

Distributed systems experience
Rust
Python
Typescript
Go
C++

Tools

Git internals
AWS
Azure
GCP
Kubernetes
Relational databases
Non-relational databases

Job description

At Hugging Face, we're on a journey to democratize good AI. We are building the fastest growing platform for AI builders with over 11 million users who collectively shared over 4 million models, 1 million datasets & 1.5 million Gradio apps. Our open-source libraries have more than 700,000 stars on Github.

Roles at Hugging Face are very fluid and dynamic -- we're looking for someone who is comfortable taking on different challenges that evolve over time. This role sits on the Xet Storage team, the group responsible for the storage system behind all of Hugging Face. Today we store over 200PB (and growing rapidly!) of the world's most important ML & AI assets. Xet is the underlying storage architecture for the entire Hugging Face platform and community -- from the largest model repositories to the datasets and Spaces that millions of people build on every day.

In this role, you'll work across two closely connected surfaces. You'll contribute to xet-core, our open-source project written in Rust that powers hf-xet -- the Python library underpinning the Hugging Face Hub client and the wider ecosystem of open-source tools the community relies on. And you'll design, build, and operate meaningful and challenging features in the Xet Storage backend, contributing to the broader Infrastructure organization at Hugging Face. We're a small team building and operating incredible things at enormous scale, in a high-trust, low-process, async, and remote environment -- and we lean heavily on the latest AI tools to move fast.

If you love writing low-level, high-performance code and are energized by the challenge of large-scale, scalable services, this is the team for you. This usually means proven experience building and operating production systems software, but we consider every applicant on an individual basis.

Requirements
  • 8+ years building and scaling distributed systems, storage, or networking infrastructure
  • Proficiency in a low-level systems language, with Rust strongly preferred (we also work in Python, Typescript, Go, and C++ across the stack)
  • A track record of working independently in a high-trust, low-process environment -- you're comfortable with ambiguity and take ownership end to end
  • Comfort operating in a fast-moving, async, and fully remote environment
  • A passion for building simple, robust, and scalable systems relied on by engineering and science teams around the world
Bonus points if you have
  • Experience designing efficient, high-performance, fault-tolerant data storage and retrieval systems
  • Experience operating production systems -- monitoring, alerting, and distributed debugging and recovery
  • Familiarity with git internals, cloud infrastructure (AWS, Azure, GCP, Kubernetes), databases (relational and non-relational), or networking
About You

If you love open-source, are excited by the intersection of low-level performance work and large-scale services, and want your code to sit at the foundation of the world's largest platform for AI builders, then we can't wait to see your application!

Benefits
More about Hugging Face
  • We are actively working to build a culture that values diversity, equity, and inclusivity. We are intentionally building a workplace where people feel respected and supported—regardless of who you are or where you come from. We believe this is foundational to building a great company and community. Hugging Face is an equal opportunity employer and we do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
  • We value development. You will work with some of the smartest people in our industry. We are an organization that has a bias for impact and is always challenging ourselves to continuously grow. We provide all employees with reimbursement for relevant conferences, training, and education.
  • We care about your well-being. We offer flexible working hours and remote options. We offer health, dental, and vision benefits for employees and their dependents. We also offer parental leave and flexible paid time off.
  • We support our employees wherever they are. While we have office spaces in NYC and Paris, we're very distributed and all remote employees have the opportunity to visit our offices. If needed, we'll also outfit your workstation to ensure you succeed.
  • We want our teammates to be shareholders. All employees have company equity as part of their compensation package. If we succeed in becoming a category-defining platform in machine learning and artificial intelligence, everyone enjoys the upside.
  • We support the community. We believe major scientific advancements are the result of collaboration across the field. Join a community supporting the ML/AI community.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud ML DevRel Engineer - US remote
Cloud ML DevRel Engineer - US remote

Hugging Face • New York (NY)

On-site
USD 120,000 - 160,000
Health, dental, and vision benefits
Flexible working hours
Company equity
+1
Senior Rust Engineer — Scalable, Remote Storage Systems
Senior Rust Engineer — Scalable, Remote Storage Systems

Hugging Face • New York (NY)

On-site
USD 190,000 - 260,000
Equity in company
Remote options
Health, dental and vision benefits
+1
Open-Source Machine Learning Engineer - US Remote
Open-Source Machine Learning Engineer - US Remote

Hugging Face • New York (NY)

Remote
USD 100,000 - 140,000
Health, dental, and vision benefits
Flexible working hours
Parental leave
+2
Software Engineer, Platform Engineering
Software Engineer, Platform Engineering

Helsing • Washington

Hybrid
USD 180,000 - 230,000
Remote or hybrid work option
Competitive compensation
Software Engineer - X Developer Platform
Software Engineer - X Developer Platform

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Equity
401(k) retirement plan
Medical, Vision, Dental coverage
+2
Senior Software Engineer - Data Plane
Senior Software Engineer - Data Plane

Hume AI • New York (NY)

On-site
USD 180,000 - 260,000
Storage Infra Engineer for AI/ML at Scale (Equity)
Storage Infra Engineer for AI/ML at Scale (Equity)

Lightning AI • New York (NY)

Hybrid
USD 180,000 - 220,000
Comprehensive health coverage
Meaningful equity
401(k) matching and pension
Member of Technical Staff - Media
Member of Technical Staff - Media

SpaceXAI • Seattle (WA)

On-site
USD 180,000 - 440,000
Equity
Medical coverage
Vision coverage
+6
Senior Software Engineer (AI Inference & Runtime Platform) at AZX
Senior Software Engineer (AI Inference & Runtime Platform) at AZX

Matcha • Northern (KY)

Hybrid
USD 150,000 - 190,000
Health insurance with dependents
Bonus eligibility
Equity
+2
Software Engineer - X Developer Platform
Software Engineer - X Developer Platform

SpaceXAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Equity
Medical, Vision & Dental
401(k) retirement plan
+2