Staff Engineer, Hardware Reliability & Fleet Automation

LinkedIn

Sunnyvale (CA)

Hybrid

USD 156,000 - 255,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

LinkedIn seeks a Staff Engineer for the Hardware Capacity Engineering team in Sunnyvale, CA. The role focuses on scaling and sustaining on‑prem data center hardware, building automation, and collaborating with SRE, software, and hardware vendors. Hybrid work model with office days.

Responsibilities include designing tests, benchmarking, onboarding new platforms, cost analysis, and leading fleet upgrades to ensure high reliability and observability at scale.

Qualifications

  • BS in Computer Science, Computer Engineering, or related field, or equivalent practical experience.
  • 6+ years working in Linux-based infrastructure, systems, or hardware engineering.
  • 4+ years of hardware troubleshooting, systems engineering, and performance analysis.
  • Experience developing software or automation for infrastructure at scale.

Responsibilities

  • Collaborate with LinkedIn engineering teams to select hardware platforms for apps.
  • Design test environments; benchmark compute, storage, and power; report results.
  • Qualify and integrate new server platforms end-to-end with vendors.
  • Drive cost analysis and define SLAs with partner teams.
  • Qualify BIOS, BMC, and firmware; lead fleet upgrade programs.
  • Improve fleet reliability via fault detection and remediation.
  • Design and own automation for hardware qualification and fleet health monitoring.
  • Contribute to AI/ML infrastructure performance for GPU platforms.
  • Troubleshoot complex hardware issues and lead incident response.

Skills

Linux
Hardware Troubleshooting
Infrastructure Automation
Software Development

Education

BS in Computer Science, Computer Engineering, or related field

Job description

LinkedIn seeks a Staff Engineer for the Hardware Capacity Engineering team in Sunnyvale, CA. The role focuses on scaling and sustaining on‑prem data center hardware, building automation, and collaborating with SRE, software, and hardware vendors. Hybrid work model with office days.

Responsibilities include designing tests, benchmarking, onboarding new platforms, cost analysis, and leading fleet upgrades to ensure high reliability and observability at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Hardware SRE — Firmware & Datacenter Reliability
Hardware SRE — Firmware & Datacenter Reliability

SpaceXAI • Memphis (TN)

On-site
USD 110,000 - 160,000
Hardware SRE - Data Center Reliability & Failure Analysis
Hardware SRE - Data Center Reliability & Failure Analysis

Socket.dev • Memphis (TN)

On-site
USD 110,000 - 160,000
Hardware Sustaining Engineering Manager
Hardware Sustaining Engineering Manager

Hobbsnews • Sunnyvale (CA)

On-site
USD 153,000 - 311,000
Health & Wellbeing benefits
Professional development programs
Inclusive work environment
Senior Site Reliability Engineer: Scalable Hybrid Infra
Senior Site Reliability Engineer: Scalable Hybrid Infra

Redwood Materials • Nevada (IA)

On-site
USD 140,000 - 180,000
Staff Hardware Engineer – System Integration & Test
Staff Hardware Engineer – System Integration & Test

Socket.dev • Mountain View (CA)

On-site
USD 162,000 - 234,000
Medical,dental,vision benefits
401k with employer match
Vehicle lease program
Head of Fleet Reliability & Automation
Head of Fleet Reliability & Automation

CoreWeave • Sunnyvale (CA)

On-site
USD 180,000 - 230,000
Head of Hardware Engineering: Build & Scale Teams
Head of Hardware Engineering: Build & Scale Teams

Carbon • Sunnyvale (CA)

On-site
USD 277,000 - 417,000
Hardware Reliability SRE, Data Center
Hardware Reliability SRE, Data Center

Pantera Capital • Southaven (MS)

On-site
USD 110,000 - 170,000
Senior SRE - Data Offload & CI for Autonomous Fleet (Onsite)
Senior SRE - Data Offload & CI for Autonomous Fleet (Onsite)

Booster • Mountain View (CA)

On-site
USD 180,000 - 260,000
Hybrid SRE Engineer for Scalable Reliability & Equity
Hybrid SRE Engineer for Scalable Reliability & Equity

EarnIn • Mountain View (CA)

Hybrid
USD 139,000 - 232,000