AI Infrastructure Performance & Modeling Lead

Renice AI

Mountain View (CA)

Hybrid

USD 180,000 - 280,000

Full time

12 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Renice AI is seeking a Performance Lead to drive architecture decisions for AI inference workloads across the full system stack in Mountain View, CA. You will build modeling frameworks, analyze tradeoffs across compute, memory, networking, and storage, and translate results into actionable guidance for internal teams and hardware partners.

This role leads a small team of engineers, collaborates with ML, systems, and hardware groups, and shapes reference designs and long-term infrastructure

Qualifications

  • Experience owning or building performance modeling frameworks used to drive real system design decisions.
  • Deep knowledge of AI and machine learning workloads, including training and/or inference at scale.
  • Understanding of system-level tradeoffs across compute, memory, and networking in large-scale distributed systems.
  • Comfort working across abstraction layers, from workload behavior to hardware implementation.
  • Experience using analytical modeling or simulation to inform architectural decisions.
  • Ability to operate in ambiguous problem spaces and turn open-ended questions into structured analysis.
  • Clear communication and the ability to influence internal teams and external partners.

Responsibilities

  • Build and own a performance modeling framework and toolchain to evaluate AI systems across multiple levels of abstraction.
  • Analyze and quantify architectural tradeoffs across compute, memory, networking, storage, and system topology.
  • Develop performance models to guide decisions on scale-up versus scale-out architectures, interconnect and network design, and memory hierarchy and system balance.
  • Translate modeling outputs into clear recommendations for internal teams and external hardware vendors.
  • Influence reference designs and vendor roadmaps through data-driven insights.
  • Partner closely with machine learning, systems, and hardware teams to understand workload characteristics and requirements.
  • Lead and grow a small team of two to three engineers, setting technical direction and maintaining high standards for modeling rigor.
  • Continuously improve modeling fidelity by validating against real system behavior and measurements.

Skills

Performance modeling
AI workloads
System tradeoffs
Analytical modeling
Cross-functional communication

Job description

Renice AI is seeking a Performance Lead to drive architecture decisions for AI inference workloads across the full system stack in Mountain View, CA. You will build modeling frameworks, analyze tradeoffs across compute, memory, networking, and storage, and translate results into actionable guidance for internal teams and hardware partners.

This role leads a small team of engineers, collaborates with ML, systems, and hardware groups, and shapes reference designs and long-term infrastructure

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Infrastructure Performance Lead
AI Infrastructure Performance Lead

Renice AI • Mountain View (CA)

Hybrid
USD 180,000 - 280,000
AI Infrastructure Performance Modeling Lead
AI Infrastructure Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
AI Infrastructure Performance Engineer
AI Infrastructure Performance Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance
Performance Modeling Engineer - AI Systems & Infrastructure
Performance Modeling Engineer - AI Systems & Infrastructure

OpenAI • Seattle (WA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model
Performance Modeling Lead for AI Infrastructure
Performance Modeling Lead for AI Infrastructure

OpenAI • San Francisco (CA)

Hybrid
USD 342,000 - 555,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Seattle (WA)

On-site
USD 342,000 - 555,000
Performance Modeling Lead for AI Infrastructure
Performance Modeling Lead for AI Infrastructure

OpenAI • Seattle (WA)

Hybrid
USD 342,000 - 555,000
AI Infrastructure Performance Modeling Engineer
AI Infrastructure Performance Modeling Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 266,000 - 445,000
Senior AI Infrastructure Engineer - Real-Time Inference
Senior AI Infrastructure Engineer - Real-Time Inference

Didi Labs • San Jose (CA)

Hybrid
USD 170,000 - 351,000