GPU Stack Unified Build & release platform Engineer

AMD

San Jose (CA)

On-site

USD 250,000 - 420,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

AMD in San Jose, CA seeks a Director-level leader to own a large-scale build and release platform across firmware, kernel, and ROCm. You will drive architecture decisions, coordinate executives, and guide a tiger team of engineers to accelerate validated GPU stack releases while maintaining stability.

You will spend ~50% coding, mentoring senior engineers, and shaping cross-team workflows. This role emphasizes strategic change with executive sponsorship and open-source alignment.

Qualifications

  • Strong leadership of director-level or higher with technical credibility.
  • Experience shaping large-scale infrastructure or platform modernization across teams.
  • Ability to sequence technical change while protecting active product commitments.
  • Executive-level communication and stakeholder alignment.

Responsibilities

  • Technical direction for the unified build and release platform from PoC to production.
  • Align executives and product owners on phased rollout to preserve shipments.
  • Drive organizational change across firmware, kernel, and ROCm teams.
  • Lead the tiger team and unblock senior engineers, coordinating cross-workstreams.
  • Hands-on coding ~50% to exemplify leadership and drive critical work.
  • Advance agentic AI workflows as a productivity multiplier for the org.
  • Own end-to-end strategy and ensure customer impact and timely releases.

Skills

Executive leadership
Software engineering
Architecture decisions
CI/CD infrastructure
Agentic AI fluency
Cross-team collaboration

Education

Master’s degree or PhD in related discipline

Tools

GitHub Actions
Self-hosted CI
AWS cloud environments

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE:

AMD's AI software stack is moving fast — and keeping pace means shipping complete, validated GPU stack releases to customers as quickly as the software can evolve. Today, that release velocity is limited by the coordination overhead between layers: firmware, kernel driver, and ROCm each have their own build systems, their own workflows, and no shared release baseline. Every customer delivery requires manual effort to assemble and validate a coherent recipe across all three, and lower-level components — firmware and kernel — are increasingly the bottleneck to getting the full stack out the door.

We're building the infrastructure to change that: a unified build and CI platform that treats the complete GPU stack as a single deliverable, built from source, validated together, and releasable on demand. The result is faster time-to-customer for validated, fully-supported GPU stack releases — without sacrificing the stability that product lines depend on.

THE PERSON:

We need an AMD Director to own this area end-to-end. You'll lead a tiger team of 5–10 engineers while spending roughly half your time in the code yourself — making architecture calls, driving the hardest technical problems, and setting the bar for how the team works. The other half is working directly with executives to sequence the rollout in a way that accelerates release velocity without disrupting active product lines. This is an organizational change effort as much as it is an engineering one — and we need someone who can operate credibly at both levels without giving up either.

KEY RESPONSIBILITIES:
  • Technical direction for the unified build and release platform— from PoC architecture decisions through phased production rollout, you'll set the direction and be accountable for the outcome. You'll have senior ICs executing against that direction; your job is to make sure the strategy is right and stays right as the platform matures.
  • Executive alignment on phased rollout— the transition to a unified platform cannot delay active product shipments. You'll work directly with executives and product line owners to design a rollout sequence that delivers value incrementally, identifies and mitigates risk at each phase, and builds organizational confidence in the new platform before dependencies on the old one are removed.
  • Organizational change across firmware, kernel, and ROCm teams— unifying the build and release process for 1,000+ developers requires more than good tooling. You'll work with component team leads and engineering directors to align on new workflows, build the cultural shift toward shift-left development and trunk health, and get organizational buy-in at the right level.
  • Tiger team leadership— you'll directly lead the senior engineers building the platform, including the Build Architect and CI Infrastructure Engineer. You'll keep them unblocked, make the cross-cutting technical calls, and ensure the build and CI workstreams converge into a coherent system.
  • Code at ~50% — and mean it— for this role, this is a real commitment, not a line in a job description. You'll own meaningful parts of the platform yourself, not just review others' work. We're at a pivotal moment in the industry: agentic AI tools are genuinely changing what it means to be a technical leader, and the best engineering teams are being led by people who are in it alongside them. If you've drifted away from coding over the years but is ready to re-engage — especially with AI as a force multiplier getting you back up to speed — we want to talk. What matters is the commitment, not whether you've been shipping code every week.
  • Sharpen your agentic AI engineering leadership— this platform has a scope (65+ firmware components, 1,000+ developers, multi-layer GPU stack) that demands agentic AI as a genuine force multiplier, not a productivity experiment. You'll be setting the example for how senior technical leaders use these tools, developing that capability in your team, and building an organization that moves faster because of it.
  • Own a strategic area end-to-end— from PoC through production rollout, across build systems and CI, across firmware and kernel and ROCm, across the tiger team and the executive stakeholders. Full ownership, full accountability.
  • Direct line to customer impact— the platform you build determines how quickly AMD can put validated GPU stack releases in front of customers. That's a board-level concern and you'll be the person driving it.
  • Organizational change at scale— this is a rare opportunity to shift how a major hardware company builds and releases its GPU stack, with the executive support and mandate to make it stick.
  • Open-source aligned— this work follows the same shift-left, trunk-health principles driving AMD's ROCm open-source direction, with visibility into AMD's most strategic hardware and software programs.
PREFERRED EXPERIENCE:
  • Deep software engineering experience, with demonstrated technical leadership at the director or fellow level
  • Deep technical range across build systems, CI/CD infrastructure, and software release pipelines — credible with senior ICs, fluent with engineering executives
  • Track record of driving large-scale infrastructure or platform modernization across multiple teams with competing priorities
  • Experience sequencing technical change in a way that protects active product commitments — you know how to be bold about direction while being careful about disruption
  • Executive-level communication: able to frame technical tradeoffs as business decisions, build alignment across stakeholders, and hold accountability for outcomes at an organizational level
  • Demonstrated ability to build and lead high-performing small teams in a greenfield or PoC context
  • Fluency with agentic AI workflows (Cursor, Claude, Copilot, etc.) — both as a personal productivity multiplier and as a capability you actively develop in your team
  • Experience with firmware, kernel, or embedded software release pipelines and their specific constraints
  • Familiarity with GitHub Actions, self-hosted CI infrastructure, and cloud-based build environments (AWS)
  • Experience driving adoption of shift-left development practices across large engineering organizations

PREFERRED ACADEMIC CREDENTIALS:

  • Master’s degree or PhD in related discipline preferred. Equivalent professional experience demonstrating engineering expertise will also be considered.

LOCATION: San Jose, CA

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GPU Stack Unified Build & release platform Engineer
GPU Stack Unified Build & release platform Engineer

Advanced Micro Devices, Inc. • San Jose (CA), Northern (KY)

Hybrid
USD 260,000 - 380,000
AMD benefits
Hybrid work arrangement
AI Infrastructure Engineer
AI Infrastructure Engineer

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 140,000 - 190,000
AI Infrastructure Engineer
AI Infrastructure Engineer

AMD • San Jose (CA)

On-site
USD 140,000 - 170,000
AMD benefits
Sr. Product Manager - AI Infrastructure and Solutions
Sr. Product Manager - AI Infrastructure and Solutions

AMD • Austin (TX)

On-site
USD 170,000 - 230,000
Sr. Director Data Center GPU Platform & System Validation GPU
Sr. Director Data Center GPU Platform & System Validation GPU

Advanced Micro Devices • Austin (TX)

On-site
USD 150,000 - 200,000
Principal System Software Architect, AI/GPU Platforms
Principal System Software Architect, AI/GPU Platforms

Advanced Micro Devices, Inc. • Austin (TX)

Hybrid
USD 180,000 - 240,000
Principal Software Development Engineer
Principal Software Development Engineer

Advanced Micro Devices • Santa Clara (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Senior Manager, GPU Application Engineering
Senior Manager, GPU Application Engineering

AMD • Santa Clara (CA)

On-site
USD 130,000 - 160,000
Sr. Product Manager - AI Infrastructure and Solutions
Sr. Product Manager - AI Infrastructure and Solutions

Advanced Micro Devices • Austin (TX)

On-site
USD 140,000 - 210,000
Sr. Director Data Center GPU Platform & System Validation GPU
Sr. Director Data Center GPU Platform & System Validation GPU

AMD • Austin (TX)

On-site
USD 140,000 - 200,000
AMD benefits at a glance.