Principal AI System Architect

Socket.dev

San Jose (CA)

On-site

USD 220,000 - 280,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Credo in San Jose, CA seeks a Principal AI System Architect to tackle system-level interconnection challenges connecting XPUs into scalable AI infrastructure for efficient large-scale training and inference.

The role emphasizes hands-on NPU hardware experience, architecture definition through bring-up, and collaboration across silicon, memory, networking and software teams to define interconnect fabric requirements. A base salary range is offered with bonus and equity.

Qualifications

  • Ten years of experience in XPU or AI interconnect system design, architecture and micro-architecture.
  • Five cycles of complete ASIC tapeouts.
  • Strong understanding of GPU, NPU, and/or CPU design principles; deep NPU hardware architecture expertise.
  • Familiarity with large language model architectures and distributed training/inference concepts.
  • Demonstrated ability to analyze and synthesize complex technical material.
  • Strong cross-functional leadership and communication skills with compiler, software, SOC & back-end teams.

Responsibilities

  • Define the memory and interconnection architecture linking XPUs to memory and across node, rack, and cluster boundaries.
  • Solve communication and topology bottlenecks during distributed LLM training and inference.
  • Define how XPUs utilize memory and interconnect fabric for workloads across network topology.
  • Collaborate with silicon, memory, networking, and software teams to specify interconnect fabric requirements.
  • Provide hands-on NPU hardware architecture judgment for interconnect and system design decisions.
  • Track leading open LLM architectures to anticipate stress on interconnect and topology.
  • Lead and mentor a team of architects focused on system interconnection and present strategy to leadership.

Skills

XPU interconnect design
ASIC tapeouts
Computer architecture
NPU hardware architecture
Distributed training/inference
Cross-functional leadership
Communication with compiler/SOC

Education

Bachelor's degree in Computer Engineering, Electrical Engineering, Computer Science
Master's degree or PhD in a relevant field

Job description

Credo is looking for a Principal AI System Architect to join our team in San Jose, CA, reporting to AVP, XPU system AI interface. This role needs to solve the system-level interconnection challenges that connect XPUs (NPU/GPU/custom silicon) into working AI infrastructure. This role needs to follow and define how compute elements are wired, networked, and orchestrated together at node, rack, and cluster scale so that large-scale LLM training and inference run efficiently across the platform. We strongly prefer candidates with a hands‑on NPU hardware background and real experience taking NPU silicon into production deployment at scale.

Base salary range

is $220,000 - $280,000 a year. The base salary offer will depend on factors such as education, experience, training, skills, qualifications, and location. This position is also eligible for a discretionary bonus, equity and a full range of medical and other benefits.

Why Credo
  • Purpose: We invest in what matters. From meaningful-future shaping projects to competitive compensation, we empower you to grow your career while making a lasting impact.
  • People: Connection starts within. We collaborate, celebrate wins, and create an environment where everyone can do their best work.
  • Possibilities: Our belief shapes what's next. Our technology powers the most reliable and energy‑efficient connections around the world – and our team powers new products and markets that come next.
Basic Qualifications
  • Bachelor's degree in Computer Engineering, Electrical Engineering, Computer Science, or related field required
  • Ten years in XPU or AI interconnect system design, architecture and micro architecture working experience.
  • Five cycles of complete ASIC tapeouts.
  • Strong foundational understanding of computer architecture, including GPU, NPU, and/or CPU design principles; deep, hands‑on NPU hardware architecture expertise.
  • Familiarity with large language model architectures and distributed training/inference concepts (e.g., parallelism strategies, model serving) through coursework, research, or personal projects.
  • Demonstrated analytical ability, for example, through published research, thesis work, or quantitative project work- with a track record of independently studying and synthesizing technical material.
  • Strong cross-functional leadership and communication skills, daily experience working with compiler, software, SOC & backend team.
Preferred Qualifications
  • Master's degree or PhD in a relevant technical field.
  • Direct NPU tapeout and production ramp experience.
  • Hands‑on NPU hardware background, with direct experience carrying an NPU design from architecture definition through silicon bring‑up, validation, and volume production deployment.
  • Working knowledge of high‑speed interconnect or networking concepts (e.g., PCIe, Ethernet, RDMA).
  • Define the memory and interconnection architecture linking XPUs to memory and across node, rack, and cluster boundaries -not the internal design of the NPUs/GPUs/CPUs themselves.
  • Solve communication and topology bottlenecks that arise from all kinds of parallelism strategies during distributed LLM training and inference.
  • Define how XPUs utilizes the memory and interconnect fabric, ensuring workloads map cleanly onto the network topology.
  • Work with silicon, memory, networking, and software teams to specify interconnect fabric requirements and resolve integration issues between compute nodes.
  • Bring hands‑on NPU hardware architecture judgment to interconnect and system design decisions, informed by direct experience taking NPU silicon from definition through bring‑up, validation, and production deployment.
  • Track leading open LLM architectures to anticipate how model structure will stress interconnect and system topology.
  • Lead and mentor a team of architects focused on system interconnection; represent this strategy to leadership and partners.

Credo’s mission is to transform connectivity at scale through fast, reliable, and energy‑efficient system solutions. Our high‑speed copper and optical interconnect products deliver industry‑leading power and performance at up to 1.6T to meet the ever‑expanding data infrastructure demands of AI.

Our product portfolio includes ZeroFlap (ZF) Active Electrical Cables (AECs) and ZF optical transceivers, OmniConnect memory solutions, and a suite of retimers and DSPs for optical and copper Ethernet and PCIe, all leveraging the PILOT diagnostic and analytics software platform. Credo innovations enable our customers to connect the systems that connect the world.

Credo is committed to creating an inclusive environment for all employees and welcome applicants from diverse backgrounds without regard to race, color, religion, gender, sex, gender identity, sexual orientation, pregnancy, marital status, national origin, ethnicity, genetic information, age, disability, veteran status, or any other legally protected basis. If you have a disability or special need that requires accommodation to navigate our website or complete the application process, email people@credosemi.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal AI System Architect
Principal AI System Architect

Credo Semiconductor, Inc. • San Jose (CA)

On-site
USD 220,000 - 280,000
Principal AI System Architect
Principal AI System Architect

Credo • San Jose (CA)

On-site
USD 220,000 - 280,000
Principal Design Engineer
Principal Design Engineer

Credo • San Jose (CA)

On-site
USD 180,000 - 250,000
Discretionary bonus
Equity
Medical benefits
Director, System Validation Engineering
Director, System Validation Engineering

Socket.dev • San Jose (CA)

On-site
USD 200,000 - 250,000
Discretionary bonus
Equity
Medical benefits
System Application Engineer
System Application Engineer

Credo • San Jose (CA)

On-site
USD 100,000 - 130,000
Discretionary bonus
Equity
Full range of medical and other benefits
Senior Application Engineer
Senior Application Engineer

Credo • San Jose (CA)

On-site
USD 120,000 - 200,000
Discretionary bonus
Equity
Medical benefits
+1
Field Application Engineer
Field Application Engineer

Credo • San Jose (CA)

On-site
USD 65,000 - 100,000
Discretionary bonus
Equity
Medical benefits
Software Field Application Engineer
Software Field Application Engineer

Credo • San Jose (CA)

On-site
USD 65,000 - 110,000
Discretionary bonus
Equity
Medical benefits
Staff Application Engineer
Staff Application Engineer

Credo • San Jose (CA)

On-site
USD 140,000 - 210,000
Discretionary bonus
Equity
Medical benefits
Senior Software Engineer
Senior Software Engineer

Credo • San Jose (CA)

On-site
USD 150,000 - 175,000