Principal Engineer, AI Serving Framework Architect (Software)

Samsung Semiconductor

San Jose (CA)

On-site

USD 219,000 - 351,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

4+ weeks of paid time off
Medical/Dental/Vision/401k
Flexible work environment
Onsite gym and café
Charitable giving match

Job summary

A technology company in San Jose is seeking a Principal Engineer, AI Serving Framework Architect. This role involves leading research teams and leveraging expertise in AI workloads. Ideal candidates will have a PhD and over 15 years of experience in large-scale computing, taking the lead on projects around AI inference performance. An inclusive workplace offers a competitive salary and extensive benefits, emphasizing professional and personal growth.

Qualifications

  • 15+ years of experience in AI Serving Framework for large-scale computing.
  • Led project to build and optimize a Large Language Model (LLM) inference software stack.
  • In-depth understanding of inference engines such as vLLM.

Responsibilities

  • Lead research teams and propose technical direction.
  • Investigate dynamic scheduling methodologies for AI inference performance.
  • Propose software design for optimization algorithms on open-source platforms.

Skills

PyTorch
Python
C++
Collaboration
Communication

Education

PhD in Computer Science or a related field

Job description

Principal engineer, AI Serving Framework Architect (Software)

San Jose, California, United States

Please Note:

To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period.

Advancing the World’s Technology Together

Our technology solutions power the tools you use every day–including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.

We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.

Job Title: Principal engineer, AI Serving Framework Architect (Software)

What You’ll Do

The Architecture Research Lab (ARL) focuses on addressing fundamental system‑level bottlenecks in modern AI, particularly in memory capacity/bandwidth and system‑scale communication. By leveraging Samsung’s world‑class memory technologies, ARL explores and defines next‑generation AI system architectures that deliver step‑function improvements in performance, efficiency, and scalability.

We are seeking a Principal AI System Architect who will play a key role in bridging AI workloads, system architecture, and hardware design. In this role, you will develop system‑level performance models, drive architecture‑level design decisions, and propose forward‑looking AI system architectures that shape Samsung’s long‑term AI platform strategy.

Responsibilities
  • Lead research teams in Korea and propose technical direction as a Tech Lead.
  • Research dynamic scheduling methodologies for maximizing AI inference performance in multi‑rack scale memory‑centric systems composed of heterogeneous compute‑capable memory and hierarchical memory.
  • Investigate methods to accelerate search operations in RAG’s vector DB and AI Agent’s knowledge‑graph by leveraging compute‑capable memory.
  • Study strategies for optimally placing KVCache and a vector DB in hierarchical memory to minimize frequent SSD accesses and reduce IO stalls.
  • Propose SW design for implementing the derived optimization algorithms on open‑source platforms such as vLLM.
What You Bring
  • PhD in Computer Science or a related field with 15+ years of experience in AI Serving Framework for large‑scale computing, focusing on AI workloads.
  • Led a project to build and optimize a Large Language Model (LLM) inference software stack on a multi‑rack scale system delivering AI inference services to over 100,000 users.
  • Extensive experience designing AI inference software stacks for heterogeneous devices.
  • In‑depth understanding of the internal architecture and operation mechanisms of inference engines such as vLLM.
  • Proficiency in AI inference system profiling and optimization.
  • Knowledge and practical experience with future AI workloads, including reasoning models, multi‑modal solutions, AI agents, and world models.
  • Strong understanding of compute, memory, and networking bottlenecks in AI systems.
  • Required skillsets: PyTorch, Python, and C++.
  • A collaborative mindset, curiosity, and resilience in solving complex challenges.
  • Excellent verbal, presentation, and written communication skills.
  • (Nice to have) Native or fluent Korean speakers are preferred.
  • Inclusive, adaptable to diverse global norms and situations.
  • Approach challenges with curiosity and resilience, seeking data to build solutions.
  • Innovative and creative, proactively exploring new ideas and adapting quickly to change.
Location

Daily onsite presence at our San Jose office in alignment with our Flexible Work policy.

What We Offer

The pay range below is for all roles at this level across all US locations and functions. Individual pay rates depend on a number of factors—including the role’s function and location, as well as the individual’s knowledge, skills, experience, education, and training. We also offer incentive opportunities that reward employees based on individual and company performance.

This is in addition to our diverse package of benefits centered around the wellbeing of our employees and their loved ones. In addition to the usual Medical/Dental/Vision/401k, our inclusive rewards plan empowers our people to care for their whole selves. An investment in your future is an investment in ours.

Give Back—With a charitable giving match and frequent opportunities to get involved, we take an active role in supporting the community.

Enjoy Time Away—You’ll start with 4+ weeks of paid time off a year, plus holidays and sick leave, to rest and recharge.

Care for Family—Whatever family means to you, we want to support you along the way—including a stipend for fertility care or adoption, medical travel support, and virtual vet care for your fur babies.

Prioritize Emotional Wellness—With on‑demand apps and free confidential therapy sessions, you’ll have support no matter where you are.

Stay Fit—Eating well and being active are important parts of a healthy life. Our onsite café and gym, plus virtual classes, make it easier.

Embrace Flexibility—Benefits are best when you have the space to use them. That’s why we facilitate a flexible environment so you can find the right balance for you.

Base Pay Range

$219,000 – $351,000 USD

Equal Opportunity Employment Policy

Samsung Semiconductor takes pride in being an equal opportunity workplace dedicated to fostering an environment where all individuals feel valued and empowered to excel, regardless of race, religion, color, age, disability, sex, gender identity, sexual orientation, ancestry, genetic information, marital status, national origin, political affiliation, or veteran status.

When selecting team members, we prioritize talent and qualities such as humility, kindness, and dedication. We extend comprehensive accommodations throughout our recruiting processes for candidates with disabilities, long‑term conditions, neurodivergent individuals, or those requiring pregnancy‑related support. All candidates scheduled for an interview will receive guidance on requesting accommodations.

We do not accept unsolicited resumes. Only authorized recruitment agencies that have a current and valid agreement with Samsung Semiconductor, Inc. are permitted to submit resumes for any job openings.

Applicant Privacy Policy

https://semiconductor.samsung.com/about-us/careers/us/privacy/

Job ID: 42853

#LI-SF1

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Engineer, AI System Architect (Hardware)
Principal Engineer, AI System Architect (Hardware)

Samsungsemiconductor • San Jose (CA)

On-site
USD 219,000 - 351,000
Medical/Dental/Vision
401k
Paid time off
+2
Principal Engineer, AI System Architect (Hardware)
Principal Engineer, AI System Architect (Hardware)

Samsung Semiconductor • San Jose (CA)

On-site
USD 219,000 - 351,000
Medical/Dental/Vision benefits
Flexible work environment
401(k)
+2
Staff Engineer, AI System Architect (Hardware)
Staff Engineer, AI System Architect (Hardware)

Conductor • San Jose (CA)

On-site
USD 163,000 - 253,000
4+ weeks of paid time off
Flexible work environment
Comprehensive wellbeing benefits
Senior Director, Architecture Research Lab
Senior Director, Architecture Research Lab

Conductor • San Jose (CA)

On-site
USD 246,000 - 430,000
Medical/Dental/Vision benefits
401k plan
4+ weeks of paid time off
+2
Senior Director, Architecture Research Lab
Senior Director, Architecture Research Lab

Samsungsemiconductor • San Jose (CA)

On-site
USD 246,000 - 430,000
4+ weeks paid time off
Charitable giving match
Onsite gym and café
+1
Principal Engineer, AI Serving Framework Architect (Software)
Principal Engineer, AI Serving Framework Architect (Software)

Samsungsemiconductor • San Jose (CA)

On-site
USD 219,000 - 351,000
4+ weeks of paid time off per year
Medical/Dental/Vision/401k
Flexible work environment
+1
Staff Software Engineer AI/ML
Staff Software Engineer AI/ML

Conductor • San Jose (CA)

On-site
USD 163,000 - 253,000
4+ weeks of PTO
Paid holidays and sick leave
Fertility care or adoption stipend
+2
Staff Software Engineer AI/ML
Staff Software Engineer AI/ML

Samsungsemiconductor • San Jose (CA)

On-site
USD 163,000 - 253,000
4+ weeks paid time off
Medical/Dental/Vision/401k
Flexible environment
+5
Senior Engineer, Performance Architecture
Senior Engineer, Performance Architecture

Samsungsemiconductor • San Jose (CA)

On-site
USD 138,000 - 206,000
4+ weeks of paid time off
Medical/Dental/Vision benefits
Flexible work environment
+1
Staff Software Engineer AI/ML
Staff Software Engineer AI/ML

Socket.dev • San Jose (CA)

On-site
USD 163,000 - 253,000
Charitable giving match
4+ weeks paid time off
Fertility/adoption stipend
+2