GenAI Platform Architect for High-Scale LLM Inference

Socket.dev

New Jersey

On-site

USD 180,000 - 240,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

JPMorganChase, within the Chief Data and Analytics Office, seeks a Principal Software Engineer to lead the GenAI serving platform, focusing on high-performance LLM inference, intelligent routing, and GPU efficiency. You will influence senior stakeholders, drive secure scaling, and deliver enterprise-grade AI capabilities across portfolios.

You will own design and implementation of distributed inference architectures, conduct performance benchmarking, and ensure operational excellence with robust

Qualifications

  • Formal training or certification in software engineering with 7+ years experience in large-scale platforms and services.
  • Strong proficiency in Python, Java, Scala, or Go with focus on code quality and testing.
  • Experience leading agentic AI development practices and secure handling of data.

Responsibilities

  • Design and operate a high-throughput LLM serving platform with low latency and autoscaling.
  • Build GenAI Gateway / inference API for authentication, routing, and observability.
  • Develop and optimize multi-backend model routing balancing quality, latency, and cost.
  • Drive GPU serving optimization including memory management and KV-cache strategies.
  • Implement quantization and safe rollout practices to reduce cost while preserving quality.
  • Define and govern disaggregated serving architectures to improve tail latency.
  • Write secure, production-grade code and review others’ work; create reusable frameworks.
  • Own SDK and service integrations ensuring reliability and performance.
  • Establish SLOs/SLAs and build performance benchmarking and observability.

Skills

Programming languages
Agentic AI adoption
Distributed systems
GPU inference
Security by design
Communication
SLOs/SLAs

Tools

Kubernetes
TensorRT-style

Job description

JPMorganChase, within the Chief Data and Analytics Office, seeks a Principal Software Engineer to lead the GenAI serving platform, focusing on high-performance LLM inference, intelligent routing, and GPU efficiency. You will influence senior stakeholders, drive secure scaling, and deliver enterprise-grade AI capabilities across portfolios.

You will own design and implementation of distributed inference architectures, conduct performance benchmarking, and ensure operational excellence with robust

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal LLM Inference Engineer — AI Platform
Principal LLM Inference Engineer — AI Platform

JPMorganChase • Jersey City (NJ)

On-site
USD 210,000 - 320,000
Senior LLM Inference Engineer – AI Platform Lead
Senior LLM Inference Engineer – AI Platform Lead

Fairygodboss • Jersey City (NJ)

On-site
USD 190,000 - 240,000
Senior Principal LLM Engineer — AI Platform & Scale
Senior Principal LLM Engineer — AI Platform & Scale

JPMorgan Chase & Co. • Palo Alto (CA)

On-site
USD 233,000 - 325,000
Principal AI Engineer - Lead LLM Platform & Innovation
Principal AI Engineer - Lead LLM Platform & Innovation

Next Frontier Capital • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Senior GenAI Platform Lead — Enterprise AI Systems
Senior GenAI Platform Lead — Enterprise AI Systems

JPMorganChase • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Lead AI Platform Engineer for GenAI & Enterprise
Lead AI Platform Engineer for GenAI & Enterprise

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 180,000 - 260,000
Senior Principal AI/ML Engineer — LLM & GNN Architect
Senior Principal AI/ML Engineer — LLM & GNN Architect

Fairygodboss • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Senior AI/ML Platform Architect
Senior AI/ML Platform Architect

JPMorganChase • Jersey City (NJ)

On-site
USD 130,000 - 160,000
Senior AI Platform Engineer - Enterprise LLMs
Senior AI Platform Engineer - Enterprise LLMs

Next Frontier Capital • Jersey City (NJ)

On-site
USD 170,000 - 210,000
Lead AI Engineer for Enterprise LLM Platform
Lead AI Engineer for Enterprise LLM Platform

Next Frontier Capital • Palo Alto (CA)

On-site
USD 180,000 - 260,000