Software Engineer III - AI/ML Platform

Fairygodboss

Greater London

On-site

GBP 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorganChase in London seeks a Software Engineer III to join the Firmwide LLM Serving Platform team. You will build core backend services for LLM inference, optimize latency, and operate production systems at scale.

This role blends infrastructure and ML concepts: you won't need to be an ML researcher but will work across APIs, CI/CD, and tooling to deliver reliable services with strong observability and secure practices.

Qualifications

  • Formal training or certification in software engineering.
  • Bachelor's degree in Computer Science or equivalent.
  • Solid programming fundamentals and debugging skills.
  • Comfort with Go, Python, or TypeScript and quick ramp-up on others.
  • Interest in distributed systems and learning about LLMs.

Responsibilities

  • Build core backend services for LLM inference, including routing and batching.
  • Implement and maintain APIs and SDKs used across teams.
  • Profile and optimize end-to-end performance across CPU, memory, and GPU.
  • Improve reliability with health checks, autoscaling, and runbooks.
  • Add observability with metrics, tracing, and dashboards tied to SLOs.
  • Support safe deployments through CI/CD improvements and feature flags.
  • Learn LLM serving concepts and responsible AI practices.
  • Contribute to automation and reusable patterns across the team.

Skills

Go
Python
TypeScript
Distributed systems
Kubernetes
gRPC
Redis

Education

Bachelor's Degree in Computer Science or equivalent

Tools

Docker
Kubernetes
APIs/SDKs
CI/CD tooling

Job description

We have an opportunity to impact your career and provide an adventure where you can push the limits of what's possible. As a Software Engineer III at JPMorganChase within the Firmwide LLM Serving Platform team, you are an integral part of an agile team that designs, builds, and operates the services that make large language models usable at scale. This is an infrastructure-meets-ML role: you don't need to be an ML researcher, but you should be excited to learn how model architectures and inference constraints translate into real production systems. You will contribute to a living platform where we optimize performance - pushing down latency, increasing throughput, maximizing GPU utilization, and eliminating waste across the request lifecycle.

Job responsibilities
  • Build core backend services for LLM inference, including request routing, batching, scheduling, streaming responses, and quota/limits.
  • Implement and maintain APIs and SDKs used by product and application teams across the firm.
  • Profile and optimize performance end-to-end across CPU, memory, network, serialization, concurrency, GPU utilization, and caching.
  • Improve reliability and operability through health checks, graceful degradation, autoscaling behaviors, incident follow-ups, and runbooks.
  • Contribute to system design by breaking down ambiguous problems, proposing approaches, and making pragmatic tradeoffs.
  • Add observability with metrics, tracing, logging, dashboards, and actionable alerts tied to SLOs.
  • Support safe deployments through CI/CD improvements, canarying, feature flags, backward compatibility, and rollback plans.
  • Learn LLM serving fundamentals - tokenization costs, KV cache, quantization, context length tradeoffs, throughput vs. latency.
  • Leverages enterprise-authorized AI coding assist tools within the work environment to improve code quality, delivery speed, and productivity across complex deliverables (e.g., code generation/refactoring, unit test creation, documentation), while validating outputs through peer review, automated testing, and secure coding standards; contributes learnings and reusable patterns to improve broader team effectiveness.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.
  • Add to team culture of diversity, opportunity, inclusion, and respect.
Required qualifications, capabilities, and skills
  • Formal training or certification on software engineering concepts and applied experience.
  • Bachelor's Degree in Computer Science or equivalent.
  • Solid programming fundamentals: data structures, concurrency basics, debugging, testing.
  • Comfort working in one or more of Go, Python, or TypeScript, with the ability to ramp up quickly on the others.
  • Interest in distributed systems and system design, even if you haven't built large systems yet.
  • Curiosity about LLMs and AI model architecture, with willingness to learn quickly.
  • A measurement-driven mindset: you like profiling, benchmarking, and proving improvements with data.
  • Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, test creation, troubleshooting, or documentation) with demonstrated ability to critically evaluate, validate, and refine AI-generated outputs for correctness, performance, and security.
  • Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; ability to guide peers on safe and effective usage within team practices.
Preferred qualifications, capabilities, and skills
  • Experience with performance profiling tools such as pprof, flamegraphs, or distributed tracing systems.
  • Familiarity with containers and orchestration (Docker, Kubernetes) and service-to-service networking.
  • Understanding of inference concepts: batching, streaming tokens, GPU memory constraints, KV cache.
  • Experience with high-throughput APIs (gRPC/HTTP), eventing/queues, or caching layers such as Redis.
  • Exposure to reliability practices: SLOs/SLIs, on-call rotations, incident reviews.
ABOUT US

J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world's most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.

ABOUT THE TEAM

J.P. Morgan's Commercial & Investment Bank is a global leader across banking, markets, securities services and payments. Corporations, governments and institutions throughout the world entrust us with their business in more than 100 countries. The Commercial & Investment Bank provides strategic advice, raises capital, manages risk and extends liquidity in markets around the world.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer III - AI/ML Platform
Software Engineer III - AI/ML Platform

JPMorganChase • Greater London

On-site
GBP 90,000 - 115,000
Software Engineer III - AI/ML Platform
Software Engineer III - AI/ML Platform

United States Digital Space LLC • Greater London

On-site
GBP 120,000 - 180,000
Software Engineer III - AI/ML Platform
Software Engineer III - AI/ML Platform

JPMorgan Chase & Co. • Greater London

On-site
GBP 110,000 - 150,000
Senior Lead Software Engineer - LLM Ops Platform Reliability
Senior Lead Software Engineer - LLM Ops Platform Reliability

JPMorganChase • Glasgow

On-site
GBP 90,000 - 130,000
Software Engineer III - Cloud/Java-Python/AI-ML
Software Engineer III - Cloud/Java-Python/AI-ML

JPMorganChase • Bournemouth

On-site
GBP 70,000 - 110,000
Sr Lead AI Platform Engineer
Sr Lead AI Platform Engineer

J.P. MORGAN • Glasgow

On-site
GBP 150,000 - 190,000
Software Engineer III - Python
Software Engineer III - Python

JPMorganChase • Glasgow

On-site
GBP 75,000 - 120,000
Sr Lead AI Platform Engineer
Sr Lead AI Platform Engineer

JPMorganChase • Glasgow

On-site
GBP 180,000 - 240,000
Software Engineer III - Python, AIML, AWS, EKS
Software Engineer III - Python, AIML, AWS, EKS

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 120,000
Senior Lead Software Engineer - LLM Ops Platform Reliability
Senior Lead Software Engineer - LLM Ops Platform Reliability

J.P. MORGAN • Greater London

On-site
GBP 140,000 - 210,000