Principal Software Engineer- Core AI Platform

JPMorganChase

Seattle (WA)

On-site

USD 180,000 - 240,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

JPMorganChase seeks a Principal Engineer to advance the Core AI Infrastructure Platform. You will design and promote a shared ecosystem unifying training and inference pipelines across hybrid-cloud environments, enabling thousands of engineers to deploy AI models safely and efficiently in production.

Responsibilities include building scalable platform services, applying observability and SRE practices, and guiding teams on reliable, secure AI delivery.

Qualifications

  • Formal training or certification on software engineering concepts and 10+ years applied experience.
  • Strong hands-on coding experience in Python with production-grade services.
  • Experience leading AI-assisted software development tooling and validating AI outputs for correctness, performance, and security.
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity and secure handling.
  • Hands-on experience with system design, automated testing, debugging, and operational stability for production software.
  • Experience implementing observability, logging, metrics, alerts, SLIs, incident response and root-cause analysis for services in production.
  • Working knowledge of cloud platforms, AI/ML platforms, distributed systems, or infrastructure engineering.
  • Ability to break down technical requirements into executable tasks and meet milestones.
  • Strong written and verbal communication skills.

Responsibilities

  • Architect and design solutions to enhance reliability and scalability of AI/ML platforms and applications.
  • Build and enhance reusable platform services, APIs, SDKs, agents, and libraries for model hosting and inference.
  • Partner with AI infrastructure teams to implement Foundation Services capabilities.
  • Own and evolve non-functional requirements and tooling for observability, security, and cost optimization.
  • Establish standards and reference architectures for reliability and operational readiness.
  • Partner with product teams to meet service reliability targets including performance and recoverability.
  • Participate in on-call rotations and resolve complex production issues.
  • Mentor engineers and raise engineering quality and operational rigor.

Skills

Python
Production-grade services
AI infrastructure
Observability & SRE
System design
Mentoring

Job description

Job Description

If you are looking for a game-changing career, working for one of the world's leading financial institutions, you've come to the right place.

If you are looking for a game-changing career, working for one of the world's leading financial institutions, you've come to the right place. As a Principal Engineer at JPMorgan Chase on the Core AI Infrastructure Platform team, you will design and promote the shared ecosystem that unifies our training and inference pipelines across hybrid-cloud and Neo-cloud environments. If you thrive on solving deep infrastructure challenges, establishing resilient SRE standards, and building foundational tech stack that enables thousands of engineers to safely and efficiently deploy cutting-edge AI models into production, this is your opportunity to shape the future of AI infrastructure at scale.

Job Responsibilities
  • Architect and Design solutions to enhance the reliability and scalability of AI/ML platforms and applications to accommodate fast-growing demand
  • Build and enhance reusable platform services, APIs, SDKs, agents, skills, and libraries that standardize how application teams consume model hosting, inference, and AI/ML managed services
  • Partner with AI infrastructure training, inference, and architecture teams to implement AI Foundation Services capabilities that unblock AI use cases, supporting delivery from technical design through build, launch, and early operational support
  • Own and evolve non-functional requirements and build/enhance tooling for observability, resilience, security controls, infrastructure management, and cost optimization
  • Establish and enforce standards and reference architectures for reliability, observability, automation, and operational readiness across services
  • Partner with product and platform engineering teams to define and meet service reliability targets, including performance, availability, and recoverability
  • Participate in on-call rotations, debug and resolve complex production issues; identify systemic gaps and drive durable remediation
  • Mentor and guide engineers; raise the bar on engineering quality, documentation, and operational rigor
Required Qualifications, Capabilities, And Skills
  • Formal training or certification on software engineering concepts and 10+ years applied experience
  • Strong hands‑on coding experience in Python with experience delivering production‑grade services
  • Demonstrated experience leading effective use of approved AI-assisted software development tools (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching engineers on safe, compliant adoption within delivery practices
  • Hands‑on practical experience with system design, automated testing, debugging, and operational stability for production software
  • Experience implementing observability, logging, metrics, alerts, Service Level Objectives, incident response practices, and root‑cause analysis for services in production
  • Working knowledge of software application development and technical processes, with depth in one or more areas such as cloud platforms, artificial intelligence, machine learning platforms, distributed systems, or infrastructure engineering
  • Ability to break down technical requirements into executable engineering tasks, manage dependencies, and deliver against milestones in partnership with product and application teams
  • Strong written and verbal communication skills, with the ability to explain technical decisions, trade‑offs, issues, and risks to engineering teams and stakeholders
Preferred Qualifications, Capabilities, And Skills
  • Proven skills in managing AI infrastructure on cloud platforms including deployment, scaling, monitoring, and optimizing machine learning workloads
  • Experience building reusable "golden path" assets such as templates, reference implementations, SDKs, automated tests, onboarding guides, and deployment patterns
  • Experience developing generative AI applications/AI agents and/or implementing AI‑assisted operations with appropriate guardrails
About Us

JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management. We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission‑based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on‑site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation. JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans.

About The Team

Our Global Technology Infrastructure group is a team of innovators who love technology as much as you do. Together, you'll use a disciplined, innovative and a business focused approach to develop a wide variety of high-quality products and solutions. You'll work in a stable, resilient and secure operating environment where you—and the products you deliver—will thrive.

High Risk Roles (HRR) are sensitive roles within the technology organization that require high assurance of the integrity of staff by virtue of 1) sensitive cybersecurity and technology functions they perform within systems or 2) information they receive regarding sensitive cybersecurity or technology matters. Users in these roles are subject to enhanced pre‑hire screening which includes both criminal and credit background checks (as allowed by law). The enhanced screening will need to be successfully completed prior to commencing employment or assignment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Software Engineer - Infrastructure Platforms
Lead Software Engineer - Infrastructure Platforms

JPMorganChase • Seattle (WA)

On-site
USD 180,000 - 240,000
Senior Lead Software Engineer- AI Platform engineer
Senior Lead Software Engineer- AI Platform engineer

Fairygodboss • Seattle (WA)

On-site
USD 140,000 - 190,000
Senior Lead Software Engineer, AI Platforms
Senior Lead Software Engineer, AI Platforms

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
Senior Lead Software Engineer - Full Stack/Infrastructure
Senior Lead Software Engineer - Full Stack/Infrastructure

Fairygodboss • Plano (TX)

On-site
USD 180,000 - 240,000
Health care
On-site health centers
Retirement plan
+2
Lead Software Engineer - Private Cloud
Lead Software Engineer - Private Cloud

JPMorganChase • Plano (TX)

On-site
USD 140,000 - 210,000
Senior Lead Software Engineer- AI/ML Training Infrastructure
Senior Lead Software Engineer- AI/ML Training Infrastructure

Fairygodboss • Seattle (WA)

On-site
USD 180,000 - 240,000
Principal Software Engineer - AI Foundations
Principal Software Engineer - AI Foundations

JPMorganChase • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Lead Software Engineer - AI, Python/Java
Lead Software Engineer - AI, Python/Java

Next Frontier Capital • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Lead Software Engineer - Private Cloud
Lead Software Engineer - Private Cloud

Fairygodboss • Plano (TX)

On-site
USD 140,000 - 190,000
Senior Principal Software Engineer - AI Development | Engineering Services & Platforms
Senior Principal Software Engineer - AI Development | Engineering Services & Platforms

JPMorganChase • Columbus (OH)

On-site
USD 180,000 - 240,000