Senior Lead Software Engineer- AI/ML Platform

JPMorganChase

Wilmington (DE)

On-site

USD 150,000 - 210,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorganChase is seeking a Senior Lead Software Engineer to design, build, and operate foundational cloud infrastructure enabling data scientists and ML engineers to develop, train, and deploy AI solutions across the firm. You will lead platform reliability, scalability, and automation while collaborating with cross-functional teams to solve complex infrastructure challenges.

The role accelerates the firm’s AI/ML capabilities, delivering production-grade deployments and measurable business

Qualifications

  • Formal training or certification in software engineering concepts with 5+ years of applied experience.
  • Experience delivering secure, production-quality code in Python or Java.
  • Strong foundations in distributed systems, microservices, and platform architecture/design principles.
  • Proven ability to architect and operate cloud-native infrastructure on AWS (compute, networking, storage, security) and other major clouds.
  • Demonstrated expertise with infrastructure-as-code tooling, specifically Terraform, in large-scale cloud environments.
  • Hands-on experience with Docker and Kubernetes, including AWS EKS operations.
  • Experience building or supporting production AI/ML platforms (training, deployment, and model serving/inference), including GPU infrastructure/tooling.
  • Strong DevOps/platform engineering practices: CI/CD, release automation, automated testing, and observability.
  • Experience with SQL/NoSQL databases and data integration; strong Linux, scripting, and networking fundamentals.
  • Demonstrated experience leading effective use of enterprise-authorized AI-assisted software development tools within the work environment.

Responsibilities

  • Builds and maintains reusable AI/ML platform infrastructure and shared services to support development, deployment, and operations at scale.
  • Architects, deploys, and operates secure cloud and container-based environments for training and inference, including GPU-intensive workloads.
  • Design and implement platform tooling, automation, and infrastructure-as-code solutions to streamline model deployment, environment provisioning, release management, and operational support.
  • Develops and maintains production-grade services, APIs, SDK integrations, and workflows that support model training, serving, evaluation pipelines, and AI application lifecycle management.
  • Partners with data science, ML engineering, and application teams to translate model and compute requirements into platform standards and deployment patterns.
  • Optimizes platform reliability, scalability, latency, and cost through orchestration, scheduling, and hardware acceleration.
  • Establishes operational best practices including monitoring, logging, observability, access controls, incident response, and production troubleshooting.
  • Supports enterprise LLM operationalization, including fine-tuning workflows, inference optimization, and evaluation; contribute to documentation and engineering standards.

Skills

Python
Java
Distributed systems
Cloud-native infra
DevOps practices
Secure coding
SQL/NoSQL
Linux scripting
AI-assisted development

Education

Software engineering certification

Tools

Terraform
Docker
Kubernetes
AWS EKS

Job description

Job Description

Be an integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch technology products.

As a Senior Lead Software Engineer at JPMorgan Chase within Corporate - AIML Data Platforms team , you will design, build, and operate the foundational cloud infrastructure that enables data scientists and machine learning engineers to develop, train, and deploy intelligent solutions across the firm. In this role you will serve as a technical leader, driving platform reliability, scalability, and automation while collaborating with cross-functional teams to solve complex infrastructure challenges. Your work will directly accelerate the firm's AI/ML capabilities—enabling faster experimentation and production-grade deployments that create measurable business impact.

Job Responsibilities
  • Builds and maintains reusable AI/ML platform infrastructure and shared services to support development, deployment, and operations at scale.
  • Architects, deploys, and operates secure cloud and container-based environments for training and inference, including GPU-intensive workloads.
  • Design and implement platform tooling, automation, and infrastructure-as-code solutions to streamline model deployment, environment provisioning, release management, and operational support.
  • Develops and maintains production-grade services, APIs, SDK integrations, and workflows that support model training, serving, evaluation pipelines, and AI application lifecycle management.
  • Partners with data science, ML engineering, and application teams to translate model and compute requirements into platform standards and deployment patterns.
  • Optimizes platform reliability, scalability, latency, and cost through orchestration, scheduling, and hardware acceleration.
  • Establishes operational best practices including monitoring, logging, observability, access controls, incident response, and production troubleshooting.
  • Supports enterprise LLM operationalization, including fine-tuning workflows, inference optimization, and evaluation; contribute to documentation and engineering standards.
  • Drives adoption and governance of approved AI-assisted engineering practices across teams to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test acceleration, release readiness, incident/root-cause analysis), while establishing measurable validation standards (secure coding, peer review, automated testing) and promoting reuse of proven patterns and automation within the SDLC/TLM toolchain.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including approved AI-assisted development and automation capabilities, to improve the value realized by automation at scale.
Required Qualifications, Capabilities, And Skills
  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • Experience delivering secure, production-quality code in Python or Java.
  • Strong foundations in distributed systems, microservices, and platform architecture/design principles.
  • Proven ability to architect and operate cloud-native infrastructure on AWS (compute, networking, storage, security) and other major clouds.
  • Demonstrated expertise with infrastructure-as-code tooling, specifically Terraform, in large-scale cloud environments.
  • Hands-on experience with Docker and Kubernetes, including AWS EKS operations.
  • Experience building or supporting production AI/ML platforms (training, deployment, and model serving/inference), including GPU infrastructure/tooling.
  • Strong DevOps/platform engineering practices: CI/CD, release automation, automated testing, and observability (monitoring/logging/tracing).
  • Experience with SQL/NoSQL databases and data integration; strong Linux, scripting, and networking fundamentals.
  • Demonstrated experience leading effective use of enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching senior engineers/leads on compliant usage patterns and controls.
Preferred Qualifications, Capabilities, And Skills
  • Proficiency in Go or Python for automation, tooling development, or platform service implementation.
  • Experience with MLOps frameworks and tools such as Kubeflow, MLflow, or similar AI/ML lifecycle management platforms.
  • Working knowledge of ML frameworks (PyTorch, TensorFlow, Hugging Face, scikit-learn) for model integration and operationalization.
  • Exposure to multi-cloud or hybrid cloud architectures and platform portability strategies.
ABOUT US

JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.

We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.

JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans

About The Team

Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we're setting our businesses, clients, customers and employees up for success.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Lead Software Engineer- AI/ML Platform
Senior Lead Software Engineer- AI/ML Platform

Fairygodboss • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Sr Lead Software Engineer - AWS - Lead AI/ML Platform Engineer
Sr Lead Software Engineer - AWS - Lead AI/ML Platform Engineer

JPMorganChase • Jersey City (NJ)

On-site
USD 170,000 - 210,000
Comprehensive health care coverage
On-site health and wellness centers
Retirement savings plan
+2
Senior Lead Software Engineer-AI Foundation Services
Senior Lead Software Engineer-AI Foundation Services

JPMorganChase • Plano (TX)

On-site
USD 170,000 - 210,000
Health care coverage
On-site health centers
Retirement savings plan
+1
Lead Software Engineer - AI Platform Reliability
Lead Software Engineer - AI Platform Reliability

JPMorganChase • Seattle (WA)

On-site
USD 180,000 - 240,000
Lead Software Engineer - Machine Learning Platform
Lead Software Engineer - Machine Learning Platform

JPMorganChase • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Health care coverage
On-site health centers
Retirement savings plan
+2
Sr Lead Software Engineer - AWS - Lead AI/ML Platform Engineer
Sr Lead Software Engineer - AWS - Lead AI/ML Platform Engineer

慨正橡扯 • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Senior Lead Software Engineer- AI Platform engineer
Senior Lead Software Engineer- AI Platform engineer

Next Frontier Capital • United States

On-site
USD 120,000 - 160,000
Comprehensive health care coverage
Retirement savings plan
Tuition reimbursement
Lead Software Engineer - Applied AI ML Lead
Lead Software Engineer - Applied AI ML Lead

JPMorganChase • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Principal AI & Platform Architect for Scalable Finance
Principal AI & Platform Architect for Scalable Finance

JPMorganChase • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Senior Lead Software Engineer
Senior Lead Software Engineer

JPMorganChase • Seattle (WA)

On-site
USD 130,000 - 160,000
Comprehensive health care coverage
Retirement savings plan
Tuition reimbursement
+1