Lead Software Engineer - Databricks/PySpark/AI

JPMorgan Chase & Co.

Fairfax (DE)

On-site

USD 180,000 - 230,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

JPMorganChase is seeking a Lead Software Engineer - Databricks/PySpark/AI to lead a senior, hands-on team within Corporate Sector-Global Finance. You will build, optimize, and deploy data products powering agentic AI systems, mentor junior engineers, and drive enterprise data accessibility and governance.

The role emphasizes AWS-based production delivery, CI/CD, and cross-functional collaboration. You will champion engineering excellence, design scalable data infrastructure, and integrate tools

Qualifications

  • 5+ years of professional software engineering experience.
  • Strong Python/PySpark production-grade coding skills.
  • Experience leading hands-on engineering teams in an agile environment.
  • Proven ability to build and optimize data pipelines for AI-enabled systems.

Responsibilities

  • Build and optimize data pipelines and workflows for agentic AI systems.
  • Promote AI-assisted engineering practices, code quality, and automated testing.
  • Develop data retrieval and indexing layers for cross-source AI access.
  • Create APIs and data services for AI agents to interact with enterprise data.
  • Mentor junior engineers and collaborate with product and data science teams.

Skills

Python
PySpark
Databricks
AWS
CI/CD
Leadership

Tools

Databricks
S3
Lambda
Glue
Snowflake
Terraform

Job description

We have an exciting and rewarding opportunity for you to take your data engineering career to the next level.

As a Lead Software Engineer - Databricks/PySpark/AI at JPMorganChase within the Corporate Sector-Global Finance team, you will serve as a senior hands-on developer and technical leader within an agile team, responsible for building, delivering, and optimizing cutting-edge data products that power agentic AI systems — autonomous AI agents capable of planning, reasoning, and executing multi-step tasks. In this role, you will write production-quality code daily, drive implementation of essential technology solutions including data infrastructure, tool integrations, and retrieval systems that enable AI agents to access, interpret, and act on enterprise data in support of the firm’s business goals. You will be expected to mentor junior engineers, collaborate with cross-functional stakeholders, and champion engineering excellence through hands-on delivery.

Job Responsibilities
  • Building and optimizing data pipelines and workflows that serve as the backbone for agentic AI systems, ensuring agents have reliable, real-time access to high-quality, structured and unstructured data
  • Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test strategy acceleration, incident/root-cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.
  • Developing data retrieval and indexing layers that enable AI agents to autonomously search, query, and synthesize information across multiple data sources
  • Building and maintaining tool-use infrastructure — APIs, data services, and function endpoints — that AI agents invoke to execute tasks, retrieve data, and interact with enterprise systems
  • Implementing and enforcing best practices for data management, ensuring data quality, security, and compliance, including governance of data consumed and generated by autonomous AI agents
  • Hands‑on development of secure, high-quality production code following AWS best practices, and deploying efficiently using CI/CD pipelines;

    Building orchestration and state management layers that support multi‑step agent workflows, including memory, context persistence, and task chaining

  • Writing and reviewing code daily, conducting thorough code reviews, and raising the technical bar across the team;

    Mentoring and guiding junior and mid‑level engineers through pairing, code reviews, and technical coaching

  • Collaborating with product owners, data scientists, and business stakeholders to translate business requirements into working, production‑ready agentic AI solutions;

    Evaluating and adopting emerging agentic AI frameworks, tools, and data engineering practices to continuously improve the team’s development capabilities

Required Qualifications, Capabilities, and Skills
  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • Demonstrated experience leading effective use of approved AI‑assisted software development tools (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security.
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching engineers on safe, compliant adoption within delivery practices
  • Expert‑level programming skills in Python/PySpark with a strong portfolio of production‑grade code
  • Extensive hands‑on experience with Databricks and the AWS cloud ecosystem, including AWS Glue, S3, SQS/SNS, Lambda,

    Spark and SQL

  • Strong hands‑on experience with Lakehouse/Delta Lake architecture, application development, testing, and ensuring operational stability; Snowflake, Terraform and LLMs; Data Observability, Data Quality, Query Optimization & Cost Optimization
  • In‑depth knowledge of Big Data and data warehousing concepts at enterprise scale
  • Extensive experience with CI/CD processes and automated testing frameworks
  • Solid understanding of agile methodologies, including DevOps practices, application resiliency, and security measures
  • Understanding of agentic AI concepts — how autonomous AI agents plan, reason, use tools, and execute multi‑step workflows — and the data infrastructure required to support them
  • Experience building APIs, data services, and retrieval systems that serve as the connective tissue between AI agents and enterprise data

Preferred Qualifications, Capabilities, and Skills
  • Experience with agentic AI frameworks (e.g., LangGraph, AutoGen, CrewAI, OpenAI Assistants API) and understanding of how data engineering underpins agent orchestration
  • Familiarity with tool‑use and function‑calling patterns for LLM‑based agents, including building and exposing APIs and data endpoints that agents can invoke autonomously
  • Experience with vector databases (e.g., Pinecone, FAISS, Chroma) and embedding workflows for powering agent memory, semantic search, and retrieval‑augmented generation (RAG)
  • Exposure to agent memory and state management patterns — short‑term context windows, long‑term persistent memory stores, and conversation/task history management
  • Familiarity with guardrails and safety frameworks for autonomous AI systems, including input/output validation, action approval workflows, and human‑in‑the‑loop controls
  • Understanding of observability and monitoring for agentic systems — tracing agent decision paths, logging tool invocations, and debugging multi‑step autonomous workflows
  • Understanding of responsible AI principles, particularly around autonomous decision‑making, data provenance, and auditability of agent actions
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Software Engineer - Databricks/PySpark/AI
Lead Software Engineer - Databricks/PySpark/AI

JPMorgan Chase & Co. • Wilmington (DE)

On-site
USD 140,000 - 210,000
Sr. Lead Software Engineer
Sr. Lead Software Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 140,000 - 190,000
Lead Software Engineer - AI Application
Lead Software Engineer - AI Application

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 180,000 - 260,000
Lead Software Engineer- Financial Services Data Engineering: Pyspark / Java / BigData / Datalak[...]
Lead Software Engineer- Financial Services Data Engineering: Pyspark / Java / BigData / Datalak[...]

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 170,000 - 210,000
Lead Software Engineer- Financial Services Data Engineering: Pyspark / Java / BigData / Datalake / AI
Lead Software Engineer- Financial Services Data Engineering: Pyspark / Java / BigData / Datalake / AI

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Sr Lead Software Engineer
Sr Lead Software Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 230,000
Lead Software Engineer-Data Engineer, Pyspark, Databricks
Lead Software Engineer-Data Engineer, Pyspark, Databricks

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 115,000 - 170,000
Lead Software Engineer - Data Analytics
Lead Software Engineer - Data Analytics

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 230,000
Lead Software Engineer - Data Engineering & Applied AI
Lead Software Engineer - Data Engineering & Applied AI

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 150,000 - 190,000
Lead Software Engineer - Cloud/AI Engineer
Lead Software Engineer - Cloud/AI Engineer

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 150,000 - 210,000