Lead Software Engineer- Financial Services Data Engineering: Pyspark / Java / BigData / Datalak[...]

JPMorgan Chase & Co.

Jersey City (NJ)

On-site

USD 170,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorgan Chase & Co. in Jersey City seeks a Lead Software Engineer to shape AI‑first analytics and data products within Global Prime Brokerage. You will lead an agile team, drive modernization of the data mesh, and build scalable, governed data platforms using Python, PySpark, and Java.

You will deliver secure, reliable analytics solutions across business functions while advancing AI capabilities and ecosystem integration.

Qualifications

  • Formal training or certification on software engineering concepts and 5+ years applied experience.
  • Proven leadership delivering AI-first analytics and data engineering at scale.
  • Experience developing data ingestion and integration processes for analytics-ready datasets.
  • Deep hands-on experience with big data platforms (Spark, Databricks, Snowflake, Iceberg).
  • Strong programming in Python and PySpark, or Java, with CI/CD and containerization practices.
  • Applied AI expertise in ML pipelines, NLP/LLMs, and autonomous agents.
  • Governance across lineage, data quality, and access control with regulatory alignment.
  • Comfortable in an agile environment with strong communication.
  • Proficient in SDLC tools and AI-assisted development.

Responsibilities

  • Collaborate with business and technology teams to develop AI‑first analytics and data product solutions.
  • Define and enforce architecture for an AI‑driven data product lifecycle: semantic extraction/alignment, automated lineage and DQ, pipeline code generation, mesh registration, and self‑service consumption.
  • Build and operate autonomous agents for data engineering that detect schema drift, propose transformations, reconcile semantics, triage data incidents, and maintain governance evidence under human‑in‑the‑loop controls.
  • Design analytics platforms capable of running reporting and other analytics; explore innovative ideas by building real‑time and batch analytics solutions.
  • Establish monitoring and alerting of solution events related to performance, scalability, availability, and reliability.
  • Provide technical leadership, guidance, and direction to other team members; build prototypes for demonstrations for peer groups, business partners, and senior leaders.
  • Drive team adoption of enterprise‑authorized AI‑assisted engineering practices to improve code quality, delivery speed, and operational outcomes.
  • Apply knowledge of SDLC tools to improve automation and value of AI outputs.
  • Lead migration and modernization from legacy stacks to the strategic mesh ecosystem (Databricks/Iceberg/common services).
  • Industrialize entity resolution with ML/LLM solutions and standardize analytical product packaging for reuse and monetization.
  • Embed governance, lineage, and DQ by design across critical domains and regulatory reporting.

Skills

Leadership
AI-first analytics
Python
PySpark
Java
Big data
Data governance
CI/CD
Secure coding
Agile
Communication
ML/LLMs

Tools

Spark
Databricks
Snowflake
Iceberg
Kubernetes
Docker

Job description

We have an opportunity to impact your career and provide an adventure where you can push the limits of what's possible.

As a Lead Software Engineer- Python / Pyspark / Java / BigData / Data Modernization / AI, at JPMorganChase within the Asset and Wealth Management- Global Prime Brokerage Team, youare an integral part of an agile team that works to enhance, build, and deliver trusted market-leading technology products in a secure, stable, and scalable way. As a core technical contributor, you are responsible for conducting critical technology solutions across multiple technical areas within various business functions in support of the firm’s business objectives.

We look for people who are passionate about solving business problems through innovation, analytics, and an AI‑first engineering mindset—building reusable, governed analytical data products and accelerating delivery of regulatory and CEO‑priority analytics. You will define and enforce an AI-driven data product lifecycle (semantic alignment, automated lineage and data quality, pipeline/code generation, mesh registration, and self‑service consumption) and will build human‑in‑the‑loop autonomous agents to detect schema drift, propose transformations, reconcile semantics, triage data incidents, and generate governance evidence. You’ll be required to apply your depth of knowledge and expertise to all aspects of the analytics development lifecycle, and partner continuously with stakeholders across product, platform, risk, and domain teams. You will lead an AI‑first transformation of data engineering and analytics by productizing the data product lifecycle (semantics, lineage, DQ, governance) and building autonomous agents (human‑in‑the‑loop) that reduce manual toil, improve auditability, and enable self‑service consumption on the strategic data mesh. The role also owns modernization of the strategic data mesh.

Job responsibilities
  • Collaborate with business and technology teams to develop AI‑first analytics and data product solutions
  • Define and enforce architecture for an AI‑driven data product lifecycle: semantic extraction/alignment, automated lineage and DQ, pipeline code generation, mesh registration, and self‑service consumption
  • Build and operate autonomous agents for data engineering that detect schema drift, propose transformations, reconcile semantics, triage data incidents, and maintain governance evidence under human‑in‑the‑loop controls
  • Design analytics platforms capable of running reporting and other analytics; explore innovative ideas by building real‑time and batch analytics solutions
  • Establish appropriate monitoring and alerting of solution events related to performance, scalability, availability, and reliability
  • Provide technical leadership, guidance, and direction to other team members; build prototypes for demonstrations for peer groups, business partners, and senior leaders
  • Drive team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test strategy acceleration, incident/root‑cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team
  • Apply knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation
  • Lead migration and modernization from legacy analytics/reporting stacks to the strategic mesh ecosystem (e.g., Databricks/Iceberg/common services), reducing fragmentation and duplicated data products
  • Industrialize entity resolution and parent identification with ML/LLM solutions and standardize analytical product packaging to enable reuse and monetization
  • Embed governance, lineage, and DQ by design across critical domains and regulatory reporting, improving auditability and control posture
Required qualifications, capabilities, and skills
  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • Proven leadership delivering AI‑first analytics and data engineering at scale, including productized data mesh patterns, semantic layers, and analytical data product lifecycle ownership
  • Experience developing data ingestion and integration processes, sourcing data from multiple platforms, and applying data cleansing/transformation rules for analytics‑ready datasets
  • Deep hands‑on experience with big data and modern data platforms (e.g., Spark, Databricks, Snowflake, Iceberg) and building robust pipelines and data lake/lakehouse frameworks
  • Strong programming capability in Python and PySpark, or Java, with strong CI/CD and containerization practices
  • Applied AI expertise in ML pipelines, NLP/LLMs, and agentic frameworks to build autonomous agents for engineering tasks (schema drift detection, semantic reconciliation, incident triage, governance evidence generation) under human‑in‑the‑loop controls
  • Governance proficiency across lineage, data quality, and access control with evidence generation aligned to regulatory expectations (e.g., BCBS 239‑class lineage/DQ)
  • Comfortable working in an agile and collaborative environment; strong written and verbal communication skills
  • Demonstrated experience leading effective use of approved AI‑assisted software development tools (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security.
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching engineers on safe, compliant adoption within delivery practices
  • Proficient in all aspects of the Software Development Life Cycle

Preferred qualifications, capabilities, and skills
  • Python and Java
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Software Engineer - Data Engineering & Applied AI
Lead Software Engineer - Data Engineering & Applied AI

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 150,000 - 190,000
Lead Software Engineer - Data Engg/ Data Analytics
Lead Software Engineer - Data Engg/ Data Analytics

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 140,000 - 210,000
Lead Software Engineer- Big Data Python /Java , Databricks
Lead Software Engineer- Big Data Python /Java , Databricks

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 140,000 - 210,000
Lead Software Engineer - Python, AWS, Big Data
Lead Software Engineer - Python, AWS, Big Data

JPMorgan Chase & Co. • Wilmington (DE)

On-site
USD 140,000 - 190,000
Lead Software Engineer - Data & AI Platform Engineer
Lead Software Engineer - Data & AI Platform Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 230,000
Lead Software Engineer - Python/ Data
Lead Software Engineer - Python/ Data

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 120,000 - 150,000
Lead Software Engineer-Big Data Python / Java , Databricks
Lead Software Engineer-Big Data Python / Java , Databricks

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 140,000 - 170,000
Lead Software Engineer - Data & AI Platform Engineer
Lead Software Engineer - Data & AI Platform Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 140,000 - 210,000
Lead Software Engineer - AI Platforms
Lead Software Engineer - AI Platforms

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 150,000 - 210,000
Senior Lead Software Engineer-Big Data Python /Java , Databricks
Senior Lead Software Engineer-Big Data Python /Java , Databricks

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 150,000 - 230,000