Knowledge Graph Engineer

BNY Mellon

Pittsburgh (Allegheny County)

Hybrid

USD 140,000 - 210,000

Full time

12 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

BNY Mellon seeks a Knowledge Graph Engineer to design and build data pipelines and graph-enabled structures powering the Investment Data Standard (IDS) and enterprise knowledge graph. You will collaborate with the Ontology and Knowledge Architecture lead across platform, product, and data teams to deliver scalable, production-ready solutions that enable analytics, AI, and client platforms.

The role involves transforming diverse data formats, implementing data quality controls, and supporting

Qualifications

  • Bachelor's degree in a related discipline or equivalent work experience.
  • 8–12 years of experience with at least 4 years focused on data analysis and business intelligence.

Responsibilities

  • Partner with Ontology and Knowledge Architecture to perform entity resolution and map data to IDS entities.
  • Design and build scalable pipelines to ingest and process data from internal platforms and external vendors.
  • Transform diverse data formats into clean, standardized datasets aligned to IDS models.
  • Develop frameworks to normalize identifiers, units, hierarchies, and event data.
  • Implement data quality controls including lineage and provenance from source to IDS downstream products.
  • Support multi-vendor data ingestion, comparison, and reconciliation with analytics.
  • Build modular, cloud-native pipelines optimized for scalability and cost efficiency.
  • Collaborate to translate business and data requirements into production-ready solutions.

Skills

RDF/OWL/SHACL
SPARQL
LPG/Cypher/GQL
Python
Data pipelines
Data modeling
ETL/ELT
SQL
Data quality

Education

Bachelor's degree
Advanced degree preferred

Tools

Snowflake
AWS
Databricks

Job description

We are seeking a Knowledge Graph Engineer to join our Data Innovation team. In this role, you will design and build the data pipelines and graph-enabled data structures that power our Investment Data Standard (IDS) and enterprise knowledge graph, enabling a unified, high-quality data ecosystem that supports analytics, AI, and client-facing solutions.

You will work closely with the Ontology and Knowledge Architecture lead and collaborate across platform, product, and data teams to deliver scalable, production-ready solutions aligned with our broader data transformation strategy. This role is located in New York, NY, Pittsburgh, PA or Lake Mary, FL.

In this role, you'll make an impact in the following ways:

  • Partner with the Ontology and Knowledge Architecture team to perform entity resolution, map source data to IDS entities, relationships, and attributes, and integrate it into the enterprise knowledge graph.
  • Design and build scalable pipelines to ingest and process data from internal platforms and external vendors across batch, streaming, and near-real-time patterns.
  • Transform diverse data formats, including APIs, flat files, streaming data, and unstructured content, into clean, standardized datasets aligned to IDS entity models.
  • Develop reusable frameworks to normalize identifiers, symbology, units, hierarchies, and event data such as corporate actions and transactions.
  • Implement robust data quality controls, including completeness, accuracy, consistency, schema validation, anomaly detection, lineage, provenance, and traceability from source systems through IDS to downstream products.
  • Support multi-vendor data ingestion, comparison, and reconciliation, including source prioritization, hierarchy logic, and coverage and quality analytics.
  • Build modular, reusable, cloud-native pipelines optimized for scalability, performance, reliability, and cost efficiency.
  • Collaborate cross-functionally to translate business and data requirements into production-ready solutions and support downstream distribution through APIs, data products, and client platforms.

To be successful in this role, we're seeking the following:

  • Bachelor's degree in a related discipline or equivalent work experience required. An advanced degree with a preference in statistics/statistical analysis is preferred.
  • Typically, 8-12 years of experience, with at least 4years' experience with a strong focus on data analysis and business intelligence is preferred.
  • Expert command of both RDF/OWL/SHACL/SPARQL and LPG/Cypher/GQL. Bonus points if you can round-trip data between the two while avoiding semantic drift.
  • Experience in building ontology-based knowledge graphs. Understand the options for graph data persistence and virtualization and be able to elucidate the tradeoffs. Experience beyond R2RML (NoSQL, API, unstructured data, etc.) is a plus.
  • Strong perspective on graph modularization, versioning, and temporality.
  • Solid understanding of AI-to-KG integration patterns (MCP, Graph RAG, AI-assisted Identity resolution and Entity/Relationship extraction, hybrid KG/Vector retrieval, text-to-query, agentic workflows).
  • Familiarity with the current knowledge graph technology landscape, including vendor solutions and open-source alternatives; broader awareness of adjacent technologies such as data catalogs and semantic layers is a plus.
  • Understanding data entitlements, licensing, and usage tracking.
  • Experience in data engineering, building and scaling production-grade data pipelines (Python, Spark, and SQL), with strong understanding of ETL/ELT frameworks and orchestration tools.
  • Proven ability to design and operate high-volume, resilient pipelines across batch, streaming, and distributed environments.
  • Experience designing data transformation and normalization layers, including schema evolution and backward compatibility.
  • Expertise with modern data platforms (e.g., Snowflake, AWS, Databricks), lakehouse architectures, and API-based data integration.
  • Strong capabilities in performance tuning, cost optimization, and implementing data quality, monitoring, logging, and lineage frameworks.
  • Domain experience with financial datasets (market data, pricing, reference data, portfolio holdings, transactions, corporate actions) and familiarity with key vendors (e.g., Bloomberg, ICE, MSCI).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Knowledge Graph Engineer – Entity Resolution & Graph Analytics
Knowledge Graph Engineer – Entity Resolution & Graph Analytics

3GIMBALS • Virginia (MN)

On-site
USD 120,000 - 190,000
Knowledge Graph Engineer – Entity Resolution & Graph Analytics
Knowledge Graph Engineer – Entity Resolution & Graph Analytics

3GIMBALS • United States

On-site
USD 130,000 - 180,000
Knowledge graph Engineer
Knowledge graph Engineer

Intuitive.ai • United States

Hybrid
USD 90,000 - 120,000
Senior Knowledge Graph Engineer
Senior Knowledge Graph Engineer

BNY Mellon • Pittsburgh

Hybrid
USD 140,000 - 210,000
Graph Data Architect: Knowledge Graphs & Ontology Expert
Graph Data Architect: Knowledge Graphs & Ontology Expert

VOLTO Consulting • Tampa (FL)

On-site
Data Architect with Fabric, Graph and Ontology
Data Architect with Fabric, Graph and Ontology

VOLTO Consulting • Tampa (FL)

On-site
USD 100,000 - 130,000
Graph AI Platform Engineer
Graph AI Platform Engineer

Photon • United States

On-site
USD 140,000 - 190,000
Sr. Consultant Machine Learning & Knowledge Graph Engineer
Sr. Consultant Machine Learning & Knowledge Graph Engineer

Dell Technologies • Austin (TX)

On-site
USD 170,000 - 250,000
Graph Data Engineer
Graph Data Engineer

NextGenEnergyJobs • Arlington (VA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Sr. Consultant Machine Learning & Knowledge Graph Engineer
Sr. Consultant Machine Learning & Knowledge Graph Engineer

Socket.dev • Town of Texas (WI)

On-site
USD 140,000 - 230,000