Staff Data Engineer

Newmark

Dallas (TX)

Hybrid

USD 190,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Newmark seeks a senior data platform architect to own the end-to-end data architecture for enterprise-scale initiatives. You will drive ingestion, transformation, and serving layers while leading MDM and AI-enabled workflows across cross-functional teams.

The role emphasizes designing scalable data pipelines, cloud-native architectures, and governance frameworks, with mentoring responsibilities and a focus on AI-enabled data capabilities. Hybrid work and competitive compensation are offered.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, MIS, or related field preferred.
  • 12+ years of experience in data engineering or software engineering, with demonstrated experience architecting data platforms and pipelines at scale.
  • Expert-level SQL and strong proficiency in Python (Scala or Java a plus) for large-scale data processing and transformation.
  • Deep experience with cloud data platforms (Databricks, Snowflake, Synapse, BigQuery, Redshift) and cloud-native architecture patterns.
  • Deep understanding of distributed systems, data modeling (dimensional, data vault, lakehouse), and ETL/ELT architecture.
  • Hands-on experience designing and implementing Master Data Management (MDM) solutions, including entity resolution, match/merge, golden records, and reference/hierarchy management (e.g., Informatica, Reltio, Profisee, or similar).
  • Hands-on experience building or integrating agentic AI systems, LLM-powered applications, RAG pipelines, or AI agent orchestration frameworks (e.g., LangChain, AutoGen, Semantic Kernel, MCP).
  • Experience building backend data services and APIs (REST/GraphQL), with comfort working across the full stack.
  • Strong background with both relational (SQL) and NoSQL data stores, plus data lake/lakehouse formats (Delta, Iceberg, Parquet).
  • Deep understanding of CI/CD pipelines, infrastructure as code, and DevOps/DataOps practices.
  • Proven track record of leading large-scale technical initiatives across multiple teams.
  • Demonstrated ability to mentor engineers and influence technical direction without direct reporting authority.

Responsibilities

  • Own and drive the technical architecture for complex, cross-team data initiatives spanning ingestion, transformation, storage, and serving layers.
  • Design, build, and maintain scalable, high-performance data pipelines and distributed data platforms in a cloud-native environment (Azure, AWS, or GCP).
  • Architect and lead enterprise MDM, including golden records, entity resolution, data domains, reference and hierarchy management, and stewardship, to create trusted data.
  • Architect and integrate agentic AI and LLM-driven workflows into data platforms and pipelines to drive efficiency and new capabilities.
  • Build and support the data foundations for machine learning and AI, including feature stores, vector stores, embeddings, and ML/LLMOps pipelines.
  • Design and deliver backend data services and APIs (REST/GraphQL), and expose curated datasets to applications, analytics, and BI consumers.
  • Set engineering standards and best practices for data quality, modeling, testing, observability, and deployment, including AI tools usage.
  • Establish data governance, lineage, cataloging, and quality frameworks across the data estate.
  • Lead technical design reviews and provide architectural guidance to multiple engineering and data teams.
  • Partner with product, analytics, and engineering leadership to translate business strategy into scalable data roadmaps, including AI-driven capabilities.
  • Identify and resolve systemic performance, reliability, and scalability issues across the data stack.
  • Mentor and coach senior and mid-level engineers, raising the technical bar across data engineering, MDM, and AI practices.
  • Drive adoption of modern frameworks, tools, and engineering practices, including agentic AI and LLM tooling, to improve delivery velocity and platform resilience.
  • Maintain awareness of emerging technologies and industry trends, particularly in agentic AI, master data management, and modern data platforms, and assess their applicability to the business.

Skills

SQL
Python
MDM
Distributed systems
Data modeling
ETL/ELT
LLM apps
RAG pipelines
APIs

Education

Bachelor's degree in CS/Engineering/MIS or related

Tools

Databricks
Snowflake
Synapse
BigQuery
Redshift
LangChain
AutoGen
Semantic Kernel
MCP
Docker
Kubernetes

Job description

Responsibilities
  • Own and drive the technical architecture for complex, cross-team data initiatives spanning ingestion, transformation, storage, and serving layers.
  • Design, build, and maintain scalable, high-performance data pipelines and distributed data platforms in a cloud-native environment (Azure, AWS, or GCP).
  • Architect and lead enterprise Master Data Management (MDM), including golden records, entity resolution, data domains, reference and hierarchy management, and stewardship, to create trusted, authoritative data across the business.
  • Architect and integrate agentic AI and LLM-driven workflows (autonomous agents, RAG pipelines, AI copilots) into data platforms and pipelines to drive efficiency and new capabilities.
  • Build and support the data foundations for machine learning and AI, including feature stores, vector stores, embeddings, and ML/LLMOps pipelines.
  • Design and deliver backend data services and APIs (REST/GraphQL), and contribute across the stack to expose curated datasets to applications, analytics, and BI consumers.
  • Set engineering standards and best practices for data quality, modeling, testing, observability, and deployment across the organization, including responsible use of AI-assisted development tools.
  • Establish data governance, lineage, cataloging, and quality frameworks across the data estate.
  • Lead technical design reviews and provide architectural guidance to multiple engineering and data teams.
  • Partner with product, analytics, and engineering leadership to translate business strategy into scalable data roadmaps, including AI-driven capabilities.
  • Identify and resolve systemic performance, reliability, and scalability issues across the data stack.
  • Mentor and coach senior and mid-level engineers, raising the technical bar across the organization on data engineering, MDM, and AI practices.
  • Drive adoption of modern frameworks, tools, and engineering practices, including agentic AI and LLM tooling, to improve delivery velocity and platform resilience.
  • Maintain awareness of emerging technologies and industry trends, particularly in agentic AI, master data management, and modern data platforms, and assess their applicability to the business.
Qualifications
Basic Qualifications
  • Bachelor's degree in Computer Science, Engineering, MIS, or related field preferred.
  • 12+ years of experience in data engineering or software engineering, with demonstrated experience architecting data platforms and pipelines at scale.
  • Expert-level SQL and strong proficiency in Python (Scala or Java a plus) for large-scale data processing and transformation.
  • Deep experience with cloud data platforms (e.g., Databricks, Snowflake, Synapse, BigQuery, Redshift) and cloud-native architecture patterns.
  • Deep understanding of distributed systems, data modeling (dimensional, data vault, lakehouse), and ETL/ELT architecture.
  • Hands-on experience designing and implementing Master Data Management (MDM) solutions, including entity resolution, match/merge, golden records, and reference/hierarchy management (e.g., Informatica, Reltio, Profisee, or similar).
  • Hands-on experience building or integrating agentic AI systems, LLM-powered applications, RAG pipelines, or AI agent orchestration frameworks (e.g., LangChain, AutoGen, Semantic Kernel, MCP).
  • Experience building backend data services and APIs (REST/GraphQL), with comfort working across the full stack.
  • Strong background with both relational (SQL) and NoSQL data stores, plus data lake/lakehouse formats (Delta, Iceberg, Parquet).
  • Deep understanding of CI/CD pipelines, infrastructure as code, and DevOps/DataOps practices.
  • Proven track record of leading large-scale technical initiatives across multiple teams.
  • Demonstrated ability to mentor engineers and influence technical direction without direct reporting authority.
Preferred Qualifications
  • Experience with data governance, lineage, and cataloging tools (e.g., Unity Catalog, Microsoft Purview, Collibra, Alation).
  • Experience designing multi-agent systems, tool-calling architectures, or retrieval-augmented generation (RAG) pipelines.
  • Experience with event-driven architectures and streaming/real-time data processing (e.g., Kafka, Event Hubs, Kinesis, Flink, Spark Structured Streaming).
  • Experience building the data layer for ML/AI, including feature stores, vector databases, embeddings, and ML/LLMOps.
  • Familiarity with containerization and orchestration (Docker, Kubernetes) and workflow orchestration (Airflow, Dagster, dbt).
  • Prior experience in commercial real estate, fintech, or operations/transaction systems.
  • Track record of speaking, writing, or open-source contributions that demonstrate technical thought leadership, especially in applied AI or data.
Why Join Us?
  • Shape the technical direction of business-critical data platforms at enterprise scale, including master data management and next-generation agentic AI initiatives.
  • Be part of a high-impact team where ownership, innovation, and technical excellence drive success.
  • Competitive compensation, growth opportunities, and access to world-class engineering, data, and AI resources.
  • Collaborative Culture: Join a high-caliber team with deep expertise across data engineering, cloud, MDM, agentic AI, and distributed systems.
  • Growth & Learning: Access world-class learning resources and mentorship to advance your career.
  • Work-Life Balance: Flexible working hours and hybrid options.
  • Benefits: Comprehensive health, dental and vision insurance.
Salary

The expected base salary for this position ranges from $190,000 to $250,000 annually. The actual base salary will be determined on an individualized basis taking into account a wide range of factors including, but not limited to, relevant skills, experience, education, and, where applicable, licenses or certifications held. In addition to base salary and a competitive benefits package, this position may be eligible for additional types of compensation including discretionary bonuses and other short- and long-term incentives (e.g., deferred cash, equity, etc.).

About Us

Newmark Group, Inc. (Nasdaq: NMRK), together with its subsidiaries ("Newmark"), is a world leader in commercial real estate, seamlessly powering every phase of the property life cycle. Newmark's comprehensive suite of services and products is uniquely tailored to each client, from owners to occupiers, investors to founders, and startups to blue-chip companies. Combining the platform’s global reach with market intelligence in both established and emerging property markets, Newmark provides superior service to clients across the industry spectrum. For the twelve months ended March 31, 2026, Newmark generated revenues of more than $3.4 billion. As of March 31, 2026, Newmark and its business partners together operated from over 185 offices with more than 9,600 professionals across four continents. To learn more, visit nmrk.com or follow @newmark .

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer/Full Stack
Senior Data Engineer/Full Stack

Newmark • Chicago (IL)

Hybrid
USD 190,000 - 230,000
Managing Director
Managing Director

Newmark Group • San Francisco (CA), Northern (KY)

Hybrid
USD 90,000 - 110,000
Industry leading Parental Leave Policy
Bright Horizons back-up care program
Generous paid time off
+2
Managing Director
Managing Director

Cantor Fitzgerald • San Francisco (CA)

On-site
USD 90,000 - 110,000
Parental Leave Policy (up to 16 weeks)
Back-up care program
Generous paid time off
+2
Research Coordinator
Research Coordinator

Newmark • Cleveland (OH)

On-site
USD 40,000 - 55,000
Staff Data Engineer
Staff Data Engineer

Newmark Group • Dallas (TX)

Hybrid
USD 190,000 - 250,000
Senior Associate, Newmark Consulting Group
Senior Associate, Newmark Consulting Group

Newmark • New York (NY)

On-site
USD 120,000 - 180,000
Associate Human Resources Business Partner
Associate Human Resources Business Partner

Newmark • United States

On-site
USD 55,000 - 85,000
Associate Human Resources Business Partner
Associate Human Resources Business Partner

Cantor Fitzgerald • Jacksonville (FL)

On-site
USD 65,000 - 90,000
Regional Workplace Manager
Regional Workplace Manager

Newmark • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Building Engineer, Critical Data Centers
Building Engineer, Critical Data Centers

Newmark • Olde West Chester (OH)

On-site
USD 70,000 - 100,000
Parental Leave
Healthcare
Backup care
+4