Senior Software Engineer- Big Data & MCP, Data Foundations

RevSpring

Salt Lake City (UT)

On-site

USD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

RevSpring is looking for a Senior Software Engineer specializing in Big Data within our Data Foundations team in Salt Lake City, Utah. The ideal candidate will design and optimize distributed data pipelines and develop resilient Python services. A bachelor's degree and 5+ years of experience in Python, along with strong SQL skills, are essential for success in this role.

Candidates will work in a collaborative environment, irrespective of their backgrounds, as we value diverse perspectives.

Qualifications

  • Proven experience designing and orchestrating large-scale ETL/ELT pipelines using Apache Beam/Google Cloud Dataflow.
  • 4+ years of experience working with relational databases and analytical data warehouses.
  • Experience building scalable Python services and high-performance data APIs.

Responsibilities

  • Design, build, and optimize large-scale distributed data pipelines.
  • Develop resilient Python services and DBT models.
  • Write clean, maintainable, well-tested code and mentor others.

Skills

Data Engineering
SQL
Elasticsearch
Python services
Git
Containerization

Education

Bachelor’s Degree

Tools

Apache Beam
Google Cloud Platform
Docker

Job description

Senior Software Engineer- Big Data & MCP, Data Foundations
Essential Functions
  • Collaborate and Innovate: Partner with product managers, data engineers, and business leaders to translate complex product and data requirements into scalable, reliable data pipelines and the search experiences they power.
  • Architect Data Pipelines: Design, build, and optimize large-scale distributed batch and streaming pipelines (using Apache Airflow, Apache Beam/Dataflow, and BigQuery) to ingest, model, and transform high-volume healthcare data into clean, well‑tested, query‑ready datasets and search indices.
  • Build Data Models & Backend Services: Develop resilient Python services and DBT models that power data delivery and self‑service analytics, including Model Context Protocol (MCP) servers that expose curated data and tooling to downstream and AI consumers, and integrate with external REST/SOAP APIs and third‑party data sources.
  • Optimize Data & Search Performance: Deeply tune pipeline throughput, data warehouse performance, and search indexing—optimizing BigQuery cost and query performance and Elasticsearch index design to ensure data freshness, relevance, and scalability across high‑volume datasets.
  • Drive Engineering Excellence: Write clean, maintainable, well‑tested code and lead by example through rigorous code reviews, architectural and data‑modeling design discussions, and mentoring, driving a culture of high‑quality software and trustworthy data.
  • Pioneer New Technologies: Stay at the forefront of modern data engineering, the analytics‑engineering ecosystem (e.g., DBT, BigQuery), and information retrieval, proactively applying these advancements to strengthen our data platform and the products it powers.
Minimum Requirements
  • Data Engineering: Proven experience designing and orchestrating large‑scale ETL/ELT pipelines using Apache Beam/Google Cloud Dataflow (or similar), and DBT, built on modern cloud data warehouses. BigQuery experience is a plus.
  • Databases & SQL: 4+ years of experience working with relational databases and analytical data warehouses, with deep, advanced SQL skills and solid data‑modeling fundamentals (e.g., dimensional and normalized modeling).
  • Search & Indexing: Working experience with search indexing and Elasticsearch, including index management, mappings, and building and maintaining search indices from pipeline output. Familiarity with hybrid (BM25 + semantic/vector) search is a plus.
  • Backend & Data Services: Experience building scalable Python services and high‑performance data APIs, including developing Model Context Protocol (MCP) servers that expose data and tooling to downstream and AI consumers.
  • Infrastructure & DevOps: Strong understanding of containerization (Docker), CI/CD methodologies (e.g., GitHub Actions), Git, Infrastructure as Code (e.g., Terraform/Pulumi), and managing services within cloud platforms (3+ years of GCP experience preferred).
  • Familiarity with healthcare data standards (e.g., NPPES/NPI registries, NUCC Provider Taxonomy, machine‑readable files (MRFs) for cost transparency, and FHIR).
  • Experience with data quality and pipeline testing frameworks (e.g., dbt tests, Great Expectations) and streaming/event ingestion (e.g., Pub/Sub, Kafka).
  • Experience integrating graph‑based data and healthcare taxonomy ontologies to enrich datasets and search query context.
  • Experience with observability and logging platforms (e.g., DataDog) for monitoring pipeline health and data freshness.
Education

Bachelor’s Degree

Experience

5+ years of professional experience with Python, with strong software‑engineering fundamentals (testing, code review, design). 3+ years experience with Java or another JVM language is also highly desired, particularly for Beam/Dataflow.

Language Skills

Ability to read, analyze and interpret general business periodicals, professional journals, technical procedures or governmental regulations. Ability to write reports, business correspondence and procedure manuals. Ability to effectively present information and respond to questions from a variety of both internal and external sources.

Physical Capabilities

Standard categories. The physical capabilities described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

EEO Statement

RevSpring is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer- Big Data & MCP, Data Foundations
Senior Software Engineer- Big Data & MCP, Data Foundations

RevSpring • Portland (OR)

On-site
USD 100,000 - 130,000
Senior Software Engineer- Big Data & MCP, Data Foundations
Senior Software Engineer- Big Data & MCP, Data Foundations

RXinsider LTD. • Salt Lake City (UT)

On-site
USD 140,000 - 190,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RevSpring • Columbus (OH)

On-site
USD 90,000 - 120,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RXinsider LTD. • Boston (MA)

On-site
USD 120,000 - 170,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RXinsider LTD. • Columbus (OH)

On-site
USD 120,000 - 160,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RevSpring • Boston (MA)

On-site
USD 140,000 - 190,000
Senior Software Engineer (Python, DBT, ETL)
Senior Software Engineer (Python, DBT, ETL)

RXinsider LTD. • Portland (OR)

On-site
USD 120,000 - 160,000
Senior Software Engineer (Python, DBT, ETL)
Senior Software Engineer (Python, DBT, ETL)

RevSpring Inc • Phoenix (AZ)

On-site
USD 130,000 - 170,000
Senior Data Engineer
Senior Data Engineer

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Data Architect (Data Platform)
Data Architect (Data Platform)

RXinsider LTD. • Salt Lake City (UT)

On-site
USD 120,000 - 180,000