Senior Software Engineer- Big Data & MCP, Data Foundations

RevSpring

Portland (OR)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

RevSpring is looking for a Senior Software Engineer specialized in Big Data & MCP to develop and optimize robust data pipelines. You will work with product managers and data engineers to enhance our data platform through innovative techniques.

The ideal candidate has a Bachelor's degree, 5+ years of experience in Python, and strong skills in cloud data management and SQL. Familiarity with Elasticsearch and containerization tools like Docker are key for success in this role.

Qualifications

  • Proven experience with large-scale ETL/ELT pipeline design and orchestration.
  • 4+ years of advanced SQL skills and strong data-modeling fundamentals.
  • Experience building scalable Python services and APIs.

Responsibilities

  • Design and optimize data pipelines using modern tools.
  • Collaborate with teams to translate requirements into data solutions.
  • Drive engineering excellence through clean code and rigorous reviews.

Skills

ETL/ELT pipeline design
Advanced SQL
Python services development
Cloud infrastructure management (GCP)
Search indexing with Elasticsearch

Education

Bachelor’s Degree

Tools

Apache Beam
Google Cloud Dataflow
Docker
Terraform

Job description

Senior Software Engineer- Big Data & MCP, Data Foundations
Essential Functions
  • Collaborate and Innovate: Partner with product managers, data engineers, and business leaders to translate complex product and data requirements into scalable, reliable data pipelines and the search experiences they power.
  • Architect Data Pipelines: Design, build, and optimize large-scale distributed batch and streaming pipelines (using Apache Airflow, Apache Beam/Dataflow, and BigQuery) to ingest, model, and transform high-volume healthcare data into clean, well‑tested, query‑ready datasets and search indices.
  • Build Data Models & Backend Services: Develop resilient Python services and DBT models that power data delivery and self‑service analytics, including Model Context Protocol (MCP) servers that expose curated data and tooling to downstream and AI consumers, and integrate with external REST/SOAP APIs and third‑party data sources.
  • Optimize Data & Search Performance: Deeply tune pipeline throughput, data warehouse performance, and search indexing—optimizing BigQuery cost and query performance and Elasticsearch index design to ensure data freshness, relevance, and scalability across high‑volume datasets.
  • Drive Engineering Excellence: Write clean, maintainable, well‑tested code and lead by example through rigorous code reviews, architectural and data‑modeling design discussions, and mentoring, driving a culture of high‑quality software and trustworthy data.
  • Pioneer New Technologies: Stay at the forefront of modern data engineering, the analytics‑engineering ecosystem (e.g., DBT, BigQuery), and information retrieval, proactively applying these advancements to strengthen our data platform and the products it powers.
Minimum Requirements
  • Data Engineering: Proven experience designing and orchestrating large‑scale ETL/ELT pipelines using Apache Beam/Google Cloud Dataflow (or similar), and DBT, built on modern cloud data warehouses. BigQuery experience is a plus.
  • Databases & SQL: 4+ years of experience working with relational databases and analytical data warehouses, with deep, advanced SQL skills and solid data‑modeling fundamentals (e.g., dimensional and normalized modeling).
  • Search & Indexing: Working experience with search indexing and Elasticsearch, including index management, mappings, and building and maintaining search indices from pipeline output. Familiarity with hybrid (BM25 + semantic/vector) search is a plus.
  • Backend & Data Services: Experience building scalable Python services and high‑performance data APIs, including developing Model Context Protocol (MCP) servers that expose data and tooling to downstream and AI consumers.
  • Infrastructure & DevOps: Strong understanding of containerization (Docker), CI/CD methodologies (e.g., GitHub Actions), Git, Infrastructure as Code (e.g., Terraform/Pulumi), and managing services within cloud platforms (3+ years of GCP experience preferred).
  • Familiarity with healthcare data standards (e.g., NPPES/NPI registries, NUCC Provider Taxonomy, machine‑readable files (MRFs) for cost transparency, and FHIR).
  • Experience with data quality and pipeline testing frameworks (e.g., dbt tests, Great Expectations) and streaming/event ingestion (e.g., Pub/Sub, Kafka).
  • Experience integrating graph‑based data and healthcare taxonomy ontologies to enrich datasets and search query context.
  • Experience with observability and logging platforms (e.g., DataDog) for monitoring pipeline health and data freshness.
Education

Bachelor’s Degree

Experience

5+ years of professional experience with Python, with strong software‑engineering fundamentals (testing, code review, design). 3+ years experience with Java or another JVM language is also highly desired, particularly for Beam/Dataflow.

Language Skills

Ability to read, analyze and interpret general business periodicals, professional journals, technical procedures or governmental regulations. Ability to write reports, business correspondence and procedure manuals. Ability to effectively present information and respond to questions from a variety of both internal and external sources.

Physical Capabilities

Standard categories. The physical capabilities described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

EEO Statement

RevSpring is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer- Big Data & MCP, Data Foundations
Senior Software Engineer- Big Data & MCP, Data Foundations

RevSpring • Salt Lake City (UT)

On-site
USD 90,000 - 120,000
Senior Software Engineer- Big Data & MCP, Data Foundations
Senior Software Engineer- Big Data & MCP, Data Foundations

RXinsider LTD. • Salt Lake City (UT)

On-site
USD 140,000 - 190,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RevSpring • Columbus (OH)

On-site
USD 90,000 - 120,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RXinsider LTD. • Boston (MA)

On-site
USD 120,000 - 170,000
Senior Data Engineer - Analytics
Senior Data Engineer - Analytics

RXinsider LTD. • Columbus (OH)

On-site
USD 120,000 - 160,000
Senior Software Engineer (Python, DBT, ETL)
Senior Software Engineer (Python, DBT, ETL)

RXinsider LTD. • Portland (OR)

On-site
USD 120,000 - 160,000
Senior Software Engineer (Python, DBT, ETL)
Senior Software Engineer (Python, DBT, ETL)

RevSpring Inc • Phoenix (AZ)

On-site
USD 130,000 - 170,000
Senior Data Engineer
Senior Data Engineer

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Data Architect (Data Platform)
Data Architect (Data Platform)

RXinsider LTD. • Salt Lake City (UT)

On-site
USD 120,000 - 180,000
Senior Big Data Engineer (Python + GCP)
Senior Big Data Engineer (Python + GCP)

SoftServe • Town of Poland (NY)

On-site
USD 120,000 - 160,000