Senior Data Engineer

Pantheon-Data

Reston (VA)

Hybrid

USD 140,000 - 160,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

SmartBenefits program
Transportation benefits
Tuition assistance may be available

Job summary

Pantheon Data seeks a hands-on Data Engineer to design, build, and operate data foundations supporting analytics, AI/ML, and document processing. You will handle data across structured, semi-structured, and unstructured sources with an emphasis on reliability and maintainability.

The role requires strong Python and SQL skills, data modeling, orchestration, and data quality expertise. Collaboration with ML, software, and cloud engineers is essential to deliver production-ready data services for

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a related technical field from an ABET accredited university.
  • 5+ years of professional hands-on data engineering, software engineering, analytics engineering, or closely related experience.
  • Strong Python programming skills with production-oriented code.
  • Strong SQL skills with data modeling, performance optimization, joins, indexing, and data quality.
  • Understanding of batch processing, event-driven workflows, ETL/ELT, orchestration, idempotency, retries, backfills, lineage, and failure handling.
  • Knowledge of OLTP vs OLAP systems and different storage/analysis patterns.
  • Experience building/supporting data pipelines moving data between systems (APIs, databases, files, object storage, queues, warehouses).
  • Ability to reason about data correctness, schema changes, validation, reconciliation, and operational recovery.
  • Proficiency with Git, PRs, code review, documentation, and collaborative practices.
  • Strong communication skills to explain data flow and tradeoffs to technical and non-technical stakeholders.
  • Ability to work in a distributed, cross-functional environment and meet deadlines.
  • Proficiency with Microsoft Office suite.

Responsibilities

  • Design, build, and maintain reliable data pipelines for structured, semi-structured, and unstructured data sources.
  • Develop Python-based processing workflows for data ingestion, normalization, validation, enrichment, and transformation.
  • Work with SQL and relational data stores to support transactional, analytical, and application-facing use cases.
  • Design data models and storage patterns for OLTP, OLAP, object storage, document-oriented, search, vector, or graph patterns where applicable.
  • Implement orchestration and scheduling for repeatable data workflows using tools like Airflow, AWS Step Functions, Dagster, Prefect, Glue, or similar.
  • Build automated quality checks, reconciliation logic, validation reports, and operational alerts.
  • Support data pipelines feeding AI/ML, retrieval, document intelligence, analytics, and application workflows.
  • Collaborate with ML engineers, software engineers, cloud engineers, and product stakeholders to turn ambiguous data problems into working software.
  • Write maintainable code, participate in code reviews, document data flows, and contribute to engineering standards for testing, deployment, observability, and version control.
  • Help improve velocity of a growing engineering team by owning well-scoped data engineering work while growing into broader system ownership.

Skills

Python
SQL
Data engineering
Batch processing
Orchestration
Data quality
Data modeling
Git
Communication
AWS

Education

Bachelor's degree in Computer Science/Engineering or related technical field

Tools

Airflow
Dagster
Prefect
Glue
AWS services (S3, Lambda, Redshift, etc.)
Spark/Databricks

Job description

Company Overview

Pantheon Data (a Kenific Holding company) is a private, small business based in the Washington, DC, area. Pantheon Data was founded in 2011, initially providing acquisition and supply chain management services to the US Coast Guard. Our service offerings have grown in the past ten years, including infrastructure resiliency, contact center operations, information technology, software engineering, program management, strategic communications, engineering, and cybersecurity. We have also grown our customer base to include commercial clients. The company has used this experience to expand our service offerings to other agencies within the Department of Homeland Security (DHS), the Department of Defense (DoD), and other Federal Civilian Agencies.

Position Overview

We are seeking a hands-on Data Engineer to help design, build, and operate the data foundations that support advanced analytics, AI/ML, and intelligent document processing solutions. The right candidate is a strong engineer who understands how data moves through real systems: ingestion, orchestration, transformation, quality checks, storage, query patterns, operational monitoring, and delivery to downstream applications. This person should be comfortable working across structured, semi-structured, and unstructured data, and should bring the judgment to build pipelines that are reliable, explainable, maintainable, and useful to the engineering teams and products that depend on them.

The ideal candidate has strong Python and SQL skills, understands when data should be modeled for operational use versus analytical use, and can reason clearly about batch processing, event-driven pipelines, data quality, lineage, and downstream consumption. They should be able to become productive quickly in a complex engineering environment, ask good questions, and build systems that other engineers can trust and extend. Experience with AWS, vector search, document data, or AI/ML data pipelines is valuable, but the core requirement is strong data engineering judgment: knowing how to move, structure, validate, and serve data reliably in support of real products and mission needs.

Responsibilities
  • Design, build, and maintain reliable data pipelines for structured, semi-structured, and unstructured data sources.
  • Develop Python-based processing workflows for data ingestion, normalization, validation, enrichment, and transformation.
  • Work with SQL and relational data stores to support transactional, analytical, and application-facing use cases.
  • Help design data models and storage patterns appropriate to the workload, including OLTP, OLAP, object storage, document-oriented, search, vector, or graph-oriented patterns when applicable.
  • Implement orchestration and scheduling for repeatable data workflows using tools such as Airflow, AWS Step Functions, Dagster, Prefect, Glue workflows, or similar technologies.
  • Build automated quality checks, reconciliation logic, validation reports, and operational alerts so data issues are detected early and can be diagnosed quickly.
  • Support data pipelines that feed AI/ML, retrieval, document intelligence, analytics, and application workflows.
  • Collaborate with machine learning engineers, software engineers, cloud engineers, and product stakeholders to turn ambiguous data problems into working software.
  • Write maintainable code, participate in code reviews, document data flows, and contribute to engineering standards for testing, deployment, observability, and version control.
  • Help improve the velocity of a growing engineering team by taking ownership of well-scoped data engineering work while continuing to grow into broader system ownership.
Required Skills and Experience
  • Bachelor's degree in Computer Science,Engineering, or a related technical field from an ABET accredited university.
  • 5+ years of professional hands-on data engineering, software engineering, analytics engineering, or closely related experience.
  • Strong Python programming skills, including experience writing maintainable production-oriented code rather than only notebooks or one-off scripts.
  • Strong SQL skills and practical understanding of data modeling, query performance, joins, indexing, schemas, normalization/denormalization, and data quality.
  • Understanding of core data engineering concepts, including batch processing, event-driven workflows, ETL/ELT, orchestration, idempotency, retries, backfills, lineage, and failure handling.
  • Working knowledge of OLTP versus OLAP systems and the tradeoffs between transactional databases, analytical stores, object storage, and search-oriented systems.
  • Experience building or supporting data pipelines that move data between systems, such as APIs, databases, files, object storage, queues, warehouses, or downstream applications.
  • Ability to reason about data correctness, schema changes, validation, reconciliation, duplicate handling, missing data, and operational recovery.
  • Comfortable working with Git, pull requests, code review, issue tracking, documentation, and collaborative software development practices.
  • Strong communication skills and the ability to explain data flow, design choices, limitations, and tradeoffs to both technical and non-technical stakeholders.
  • Ability to work effectively in a distributed, cross-functional engineering environment and produce high-quality work with limited hand-holding.
  • Ability to meet deadlines.
  • Proficient in Microsoft Suite software including Outlook, Word, Excel, SharePoint, and PowerPoint.
Preferred Skills and Experience
  • AWS data services such as S3, Lambda, Glue, Athena, Step Functions, SQS/SNS, Kinesis, EMR, RDS, DynamoDB, Redshift, OpenSearch, or CloudWatch.
  • Experience with workflow orchestration tools such as Airflow, Dagster, Prefect, AWS Step Functions, Glue, or similar.
  • Experience with PySpark, Spark, Databricks, EMR, Snowflake, Redshift, or other distributed/analytical data platforms.
  • Experience supporting AI/ML or RAG-style data workflows, including metadata enrichment, retrieval datasets, vector search, embeddings, evaluation datasets, or human validation workflows.
  • Experience with document-oriented or unstructured data pipelines, including PDFs, OCR output, tables, forms, images, extracted text, metadata, or search indexes.
  • Experience with graph databases or graph-shaped data models is a plus.
  • Experience with Docker, CI/CD, infrastructure as code, automated testing, logging, monitoring, and production support is a plus.
  • Familiarity with data governance, access control, PII handling, auditability, lineage, and compliance-sensitive environments.
Clearance Requirements

Applicants selected will be subject to a security investigation and may need to meet eligibility requirements. Secret Clearance is required for continued employment.

Work Location

Reston, VA - Remote

  • Our company prioritizes the benefits of flexibility and collaboration, whether that happens in person or remotely.
  • If the position is remote or hybrid, you may periodically work from a Pantheon Data office location or client site.
  • If this position is assigned to a Pantheon Data office location or client site, you'll work with colleagues and clients in person, as needed for specific client requirements.
Interview Requirement

Candidates who are local to the area should be prepared to participate in an in-person interview as part of the selection process. Candidates outside the local area may be considered for a virtual interview.

Compensation

The salary range for this position is $140,000 - $160,000. This is not, however, a guarantee of compensation or salary. Rather, salary will be set based on experience, geographic location and possibly contractual requirements and could fall outside of this range.

Benefits Overview
  • We are always looking for good people! Pantheon Data is committed to providing its employees with competitive salaries and benefits in order to increase employee satisfaction and productivity.
  • In addition to our benefits, we also offer SmartBenefits through the Washington Metro Area Transportation Authority, where you specify an amount of your pre-tax wages be paid directly to your SmarTrip account. In some cases, tuition assistance may be available for continuing education expenses and certifications related to their position. Additional details may be found at https://pantheon-data.com/careers/.
Pantheon Data Important Information

All qualified applicants will be considered for employment without regard to disability, status as a protected veteran, or any other status protected by applicable federal, state, local, or international law.

As part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

If you require reasonable accommodation in completing this application, interviewing, completing any pre-employment testing, or otherwise participating in the employee selection process, please direct your inquiries to our Talent Team at Recruiting@pantheon-data.com or by phone (571) 363-4020.

This company uses E-Verify to confirm each employee's work authorization. For more information, E-Verify Participation Poster

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Full Stack Developer
Senior Full Stack Developer

Pantheon-Data • Reston (VA)

On-site
USD 145,000 - 200,000
SmartBenefits (SmarTrip)
Tuition assistance
E-Verify participation
Senior Machine Learning Engineer - Intelligent Document Processing / Production AI Systems
Senior Machine Learning Engineer - Intelligent Document Processing / Production AI Systems

Pantheon-Data • Reston (VA)

On-site
USD 140,000 - 200,000
Data Engineer
Data Engineer

Pantheon Data • Washington

On-site
USD 80,000 - 120,000
SmartBenefits
Tuition assistance
SeniorDevSecOpsEngineer
SeniorDevSecOpsEngineer

Pantheon Data • Reston (VA)

On-site
USD 140,000 - 200,000
SmartBenefits through WMATA
Tuition assistance
SeniorDevSecOpsEngineer
SeniorDevSecOpsEngineer

Pantheon-Data • Reston (VA)

Hybrid
USD 140,000 - 200,000
SmartBenefits via SmarTrip
Tuition assistance
Pre-tax SmarTrip payments
Senior Cloud Engineer
Senior Cloud Engineer

Pantheon Data • Charlotte (NC)

Remote
USD 100,000 - 150,000
Senior Data Engineering Technical Expert
Senior Data Engineering Technical Expert

Pantheon Data • Washington

Hybrid
USD 100,000 - 200,000
SmartBenefits through WMATA
Tuition assistance
Data Analytics Technical Expert
Data Analytics Technical Expert

Pantheon Data • Washington

Hybrid
USD 100,000 - 200,000
SmartBenefits via WMATA SmarTrip
Tuition assistance
Senior Cloud Engineer
Senior Cloud Engineer

Pantheon-Data • Reston (VA)

On-site
USD 100,000 - 150,000
Data Scientist
Data Scientist

Pantheon Data • Washington

Hybrid
USD 80,000 - 150,000
SmartBenefits
Tuition assistance
Hybrid work schedule