Data Automation Engineer (Public Trust)

System One

Washington (District of Columbia)

Remote

USD 115,000 - 135,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

System One in Washington, DC is seeking a Data Automation Engineer (Public Trust) for contract-to-hire to design scalable data workflows in AWS, with Azure integration as needed. You will build ETL/ELT pipelines across DynamoDB, SQL Server on AWS, and Azure SQL, advancing enterprise data capabilities.

You will implement batch and near-real-time ingestion using Spark, Kafka/Flume, and align pipelines with Solr search indexing.

Qualifications

  • Bachelor's degree in Computer Science or a related field.
  • 5+ years of experience in data engineering, data automation, or a related discipline.
  • Ability to independently design, develop, test, and troubleshoot Python- and SQL-based data pipelines in AWS environments, including integrations with Azure services where required, and clearly explain personal contributions to production implementations.
  • Strong hands-on experience with Apache Spark and working knowledge of at least one streaming or ingestion technology, such as Apache Kafka or Apache Flume.
  • Hands-on experience with multiple AWS data and integration services, including several of the following: Amazon S3, AWS Glue, AWS Lambda, Amazon EMR, AWS Step Functions, and at least one AWS database service.
  • Practical experience integrating at least one LLM platform or model service, such as Amazon Bedrock, Azure OpenAI Service, or an open-source model, into a Python-based workflow.
  • Experience integrating REST APIs and external services into Python-based data pipelines and automated workflows.
  • Experience using Jira and one or more source-control, build, or CI/CD platforms, such as GitHub, Azure DevOps, or Jenkins.
  • Strong troubleshooting and performance-optimization skills across SQL, Spark, batch pipelines, and near-real-time ingestion workflows.
  • Experience supporting production data platforms, including SLA monitoring, incident resolution, root-cause analysis, data reconciliation, performance troubleshooting, vulnerability remediation, and recurring maintenance.
  • Good communication and presentation skills.
  • US Citizenship and ability to obtain Federal government Public Trust clearance.

Responsibilities

  • Design and implement scalable data automation workflows using AWS services, with integration to selected Azure data platforms where required.
  • Develop ETL/ELT processes to ingest, transform, and move data across Amazon DynamoDB, SQL Server hosted on AWS, Azure SQL, and other enterprise data sources.
  • Design, develop, and support batch and near-real-time ingestion pipelines using Apache Spark and technologies such as Kafka or Flume, and collaborate with the search engineering team to integrate those pipelines with the existing Apache Solr platform.
  • Evaluate and apply Generative AI services and frameworks, such as Amazon Bedrock, Azure OpenAI, Hugging Face, and LangChain, to prototype and evaluate selected GenAI-assisted capabilities, such as metadata enrichment, data-quality analysis, structured data extraction, anomaly identification, and natural-language access to enterprise data.
  • Recommend suitable use cases for future implementation.
  • Develop scalable data-processing solutions using Amazon EMR and containerized deployment environments such as AWS Fargate or Kubernetes.
  • Integrate Amazon Connect customer-interaction data into analytical data stores for operational reporting and analytics.
  • Apply source-control, build, containerization, and CI/CD practices using tools such as GitHub, Azure DevOps, Jenkins, and Docker.
  • Implement data solutions in accordance with established security and compliance controls, including identity and access management, KMS encryption, VPC isolation, role-based access control, and firewall policies.
  • Support Agile DevOps processes with sprint-based delivery of pipeline and AI-enabled features.

Skills

Python
SQL
AWS
Apache Spark
Apache Kafka
REST APIs
Jira
GitHub
Azure DevOps
Jenkins
Docker
Kubernetes
LLM integration
Terraform
Bedrock/Azure OpenAI

Education

Bachelor’s degree in Computer Science or related field

Tools

Docker
Kubernetes
Terraform
GitHub
Azure DevOps
Jenkins

Job description

Job Title: Data Automation Engineer (Public Trust)
Location: Washington, District Of Columbia (remote)
Type: Contract To Hire
Compensation: $57.65/HR on W2
Security Clearance: Public Trust

Responsibilities
  • Design and implement scalable data automation workflows using AWS services, with integration to selected Azure data platforms where required. Develop ETL/ELT processes to ingest, transform, and move data across Amazon DynamoDB, SQL Server hosted on AWS, Azure SQL, and other enterprise data sources.
  • Design, develop, and support batch and near-real-time ingestion pipelines using Apache Spark and technologies such as Kafka or Flume, and collaborate with the search engineering team to integrate those pipelines with the existing Apache Solr platform.
  • Evaluate and apply Generative AI services and frameworks, such as Amazon Bedrock, Azure OpenAI, Hugging Face, and LangChain, to prototype and evaluate selected GenAI-assisted capabilities, such as metadata enrichment, data-quality analysis, structured data extraction, anomaly identification, and natural-language access to enterprise data.
  • Recommend suitable use cases for future implementation.
  • Develop scalable data-processing solutions using Amazon EMR and containerized deployment environments such as AWS Fargate or Kubernetes.
  • Integrate Amazon Connect customer-interaction data into analytical data stores for operational reporting and analytics.
  • Apply source-control, build, containerization, and CI/CD practices using tools such as GitHub, Azure DevOps, Jenkins, and Docker.
  • Implement data solutions in accordance with established security and compliance controls, including identity and access management, KMS encryption, VPC isolation, role-based access control, and firewall policies.
  • Support Agile DevOps processes with sprint-based delivery of pipeline and AI-enabled features.
Requirements
  • Bachelor’s degree in Computer Science or a related field and 5+ years of experience in data engineering, data automation, or a related discipline.
  • Ability to independently design, develop, test, and troubleshoot Python- and SQL-based data pipelines in AWS environments, including integrations with Azure services where required, and clearly explain personal contributions to production implementations.
  • Strong hands-on experience with Apache Spark and working knowledge of at least one streaming or ingestion technology, such as Apache Kafka or Apache Flume.
  • Hands-on experience with multiple AWS data and integration services, including several of the following: Amazon S3, AWS Glue, AWS Lambda, Amazon EMR, AWS Step Functions, and at least one AWS database service.
  • Practical experience integrating at least one LLM platform or model service, such as Amazon Bedrock, Azure OpenAI Service, or an open-source model, into a Python-based workflow.
  • Experience integrating REST APIs and external services into Python-based data pipelines and automated workflows.
  • Experience using Jira and one or more source-control, build, or CI/CD platforms, such as GitHub, Azure DevOps, or Jenkins.
  • Strong troubleshooting and performance-optimization skills across SQL, Spark, batch pipelines, and near-real-time ingestion workflows.
  • Experience supporting production data platforms, including SLA monitoring, incident resolution, root-cause analysis, data reconciliation, performance troubleshooting, vulnerability remediation, and recurring maintenance.
  • Good communication and presentation skills.
  • US Citizenship and ability to obtain Federal government Public Trust clearance.
Preferred Qualifications
  • Relevant certifications, such as AWS Certified Data Engineer – Associate, AWS Certified Machine Learning – Specialty, Microsoft Certified: Azure AI Engineer Associate, or Databricks Certified Data Engineer.
  • Familiarity with retrieval-augmented generation pipelines, embeddings, and vector-search technologies such as Apache Solr, Amazon OpenSearch Service, pgvector, or similar platforms.
  • Experience with multi-cloud data integration (AWS and Azure).
  • Experience with Docker and Kubernetes for containerized deployment, scalable data processing, and orchestration.
  • Experience operationalizing Generative AI workflows, including prompt and model configuration, evaluation, observability, monitoring, and lifecycle management.
  • Knowledge of data lineage/governance tools (Purview, Unity Catalog, AWS Glue Catalog).
  • Familiarity with infrastructure-as-code tools, such as Terraform, AWS CloudFormation, or Bicep, for automated deployments.
  • Experience with compliance frameworks (FedRAMP, PCI-DSS, HIPAA).

System One, and its subsidiaries including Joulé and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.

System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer (onsite)
Senior Data Engineer (onsite)

System One • Atlanta (GA)

On-site
USD 140,000 - 155,000
Remote Data Automation Engineer (Public Trust)
Remote Data Automation Engineer (Public Trust)

System One • Washington

Remote
USD 115,000 - 135,000
Data Engineer IV- #26-23312
Data Engineer IV- #26-23312

US Tech Solutions • Charlotte (NC)

On-site
USD 138,000 - 152,000
Data Engineer (GovCon; Public Trust) United States - Remote
Data Engineer (GovCon; Public Trust) United States - Remote

Attain Talent • United States

Remote
USD 110,000 - 140,000
Remote Work (Hybrid)
Medical, Dental, Vision
401(k) with matching
+2
AWS AI/Data Engineer
AWS AI/Data Engineer

System One • Knoxville (TN)

Hybrid
USD 110,000 - 170,000
Health and welfare benefits
401(k) plan
Senior AWS IAM & Data Engineer
Senior AWS IAM & Data Engineer

System One • Lafayette (LA)

On-site
USD 120,000 - 160,000
Data Engineer
Data Engineer

SoTalent • McLean (VA)

On-site
USD 102,000 - 170,000
Health, dental, and vision insurance
401(k) with company matching
Tuition reimbursement
+2
Senior Data Engineer
Senior Data Engineer

Node.Digital LLC • Herndon (VA)

Hybrid
USD 120,000 - 160,000
Medical
Dental
Vision
+6
Senior Data Engineer
Senior Data Engineer

Pantheon Data LLC. • Reston (VA), Northern (KY)

Hybrid
USD 140,000 - 160,000
SmartBenefits program
Tuition assistance
Senior Data Engineer
Senior Data Engineer

Pantheon-Data • Reston (VA)

Hybrid
USD 140,000 - 160,000
SmartBenefits program
Transportation benefits
Tuition assistance may be available