Principal Data Engineer - AWS

Metis, Inc.

California (MO)

Hybrid

USD 140,000 - 200,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Metis Technology Solutions seeks a Principal Data Engineer – AWS in California to design, build, and operate data pipelines linking research platforms, external data sources, and the Sherlock data warehouse.

You will collaborate with developers, researchers, analysts, and DBAs to deliver scalable ETL/ELT solutions and robust data architectures.

Qualifications

  • Minimum 10 years of progressively responsible software engineering, data engineering, or related experience.
  • Minimum 5 years of substantial hands-on AWS experience, including design, deployment, and operation of production data solutions.
  • Demonstrated experience architecting ETL/ELT pipelines for large, heterogeneous datasets.
  • Strong hands-on experience with AWS data and compute services (S3, Lambda, Redshift, RDS, DynamoDB, Kinesis, Data Firehose, EC2, API Gateway, IAM, CloudWatch).
  • Advanced SQL skills and experience with PostgreSQL/PostGIS.
  • Experience with Airflow, Dagster, AWS Step Functions, or equivalents for workflow orchestration.
  • Terraform or similar IaC experience; security concepts: IAM, encryption, least-privilege, auditing.
  • Strong problem-solving, communication, and collaboration across engineers, researchers, and admins.

Responsibilities

  • Architects, designs, deploys, and maintains scalable data pipelines in AWS.
  • Designs ingestion and processing for structured, semi-structured, and streaming data.
  • Develops production-grade data processing in Python, SQL, and shell scripting.
  • Builds workflow orchestration using Airflow, Dagster, or AWS Step Functions.
  • Implements real-time and near-real-time streaming with Kinesis, Kafka, or equivalent.
  • Develops and maintains cloud data solutions using AWS services and data warehouses.
  • Maintains IaC with Terraform and automated CI/CD for data apps.
  • Ensures reliability, security, observability, and cost-efficiency of data infra.

Skills

AWS
ETL/ELT design
SQL
PostgreSQL/PostGIS
Data warehousing
Airflow
Dagster
Terraform
NoSQL databases
Security best practices

Education

Bachelor’s degree or higher in computer science or related technical discipline

Tools

AWS S3
AWS Lambda
Amazon Redshift
Amazon RDS
DynamoDB
Kinesis
Data Firehose
EC2
API Gateway
IAM
CloudWatch

Job description

If you are unable to complete this application due to a disability, contact this employer to ask for an accommodation or an alternative application process.

Full Time Professional Moffett Field, CA, US

14 days ago Requisition ID: 1577

Salary Range: $140,000.00 To $200,000.00 Annually

POSITION SUMMARY:

Metis Technology Solutions is seeking an experienced Principal Data Engineer – AWS to join our Software Operations team. This engineer will work closely with software developers, researchers, analysts, database engineers, and system administrators to design, develop, deploy, and operate the data infrastructure and pipelines connecting research platforms, external data sources, cloud-based services, and the project's Sherlock data warehouse.

This position:

  • Architects, designs, develops, deploys, and maintains scalable and reliable ETL/ELT data pipelines in Amazon Web Services (AWS).
  • Designs automated ingestion and processing solutions for structured, semi-structured, unstructured, batch, and streaming data.
  • Develops production-quality data-processing and pipeline software using Python, SQL, and shell scripting.
  • Designs and maintain workflow orchestration solutions using technologies such as Apache Airflow, Dagster, AWS Step Functions, or equivalent platforms.
  • Designs and implements real-time and near-real-time streaming and event-driven data pipelines using technologies such as AWS Kinesis, Amazon Data Firehose, Kafka, RabbitMQ, SQS/SNS, or equivalent technologies.
  • Develops and maintains cloud data solutions using AWS services such as Amazon S3, AWS Lambda, Amazon Redshift, Amazon RDS, Amazon DynamoDB, Amazon EC2, API Gateway, IAM, and CloudWatch, as appropriate to project requirements.
  • Administers and optimizes relational databases and cloud data warehouses, including PostgreSQL/PostGIS and Amazon Redshift or comparable technologies.
  • Develops and maintains infrastructure as code (IaC) using Terraform or equivalent technologies.
  • Develops and maintains automated deployment and CI/CD processes for data applications and supporting infrastructure.
  • Designs pipelines and supporting infrastructure for reliability, scalability, maintainability, observability, security, and efficient use of AWS resources.
  • Implements appropriate AWS security practices, including IAM roles and policies, encryption, secrets management, network controls, logging, and least-privilege access.
  • Develops monitoring, logging, metrics, and alerting that provide operational visibility into data-pipeline health and performance.
  • Documents data architectures, data flows, interfaces, infrastructure, deployment processes, operational procedures, and troubleshooting practices.
  • Collaborates with researchers, analysts, and application developers to translate research and application requirements into reliable and maintainable data-processing solutions.
MINIMUM QUALIFICATIONS:
Education:

Bachelor’s degree or higher in computer science, computer engineering, information systems, or a related technical discipline.

Required Skills and knowledge:
  • Minimum 10 years of progressively responsible software engineering, data engineering, database engineering, or closely related technical experience.
  • Minimum 5 years of substantial hands-on AWS experience, including the design, implementation, deployment, and operation of production data-processing or data-pipeline solutions.
  • Demonstrated experience architecting and developing ETL/ELT pipelines involving large, heterogeneous, or rapidly changing datasets.
  • Strong hands-on experience with AWS data and compute services. Relevant technologies may include S3, Lambda, Redshift, RDS, DynamoDB, Kinesis, Data Firehose, EC2, API Gateway, IAM, and CloudWatch.
  • Advanced SQL skills and substantial experience working with relational database systems, preferably PostgreSQL/PostGIS.
  • Demonstrated experience with cloud data warehouses such as Amazon Redshift or comparable technologies.
  • Experience with NoSQL databases or document/key-value data stores such as DynamoDB, MongoDB, or comparable technologies.
  • Demonstrated experience with data-pipeline and workflow orchestration using Apache Airflow, Dagster, AWS Step Functions, or comparable technologies.
  • Experience designing or implementing streaming, message-oriented, or event-driven data-processing systems using technologies such as Kinesis, Data Firehose, Kafka, RabbitMQ, SQS/SNS, or comparable technologies.
  • Demonstrated experience diagnosing and correcting database and data-pipeline performance problems.
  • Hands‑on experience with infrastructure as code, preferably Terraform or an equivalent technology.
  • Working knowledge of AWS security concepts including IAM, encryption, secrets management, network security, logging, and least‑privilege access.
  • Strong understanding of modern software‑engineering practices, including modular design, automated testing, version control, documentation, security, and maintainable code.
  • Strong analytical and troubleshooting skills and demonstrated ability to diagnose complex problems spanning applications, databases, data pipelines, and cloud infrastructure.
  • Proven ability to evaluate technical alternatives, make sound architectural decisions, and communicate the rationale and tradeoffs associated with those decisions.
  • Excellent written and verbal communication skills and demonstrated ability to collaborate effectively with software engineers, researchers, analysts, system administrators, and other technical stakeholders.
SECURITY CLEARANCE:

Applicant must be eligible to obtain a U.S. Government Public Trust Clearance. Must be a U.S. Citizen or Permanent Resident.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Data Engineer - AWS
Principal Data Engineer - AWS

Metis Technology Solutions Inc • California

On-site
USD 140,000 - 200,000
Principal Data Engineer - AWS
Principal Data Engineer - AWS

Metis Technology Solutions Inc • California (MO)

On-site
USD 140,000 - 200,000
Senior AWS Data Engineer - ETL & Streaming Pipelines
Senior AWS Data Engineer - ETL & Streaming Pipelines

Metis Technology Solutions Inc • California

On-site
USD 140,000 - 200,000
Senior AWS Data Engineer & Pipelines Architect
Senior AWS Data Engineer & Pipelines Architect

Metis Technology Solutions Inc • California (MO)

On-site
USD 140,000 - 200,000
Principal Data Engineer (Python, AWS, Redshift, EMR, Airflow, Databricks, Big data) | Contract [...]
Principal Data Engineer (Python, AWS, Redshift, EMR, Airflow, Databricks, Big data) | Contract [...]

Central Business Solutions, Inc • Irvine (CA)

On-site
USD 170,000 - 250,000
Data Engineer
Data Engineer

Magpie Health Analytics, Inc. • Baltimore (MD)

On-site
USD 110,000 - 165,000
Senior Data Engineer, AWS Analytics Engineering
Senior Data Engineer, AWS Analytics Engineering

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 155,000 - 209,000
AWS Data Engineer for Mission-Critical Systems
AWS Data Engineer for Mission-Critical Systems

KDA CONSULTING INC • Herndon (VA)

On-site
USD 90,000 - 130,000
Data Engineer - Python, SQL, AWS
Data Engineer - Python, SQL, AWS

Compunnel, Inc. • Durham (NC)

On-site
USD 95,000 - 120,000
Principal Data Engineer – (Hadoop/Big Data, AWS, Python, Kinesis) - Irvine, CA - Onsite - Contr[...]
Principal Data Engineer – (Hadoop/Big Data, AWS, Python, Kinesis) - Irvine, CA - Onsite - Contr[...]

Central Business Solutions, Inc • Irvine (CA)

On-site
USD 150,000 - 190,000