Data Engineer: Scalable Pipelines for AI & Analytics

SMX

Hanover (MD)

On-site

USD 103,000 - 172,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health insurance
Paid leave
Retirement plan

Job summary

SMX is seeking a Data Engineer to design, build, and maintain scalable data architecture powering AI, analytics, and reporting. You will transform raw data into high-quality pipelines and collaborate with Analysts, AI Engineers, and Software Engineers to enable enterprise-wide data usability.

The role emphasizes ELT/ETL, cloud and on-prem systems, data modeling, and data governance to ensure reliability, performance, and security across the data lifecycle.

Qualifications

  • Bachelor's degree or higher in a related field.
  • Experience with scalable data architectures is a plus.
  • Strong proficiency in Python, SQL, and Java.

Responsibilities

  • Design, construct, install, test, and maintain highly scalable data management systems and robust ELT/ETL pipelines across cloud and on-prem systems.
  • Build infrastructure for extraction, transformation, and loading of data from diverse sources using cloud technologies.
  • Implement monitoring, data quality checks, lineage tracking, and metadata management.

Skills

Python
SQL
Java
Snowflake
AWS Redshift
Spark
Databricks
Kafka
Airflow
Azure Data Factory
AWS Glue
Data modeling
Data warehousing

Education

Bachelor's degree in Computer Science/Information Technology/Data Engineering

Tools

Snowflake
AWS Redshift
Databricks
Apache Airflow
Azure Data Factory
AWS Glue
Kafka
Spark

Job description

SMX is seeking a Data Engineer to design, build, and maintain scalable data architecture powering AI, analytics, and reporting. You will transform raw data into high-quality pipelines and collaborate with Analysts, AI Engineers, and Software Engineers to enable enterprise-wide data usability.

The role emphasizes ELT/ETL, cloud and on-prem systems, data modeling, and data governance to ensure reliability, performance, and security across the data lifecycle.

Get your free, confidential resume review.

or drag and drop your file here.