Data Engineer (P-175)

smashcr

Town of Texas (WI)

On-site

USD 110,000 - 140,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

SMASH is seeking an experienced Data Engineer to design, build, and maintain scalable data pipelines and cloud-based data solutions. The role emphasizes coding excellence with Python, SQL, Spark, Spark SQL, PySpark, and Microsoft Azure.

Ideal candidates will bring hands-on experience with Microsoft Fabric and have exposure to Pharmaceutical, Life Sciences, or Insurance industries. You will collaborate with analysts, scientists, and engineers to deliver reliable data capabilities.

Qualifications

  • 5–6+ years of professional Data Engineering experience.
  • Proven hands-on experience designing, building, and maintaining production data pipelines.
  • Strong coding and software development capabilities.
  • Strong hands-on Python experience.
  • Advanced SQL skills.
  • Hands-on experience with Apache Spark.
  • Strong experience with Spark SQL.
  • Strong experience developing data solutions using PySpark.
  • Experience building data solutions within Microsoft Azure.
  • Experience developing automated workflows for data ingestion, transformation, and delivery.

Responsibilities

  • Design, build, and maintain automated data pipelines that move and transform data across systems.
  • Develop scalable data workflows to support analytics, reporting, and machine learning use cases.
  • Build data ingestion, transformation, processing, and integration solutions within Microsoft Azure.
  • Develop and maintain production-quality code using Python and SQL.
  • Use Apache Spark, Spark SQL, and PySpark to process and transform large-scale datasets.
  • Design data transformations that convert raw data into reliable and usable formats for downstream consumers.
  • Develop efficient data integration processes across multiple data sources and destinations.
  • Monitor and troubleshoot data pipelines to ensure reliability, accuracy, and performance.
  • Identify and resolve data quality, pipeline, and processing issues.
  • Optimize data workflows and code for performance, scalability, and maintainability.
  • Collaborate with Data Analysts, Data Scientists, engineering teams, and business stakeholders to understand data requirements.
  • Support the implementation and continuous improvement of cloud-based data engineering solutions.
  • Document pipeline architecture, transformations, dependencies, and technical processes.
  • Follow software engineering best practices for coding, testing, version control, and deployment.

Skills

Python
SQL
Apache Spark
Spark SQL
PySpark
Azure
Data pipelines
ETL/ELT
Data integration
Analytical thinking

Tools

Microsoft Fabric

Job description

SMASH, Who we are?

We believe in long-lasting relationships with our talent. We invest time getting to know them and understanding what they seek as their professional next step.

We aim to find the perfect match. As agents, we pair our talent with our US clients, not only by their technical skills but as a cultural fit. Our core competency is to find the right talent fast.

We purposefully move away from the “contractor” or “outsourcing” type of relationship. Our clients don’t want contractors or “just a service.” Neither does our talent.

To be eligible for this role, you must have US Citizenship or valid US work authorization.

Role summary

We are looking for an experiencedData Engineerwith a strong background in designing, building, and maintaining scalable data pipelines and cloud-based data solutions.

The ideal candidate will bring hands‑on expertise withSQL, Python, Spark, Spark SQL, PySpark, and Microsoft Azure, with a strong emphasis on coding and end-to-end pipeline development. Experience withMicrosoft Fabricand within thePharmaceutical, Life Sciences, or Insuranceindustries will be highly valued.

Responsibilities
  • Design, build, and maintain automateddata pipelinesthat move and transform data across systems.
  • Develop scalable data workflows to support analytics, reporting, and machine learning use cases.
  • Build data ingestion, transformation, processing, and integration solutions within Microsoft Azure.
  • Develop and maintain production-quality code usingPython and SQL.
  • UseApache Spark, Spark SQL, and PySparkto process and transform large-scale datasets.
  • Design data transformations that convert raw data into reliable and usable formats for downstream consumers.
  • Develop efficient data integration processes across multiple data sources and destinations.
  • Monitor and troubleshoot data pipelines to ensure reliability, accuracy, and performance.
  • Identify and resolve data quality, pipeline, and processing issues.
  • Optimize data workflows and code for performance, scalability, and maintainability.
  • Collaborate with Data Analysts, Data Scientists, engineering teams, and business stakeholders to understand data requirements.
  • Support the implementation and continuous improvement of cloud-based data engineering solutions.
  • Document pipeline architecture, transformations, dependencies, and technical processes.
  • Follow software engineering best practices for coding, testing, version control, and deployment.
Requirements – Must-haves
  • 5–6+ years of professional Data Engineering experience.
  • Proven hands‑on experiencedesigning, building, and maintaining production data pipelines.
  • Strong coding and software development capabilities.
  • Strong hands‑onPythonexperience.
  • AdvancedSQLskills.
  • Hands‑on experience withApache Spark.
  • Strong experience withSpark SQL.
  • Strong experience developing data solutions usingPySpark.
  • Experience building data solutions withinMicrosoft Azure.
  • Experience developing automated workflows for data ingestion, transformation, and delivery.
  • Experience processing and transforming large and complex datasets.
  • Strong understanding of data integration, ETL/ELT, and data pipeline architecture.
  • Ability to troubleshoot and optimize data pipelines and processing workloads.
  • Strong understanding of data quality and validation practices.
  • Strong analytical and problem‑solving skills.
  • Ability to work independently while collaborating effectively with cross‑functional technical teams.
Nice-to-haves (optional)
  • Hands‑onMicrosoft Fabric experience – strongly preferred.
  • Experience building data pipelines or engineering solutions using Microsoft Fabric.
  • Pharmaceutical industry experience.
  • Life Sciences industry experience.
  • Insurance industry experience.Experience with enterprise‑scale cloud data platforms and distributed data processing.
  • Experience supporting data solutions used for analytics, BI, reporting, or machine learning.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer (P-175)
Data Engineer (P-175)

Smash CR • Dallas (TX)

On-site
USD 110,000 - 140,000
Data Analyst (P-174)
Data Analyst (P-174)

SMASH • Town of Texas (WI)

On-site
USD 90,000 - 140,000
Data Analyst (P-174)
Data Analyst (P-174)

Smash CR • Dallas (TX)

On-site
USD 90,000 - 130,000
Data Analyst (P-174)
Data Analyst (P-174)

smashcr • Dallas (TX)

On-site
USD 90,000 - 140,000
Senior Data Engineer
Senior Data Engineer

Peyton Resource Group • Irving (TX)

On-site
USD 100,000 - 130,000
Senior Data Engineer: Azure Data Pipelines & Spark
Senior Data Engineer: Azure Data Pipelines & Spark

smashcr • Town of Texas (WI)

On-site
USD 110,000 - 140,000
Senior Data Engineer
Senior Data Engineer

DataJobs • Plano (TX)

On-site
USD 140,000 - 180,000
Employee Discount
401(k) Matching*
Growth Opportunities
+6
Data Engineer
Data Engineer

Prodigy Resources • Denver (CO)

On-site
USD 110,000 - 170,000
Senior Data Engineer
Senior Data Engineer

Foundation-Partners-Group • Orlando (FL)

On-site
USD 120,000 - 160,000
Data Engineer
Data Engineer

VTG Defense • McLean (VA)

On-site
USD 100,000 - 130,000