Junior Data Engineer (m/f/d)

Hilo

Indiana (PA)

On-site

USD 99,132 - 136,307

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote-first
Hybrid setup
Competitive compensation

Job summary

Hilo by Aktiia is seeking a Junior Data Engineer to scale a Databricks-based medallion Lakehouse from R&D to a broader company-wide data platform. You’ll ingest, centralize, and document clinical and legacy data, turning messy sources into reliable, model-ready pipelines for Data Science and ML teams.

You’ll work with Python/SQL, Databricks, and AWS in a hybrid Swiss setup, collaborating with cross-functional stakeholders to ensure data quality and observability.

Qualifications

  • Bachelor’s degree or higher in Computer Science, Data Engineering, Data Science, Software Engineering, or a related field.
  • At least 1+ year of practical post‑study experience in data engineering, data infrastructure, data ingestion, or similar role.
  • Solid hands‑on experience with Python and SQL and practical experience with Databricks; familiarity with AWS; basics of ETL/ELT pipeline design; data modelling; and Git/CI/CD practices.

Responsibilities

  • Design, build, and maintain data ingestion and processing pipelines in Databricks medallion Lakehouse.
  • Ingest, centralize, document clinical data from various sources (Castor, RedCap, databases).
  • Work with unfamiliar data, identify structures/inconsistencies, and produce reliable data assets.
  • Build preprocessing and ETL/ELT pipelines for model-ready datasets for ML and Data Science teams.
  • Define data quality, validation, traceability, and documentation standards.

Skills

Python
SQL
Databricks
Cloud (AWS)
ETL/ELT
Data modeling
Git
CI/CD
Spark

Education

Bachelor's degree

Tools

Spark
Delta Lake
Parquet
Castor
RedCap

Job description

About Hilo by Aktiia

High blood pressure is the world’s most common disease, causing 18 million deaths each year. At Hilo by Aktiia, our vision is a world where no lives are lost or damaged from the effects of high blood pressure, and our mission is to build the technology that helps people control it.

We are a venture-backed scale-up that has raised over $96M. Our technology, rooted in 18 years of research at the Swiss Center for Electronics and Microtechnology (CSEM), is the world’s only medically accurate, cuffless, continuous blood pressure monitor for daily life. It is CE Marked as a Class IIa medical device and was FDA-cleared in 2026 as the first cuffless OTC blood pressure monitor in the United States. We are remote-first, headquartered in Neuchâtel, Switzerland.

Role Overview

Imagine using your data engineering skills to improve the way cardiovascular health is monitored and managed. Hilo by Aktiia is redefining blood pressure monitoring through AI-driven optical technology built on more than 20 years of research at the Swiss Center for Electronics and Microtechnology (CSEM). Their solution combines a wearable device (Hilo), a mobile app, and a cloud-based platform for healthcare professionals — empowering users and physicians with continuous, actionable insights into blood‑pressure patterns. With more than 200,000 users, over $120M in funding, and a CE‑certified medical device already available across several markets, Aktiia is continuing to scale its technology, product ecosystem, and international presence.

As Aktiia expands its data capabilities, we’re looking for a Junior Data Engineer who will help scale a Databricks‑based medallion Lakehouse from a primarily R&D/Core Tech setup into a broader company‑wide data platform. Working closely with and learning directly from an experienced Senior Data Engineer, you will support the ingestion, centralization, documentation, and validation of clinical, legacy, and operational data. Your work will help transform complex and sometimes messy data sources into clean, reliable, and model‑ready pipelines used by Data Science, ML, Algorithm, and Core Tech teams.

Your Tasks

  • Data Ingestion & Lakehouse Development: Design, build, and maintain data ingestion and processing pipelines within Aktiia’s Databricks-hosted medallion Lakehouse, working under senior guidance while gradually taking on more ownership.
  • Clinical & Legacy Data Integration: Take an active role in ingesting, centralizing, and documenting clinical data currently spread across EDC platforms such as Castor and RedCap, databases, standalone archives, and other historically grown sources.
  • Data Exploration & Practical Data Archaeology: Work hands‑on with unfamiliar and sometimes messy datasets, identify structures and inconsistencies, and turn unstructured situations into reliable, usable data assets.
  • Pipeline Development & Preprocessing: Build preprocessing and ETL/ELT pipelines that provide clean, structured, and model‑ready datasets for Algorithm Development, Core Tech, Machine Learning, and Data Science teams.
  • Data Quality, Validation & Documentation: Define and apply practical standards for data quality, validation, traceability, and documentation — especially for sensitive and clinically relevant datasets.
  • Observability & Engineering Practices: Implement logging, validation checks, alerting, and basic observability for new pipelines, while contributing to shared codebase practices such as Git, code reviews, CI/CD, and testing.
  • Platform Scaling & Collaboration: Support the evolution of the lakehouse from selected technical use cases toward a company‑wide data infrastructure, working closely with the Senior Data Engineer, ML Engineers, Data Scientists, and cross‑functional stakeholders.
  • Academic Background: Bachelor’s degree or higher in Computer Science, Data Engineering, Data Science, Software Engineering, or a related technical field.
  • Professional Expertise: At least 1+ year of practical post‑study experience in data engineering, data infrastructure, data ingestion, or a similar hands‑on technical role. You may still be early in your career, but you have already worked with real data pipelines, production‑oriented data workflows, or collaborative data platforms.
  • Technical Experience: You have solid hands‑on experience with Python and SQL and bring practical experience with Databricks. You are familiar with cloud environments — ideally AWS. You understand the basics of ETL/ELT pipeline design, data ingestion patterns, and data modelling, and you have worked with shared codebases using Git, code reviews, CI/CD, or testing practices. Experience with Spark, Delta Lake, or Parquet is a strong plus.
  • Industry Fit: Ideally, you have gained experience in a start‑up, scale‑up, or technically demanding environment where you worked hands‑on across the data pipeline. Exposure to regulated or data‑sensitive industries such as MedTech, Pharma, FinTech, or healthcare is a plus, especially when it comes to data quality, validation, and documentation.
  • Language Skills: English at a highly proficient level is a must. French is advantageous.
  • Personality: A curious, pragmatic, and hands‑on problem‑solver, you enjoy bringing structure into complex data environments. You work independently, ask the right questions, stay well organized, and collaborate openly with Data Engineering, Data Science, ML, Algorithm, Core Tech, and Product stakeholders. You are based in Switzerland and are comfortable working in a hybrid setup with one day per week regular presence in Neuchâtel.

We’re looking for a pragmatic, hands‑on data builder who enjoys working with real‑world clinical and company data — someone who is excited to bring structure into scattered datasets, build reliable pipelines, and help create a modern data foundation for advanced ML and digital health applications.

Do you want to grow your expertise in Data Engineering and cloud‑based lakehouse architectures while contributing to health technology that has a direct impact on hundreds of thousands of people? Then we’re excited to meet you!

Why Join Us?
  • Hilo by Aktiia is becoming one of the most important med‑tech companies, and you can be part of this exciting story.
  • Work on a mission that matters: transforming cardiovascular health at global scale.
  • Be part of a high‑impact, fast‑growing, and venture‑backed scale‑up company.
  • Collaborate with a diverse, passionate, and talented team.
  • Competitive compensation and other benefits (depending on location).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist
Data Scientist

Hilobyaktiia • Germany (OH)

Hybrid
USD 90,000 - 140,000
Staff Data Engineer- Data Lake
Staff Data Engineer- Data Lake

h1 • New York (NY)

Hybrid
USD 170,000 - 190,000
Health insurance options
Generous paid time off
Flexible work hours
Staff Data Engineer
Staff Data Engineer

Iterativehealth • Town of Cambridge (NY)

On-site
USD 200,000 - 325,000
US Country Manager
US Country Manager

Hilo • United States

Hybrid
USD 165,000 - 180,000
Health, dental, vision coverage
ESOP package
Paid time off
Data Scientist Lead
Data Scientist Lead

Kontakt Micro-Location Sp. Z.o.o. • New York (NY)

Hybrid
USD 180,000 - 240,000
Health benefits
401(k)
Paid time off
+2
Staff Data Engineer
Staff Data Engineer

Iterative Health • New York (NY), Cambridge (MA)

On-site
USD 200,000 - 325,000
Senior Data Engineer
Senior Data Engineer

MediData • Woodbridge Township (NJ)

Hybrid
USD 96,000 - 128,000
Senior Data Engineer, Medical Devices / IoMT (m/f/d)
Senior Data Engineer, Medical Devices / IoMT (m/f/d)

United States Digital Space LLC • United States

Remote
USD 69,000 - 92,000
Competitive salary & benefits package
Flexible remote work arrangements
Data Engineer
Data Engineer

Circadia Health • Los Angeles (CA)

On-site
USD 150,000 - 220,000
100% company-paid medical, dental, vision
401(k) with match
Generous PTO
+2
Head of Growth US
Head of Growth US

Hilo • Northern (KY)

Hybrid
USD 170,000 - 190,000
ESOP program
Health, dental, vision, and paid time