Stand out for this role — generate a tailored resume and cover letter in about a minute.
H1 is seeking a Data Engineer II to build and operate Python, PySpark, and SQL pipelines for clinical trials data on the H1DN team. You will collaborate with clinical SMEs and Customer Success Managers to translate domain knowledge into production-ready pipeline logic that meets enterprise SLAs.
You’ll shipping changes quickly, monitor pipelines, and contribute to observability while ensuring data quality and reliability across multi-source datasets, APIs, and file formats.
At H1, we believe access to the best healthcare information is a basic human right. Our mission is to provide a platform that can optimally inform every doctor interaction globally. This promotes health equity and builds needed trust in healthcare systems. To accomplish this, our teams harness the power of data and AI-technology to unlock groundbreaking medical insights and convert those insights into action that result in optimal patient outcomes and accelerates an equitable and inclusive drug development lifecycle. Visit h1.com to learn more about us.
As part of H1’s hiring process, all candidates are required to participate in an in‑person final interview. Depending on your location, this may require travel.
H1's Data Network (H1DN) team is the client‑data mastering network at the core of how H1's products get their data. We run production ingestion for major enterprise customers. Clinical trial data is one of our highest‑visibility streams: it feeds decisions about where trials run and who runs them, and the people who depend on it are as often clinical experts as they are engineers. SLAs and customer expectations drive how we work, and we're looking for engineers who are energized by that.
As a Data Engineer II on the H1DN team, you will build and operate the pipelines behind H1's clinical trials data. You'll work primarily in Python, PySpark, and SQL, and you'll work directly with clinical subject matter experts and Customer Success Managers to turn their domain knowledge into pipeline logic that holds up in production.
You are a data engineer with a strong Python foundation and real distributed‑processing experience. You're drawn to high‑impact teams where the work is tangible: pipelines running, enterprise customers getting their data on time, clinical data that people make real decisions from. You're comfortable in an environment where recurring production runs and customer SLAs shape day‑to‑day priorities, and where a customer request can reorder your week. You'd rather sit down with a domain expert and understand why the data looks the way it does than build to a spec handed to you secondhand.
You bring experience:
This role pays $110,000 to $135,000 per year, based on experience, in addition to stock options.
Anticipated role close date: 10/20/2026
H1 is proud to be an equal opportunity employer that celebrates diversity and is committed to creating an inclusive workplace with equal opportunity for all applicants and teammates. Our goal is to recruit the most talented people from a diverse candidate pool regardless of race, color, ancestry, national origin, religion, disability, sex (including pregnancy), age, gender, gender identity, sexual orientation, marital status, veteran status, or any other characteristic protected by law.
H1 is committed to working with and providing access and reasonable accommodation to applicants with mental and/or physical disabilities. If you require an accommodation, please reach out to your recruiter once you've begun the interview process. All requests for accommodations are treated discreetly and confidentially, as practical and permitted by law.