Opportunity Overview
We are looking for a Data Engineer to build and optimize scalable data pipelines and analytical datasets on our AWS-based lakehouse platform.
- You will work with Apache Iceberg, AWS Glue, Athena, dbt, and Airflow to support reliable and efficient data processing for analytics and operational use cases.
- You will collaborate with senior engineers, analytics teams, and platform stakeholders to deliver production-grade data solutions.
What you’ll do
- Develop and maintain batch data pipelines
- Build reusable ingestion and transformation workflows
- Build and maintain Apache Iceberg datasets
- Support schema evolution and partition optimization
- Optimize Athena queries and troubleshoot performance bottlenecks
- Implement data validation checks and monitoring
- Support production issue troubleshooting
Required Qualifications
- 4–5 years of experience in data engineering
- Strong hands‑on experience with Python, SQL, PySpark/Spark
- Experience with AWS S3, Glue, Athena
- Experience with DBT and Airflow (or equivalent)
- Familiarity with Apache Iceberg / Hudi / Delta
- Understanding of partitioning and Parquet optimization
Equal Opportunity Statement
Cohere Health is an Equal Opportunity Employer. We are committed to fostering an environment of mutual respect where equal employment opportunities are available to all. To us, it’s personal.