Data Engineer (2258)

Stone Alliance Group

New York (NY)

Hybrid

USD 119,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Stone Alliance Group is seeking an experienced Data Engineer to join its Center for Data Analytics, Innovation, and Rigor in New York City. You will build scalable data infrastructure and pipelines to support AI evaluation frameworks, with exposure to clinical and NLP data.

The role requires 4+ days/week in the office, a strong background in Python, SQL, and cloud platforms, and a passion for rigorous data governance and scalable analytics in a research environment.

Qualifications

  • Master’s degree in Neuroscience, Psychology, Engineering, Computer Science or equivalent combination of education and experience is required.
  • 5+ years of experience in data analysis and data science fundamentals (e.g., algorithms, data structures, data visualization, machine learning), preferably in a clinical or research setting.
  • 5+ years of experience in at least one scientific programming language (e.g., Python/R, Matlab) and related toolboxes or frameworks (e.g., Tidyverse, Scipy, Sklearn, Polars, Pytorch) is required.
  • 5+ years of experience working in a Linux environment, using version control systems (e.g., GitHub), and software virtualization platforms (e.g., Docker).
  • 5+ years of practical experience in Extract, Transform, Load (ETL) processes and database management languages (SQL, NoSQL), and familiarity with associated cloud computing services and frameworks (AWS, Azure, Terraform).

Responsibilities

  • Create and maintain scalable data pipelines for multimodal data, including clinical, NLP, and multi-turn data.
  • Develop pipelines for data transformation, preprocessing, and management with quality and privacy compliance.
  • Perform quality assurance to maintain data integrity across the lifecycle.
  • Create interactive visualizations and dashboards to communicate data insights and pipeline performance.
  • Write documentation for scientific, clinical, or public dissemination of knowledge.
  • Perform additional job-related duties as assigned.

Skills

Data analysis
Data science fundamentals
Python
R
Matlab
ETL
Linux
GitHub
Docker
Kubernetes
SQL
NoSQL
AWS
Azure
Terraform
Data pipelines

Education

Master’s degree

Tools

Docker
Kubernetes
GitHub
Terraform

Job description

Our client is seeking an experienced Data Engineer to join their Center for Data Analytics, Innovation, and Rigor team in New York City.

As part of the Center for Data Analytics, Innovation, and Rigor team, you will report to the Rubric Engineering and Measurement Specialist. You will develop infrastructure to support large‑scale AI evaluation frameworks, design scalable data pipelines for generating and processing synthetic data, implement secure data storage solutions, and create infrastructure for real‑time model evaluation and monitoring. You will use common frameworks, platforms, and languages, such as Python, SQL, GitHub, containerization tools (e.g., Docker, Kubernetes), and cloud computing infrastructures (e.g., AWS, Azure) to build robust and scalable data infrastructure that supports our AI research initiatives.

This is an exempt, full‑time, hybrid position located in our NYC headquarters office or other relevant location. This position requires a minimum of four (4) days per week in the office, on a schedule determined by your supervisor. The in‑office requirement and schedule are subject to change based on the needs of the program and the organization.

Responsibilities
  • Create and maintain scalable data pipelines for efficient storage and retrieval of multimodal data, with particular emphasis on clinical, natural language, and multi‑turn response data.
  • Create pipelines for data transformation, preprocessing, and management. Ensure data quality, security, and compliance with privacy regulations for handling sensitive data.
  • Perform quality assurance of pipelines/processes to maintain integrity throughout the data lifecycle.
  • Create interactive visualizations and dashboards to communicate data insights and pipeline performance metrics.
  • Write documentation and relevant text for scientific, clinical, or public dissemination of knowledge.
  • Perform additional job‑related duties as assigned.
Qualifications
  • Master’s degree in Neuroscience, Psychology, Engineering, Computer Science or equivalent combination of education and experience is required.
  • 5+ years of experience in data analysis and data science fundamentals (e.g., algorithms, data structures, data visualization, machine learning), preferably in a clinical or research setting.
  • 5+ years of experience in at least one scientific programming language (e.g., Python/R, Matlab) and related toolboxes or frameworks (e.g., Tidyverse, Scipy, Sklearn, Polars, Pytorch) is required.
  • 5+ years of experience working in a Linux environment, using version control systems (e.g., GitHub), and software virtualization platforms (e.g., Docker).
  • 5+ years of practical experience in Extract, Transform, Load (ETL) processes and database management languages (SQL, NoSQL), and familiarity with associated cloud computing services and frameworks (AWS, Azure, Terraform).
Special Considerations

The anticipated salary range for this position is $119,000 - $150,000 USD annually.

Benefits

Our client’s competitive compensation and benefits include medical insurance, 401(k), paid parental leave, dependent care, flexible work schedules, discounted tickets and entertainment perks programs.

The salary range for the position is posted. Factors such as candidate’s work experience, education/training, job‑related skills, internal peer equity, as well as market and business considerations affect the salary offered within this range. In addition, this salary may be subject to a geographic adjustment (according to a specific city and state and depending on the role), if an authorization is granted to work outside of the location listed in this posting.

E‑E‑O Statement

Our client is an equal opportunity employer and does not discriminate in employment based on race, religion (including religious dress and grooming practices), color, sex/gender (including pregnancy, childbirth, breastfeeding or related medical conditions), sex stereotype, gender identity/gender expression/transgender (including whether or not you are transitioning or have transitioned) and sexual orientation; national origin (including language use restrictions and possession of a driver's license issued to persons unable to prove their presence in the United States is authorized under federal law [Vehicle Code section 12801.9]); ancestry, physical or mental disability, medical condition, genetic information/characteristics, marital status/registered domestic partner status, age (40 and over), sexual orientation, military or veteran status, or any other basis protected by federal, state or local law or ordinance or regulation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Stone Alliance Group Career Page • New York (NY)

Hybrid
USD 119,000 - 150,000
Medical insurance
401(k)
Paid parental leave
+3
Data Engineer
Data Engineer

PVH (Tommy Hilfiger/Calvin Klein) • New York (NY)

On-site
USD 80,000 - 120,000
Medical insurance
401(k) plan
Paid parental leave
+1
Data Science & Engineering Manager
Data Science & Engineering Manager

Stone Alliance Group Career Page • New York (NY)

Hybrid
USD 185,000 - 225,000
Medical insurance
401(k)
Paid parental leave
+3
Applied Data Science Specialist (2257)
Applied Data Science Specialist (2257)

Stone Alliance Group • New York (NY)

Hybrid
USD 96,000 - 120,000
Medical insurance
401(k)
Paid parental leave
+3
Data Science & Engineering Manager (2262)
Data Science & Engineering Manager (2262)

Stone Alliance Group • New York (NY)

Hybrid
USD 185,000 - 225,000
Medical insurance
401(k)
Paid parental leave
+4
Senior Data Engineer
Senior Data Engineer

Khealthcareers • New York (NY)

Hybrid
USD 150,000 - 200,000
Hybrid work schedule
18 vacation days
Stock options
+4
Post-Doctoral Fellow - Data Analytics, Innovation, and Rigor
Post-Doctoral Fellow - Data Analytics, Innovation, and Rigor

Stone Alliance Group Career Page • New York (NY)

Hybrid
USD 60,000 - 79,000
Medical insurance
401(k)
Paid parental leave
Lead Data Science Engineer
Lead Data Science Engineer

PowerToFly • New York (NY)

Hybrid
USD 135,000 - 180,000
Medical insurance
401(k) matching
Flexible paid time off
+1
Data Engineer
Data Engineer

Real Chemistry • United States

On-site
USD 90,000 - 120,000
Comprehensive medical, dental, and vision plans
Paid time off
Mental wellness support
+1
Post-Doctoral Fellow - Strategic Data Initiatives (2276)
Post-Doctoral Fellow - Strategic Data Initiatives (2276)

Stone Alliance Group • New York (NY)

Hybrid
USD 60,000 - 79,000
Medical insurance
401(k)
Paid parental leave
+3