Neural Data Infrastructure Engineer

Blackrock Neurotech

Salt Lake City (UT)

On-site

USD 120,000 - 180,000

Full time

20 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Blackrock Neurotech is seeking a Neural Data Infrastructure Engineer in Salt Lake City, UT to own the data pipeline from upload to training data loaders. You will build scalable ingestion for intracortical recordings, preserving raw data and metadata and aligning with model needs.

You will design storage, automate quality checks, and collaborate with researchers, infrastructure, and IT to support large neural datasets across local and cloud resources.

Qualifications

  • Experience designing end-to-end large-scale scientific data pipelines.
  • Bachelor’s degree in CS/EE/BME or equivalent practical experience.
  • Strong Python programming and software engineering practices.
  • Advanced understanding of extracellular neural recordings and preprocessing.
  • Experience with NEV/NSx, NWB, HDF5, and Zarr data formats.
  • Ability to bridge neuroscience, ML, and infrastructure teams.

Responsibilities

  • Own the design, implementation, and operation of the neural data pipeline from upload through training data loaders.
  • Build reliable ingestion for heterogeneous recording formats, preserving raw data, metadata, and mappings.
  • Implement and validate neural signal processing for spike events and waveforms; extend to field potentials.
  • Establish automated checks for corrupt files, missing channels, clock drift, and artifacts.
  • Design storage layouts, indexing, and caching for efficient access to large datasets across local and cloud resources.
  • Create reproducible, versioned datasets with provenance and partitioning to prevent data leakage.
  • Partner with researchers to define input representations and interfaces that meet model requirements.
  • Build parallel preprocessing and streaming loaders, ensuring GPUs are supplied as training scales.
  • Plan capacity, access controls, retention, backup, and recovery with data owners.

Skills

Python
Large-scale data pipelines
Data quality & reliability
Neural data / intracortical recordings
Version control
Debugging & profiling
Distributed processing
PyTorch familiarity

Education

Bachelor’s in CS / EE / BME
Equivalent practical experience

Tools

NEV/NSx
NWB
HDF5
Zarr
PyTorch

Job description

Build the systems that expand human capability

At Blackrock Neurotech, we’ve spent decades making the impossible possible – helping people move, speak, and reconnect with the world when they otherwise could not. We’ve seen that restoring function restores more than ability. It restores independence, identity, and agency.

Today, we are building the next generation of human capability: brain-computer interfaces that are designed to be safe, scalable, and trusted in the real world. Our work is not only about reconnecting people to what was lost, but about expanding what is possible – creating a seamless interface between human intent and technology.

This is foundational work in a category-defining field. You will help build the infrastructure for a future where neural interfaces are invisible, reliable, and deeply human-centered.

Working At Blackrock Neurotech Means
  • Owning meaningful, high-impact problems at the frontier of science and engineering
  • Building alongside experienced, thoughtful peers across disciplines
  • Solving technically complex challenges grounded in real human outcomes
  • Contributing to a culture that values rigor, clarity, and long-term thinking over noise
The Role

The Neural Data Infrastructure Engineer will own the systems that turn intracortical recordings into reliable, training-ready data for AI/ML models. You will build the data infrastructure needed to meaningfully scale model capability while preserving the scientific meaning of every recording. As a hands-on individual contributor on a small research team, you will be the technical authority on the data lifecycle from upload through training consumption. You will work closely with model researchers to understand their data requirements and translate them into reliable, scalable infrastructure. You will also partner with infrastructure and IT teams to ensure the right storage, processing, and delivery resources are in place. You will take end-to-end technical ownership of a growing intracortical BCI data corpus that sits at the foundation of Blackrock's neural foundation model work. Your work will directly shape the quality, scale, and reliability of the data these models learn from, while giving researchers the confidence and freedom to focus on model development

What You'll Do
  • Own the design, implementation, and operation of the neural data pipeline from upload and validation through preprocessing, dataset releases, and training data loaders
  • Build reliable ingestion for heterogeneous recording formats, preserving raw data, acquisition metadata, channel and electrode mappings, timestamps, units, and behavioral alignment
  • Implement and validate neural signal processing for spike events and waveforms, including detection or sorting where needed, quality assessment, binning, and normalization; extend the pipeline to field potentials as the program evolves
  • Establish automated checks for corrupt files, missing channels, clock drift, artifacts, recording discontinuities, and inconsistent metadata, with clear criteria for quarantine and recovery
  • Design storage layouts, indexing, chunking, compression, and caching that support efficient access to large neural datasets across local and cloud resources
  • Create reproducible, versioned datasets with traceable transformations, provenance, and subject and session partitions that prevent leakage between training and evaluation
  • Partner with model researchers to define input representations, sequence construction, masks, sampling policies, and interfaces that preserve neural meaning and meet architecture requirements
  • Build and profile parallel preprocessing and streaming data loaders, partnering with the training performance engineer to keep GPUs supplied as training scales
  • Work with infrastructure, IT, and data owners to plan capacity, access controls, retention, backup, and recovery appropriate to the recordings and their approved uses
  • Establish testing, monitoring, documentation, and operational practices that make pipeline failures diagnosable and dataset releases dependable
  • Communicate data quality, coverage, processing tradeoffs, and resource needs clearly, and translate evolving research requirements into maintainable engineering plans
What You Bring
  • Demonstrated experience independently designing and operating large-scale scientific data pipelines from end to end, with ownership of software quality, reliability, and delivery
  • Bachelor’s degree in Computer Science, Electrical Engineering, Biomedical Engineering, Neuroscience, or a related technical field, or equivalent practical experience
  • Exceptional programming ability in Python, with strong software design, debugging, profiling, automated testing, and version control practices
  • Advanced understanding of extracellular neural recordings, spike detection and sorting, sampling theory, filtering, aliasing, referencing, and artifact handling
  • Hands-on experience with intracortical electrophysiology data and its on-disk representation, including binary layouts, event timestamps, channel metadata, and continuous signals
  • Ability to reason about how preprocessing choices alter neural information and affect model training, evaluation, and reproducibility
  • Experience with scientific formats and storage systems such as NEV/NSx, NWB, HDF5, and Zarr, including efficient partial reads and metadata management
  • Experience with distributed or parallel processing, object storage, workflow orchestration, and restartable jobs at scale preferred
  • Familiarity with deep learning models and architectures, including how data processing decisions affect model training, inference, and evaluation
  • Strong judgment about dataset lineage, validation, access permissions, and the handling of sensitive research data
  • Ability to work across neuroscience, machine learning, and infrastructure teams and bring clarity to ambiguous requirements
  • Experience processing local field potentials, synchronizing neural and behavioral streams, or supporting longitudinal neural datasets strongly preferred
  • Experience with compiled languages, high-throughput I/O, cloud cost management, or performance-sensitive scientific computing preferred
  • Fluency with deep learning data interfaces, including tensor shapes, batching, masking, sampling, and PyTorch preferred
Working Location

This is an on-site role based at Blackrock Neurotech's headquarters in Salt Lake City, Utah. Occasional travel may be required.

How We Work
  • We take ownership of outcomes and follow through with clarity and accountability
  • We prioritize sustained, high-quality work over performative urgency
  • We value rigor, sound judgement and thoughtful decision-making
  • We collaborate deliberately: low ego, high trust and high context

This is a high-ownership role, but it is not an "always-on" one. We expect strong work and our people to have a life outside of it.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Neural Data Infrastructure Engineer
Neural Data Infrastructure Engineer

Socket.dev • Salt Lake City (UT)

On-site
USD 130,000 - 170,000
AI/ML Training Performance Engineer
AI/ML Training Performance Engineer

Blackrock Neurotech • Salt Lake City (UT)

On-site
USD 120,000 - 190,000
Director, Software Engineering
Director, Software Engineering

Blackrock Neurotech • Salt Lake City (UT)

On-site
USD 180,000 - 240,000
Neural Data Pipeline Architect
Neural Data Pipeline Architect

Socket.dev • Salt Lake City (UT)

On-site
USD 130,000 - 170,000
Electrode Manufacturing Technician I
Electrode Manufacturing Technician I

Blackrock Neurotech • Salt Lake City (UT)

On-site
USD 42,000 - 65,000
Director, Clinical Affairs
Director, Clinical Affairs

Blackrock Neurotech • Salt Lake City (UT)

On-site
USD 190,000 - 290,000
Director, Clinical Affairs
Director, Clinical Affairs

Blackrock-Neurotech • Salt Lake City (UT)

Hybrid
USD 180,000 - 260,000
Human Neuroscientist, Intracranial Recording
Human Neuroscientist, Intracranial Recording

astera • Alameda (CA)

On-site
USD 180,000 - 260,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Echo Neurotechnologies • San Francisco (CA)

On-site
USD 170,000 - 220,000
Stock options
401(k) with matching
AI/ML Training Performance Engineer
AI/ML Training Performance Engineer

Socket.dev • Salt Lake City (UT)

On-site
USD 150,000 - 200,000