Fractional Data/Full Stack Engineer, Data Storage & Ingestion Consultant

Gofractional

San Francisco (CA)

On-site

USD 137,760 - 413,280

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

A leading microscopy data solutions company based in San Francisco seeks an experienced expert to design and implement an end-to-end data pipeline. You will manage high-rate data ingest and multi-petabyte storage, ensuring uptime and optimizing performance. Candidates should have extensive experience in designing high-throughput storage systems and be present on-site during the build-out phase. This contract role offers competitive compensation ranging from $100 to $300 per hour.

Qualifications

  • 5+ years designing high-throughput storage systems.
  • Experience with NVMe RAID and HPC systems.
  • Proven track record in multi-petabyte storage deployment.

Responsibilities

  • Architect ingest and storage systems for data pipelines.
  • Optimize cost and performance for large data storage.
  • Maintain uptime for data-handling pipelines.

Skills

High-throughput storage design
HPC pipelines
NVMe RAID/striping
Linux performance tuning
Cloud storage systems (AWS S3)
Networking (25/40/100 GbE)

Job description

Role

We’re a San Francisco team collecting very large microscopy datasets and we need an expert to design and implement our end-to-end data pipeline, from high-rate ingest to multi-petabyte storage and downstream processing. You’ll own the strategy (on-prem vs. S3 or hybrid), the bill of materials, and the deployment, and you’ll be on the floor wiring, racking, tuning, and validating performance.

Our current instruments generate data at ~1+ GB/s sustained (higher during bursts) and the program will accumulate multiple petabyes total over time. You’ll help us choose and implement the right architecture considering reliability and cost controls.

Outcomes (what success looks like)
  • Within 2 weeks: Implement an immediate data-handling strategy that reliably ingests our initial data streams.

  • Within 2 weeks: Deliver a documented medium-term data architecture covering storage, networking, ingest, and durability.

  • Within 1 month: Operationalize the medium-term pipeline in production (ingest → buffer → long-term store → compute access).

  • Ongoing: Maintain ≥95% uptime for the end-to-end data-handling pipeline after setup.

Responsibilities
  • Architect ingest & storage: Choose and implement an on-prem hardware and data pipeline design or a cloud/S3 alternative with explicit cost and performance tradeoffs at multi-petabyte scale.

  • Set up a sustained-write ingest path ≥1 GB/s with adequate burst headroom (camera/frame-to-disk), including networking considerations, cooling, and throttling safeguards.

  • Optimize footprint & cost: Incorporate on-the-fly compression/downsampling options and quantify CPU budget vs. write-speed tradeoffs; document when/where to compress to control $/PB.

  • Integrate with acquisition workflows ensuring image data and metadata are compatible with downstream stitching/flat-field correction pipelines.

  • Enable downstream compute: Expose the data to segmentation/analysis stacks (local GPU nodes or cloud).

Skills
  • 5+ years designing and deploying high-throughput storage or HPC pipelines (≥1 GB/s sustained ingest) in production.

  • Deep hands-on with: NVMe RAID/striping, ZFS/MDRAID/erasure coding, PCIe topology, NUMA pinning, Linux performance tuning, and NIC offload features.

  • Proven delivery of multi-GB/s ingest systems and petabyte-scale storage in production (life-sciences, vision, HPC, or media).

  • Experience building tiered storage systems (NVMe → HDD/object) and validating real-world throughput under sustained load.

  • Practical S3/object-storage know-how (AWS S3 and/or on-prem S3-compatible systems) with lifecycle, versioning, and cost controls.

  • Data integrity & reliability: snapshots, scrubs, replication, erasure coding, and backup/DR for PB-scale systems.

  • Networking: ****25/40/100 GbE (SFP+/SFP28), RDMA/ RoCE/iWARP familiarity; switch config and path tuning.

  • Ability to spec and rack hardware: selecting chassis/backplanes, RAID/HBA cards, NICs, and cooling strategies to prevent NVMe throttling under sustained writes.

Ideal skills
  • Experience with microscopy or scientific imaging ingest at frame-to-disk speeds, including Micro-Manager-based pipelines and raw-to-containerized format conversions.

  • Experience with life science imaging data a plus.

Engagement details
  • Contract (1099 or corp-to-corp); contract-to-hire if there’s a mutual fit.

  • On-site requirement: You must be physically present in San Francisco during build-out and initial operations; local field work (e.g., UCSF) as needed.

  • Compensation: Contract, $100-300/hour

  • Timeline: Immediate start

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data/Full Stack Engineer, Data Storage & Ingestion Consultant
Data/Full Stack Engineer, Data Storage & Ingestion Consultant

Eon Systems PBC • San Francisco (CA)

On-site
USD 137,760 - 413,280
Petabyte-Scale Data Pipeline Architect — Contract (SF On-site)
Petabyte-Scale Data Pipeline Architect — Contract (SF On-site)

Gofractional • San Francisco (CA)

On-site
USD 137,760 - 413,280
Microscopy Technician
Microscopy Technician

Eonsystems • San Francisco (CA), Northern (KY)

On-site
USD 90,000 - 125,000
Microscopy Technician
Microscopy Technician

Eon Systems PBC • San Francisco (CA)

On-site
USD 90,000 - 125,000
Equity
Benefits package
Senior HPC Storage Engineer
Senior HPC Storage Engineer

University of California Los Angeles • Los Angeles (CA)

On-site
USD 150,000 - 210,000
Senior HPC Storage Engineer
Senior HPC Storage Engineer

The Regents of the University of California on behalf of their Los Angeles Campus • Los Angeles (CA)

On-site
USD 150,000 - 230,000
Staff HPC Software Engineer
Staff HPC Software Engineer

San Diego Stealth Startup • San Diego (CA)

On-site
USD 140,000 - 210,000
Data Scientist, Imaging x AI/ML
Data Scientist, Imaging x AI/ML

SupportFinity™ • Redwood City (CA)

On-site
USD 153,000 - 210,100
Generous employer match on 401(k) contributions
Paid time off for volunteering
Funding for family-forming benefits
+1
Member of Technical Staff Intern, Smart Microscopy
Member of Technical Staff Intern, Smart Microscopy

Cephla Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 34,440,000 - 48,216,000
Senior Microscope Engineer
Senior Microscope Engineer

Eon Systems PBC • San Francisco (CA)

On-site
USD 140,000 - 200,000
Equity and benefits
Cutting-edge field exposure
Collaborative team culture