Fractional Data/Full Stack Engineer, Data Storage & Ingestion Consultant

Gofractional

San Francisco (CA)

On-site

USD 137,760 - 413,280

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading microscopy data solutions company based in San Francisco seeks an experienced expert to design and implement an end-to-end data pipeline. You will manage high-rate data ingest and multi-petabyte storage, ensuring uptime and optimizing performance. Candidates should have extensive experience in designing high-throughput storage systems and be present on-site during the build-out phase. This contract role offers competitive compensation ranging from $100 to $300 per hour.

Qualifications

  • 5+ years designing high-throughput storage systems.
  • Experience with NVMe RAID and HPC systems.
  • Proven track record in multi-petabyte storage deployment.

Responsibilities

  • Architect ingest and storage systems for data pipelines.
  • Optimize cost and performance for large data storage.
  • Maintain uptime for data-handling pipelines.

Skills

High-throughput storage design
HPC pipelines
NVMe RAID/striping
Linux performance tuning
Cloud storage systems (AWS S3)
Networking (25/40/100 GbE)

Job description

Role

We’re a San Francisco team collecting very large microscopy datasets and we need an expert to design and implement our end-to-end data pipeline, from high-rate ingest to multi-petabyte storage and downstream processing. You’ll own the strategy (on-prem vs. S3 or hybrid), the bill of materials, and the deployment, and you’ll be on the floor wiring, racking, tuning, and validating performance.

Our current instruments generate data at ~1+ GB/s sustained (higher during bursts) and the program will accumulate multiple petabyes total over time. You’ll help us choose and implement the right architecture considering reliability and cost controls.

Outcomes (what success looks like)
  • Within 2 weeks: Implement an immediate data-handling strategy that reliably ingests our initial data streams.

  • Within 2 weeks: Deliver a documented medium-term data architecture covering storage, networking, ingest, and durability.

  • Within 1 month: Operationalize the medium-term pipeline in production (ingest → buffer → long-term store → compute access).

  • Ongoing: Maintain ≥95% uptime for the end-to-end data-handling pipeline after setup.

Responsibilities
  • Architect ingest & storage: Choose and implement an on-prem hardware and data pipeline design or a cloud/S3 alternative with explicit cost and performance tradeoffs at multi-petabyte scale.

  • Set up a sustained-write ingest path ≥1 GB/s with adequate burst headroom (camera/frame-to-disk), including networking considerations, cooling, and throttling safeguards.

  • Optimize footprint & cost: Incorporate on-the-fly compression/downsampling options and quantify CPU budget vs. write-speed tradeoffs; document when/where to compress to control $/PB.

  • Integrate with acquisition workflows ensuring image data and metadata are compatible with downstream stitching/flat-field correction pipelines.

  • Enable downstream compute: Expose the data to segmentation/analysis stacks (local GPU nodes or cloud).

Skills
  • 5+ years designing and deploying high-throughput storage or HPC pipelines (≥1 GB/s sustained ingest) in production.

  • Deep hands-on with: NVMe RAID/striping, ZFS/MDRAID/erasure coding, PCIe topology, NUMA pinning, Linux performance tuning, and NIC offload features.

  • Proven delivery of multi-GB/s ingest systems and petabyte-scale storage in production (life-sciences, vision, HPC, or media).

  • Experience building tiered storage systems (NVMe → HDD/object) and validating real-world throughput under sustained load.

  • Practical S3/object-storage know-how (AWS S3 and/or on-prem S3-compatible systems) with lifecycle, versioning, and cost controls.

  • Data integrity & reliability: snapshots, scrubs, replication, erasure coding, and backup/DR for PB-scale systems.

  • Networking: ****25/40/100 GbE (SFP+/SFP28), RDMA/ RoCE/iWARP familiarity; switch config and path tuning.

  • Ability to spec and rack hardware: selecting chassis/backplanes, RAID/HBA cards, NICs, and cooling strategies to prevent NVMe throttling under sustained writes.

Ideal skills
  • Experience with microscopy or scientific imaging ingest at frame-to-disk speeds, including Micro-Manager-based pipelines and raw-to-containerized format conversions.

  • Experience with life science imaging data a plus.

Engagement details
  • Contract (1099 or corp-to-corp); contract-to-hire if there’s a mutual fit.

  • On-site requirement: You must be physically present in San Francisco during build-out and initial operations; local field work (e.g., UCSF) as needed.

  • Compensation: Contract, $100-300/hour

  • Timeline: Immediate start

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data/Full Stack Engineer, Data Storage & Ingestion Consultant
Data/Full Stack Engineer, Data Storage & Ingestion Consultant

Eon Systems PBC • San Francisco (CA)

On-site
Petabyte-Scale Data Pipeline Architect — Contract (SF On-site)
Petabyte-Scale Data Pipeline Architect — Contract (SF On-site)

Gofractional • San Francisco (CA)

On-site
Founding Data Infrastructure Engineer
Founding Data Infrastructure Engineer

Breakout Ventures • San Francisco (CA)

On-site
USD 160,000 - 230,000
Comprehensive health, dental, and vision insurance
Relocation assistance
Paid time off (PTO)
+2
Senior Full Stack Engineer — Scientific Data | San Mateo, CA (On-site)
Senior Full Stack Engineer — Scientific Data | San Mateo, CA (On-site)

zella-group • San Mateo (CA)

On-site
USD 180,000 - 220,000
Principal Systems Software Engineer
Principal Systems Software Engineer

San Diego Stealth Startup • San Diego (CA)

On-site
USD 258,000 - 275,000
Microscopy Technician
Microscopy Technician

Eon Systems PBC • San Francisco (CA)

On-site
USD 90,000 - 125,000
Equity
Benefits package
Data Infrastructure Engineer
Data Infrastructure Engineer

Glyphic Biotechnologies • Berkeley (CA)

Hybrid
USD 135,000 - 179,000
Employee Stock Option Plan
Health Plan Coverage for Employees & Dependents
Employer Retirement Contributions to 401(k)
+5
Microscopy Technician
Microscopy Technician

Eon Systems • San Francisco (CA)

On-site
USD 90,000 - 125,000
Equity
Benefits package
Senior HPC Storage Engineer
Senior HPC Storage Engineer

The Regents of the University of California on behalf of their Los Angeles Campus • Los Angeles (CA)

Hybrid
USD 150,000 - 230,000
Senior HPC Storage Engineer
Senior HPC Storage Engineer

University of California Los Angeles • Los Angeles (CA)

Hybrid
USD 150,000 - 210,000