Senior Software Engineer, Data Platform

bareinsights

Los Angeles (CA)

On-site

USD 130,000 - 200,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Bareinsights is seeking a senior software engineer to own the core data platform that ingests, enriches, and delivers large-scale multimodal data for AI labs. You will design end-to-end video processing pipelines, build in-house infrastructure, and develop marketplace features for datasets, with a focus on reliability and performance.

You will work across backend, frontend, and data-heavy systems, using Python, Kafka, Kubernetes, and AWS.

Qualifications

  • 5+ years of production software engineering experience spanning backend, frontend, and data-heavy systems.
  • Hands-on experience building multimodal data ingestion and orchestration pipelines at scale.
  • Production video processing experience: FFmpeg, transcoding, HLS, and streaming formats.
  • Strong Python backend skills with FastAPI or Django, and PostgreSQL in production.
  • Production experience with Kafka and Kubernetes.
  • Microservices architecture, API design, and signal processing at scale.
  • Comfortable building custom in-house infrastructure rather than defaulting to managed services.
  • Familiarity with Next.js, AWS.
  • Background at an AI training data platform is a strong plus.
  • Combined ML/AI enrichment and full-stack engineering experience is a plus.
  • STEM degree (bachelor’s minimum; master’s preferred — non-CS disciplines such as mathematics, physics, or biology are a positive signal).

Responsibilities

  • Build high-throughput video processing pipelines covering transcoding, HLS packaging, depth-map rendering, frame-accurate synchronization, and GPU-accelerated media workflows.
  • Own ingestion and processing pipelines for video, 3D assets, depth maps, telemetry, annotations, metadata, and model-generated signals.
  • Design and build event-driven workloads on Kafka and Kubernetes spanning capture, processing, enrichment, QA, packaging, and delivery.
  • Build AI enrichment systems for scoring, metadata extraction, classification, quality review, redundancy detection, and dataset recommendations.
  • Develop internal tools for capture teams, reviewers, operators, and data managers to track, inspect, and assemble datasets.
  • Build customer-facing marketplace features enabling AI labs to search, preview, query, purchase, and receive datasets.
  • Design data models covering provenance, asset versioning, licensing state, QA state, delivery history, and customer entitlements.
  • Instrument monitoring and analytics to understand quality, throughput, bottlenecks, and dataset value.

Skills

Multimodal data ingestion
Backend & full-stack engineering
Python development
API design
Microservices architecture
Signal processing
In-house infrastructure
Data-heavy systems

Education

Bachelor’s degree in STEM
Master’s preferred (non-CS disciplines welcomed)

Tools

FFmpeg
Kafka
Kubernetes
Next.js
AWS
PostgreSQL
FastAPI
Django

Job description

About the Role

This is a senior engineering role on the core data platform at a seed-stage AI data infrastructure startup operating at the intersection of gaming and AI. You’ll own the systems that ingest, enrich, package, and deliver large-scale multimodal data to AI labs — work that sits at the heart of the product and directly shapes the company’s trajectory.

What You’ll Do
  • Build high-throughput video processing pipelines covering transcoding, HLS packaging, depth-map rendering, frame-accurate synchronization, and GPU-accelerated media workflows.

  • Own ingestion and processing pipelines for video, 3D assets, depth maps, telemetry, annotations, metadata, and model-generated signals.

  • Design and build event-driven workloads on Kafka and Kubernetes spanning capture, processing, enrichment, QA, packaging, and delivery.

  • Build AI enrichment systems for scoring, metadata extraction, classification, quality review, redundancy detection, and dataset recommendations.

  • Develop internal tools for capture teams, reviewers, operators, and data managers to track, inspect, and assemble datasets.

  • Build customer-facing marketplace features enabling AI labs to search, preview, query, purchase, and receive datasets.

  • Design data models covering provenance, asset versioning, licensing state, QA state, delivery history, and customer entitlements.

  • Instrument monitoring and analytics to understand quality, throughput, bottlenecks, and dataset value.

What We’re Looking For
  • 5+ years of production software engineering experience spanning backend, frontend, and data-heavy systems.

  • Hands-on experience building multimodal data ingestion and orchestration pipelines at scale — a must-have.

  • Production video processing experience: FFmpeg, transcoding, HLS, and streaming formats — a must-have.

  • Strong Python backend skills with FastAPI or Django, and PostgreSQL in production — a must-have.

  • Production experience with Kafka and Kubernetes.

  • Microservices architecture, API design, and signal processing at scale.

  • Comfortable building custom in-house infrastructure rather than defaulting to managed services.

  • Familiarity with the broader stack: Next.js, AWS.

  • Background at an AI training data platform is a strong plus.

  • Combined ML/AI enrichment and full-stack engineering experience is a plus.

  • STEM degree (bachelor’s minimum; master’s preferred — non-CS disciplines such as mathematics, physics, or biology are a positive signal).

Compensation & Benefits

Salary: $130,000 – $200,000 USD annually. Visa sponsorship is not available.

Location

On-site in Los Angeles, CA. Candidates based in New York City or San Francisco may also be considered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding DevOps / Systems / ML Ops Engineer - Physical AI Startup
Founding DevOps / Systems / ML Ops Engineer - Physical AI Startup

Skyrocket Ventures • San Francisco (CA)

Hybrid
USD 170,000 - 200,000
Lead Engineer, Data Platform
Lead Engineer, Data Platform

Build AI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive pay
Medical, dental, and vision packages
Housing subsidy for SF or Shenzhen
Senior AI/ML Engineer
Senior AI/ML Engineer

CB Smart Recruit • Los Angeles (CA)

On-site
USD 180,000 - 350,000
Competitive sign-on bonus
Comprehensive benefits package
Data Product Engineer
Data Product Engineer

David Joseph & Company • San Francisco (CA)

On-site
USD 230,000 - 280,000
Equity
Founding role
On-site work
Software Engineer, Data Infrastructure & Pipelining
Software Engineer, Data Infrastructure & Pipelining

Build AI • San Francisco (CA)

On-site
USD 170,000 - 290,000
Competitive pay
Medical, dental, and vision packages
Housing subsidy
+6
Senior Software Engineer, Data Platform
Senior Software Engineer, Data Platform

Talanto • San Mateo (CA)

On-site
USD 130,000 - 280,000
Healthcare
Vision
Dental
+8
Principal Software Engineer
Principal Software Engineer

bareinsights • Atlanta (GA)

Hybrid
USD 220,000 - 280,000
Technical Lead – Data Platform
Technical Lead – Data Platform

StratITech • San Francisco (CA)

On-site
USD 240,000 - 270,000
Equity
Benefits
Software Engineer, Data Infrastructure
Software Engineer, Data Infrastructure

OpenAI • United States

Hybrid
USD 185,000 - 385,000
Relocation assistance
Hybrid work model
Software Engineer - Platform/Applied AI (Fullstack)
Software Engineer - Platform/Applied AI (Fullstack)

AfterQuery • San Francisco (CA)

On-site
USD 180,000 - 220,000
Competitive salary
Equity in the company
Opportunity for significant growth