Senior Site Reliability Engineer, Robotics & Cloud Infrastructure

Bedrock Ocean Exploration

United States

On-site

USD 164,000 - 220,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Comprehensive benefits

Job summary

Bedrock Ocean Exploration is seeking a Site Reliability Engineer to own reliability from the AUV onboard compute to the customer platform. You will build automation, observability, and guardrails so continuous ocean campaigns run with minimal manual intervention.

You will drive cross-team reliability, operate in an on-call rotation spanning European and East Coast hours, and travel to field deployments as needed.

Qualifications

  • 5+ years in an SRE, DevOps, or infrastructure role with on-call ownership.
  • Experience building scalable incident management and operational processes.
  • Strong automation instincts with Python/Go and Bash.
  • Hands-on AWS across compute/storage/networking/IAM; Docker/Kubernetes.
  • Linux internals, networking, and observability tooling knowledge.

Responsibilities

  • Own reliability across vehicle, topside, cloud pipelines, and platform data delivery.
  • Extend infrastructure automation for provisioning, config, deployment, and self-recovery.
  • Improve observability with metrics, logging, tracing, and alerting for robotics and data teams.
  • Reduce on-call burden by removing single points of failure and automating manual steps.
  • Participate in 12-hour on-call rotations across European and East Coast hours; contribute to post-incident reviews.
  • Define reliability targets: availability, data yield, recovery time; partner with teams to meet them.
  • Manage AWS cloud infrastructure for data processing and platform workloads.
  • Improve fleet and vehicle-level config management and safe deployment rollback.

Skills

SRE ownership
Automation scripting
AWS experience
Docker/Kubernetes
Observability tooling
Linux fundamentals

Tools

Terraform
Prometheus/Grafana

Job description

About Bedrock Ocean Exploration

Bedrock Ocean builds and operates autonomous underwater vehicles (AUVs) that collect georeferenced ocean-floor data at commercial scale. We deliver bathymetric and imagery data products to customers through our own platform, and we’re scaling toward continuous, around-the-clock data collection campaigns spanning months at a time. Keeping vehicles in the water and data flowing reliably is a core engineering problem and this role owns the reliability of the systems on both ends of that pipeline.

About Bedrock Ocean Exploration

Headquartered in Richmond, California, Bedrock Ocean Exploration is building autonomous ocean intelligence that will enable the ocean economy to solve the world’s most pressing challenges in maritime security, infrastructure, energy, and climate. Our modular architecture, driven by Siren (autonomous underwater vehicles), Trident (command and control), and Mosaic (subsea data fusion), delivers entirely new intelligence capabilities for government and commercial partners. Missions mobilize from any vessel of opportunity in 24 to 72 hours, and our automated pipeline returns comprehensive insights in hours, not weeks like the incumbents, keeping crews safe on shore while cutting cost and time.

The Role

We’re looking for an SRE who is equally comfortable on the robotics side- compute on the vehicle, topside operator machines, field deployments- and the cloud side: data ingestion, processing pipelines, and our customer-facing platform. You’ll build the automation, observability, and operational guardrails that let a small team run continuous AUV operations without continuous heroics, turning manual recovery steps into self-healing systems and shrinking the set of failures that only one person knows how to fix.

Reports to: Head of Software.

East Coast location is required to support coverage across both European operations and the East Coast during 12-hour on-call shifts.

Travel to field deployments and Richmond HQ is expected (approximately 5–15%).

What You’ll Do
  • Own reliability across the full path from vehicle to customer: AUV onboard compute (Jetson-class modules, ROS 2), topside/operator systems, cloud data pipelines, and the platform that delivers data products.
  • Build and extend infrastructure automation- provisioning, configuration management, deployment, and self-recovery- so that routine field operations and pipeline runs require minimal manual intervention.
  • Design and improve observability: metrics, logging, tracing, and alerting that give both robotics and data teams early, actionable signal across vehicle fleets and cloud services.
  • Drive down on-call burden by identifying and eliminating single points of failure, writing runbooks, and automating the manual steps that currently require tribal knowledge.
  • Participate in a shared on-call rotation covering both robotics-side and cloud-side incidents in 12-hour shifts spanning European and East Coast business hours; lead and contribute to blameless post-incident reviews.
  • Define and track reliability targets, availability, data yield, recovery time, tied to continuous-operations goals, and partner with robotics and data teams to meet them.
  • Manage cloud infrastructure on AWS (compute, storage, networking, IaC, cost, and security posture) for data processing and platform workloads.
  • Improve fleet- and vehicle-level configuration management, deployment safety, and rollback so changes reach the field reliably and predictably.
What We’re Looking For
  • 5+ years in an SRE, DevOps, or infrastructure engineering role running production systems with real uptime and on‑call responsibilities, including senior‑level ownership of reliability outcomes.
  • Experience implementing a scalable incident management and operational excellence mechanism that treats operators as customers, building processes and tooling that serve the people running operations day to day, not just the engineering team.
  • Strong automation instincts: comfortable scripting and building tooling in Python and/or Go and Bash, and using infrastructure-as-code (Terraform or equivalent).
  • Hands‑on AWS experience across compute, storage, networking, and IAM, plus containerization and orchestration (Docker, Kubernetes or similar).
  • Working knowledge of Linux internals, networking, and observability tooling (Prometheus/Grafana or equivalents).
  • Comfort operating across environments that aren’t just cloud: embedded or edge compute, intermittent connectivity, and physical systems that fail in messy ways.
  • A reliability mindset: you instrument before you guess, you automate the second time you do something manually, and you write things down so the next person or the system can handle it without you.
  • Strong ownership and communication in a small, fast‑moving team.
Nice to Have
  • Experience with robotics or embedded systems: ROS / ROS 2, Jetson or similar edge compute, sensor integration.
  • Background supporting field operations, autonomous systems, or hardware‑in‑the‑loop environments.
  • Familiarity with data pipelines and geospatial or large‑binary data formats.
  • Experience standing up on‑call practices and incident response from an early stage.
  • Some connection to the ocean: professional, academic, or personal. You’re excited to be around people who dive, sail, build, and explore offshore.
  • Active U.S. Secret security clearance or above.
Why This Role Matters

Our biggest operational goal depends on systems that stay up and data that stays valid for long, continuous stretches with a small team and a limited rotation. The reliability and automation you build directly determines whether we can run continuous campaigns at scale. This is high‑leverage infrastructure work with a clear, measurable mission.

Not a Fit If…
  • You prefer environments where cloud and hardware never mix.
  • You’d rather build tickets than eliminate them.
  • You’re not comfortable with on‑call ownership on a small team.
  • You want to optimize existing systems, not build the reliability practice alongside the product.
Compensation

$164,000–$220,000 base salary annually, depending on location. The upper end of the range reflects compensation in the New York, NY metro. In addition, we offer comprehensive employee benefits and equity.

Work Authorization

Candidates must have legal authorization to work in the United States without visa sponsorship. Bedrock does not sponsor employment visas.

Due to the nature of our government and defense work, candidates must be eligible to obtain a U.S. Secret security clearance if requested. An active Secret or higher clearance is not required to apply, but candidates who hold one are strongly preferred.

Bedrock Ocean Exploration is an equal opportunity employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure

Bedrock Ocean Exploration • New York (NY)

On-site
USD 164,000 - 220,000
Comprehensive employee benefits
Equity program
Staff Robotics Engineer
Staff Robotics Engineer

Bedrock Ocean Exploration • Richmond (CA)

On-site
USD 190,000 - 235,000
Staff Robotics Engineer
Staff Robotics Engineer

Alumni Founders • Richmond (CA)

On-site
USD 190,000 - 235,000
Marine Robotics Assembly Technician
Marine Robotics Assembly Technician

Bedrock Ocean Exploration • Richmond (CA)

On-site
USD 65,000 - 95,000
Director of People Operations
Director of People Operations

Bedrock Ocean Exploration • Richmond (CA)

Hybrid
USD 180,000 - 280,000
Senior SRE — Robotics & Cloud Infra
Senior SRE — Robotics & Cloud Infra

Bedrock Ocean Exploration • United States

On-site
USD 164,000 - 220,000
Equity
Comprehensive benefits
Head of Survey Operations
Head of Survey Operations

Bedrock Ocean Exploration • Richmond (CA)

On-site
USD 81,000 - 122,000
Electronics Assembly Technician
Electronics Assembly Technician

Bedrock Ocean Exploration • Richmond (CA)

On-site
USD 95,000 - 110,000
Unmanned Surface Vessel (USV) Field Service Representative
Unmanned Surface Vessel (USV) Field Service Representative

Seasats • San Diego (CA)

On-site
USD 75,000 - 115,000
Competitive insurance
401k matching
Four free lunches per week
+3
Senior Software Engineer, Robotics
Senior Software Engineer, Robotics

Saildrone • California (MO)

Hybrid
USD 176,000 - 227,000
Generous Time Off
Comprehensive Health Coverage
Retirement Savings
+2