Fleet Response Engineer

MVP Ventures

Mountain View (CA)

On-site

USD 170,000 - 230,000

Full time

14 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Rhoda AI in Mountain View is seeking an experienced incident response engineer to be the first line of defense when deployed robots encounter faults. You will triage alerts across hardware, embedded software, sensors, and networking to restore operation quickly.

You will own incidents from alert to recovery, develop better monitoring, runbooks, and a reliable fleet, collaborating with AI, robot software, cloud, and hardware teams.

Qualifications

  • Strong systematic debugging from symptoms to root cause.
  • Proficient with Linux and command line; able to read logs and telemetry.
  • Proficient in Python and/or C/C++.
  • Excellent written and verbal communication; calm under pressure.
  • Willingness to participate in on-call rotations.
  • Preferred: 3+ years in robotics, automation, or related field.

Responsibilities

  • Triage and incident response across robot fleets; monitor faults, alarms and alerts.
  • Act as the first engineering responder for real-time escalations from field technicians and operations.
  • Debug complex issues across subsystems via log analysis, telemetry review and reproduction testing.
  • Drive incidents to fast resolution or mitigation to restore operation and maximize fleet uptime.
  • Coordinate with domain engineering teams (AI, robot software, cloud, hardware) to drive resolution.
  • Own each incident through to confirmed recovery and handoff; maintain clear communications.

Skills

Systematic debugging
Linux/CLI
Python
C/C++
On-call readiness

Tools

ROS/ROS2
Telemetry analysis

Job description

Location

Mountain View

Employment Type

Full time

Department

Applications Engineering- Applied AI

OverviewApplication

At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.

The Role

You’ll be the first engineer in the loop when something goes wrong on a deployed robot. When the field or operations team hits an issue they can’t resolve, they come to you. You’ll triage incoming faults and alerts, debug across the full stack: hardware, embedded, software, sensors, networking, to find the fastest safe path back to operation, and pull in the right domain experts when an issue runs deep. You’ll own incidents from first alert through recovery and root cause, and turn what you learn into better monitoring, better runbooks, and a more reliable fleet.

What You’ll Do
Triage and incident response
  • Monitor incoming faults, alarms, and alerts from robots deployed at customer sites and in internal testing; triage and prioritize by severity and operational impact
  • Act as the first engineering responder and point of contact for real-time escalations from field technicians and operations
  • Debug complex issues across subsystems through log analysis, telemetry review, and reproduction testing
  • Drive incidents to fast resolution or mitigation to restore operation and maximize fleet uptime
Escalation and coordination
  • Escalate to and coordinate with domain engineering teams (AI, robot software, cloud, hardware) to drive resolution
  • Own the on-call rotation and paging, and keep the escalation process clear and current
  • Communicate status to operations, engineering, and customer-facing teams throughout an incident
  • Own each incident through to confirmed recovery and a clean handoff
Root cause and reliability
  • Perform root-cause analysis — identify contributing factors, themes, and corrective actions — and track follow-ups to closure
  • Document investigations and fixes in a shared knowledge base and troubleshooting guide so future issues resolve faster
  • Feed field insights back to engineering to improve product reliability and issue detection
  • Track reliability trends and flag recurring or systemic failures
Tooling and prevention
  • Build (or spec, with the software team) the monitoring, alerting, and diagnostic tooling that catches issues fleet-wide
  • Establish diagnostic procedures that let technicians and operators self-serve common issues
  • Continuously reduce manual, repetitive response work through automation
Required Qualifications
  • Strong systematic debugging of complex systems — the ability to reason from symptoms to root cause and reverse-engineer unexpected behavior
  • Comfortable in Linux and the command line, and reading logs, telemetry, and system data
  • Coding proficiency in Python and/or C/C++ (Bash a plus)
  • Clear written and verbal communication, and calm judgment under pressure
  • Willingness to take part in an on-call rotation, including some nights and weekends as the fleet grows
Preferred Qualifications
  • 3+ years working with robotic, autonomous, automotive, aerospace, or industrial-automation systems in an engineering, reliability, or support capacity
  • Familiarity with ROS/ROS2 and robot subsystems (sensors, actuators, perception, networking)
  • Prior on-call, site-reliability, or release-engineering experience
  • Experience building tools that help others debug, and a track record of solving unusual bugs
  • Comfort with hardware and hardware–software interfaces; willingness to travel occasionally to deployment sites
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Fleet Response Engineer
Fleet Response Engineer

Rhoda AI • Mountain View (CA)

On-site
USD 150,000 - 190,000
Senior Fleet Software Engineer
Senior Fleet Software Engineer

MVP Ventures • Mountain View (CA)

On-site
USD 140,000 - 200,000
Fleet Engineer
Fleet Engineer

Diligent Robotics • Austin (TX)

On-site
USD 95,000 - 125,000
Robotics Fleet Incident Response Engineer
Robotics Fleet Incident Response Engineer

MVP Ventures • Mountain View (CA)

On-site
USD 170,000 - 230,000
Fleet Engineer
Fleet Engineer

Capital Factory • Austin (TX)

On-site
USD 80,000 - 120,000
Fleet Engineer Austin, Texas, United States
Fleet Engineer Austin, Texas, United States

Diligent Robotics Inc. • Austin (TX)

On-site
USD 80,000 - 100,000
Robotics Infrastructure Engineer
Robotics Infrastructure Engineer

Tutor Intelligence • City of Watertown (NY)

On-site
USD 120,000 - 160,000
Robot Technical Support Engineer (Field Service Engineer)
Robot Technical Support Engineer (Field Service Engineer)

Sudo AI • Boston (MA)

On-site
USD 70,000 - 110,000
Robot Software Engineer
Robot Software Engineer

Rhoda AI • Palo Alto (CA)

On-site
USD 130,000 - 210,000
Robot Software Engineer
Robot Software Engineer

Rhoda AI • Mountain View (CA)

On-site
USD 80,000 - 120,000