Senior Forward Deployed Engineer (DevOps/SRE)

Jobot

Pleasanton (CA)

On-site

USD 300,000 - 350,000

Full time

22 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical benefits
401(k) plan
Free meals and espresso
Equity
Commuter benefits

Job summary

Agentic AI is building a next-generation AI product for reliability engineering, backed by major investors. The founding team includes leaders from AWS, Cisco, VMware, and Gigamon, with patents and critical systems experience.

The role offers a high-impact co-founder level opportunity, comprehensive benefits, equity, and mentorship from founders. You will own customer deployments and drive platform migrations at scale in a fast-paced environment.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field or equivalent practical experience.
  • 6+ years in SRE/DevOps/Platform engineering or similar roles with technical leadership or end-to-end delivery.
  • Strong programming experience in Python, Go, or Java.

Responsibilities

  • Implement and optimize an AI-powered SRE platform across production and pre-production environments.
  • Proactively monitor deployments to maximize platform value for customers.
  • Plan, design, and maintain scalable, reliable cloud infrastructure.
  • Lead migrations from legacy alerting, AIOps, and incident management platforms.

Skills

Python/Go/Java
AWS/Azure/GCP
Kubernetes
IaC (Terraform/CF/Ansible)
CI/CD pipelines
Generative AI / ML
Observability / ITSM tools
API integration
RBAC / encryption / identity mgmt
Excellent communication

Education

Bachelor's degree in CS/Engineering

Tools

Terraform

Job description

Agentic AI Observability Tool | Co-founder took prior company from zero to $5.7B | QSBS Equity
A bit about us

Backed by over $21M in capital from leading investors, they are building a next‑generation AI product designed to transform how reliability engineering is done.

The founding team includes senior leaders and technical pioneers from industry giants like AWS, Cisco, VMware, and Gigamon - holding dozens of patents and having built critical systems at some of the most respected tech companies in the industry. This is a rare opportunity to join an early‑stage team that’s solving tough technical problems in distributed systems, observability, and automation — all while shaping a product from the ground up.

Why join us

Benefits: Comprehensive medical, vision, and dental benefits. 401 (k) plans and commuter benefits. Free lunches, snacks, and top‑of‑the‑line espressos!

Equity that could change your life.

High‑impact role with plenty of mentorship opportunities from founders and other coworkers

Collaborative coworkers with high IQ and high EQ. No politics. No bureaucracy. Open door policy.

Job Details

Salary: $300,000 - $350,000 per year

  • Implement and optimize an AI‑powered Site Reliability Engineering (SRE) platform to meet customer needs across production and pre‑production environments.
  • Proactively monitor customer deployments to ensure customers maximize value from the platform.
  • Identify latent reliability issues such as misconfigurations, deployment regressions, and scaling challenges within customer environments.
  • Recommend best practices for implementing AI‑powered SRE solutions.
  • Plan, design, build, and maintain highly scalable, reliable, and efficient cloud infrastructure.
  • Serve as the customer's technical advocate with internal engineering and product teams.
  • Conduct post‑incident reviews to identify root causes and implement preventative measures.
  • Ensure security best practices are integrated into customer deployments.
  • Train customer SRE, Operations, and Platform Engineering teams on platform usage and best practices.
  • Lead enterprise migrations from legacy alerting, AIOps, and incident management platforms, including correlation rule migration, phased cutovers, and production go‑live execution.
  • Design, build, and optimize alert normalization and correlation policies using conditions, regular expressions, field extraction, and customized workflows.
  • Integrate the platform with customer operational systems, including ITSM, collaboration, observability, source control, and documentation platforms.
  • Validate and continuously improve AI investigation quality by tuning enrichment, root cause analysis accuracy, and investigation workflows.
  • Build proactive monitoring for customer deployments to identify issues before they impact customers.
  • Own customer‑facing project communications, including executive status updates, SLA documentation, escalation management, and implementation tracking.
  • Develop long‑term technical relationships with senior engineering leadership.
  • Own customer implementations from technical discovery through solution design, implementation, user acceptance testing, production go‑live, stabilization, and ongoing optimization.
  • Translate ambiguous customer requirements into clear technical designs, milestones, acceptance criteria, and execution plans.
  • Design and implement AI‑powered investigation and automation workflows with appropriate guardrails, governance, deterministic fallbacks, and human oversight.
  • Develop reusable deployment modules, reference architectures, implementation guides, and operational runbooks to accelerate future deployments.
  • Define customer success metrics, establish baselines, measure operational improvements, and demonstrate business value through KPIs such as MTTR reduction and operational efficiency.
  • Capture customer feedback and recurring implementation learnings to influence future product development.
  • Foster a culture of continuous improvement and technical excellence.
Qualifications
  • Customer‑focused with deep empathy for SRE, DevOps, Platform Engineering, and IT Operations teams.
  • Bachelor's degree in Computer Science, Engineering, or a related technical field (or equivalent practical experience).
  • 6+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or similar infrastructure‑focused roles, including technical leadership or end‑to‑end customer delivery.
  • Experience in Forward Deployed Engineering, Solutions Engineering, Technical Customer Success, or Professional Services is highly preferred.
  • Strong programming experience in at least one language such as Python, Go, or Java.
  • Hands‑on experience with public cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Strong knowledge of Kubernetes, Infrastructure as Code (Terraform, CloudFormation, Ansible), and CI/CD pipelines.
  • Practical experience using Generative AI and machine learning technologies to improve engineering productivity.
  • Experience with observability platforms, ITSM systems, and incident management tools, including systems integration and data mapping.
  • Strong troubleshooting, analytical, and debugging skills, including alert correlation, normalization, and regular expression development.
  • Excellent written and verbal communication skills.
  • Demonstrated ownership of enterprise software implementations from discovery through production deployment.
  • Strong integration experience with APIs, webhooks, event‑driven architectures, authentication (SSO/SAML), data transformations, synchronization, and enterprise application integrations.
  • Experience designing and deploying production‑grade AI or automation workflows with governance and evaluation frameworks.
  • Understanding of enterprise security concepts including RBAC, encryption, identity management, auditing, and secure networking.
  • Ability to operate effectively in ambiguous, fast‑paced customer environments while balancing architecture with execution.
  • Self‑motivated, adaptable, and capable of managing shifting priorities while driving successful customer outcomes.
Preferred Qualifications
  • Experience supporting customers operating AI infrastructure or AI‑enabled platforms.
  • Experience migrating customers from legacy alerting, AIOps, or incident management platforms.
  • Experience building internal automation and tooling using Python, Node.js, Bash, or similar scripting languages.
Equal Opportunity & Background

Jobot is an Equal Opportunity Employer. We provide an inclusive work environment that celebrates diversity and all qualified candidates receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, age (40 and over), disability, military status, genetic information or any other basis protected by applicable federal, state, or local laws. Jobot also prohibits harassment of applicants or employees based on any of these protected categories. It is Jobot's policy to comply with all applicable federal, state and local laws respecting consideration of unemployment status in making hiring decisions.

Sometimes Jobot is required to perform background checks with your authorization. Jobot will consider qualified candidates with criminal histories in a manner consistent with any applicable federal, state, or local law regarding criminal backgrounds, including but not limited to the Los Angeles Fair Chance Initiative for Hiring and the San Francisco Fair Chance Ordinance.

Information collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal.

By applying for this job, you agree to receive calls, AI‑generated calls, text messages, or emails from Jobot, and/or its agents and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy here: jobot.com/privacy-policy

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Forward Deployed Engineer
Forward Deployed Engineer

Australia-Employment • San Francisco (CA)

On-site
USD 185,000 - 300,000
Highly competitive base salary
Equity
Medical, dental, and vision coverage
+5
Senior Forward Deployed Engineer (DevOps/SRE)
Senior Forward Deployed Engineer (DevOps/SRE)

LeoForce • Pleasanton (CA)

On-site
USD 300,000 - 350,000
Medical benefits
401(k) plan
Free meals and snacks
+2
Software Engineer
Software Engineer

Jobot • Arlington (VA)

On-site
USD 100,000 - 170,000
Health insurance
PTO and paid holidays
Equity
Manager, Software Engineering (Hands-On | Cloud SaaS | AI-Driven Development)
Manager, Software Engineering (Hands-On | Cloud SaaS | AI-Driven Development)

Jobot • Seattle (WA)

On-site
USD 160,000 - 230,000
Senior Software Engineer
Senior Software Engineer

Australia-Employment • San Francisco (CA)

Hybrid
USD 225,000 - 450,000
Equity in the company
Hybrid in SF office
Top compensation
Lead Engineer (Breach & Attack Simulation)
Lead Engineer (Breach & Attack Simulation)

Australia-Employment • Los Angeles (CA)

Remote
USD 250,000 - 400,000
Data Engineer
Data Engineer

Jobot • Addison (TX)

On-site
USD 120,000 - 160,000
Top benefits
Great culture
Staff / Senior Software Engineer, Agent Engineering
Staff / Senior Software Engineer, Agent Engineering

Australia-Employment • New York (NY)

On-site
USD 185,000 - 235,000
Equity / ownership
Direct access to founders and CTO
Very high technical ownership
+5
Senior/Lead Software Enigneer (AI-assisted development)
Senior/Lead Software Enigneer (AI-assisted development)

Jobot Consulting • West Columbia (SC)

On-site
USD 62,000 - 90,000
Lead Full Stack Software Engineer (New Application & AI Development)
Lead Full Stack Software Engineer (New Application & AI Development)

Jobot Consulting • West Columbia (SC)

On-site
USD 62,000 - 90,000