Senior Forward Deployed Engineer (DevOps/SRE)

Jobot

Pleasanton (CA)

On-site

USD 300,000 - 350,000

Full time

16 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical benefits
401(k) and commuter benefits
Free lunches and snacks
Equity participation
Mentorship from founders
Open door policy

Job summary

Jobot is seeking an experienced SRE/Platform Engineer to design and operate an AI-powered SRE platform across production and pre-production environments. You will drive reliability, security, and scalable cloud infrastructure while guiding customer implementations end-to-end.

The role requires 6+ years in SRE/DevOps, strong programming, and hands-on cloud + Kubernetes expertise. You will collaborate with founders and engineering teams to deliver high-impact outcomes and mentorship.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field; or equivalent practical experience.
  • 6+ years in SRE/DevOps or platform engineering with customer delivery experience.
  • Strong programming experience in Python, Go, or Java.
  • Hands-on with public clouds (AWS/Azure/GCP) and Kubernetes expertise.
  • Experience with IaC (Terraform, CloudFormation) and CI/CD pipelines.

Responsibilities

  • Implement and optimize AI-powered SRE platform for production and pre-production environments.
  • Monitor deployments to maximize customer value and reliability.
  • Identify resiliency issues, misconfigurations, and scaling challenges.
  • Design scalable cloud infrastructure and govern security best practices.
  • Lead customer-facing deployments and migrations with internal teams.
  • Develop automation workflows and reference architectures for future deployments.

Skills

Python/Go/Java
SRE/DevOps leadership
Kubernetes
Infrastructure as Code
Cloud platforms (AWS/Azure/GCP)
Observability/Monitoring
CI/CD pipelines
API integration

Education

Bachelor's degree in CS/Engineering

Tools

Kubernetes
Terraform
CI/CD tools
Ansible
AWS/GCP/Azure

Job description

Salary: $300,000 - $350,000 per year

This Jobot Job is hosted by: Caitlyn Hardy

A bit about us

Backed by over $21M in capital from leading investors, they are building a next-generation AI product designed to transform how reliability engineering is done.

The founding team includes senior leaders and technical pioneers from industry giants like AWS, Cisco, VMware, and Gigamon - holding dozens of patents and having built critical systems at some of the most respected tech companies in the industry. This is a rare opportunity to join an early-stage team that’s solving tough technical problems in distributed systems, observability, and automation — all while shaping a product from the ground up.

Why join us
  • Comprehensive medical, vision, and dental benefits.
  • 401(k) plans and commuter benefits.
  • Free lunches, snacks, and top-of-the-line espressos!
  • Equity that could change your life.
  • High-impact role with plenty of mentorship opportunities from founders and other coworkers
  • Collaborative coworkers with high IQ and high EQ. No politics. No bureaucracy. Open door policy.
Responsibilities
  • Implement and optimize an AI-powered Site Reliability Engineering (SRE) platform to meet customer needs across production and pre-production environments.
  • Proactively monitor customer deployments to ensure customers maximize value from the platform.
  • Identify latent reliability issues such as misconfigurations, deployment regressions, and scaling challenges within customer environments.
  • Recommend best practices for implementing AI-powered SRE solutions.
  • Plan, design, build, and maintain highly scalable, reliable, and efficient cloud infrastructure.
  • Serve as the customer's technical advocate with internal engineering and product teams.
  • Conduct post-incident reviews to identify root causes and implement preventative measures.
  • Ensure security best practices are integrated into customer deployments.
  • Train customer SRE, Operations, and Platform Engineering teams on platform usage and best practices.
  • Lead enterprise migrations from legacy alerting, AIOps, and incident management platforms, including correlation rule migration, phased cutovers, and production go-live execution.
  • Design, build, and optimize alert normalization and correlation policies using conditions, regular expressions, field extraction, and customized workflows.
  • Integrate the platform with customer operational systems, including ITSM, collaboration, observability, source control, and documentation platforms.
  • Validate and continuously improve AI investigation quality by tuning enrichment, root cause analysis accuracy, and investigation workflows.
  • Build proactive monitoring for customer deployments to identify issues before they impact customers.
  • Own customer-facing project communications, including executive status updates, SLA documentation, escalation management, and implementation tracking.
  • Develop long-term technical relationships with senior engineering leadership.
  • Own customer implementations from technical discovery through solution design, implementation, user acceptance testing, production go-live, stabilization, and ongoing optimization.
  • Translate ambiguous customer requirements into clear technical designs, milestones, acceptance criteria, and execution plans.
  • Design and implement AI-powered investigation and automation workflows with appropriate guardrails, governance, deterministic fallbacks, and human oversight.
  • Develop reusable deployment modules, reference architectures, implementation guides, and operational runbooks to accelerate future deployments.
  • Define customer success metrics, establish baselines, measure operational improvements, and demonstrate business value through KPIs such as MTTR reduction and operational efficiency.
  • Capture customer feedback and recurring implementation learnings to influence future product development.
  • Foster a culture of continuous improvement and technical excellence.
Qualifications
  • Customer-focused with deep empathy for SRE, DevOps, Platform Engineering, and IT Operations teams.
  • Bachelor's degree in Computer Science, Engineering, or a related technical field (or equivalent practical experience).
  • 6+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or similar infrastructure-focused roles, including technical leadership or end-to-end customer delivery.
  • Experience in Forward Deployed Engineering, Solutions Engineering, Technical Customer Success, or Professional Services is highly preferred.
  • Strong programming experience in at least one language such as Python, Go, or Java.
  • Hands-on experience with public cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Strong knowledge of Kubernetes, Infrastructure as Code (Terraform, CloudFormation, Ansible), and CI/CD pipelines.
  • Practical experience using Generative AI and machine learning technologies to improve engineering productivity.
  • Experience with observability platforms, ITSM systems, and incident management tools, including systems integration and data mapping.
  • Strong troubleshooting, analytical, and debugging skills, including alert correlation, normalization, and regular expression development.
  • Excellent written and verbal communication skills.
  • Demonstrated ownership of enterprise software implementations from discovery through production deployment.
  • Strong integration experience with APIs, webhooks, event-driven architectures, authentication (SSO/SAML), data transformations, synchronization, and enterprise application integrations.
  • Experience designing and deploying production-grade AI or automation workflows with governance and evaluation frameworks.
  • Understanding of enterprise security concepts including RBAC, encryption, identity management, auditing, and secure networking.
  • Ability to operate effectively in ambiguous, fast-paced customer environments while balancing architecture with execution.
  • Self-motivated, adaptable, and capable of managing shifting priorities while driving successful customer outcomes.
Preferred Qualifications
  • Experience supporting customers operating AI infrastructure or AI-enabled platforms.
  • Experience migrating customers from legacy alerting, AIOps, or incident management platforms.
  • Experience building internal automation and tooling using Python, Node.js, Bash, or similar scripting languages.

Jobot is an Equal Opportunity Employer.

We provide an inclusive work environment that celebrates diversity and all qualified candidates receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, age (40 and over), disability, military status, genetic information or any other basis protected by applicable federal, state, or local laws.

Jobot also prohibits harassment of applicants or employees based on any of these protected categories.

It is Jobot’s policy to comply with all applicable federal, state and local laws respecting consideration of unemployment status in making hiring decisions.

Sometimes Jobot is required to perform background checks with your authorization.

Jobot will consider qualified candidates with criminal histories in a manner consistent with any applicable federal, state, or local law regarding criminal backgrounds, including but not limited to the Los Angeles Fair Chance Initiative for Hiring and the San Francisco Fair Chance Ordinance.

Information collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal.

By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Jobot, and/or its agents and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy here: jobot.com/privacy-policy

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Forward Deployed Engineer
Forward Deployed Engineer

Jobot • San Francisco (CA)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision insurance
401(k) retirement plan
Paid time off and company holidays
+5
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Charleston (SC)

Hybrid
USD 160,000 - 300,000
Medical, dental, vision insurance
401(k) retirement plan
PTO and holidays
+5
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Los Angeles (CA)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision insurance
401(k) retirement plan
Paid time off and company holidays
+5
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Dallas (TX)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision insurance
401(k) retirement plan
Paid time off and company holidays
+3
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Boston (MA)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision insurance
401(k) retirement plan
Paid time off and holidays
+5
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Atlanta (GA)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision insurance
401(k) retirement plan
Paid time off and company holidays
+5
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Chicago (IL)

Hybrid
USD 160,000 - 300,000
Medical, dental, and vision
401(k) retirement plan
Paid time off and holidays
+5
Senior Sales Engineer
Senior Sales Engineer

Jobot • Pleasanton (CA)

On-site
USD 300,000 - 330,000
Stock equity
Medical benefits
401(k) plan
+1
Forward Deployed Engineer
Forward Deployed Engineer

Jobot • Santa Clara (CA)

On-site
USD 150,000 - 225,000
Competitive compensation
Equity + benefits
Gym membership
+2
DevOps Engineer
DevOps Engineer

Jobot • Fort Worth (TX)

On-site
USD 120,000 - 140,000
Competitive compensation
Comprehensive benefits
Career growth opportunities