Founding SRE for AI-Native Cloud Platform — Remote

Gradle Technologies

North Township (IN)

On-site

USD 150,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote-first culture
Competitive salary + equity
Annual company offsite

Job summary

Gradle Technologies is hiring a founding SRE/DevOps engineer to own reliability across Develocity instances. You will manage production services, run on-call rotations, automate deployments, and build observability across the stack.

You’ll collaborate with Cloud Platform and Engineering to embed reliability from design onward. The role demands 5+ years in SRE/DevOps, strong Kubernetes and AWS experience, and proficiency with Python/Bash scripting.

Qualifications

  • 5+ years in SRE, DevOps, or equivalent role operating production services at scale.
  • Strong Kubernetes experience in production environments.
  • Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).
  • Proficiency with observability tools (Prometheus, Grafana) and Terraform.
  • Track record of incident management and response.
  • Knowledge of SRE practices (SLAs, SLOs) and scripting (Python, Bash).
  • Experience with 24/7 on-call rotations and strong English communication.

Responsibilities

  • Operate and maintain Develocity instances and supporting services.
  • Participate in follow-the-sun on-call rotation; own incident response.
  • Drive automation across deployment, upgrades, monitoring, and self-healing.
  • Build and maintain observability for all managed services (logging, metrics, tracing, alerts).
  • Collaborate with engineering to bake reliability into features from the start.
  • Run incident retrospectives and implement lessons learned.
  • Own disaster recovery, backups, and business continuity.
  • Communicate with customers during incidents and maintenance windows.
  • Optimize performance, resource usage, and costs.
  • Help evolve SaaS operations as we scale.

Skills

Kubernetes
AWS
Python
Bash
English comms

Tools

Prometheus
Grafana
Terraform
CI/CD tooling

Job description

Gradle Technologies is hiring a founding SRE/DevOps engineer to own reliability across Develocity instances. You will manage production services, run on-call rotations, automate deployments, and build observability across the stack.

You’ll collaborate with Cloud Platform and Engineering to embed reliability from design onward. The role demands 5+ years in SRE/DevOps, strong Kubernetes and AWS experience, and proficiency with Python/Bash scripting.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Gradle Inc. • United States

Remote
USD 150,000 - 200,000
Competitive salaries
Equity grants
Remote work
+2
Senior SRE - Remote-First, Reliability & Automation
Senior SRE - Remote-First, Reliability & Automation

Gradle Inc. • United States

Remote
USD 150,000 - 200,000
Senior AI-Powered SRE / FDE — Cloud, Automation
Senior AI-Powered SRE / FDE — Cloud, Automation

Jobot • Pleasanton (CA)

On-site
USD 300,000 - 350,000
Medical benefits
401(k) and commuter benefits
Free lunches and snacks
+3
Senior DevOps Engineer & SRE: Cloud-Native Reliability Lead
Senior DevOps Engineer & SRE: Cloud-Native Reliability Lead

Front Door Defense • New York (NY)

On-site
USD 165,000 - 215,000
Senior Staff DevOps Engineer - Remote (AI Data Platform)
Senior Staff DevOps Engineer - Remote (AI Data Platform)

Stealth AI Startup • United States

On-site
USD 140,000 - 190,000
SRE / Dev Ops
SRE / Dev Ops

Enid • Las Vegas (NV)

On-site
USD 140,000 - 170,000
Staff SRE Engineer: AI-Driven Platform Reliability
Staff SRE Engineer: AI-Driven Platform Reliability

AI Chopping Block • Costa Mesa (CA), Northern (KY)

Hybrid
USD 191,000 - 253,000
CloudDevs: Senior Site Reliability Engineer (SRE)
CloudDevs: Senior Site Reliability Engineer (SRE)

Breakout Tools • San Francisco (CA)

On-site
USD 120,000 - 160,000
Remote Senior DevOps & SRE Engineer: Cloud Reliability
Remote Senior DevOps & SRE Engineer: Cloud Reliability

United States Digital Space LLC • United States

Remote
USD 120,000 - 180,000
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform

Instrumental Inc. • Palo Alto (CA)

On-site
USD 175,000 - 229,000
Health benefits
Commuter plans
Parental leave