Senior Site Reliability Engineer

Intuition Machines

United States

Remote

PHP 7,299,000 - 10,949,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote position
Flexible working hours
Global team
Modern development workflows
High impact projects

Job summary

Intuition Machines is hiring a Senior Site Reliability Engineer to build scalable, secure, cost-efficient systems across multiple clouds.

You will focus on performance, availability and reliability, work with Kubernetes, monitoring, and backend services in distributed, high-scale environments, and collaborate with global teams to ship reliable software quickly.

We offer a fully remote role with flexible hours, a global team, modern workflows, and high-impact work at scale.

Qualifications

  • Expert in Kubernetes
  • Strong monitoring of applications, infrastructure and networks
  • Background in backend development with Kubernetes-based systems
  • Strong programming skills in Python, JavaScript, Go, C++, or Rust
  • Strong networking knowledge including proxies and CDNs (Cloudflare)
  • Multi-cloud experience with virtual networking, load balancing, WAF
  • Extensive CI/CD experience
  • Hands-on with high-scale, high-uptime environments
  • Minimum of six years of hands-on experience
  • Experience with distributed systems and queue-first architectures

Responsibilities

  • Work with large-scale systems handling millions of requests per second, across multiple cloud providers
  • Develop solutions to enhance performance, availability, security, and cost‑effectiveness
  • Keep us fast and our dev teams productive ensuring peer releases improve performance across spectrum
  • Source improvement ideas from customers, internal community, metrics; make rapid decisions
  • Be creative and create value, improving the experience for our customers

Skills

Kubernetes
Monitoring
Backend development
Python
Go
JavaScript
C++
Rust
Networking
CI/CD
Distributed systems
Multi-cloud
6+ years experience

Tools

Cloudflare

Job description

Intuition Machines uses AI/ML to build enterprise security products. We apply our research to systems that serve hundreds of millions of people, with a team distributed around the world. You are probably familiar with our best‑known product, the hCaptcha security suite. Our approach is simple: low overhead, small teams, and rapid iteration.

As a Senior Site Reliability Engineer, you will focus on engineering solutions related to performance, availability, security, and cost‑effectiveness. We consider these non‑functional features to be core requirements for us and our customers. You will work at multiple layers of our internet‑scale system (infrastructure, data, application logic) and build the solutions.

Using AI: Coding agents are indisputably useful tools. We provide access to the top 3 models, and were early adopters of evals-first development flows. Familiarity with coding using agents is part of all interviews. However, reliability and correctness are critical for us. You will need to read and understand every line of code with your name on it, and it will be reviewed by both people and machines.

What you will do:
  • Work with large-scale systems (handling millions of requests per second, serving millions of users, across multiple cloud providers)
  • Develop solutions to enhance performance, availability, security, and cost‑effectiveness
  • Keep us up, keep us fast, and keep our dev teams productive ensuring that every peer release improves performance across the spectrum including quality, security, uptime, speed‑to‑deliver, threat detection, and customer engagement
  • Source improvement ideas, priority and capabilities from customers, the internal community, new and existing system metrics. Make decisions rapidly
  • Be creative and desire an environment where you can directly create value and be a force to improve the experience for our customers
What we are looking for:
  • Expert in Kubernetes
  • Expert in monitoring applications, infrastructure and network.
  • Background in software engineering with expertise in backend development within Kubernetes‑based systems
  • Strong programming skills in one or more of the following languages: Python, JavaScript, Go, C++, Rust
  • Strong understanding and experience in networking, proxies, content delivery networks (Cloudflare)
  • Multi cloud experience including virtual networking, load balancing, web application firewall
  • Strong experience with CI/CD
  • Hands‑on experience in development and orchestration within high‑scale, high‑uptime, and high‑reliability environments
  • Minimum of six years of hands‑on experience in related roles (engineering, DevOps, SRE)
  • Familiarity with distributed systems, including queue‑first architectures and sharding
  • Demonstrated engineering expertise, including gathering requirements, problem‑solving, and making recommendations
  • Preferred: Familiarity with security frameworks, attack vectors, botnets, and impact analysis.
What we offer:
  • Fully remote position with flexible working hours
  • An inspiring team of colleagues spread all over the world
  • Pleasant, modern development and deployment workflows: ship early, ship often
  • High impact: lots of users, happy customers, high growth, and cutting edge R&D
  • Flat organization, direct interaction with customer teams

We celebrate equality of opportunity and are committed to creating an inclusive environment for all team members. Join us as we transform cybersecurity, user privacy, and machine learning online!

Please note that all positions require pre-employment screening, including third-party verification of work history, education, and identity, as well as a final in‑person interview and identity verification step, which will be conducted in your country of residence.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Limelight Health • Austin (TX)

On-site
USD 152,000 - 195,000
Competitive salary
Stock options
Health benefits
+2
Site Reliability Engineer Austin, TX
Site Reliability Engineer Austin, TX

Future Secure AI Pty • Austin (TX)

On-site
USD 100,000 - 140,000
Flexible work environment
Competitive salary
Growth trajectory
Senior AI Platform Engineer
Senior AI Platform Engineer

hackerone • Washington

On-site
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AlleyCorp • United States

On-site
USD 152,000 - 195,000
Health benefits
Unlimited PTO
Parental leave
+1
Senior Software Engineer, Applied AI
Senior Software Engineer, Applied AI

hackerone • Boston (MA)

Hybrid
USD 190,000 - 230,000
Health (medical, vision, dental) insurance
Equity stock options
Retirement plans
+1
Software Engineer (Agent Infra)
Software Engineer (Agent Infra)

Hinoki • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Software Engineer, Applied AI
Software Engineer, Applied AI

hackerone • Boston (MA)

Hybrid
USD 166,000 - 203,000
Health insurance
Paid public holidays and unlimited PTO
Equity stock options
+1
Senior Software Engineer, AI Platform
Senior Software Engineer, AI Platform

HackerOne • Austin (CO)

Hybrid
USD 190,000 - 230,000
Health insurance
Equity stock options
Unlimited PTO
+2
Infrastructure Engineer, Security
Infrastructure Engineer, Security

Thinkingmachines • San Francisco (CA)

On-site
USD 200,000 - 475,000
Health insurance
Unlimited PTO
Paid parental leave
+1
Senior Software Engineer, Intelligence Services (US)
Senior Software Engineer, Intelligence Services (US)

Centripetal • Reston (VA)

Hybrid
USD 100,000 - 130,000