Senior Site Reliability Engineer

Imachines

Chile

Remote

CLP 115,385,000 - 173,077,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Fully remote
Global team
Ship early, ship often
High impact
Flat organization

Job summary

Intuition Machines is seeking a Senior Site Reliability Engineer to focus on performance, availability, security, and cost-effectiveness across its internet-scale systems. The role involves working across infrastructure, data, and application logic in a highly reliable environment with rapid iteration.

You will read and understand lines of code with your name on it, collaborate with distributed teams worldwide, and contribute to improvements in latency, uptime, and security.

Qualifications

  • Minimum of six years of hands-on experience in related roles (engineering, DevOps, SRE).
  • Expert in Kubernetes and monitoring applications, infrastructure and network.
  • Strong programming skills in one or more of the following languages: Python, JavaScript, Go, C++, Rust.

Responsibilities

  • Work with large-scale systems (handling millions of requests per second) across multiple cloud providers.
  • Develop solutions to enhance performance, availability, security, and cost-effectiveness.
  • Keep systems fast and teams productive, ensuring every peer release improves performance and security.
  • Source improvement ideas, priority and capabilities from customers and the internal community, and from system metrics.
  • Be creative and a force to improve the experience for customers.

Skills

Kubernetes
Monitoring
Networking
CI/CD
Python
Go
JavaScript
C++
Rust

Tools

Docker
Git
Terraform
Prometheus
Cloudflare

Job description

Intuition Machines uses AI/ML to build enterprise security products. We apply our research to systems that serve hundreds of millions of people, with a team distributed around the world. You are probably familiar with our best-known product, the hCaptcha security suite. Our approach is simple: low overhead, small teams, and rapid iteration.

As a Senior Site Reliability Engineer, you will focus on engineering solutions related to performance, availability, security, and cost-effectiveness. We consider these non-functional features to be core requirements for us and our customers. You will work at multiple layers of our internet-scale system (infrastructure, data, application logic) and build the solutions.

Using AI: Coding agents are indisputably useful tools. We provide access to the top 3 models, and were early adopters of evals-first development flows. Familiarity with coding using agents is part of all interviews. However, reliability and correctness are critical for us. You will need to read and understand every line of code with your name on it, and it will be reviewed by both people and machines.

What you will do:
  • Work with large-scale systems (handling millions of requests per second, serving millions of users, across multiple cloud providers).
  • Develop solutions to enhance performance, availability, security, and cost-effectiveness.
  • Keep us up, keep us fast, and keep our dev teams productive ensuring that every peer release improves performance across the spectrum including quality, security, uptime, speed-to-deliver, threat detection, and customer engagement.
  • Source improvement ideas, priority and capabilities from customers, the internal community, new and existing system metrics. Make decisions rapidly.
  • Be creative and desire an environment where you can directly create value and be a force to improve the experience for our customers.
What we are looking for:
  • Expert in Kubernetes.
  • Expert in monitoring applications, infrastructure and network.
  • Background in software engineering with expertise in backend development within Kubernetes-based systems.
  • Strong programming skills in one or more of the following languages: Python, JavaScript, Go, C++, Rust.
  • Strong understanding and experience in networking, proxies, content delivery networks (Cloudflare)
  • Multi cloud experience including virtual networking, load balancing, web application firewall.
  • Strong experience with CI/CD.
  • Hands-on experience in development and orchestration within high-scale, high-uptime, and high-reliability environments.
  • Minimum of six years of hands-on experience in related roles (engineering, DevOps, SRE).
  • Familiarity with distributed systems, including queue-first architectures and sharding.
  • Demonstrated engineering expertise, including gathering requirements, problem-solving, and making recommendations.
  • Preferred: Familiarity with security frameworks, attack vectors, botnets, and impact analysis.
What we offer:
  • Fully remote position with flexible working hours.
  • An inspiring team of colleagues spread all over the world.
  • Pleasant, modern development and deployment workflows: ship early, ship often.
  • High impact: lots of users, happy customers, high growth, and cutting edge R&D.
  • Flat organization, direct interaction with customer teams.

We celebrate equality of opportunity and are committed to creating an inclusive environment for all team members.Join us as we transform cybersecurity, user privacy, and machine learning online!

Please note that all positions require pre-employment screening, including third-party verification of work history, education, and identity, as well as a final in-person interview and identity verification step, which will be conducted in your country of residence.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Senior SRE — Scale, Security & Uptime
Remote Senior SRE — Scale, Security & Uptime

Imachines • Chile

Remote
CLP 115,385,000 - 173,077,000
Fully remote
Global team
Ship early, ship often
+2
Forward Deployed AI Engineer (Python) - Remote -Latin America
Forward Deployed AI Engineer (Python) - Remote -Latin America

FullStack • Valparaíso

Remote
CLP 115,607,000 - 173,410,000
100% remote work
Continuing education opportunities
Grow your career with leading clients
Forward Deployed Engineer (TypeScript + AI) - Remote - Latin America
Forward Deployed Engineer (TypeScript + AI) - Remote - Latin America

FullStack • Concepcion

Remote
CLP 117,417,000 - 176,125,000
Competitive pay
100% remote work
Continuing education opportunities
+1
Senior Software Engineer (CRM+ Platform) - Remote - Latin America
Senior Software Engineer (CRM+ Platform) - Remote - Latin America

FullStack • Santiago

Remote
CLP 115,607,000 - 173,410,000
100% remote work
Competitive pay
Continuing education opportunities
+1
Forward Deployed Engineer (TypeScript + AI) - Remote - Latin America
Forward Deployed Engineer (TypeScript + AI) - Remote - Latin America

FullStack • Valparaíso

Remote
CLP 97,847,000 - 156,556,000
100% remote work
Competitive pay
Continuing education opportunities
+1
AI Computer Vision Engineer - Remote - Latin America
AI Computer Vision Engineer - Remote - Latin America

FullStack • Concepcion

Remote
CLP 117,417,000 - 146,771,000
Competitive pay
Remote work
Work with startups
+2
Senior Customer Engineer, (Santiago, Chile).
Senior Customer Engineer, (Santiago, Chile).

CloudFlare • Santiago

On-site
CLP 12,000,000 - 20,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Applydigital • Santiago

On-site
CLP 56,075,000 - 84,112,000
Generous vacation policy
Flexible work arrangements
AI upskilling & training budgets
+1
AI Machine Learning Engineer - Remote - Latin America
AI Machine Learning Engineer - Remote - Latin America

FullStack • Valparaíso

Remote
CLP 117,417,000 - 176,125,000
100% remote work
Continuing education opportunities
AI Machine Learning Engineer - Remote - Latin America
AI Machine Learning Engineer - Remote - Latin America

FullStack • Concepcion

Remote
CLP 117,417,000 - 176,125,000
100% remote work
Growth opportunities