Evening SRE: Incident Commander & Automation Lead

Infojini Inc

Mexico

Hybrid

MXN 900,000 - 1,300,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Infojini Inc is expanding its Site Reliability Engineering team to ensure reliability, availability, and performance of production systems. You will be part of offshore SRE squads collaborating with US-based engineers to respond to incidents in real time and drive proactive reliability improvements across cloud platforms.

The role emphasizes real-time incident response, automation, and strong collaboration across engineering, product, and leadership to protect customer experience in a

Qualifications

  • Strong incident management in production environments.
  • Troubleshoot complex distributed systems under pressure.
  • Experience with multi-cloud environments (AWS, GCP, Azure).
  • Hands-on with Linux systems administration and networking fundamentals.
  • Familiarity with CI/CD pipelines and DevOps practices.

Responsibilities

  • Act as first responder to alerts and production incidents, assessing severity and initiating mitigation actions.
  • Lead major incidents as Incident Commander, guiding bridge calls with clarity and urgency.
  • Drive root cause analysis within 30 minutes for critical incidents when possible.
  • Communicate across engineering, product, and leadership during high-pressure situations.
  • Reduce toil by automating operational tasks and building tooling for incident response.

Skills

Incident management
Linux administration
AWS
GCP
Azure
Multi-cloud
CI/CD
GitHub/GitLab
Containerization
Networking
Databases
Java
Node.js
React
Load balancing
AI-assisted tooling

Job description

Infojini Inc is expanding its Site Reliability Engineering team to ensure reliability, availability, and performance of production systems. You will be part of offshore SRE squads collaborating with US-based engineers to respond to incidents in real time and drive proactive reliability improvements across cloud platforms.

The role emphasizes real-time incident response, automation, and strong collaboration across engineering, product, and leadership to protect customer experience in a

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Infojini Inc • Mexico

Hybrid
MXN 900,000 - 1,300,000
Hybrid SRE - Automation & Incident Response
Hybrid SRE - Automation & Incident Response

8x8, Inc. • Manila

Hybrid
PHP 781,200 - 1,674,000
Onboarding program
Global-scale production exposure
Blameless post-mortems culture
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Acquire Intelligence • Taguig

On-site
PHP 900,000 - 1,500,000
JD IC4 - Sr Infra Engineer - SRE
JD IC4 - Sr Infra Engineer - SRE

Spin Careers • Philippines

Remote
PHP 1,200,000 - 2,100,000
Senior SRE Engineer: Reliability, Automation & Leadership
Senior SRE Engineer: Reliability, Automation & Leadership

Spin Careers • Philippines

Remote
PHP 1,200,000 - 2,100,000
Senior SRE: Scale, Automate & Self-Healing Systems
Senior SRE: Scale, Automate & Self-Healing Systems

Replit • España

On-site
PHP 2,461,000 - 4,308,000
Competitive Salary & Equity
Health, Dental, Vision and Life Insurance
Flexible Time Off (FTO) + Holidays
+2
Associate SRE — Platform Reliability & Automation
Associate SRE — Platform Reliability & Automation

Railway Corp • Mexico

On-site
PHP 5,474,000 - 7,908,000
SRE Manager — Incident Response & Reliability Leader
SRE Manager — Incident Response & Reliability Leader

Procter & Gamble • Philippines

On-site
PHP 1,200,000 - 1,800,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Senior SRE: Platform Reliability & Incident Mastery
Senior SRE: Platform Reliability & Incident Mastery

AMADEUS MARKETING PHILS, INC. • Pateros

Hybrid
PHP 900,000 - 1,500,000