Senior Site Reliability Engineer (SRE)

Oowlish

São Paulo

Teletrabalho

BRL 120 000 - 180 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Home office setup
Competitive compensation
Career growth plans
International projects
Oowlish English Program
Oowlish Fitness with Total Pass
Internal games and competitions

Resumo da oferta

Oowlish, based in São Paulo, Brazil, is seeking a Senior Site Reliability Engineer (SRE) to enhance system reliability and operational excellence for critical production systems. This role focuses on ensuring availability, leading incident response, and refining operational practices.

Successful candidates will have extensive SRE experience, strong programming skills in Python, Go, or TypeScript, and expertise in observability strategies. Offering flexible remote work opportunities, Oowlish promotes professional growth in an innovative environment.

Qualificações

  • 5+ years of experience in SRE, production engineering, or similar roles.
  • Proven experience operating production systems in high-availability environments.
  • Strong software engineering skills using Python, Go, or TypeScript.

Responsabilidades

  • Define, implement, and continuously improve SLIs, SLOs, and Error Budgets.
  • Lead Incident Command during production incidents.
  • Automate operational processes and reliability improvements.

Conhecimentos

Production engineering
Incident response
Observability
Python
Go
TypeScript
Cloud platforms
Monitoring
Logging
Alerting

Ferramentas

Datadog
AWS
Kubernetes
PostgreSQL
SQL Server

Descrição da oferta de emprego

Join Our Team

Oowlish, one of Latin America's rapidly expanding software development companies, is seeking experienced technology professionals to enhance our diverse and vibrant team.

As a valued member of Oowlish, you will collaborate with premier clients from the United States and Europe, contributing to pioneering digital solutions. We are certified as a Great Place to Work and offer opportunities for professional development, growth, and a chance to make a significant international impact. We provide remote work flexibility and require proficiency in English.

About the Role

Senior Site Reliability Engineer (SRE) responsible for the reliability, availability, and operational excellence of business‑critical production systems.

This is a dedicated Site Reliability Engineering role—not a general DevOps or infrastructure position. You will define how reliability is measured, lead incident response during production outages, drive observability strategy, and continuously improve operational practices across high‑availability environments.

Responsibilities
  • Define, implement, and continuously improve SLIs, SLOs, and Error Budgets
  • Develop and maintain observability strategies, including monitoring, logging, tracing, and alerting
  • Own observability configuration, instrumentation, and alert optimization
  • Lead Incident Command during production incidents and coordinate cross‑functional response efforts
  • Drive blameless post‑mortems and ensure corrective actions are completed
  • Own and continuously improve the on‑call program, including rotations, escalation policies, runbooks, and alert tuning
  • Establish production readiness standards for new services
  • Partner with engineering teams on capacity planning, scalability, and disaster recovery initiatives
  • Automate operational processes and reliability improvements using software engineering best practices
  • Continuously improve system reliability, availability, and operational efficiency
Requirements
  • 5+ years of experience in SRE, production engineering, reliability engineering, or similar roles
  • Proven experience operating production systems in high‑availability environments
  • Hands‑on experience defining and managing SLOs, SLIs, and Error Budgets
  • Experience leading production incident response and Incident Command
  • Strong observability and monitoring experience
  • Strong software engineering skills using Python, Go, or TypeScript
  • Experience working with cloud platforms
  • Excellent written and verbal English communication skills
Must Have
  • Proven SRE experience
  • Experience defining and managing:
    • Service Level Indicators (SLIs)
    • Service Level Objectives (SLOs)
    • Error Budgets
  • Experience leading Incident Command during major production incidents
  • Experience conducting blameless post‑mortems and driving follow‑up actions
  • Experience designing, maintaining, and improving on‑call programs
  • Experience developing runbooks and escalation policies
  • Strong observability experience, including:
    • Monitoring
    • Logging
    • Alerting
    • Distributed Tracing
  • Experience tuning alerts to reduce operational noise
  • Strong automation skills using Python, Go, or TypeScript
  • Experience supporting mission‑critical production systems
  • Experience working in high‑availability production environments
Nice to Have
  • Experience with Datadog
  • Experience with AWS
  • Experience with Heroku
  • Experience working in regulated industries (Healthcare, HIPAA, Financial Services, etc.)
  • Experience establishing or maturing an SRE practice
  • Capacity planning experience
  • Disaster recovery planning and execution
  • Experience with Kubernetes
  • Experience with PostgreSQL or SQL Server
  • Experience supporting modern TypeScript‑based applications
Benefits & Perks
  • Home office setup
  • Competitive compensation based on experience
  • Career plans to allow for extensive growth within the company
  • International projects
  • Oowlish English Program (technical and conversational)
  • Oowlish Fitness with Total Pass
  • Internal games and competitions
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

DevOps & Site Reliability Engineer
DevOps & Site Reliability Engineer

Oowlish • Vila Velha

Híbrido
Home office
Competitive compensation
Career growth plans
+4
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Riograndina

Híbrido
BRL 385 000 - 551 000
Professional growth
Competitive compensation
Flextime
+1
Site Reliability Engineer ID55632
Site Reliability Engineer ID55632

AgileEngine • São Paulo

Híbrido
Professional growth opportunities
Competitive USD-based compensation
Exciting projects with top companies
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Recife

Híbrido
BRL 298 000 - 399 000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)

Oowlish Technology • Fortaleza

Presencial
BRL 100 000 - 140 000
Home office
Competitive compensation based on experience
Career plans for extensive growth
+4
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)

Oowlish Technology • Espumoso

Presencial
BRL 120 000 - 160 000
Home office
Competitive compensation based on experience
Career plans for extensive growth
+4
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)

Oowlish • Rio de Janeiro

Presencial
BRL 180 000 - 260 000
Remote work
Competitive pay
Career growth
+4
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)
Senior Software Engineer (Full Stack | Node.js & React | AI Focus)

Oowlish • Curitiba

Presencial
BRL 250 000 - 380 000
Home office
Competitive compensation
Career growth
+3
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • São Bernardo do Campo

Híbrido
BRL 250 000 - 360 000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Full-Stack Engineer (Node.js + TypeScript)
Senior Full-Stack Engineer (Node.js + TypeScript)

Oowlish • Brasília

Presencial
BRL 399 000 - 500 000
Home office
Flexible hours
Competitive compensation
+5