Site Reliability Engineer Id60188 Em Recife (pe)

AgileEngine

Pernambuco

Presencial

BRL 348 432 - 472 872

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Professional growth: Mentorship, Tech Talks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
Flextime: Flexible schedule with remote and office options.

Resumo da oferta

AgileEngine in Recife, Pernambuco is seeking a Site Reliability Engineer to maintain and improve the reliability of production and staging environments on a cloud-based SaaS platform. Responsibilities include responding to live incidents, enhancing observability practices using Kubernetes and Terraform, and automating operational tasks. The ideal candidate will have 2+ years of experience in Site Reliability Engineering, strong skills in AWS, and a good understanding of CI/CD systems. Competitive compensation and flexible schedules are offered.

Qualificações

  • 2+ years of experience in Site Reliability Engineering, Dev Ops, or Production Operations.
  • Experience with AWS supporting production environments.
  • Strong understanding of CI/CD systems such as GitHub Actions, Jenkins, or CircleCI.
  • Experience with Kubernetes such as EKS or kOps.
  • Ability to work within structured operational processes and SLAs.

Responsabilidades

  • Monitor and support production and staging environments in real time.
  • Respond to incidents, perform triage and root cause analysis.
  • Participate in an on-call rotation with defined SLAs.
  • Maintain and enhance monitoring, alerting, dashboards, logs, and metrics.
  • Support CI/CD pipelines, production releases, and Git Ops workflows.

Conhecimentos

Site Reliability Engineering
Dev Ops
Production Operations
AWS
CI/CD systems
Git Ops
Kubernetes
Docker
Observability tools
Scripting languages (Bash, Python, Go)
Infrastructure as Code
English communication

Ferramentas

GitHub
Jira
Confluence
Grafana
Prometheus
Terraform

Descrição da oferta de emprego

Cargo:

Site reliability engineer id60188 em recife (PE) – recife

Requisitos:

Job Description Agile Engine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

ABOUT THE ROLE

We are looking for an SRE Operations Engineer to keep production and staging environments running reliably across a cloud-based SaaS platform. You’ll respond to live incidents, reduce operational toil through automation, and improve observability using Kubernetes, Terraform, Grafana, and AWS. A hands-on role with real ownership across CI/CD pipelines, Git Ops workflows, and on-call rotations.

WHAT YOU WILL DO
  • Monitor and support production and staging environments in real time, ensuring high availability, performance, and stability.
  • Respond to incidents, perform triage and root cause analysis, and contribute to post-incident reviews and remediation efforts.
  • Participate in an on-call rotation with defined SLAs.
  • Handle ad-hoc and unplanned operational requests from Product, Support, and internal teams.
  • Maintain and enhance monitoring, alerting, dashboards, logs, and metrics, and improve observability practices.
  • Support CI/CD pipelines, production releases, and Git Ops workflows.
  • Contribute to automation efforts to reduce operational toil.
  • Maintain and improve Kubernetes-based infrastructure and containerized workloads.
  • Support Infrastructure as Code practices and ongoing environment improvements.
MUST HAVES
  • 2+ years of experience in Site Reliability Engineering, Dev Ops, or Production Operations.
  • Experience with AWS supporting production environments.
  • Experience supporting production SaaS applications.
  • Strong understanding of CI/CD systems such as GitHub Actions, Jenkins, or CircleCI.
  • Experience with Git Ops and strong Git fundamentals.
  • Experience using GitHub, Jira, and Confluence in collaborative environments.
  • Experience with Kubernetes such as EKS or kOps.
  • Experience with Docker and containerization.
  • Experience with observability tools such as Grafana, Prometheus, Loki, or PagerDuty.
  • Experience with scripting languages such as Bash, Python, or Go.
  • Experience with Infrastructure as Code such as Terraform or Helm.
  • Ability to work within structured operational processes and SLAs.
  • Strong written and verbal English communication skills.
  • Self-driven with a growth mindset.
NICE TO HAVES
  • AWS certifications such as Solutions Architect, Dev Ops Engineer, or Sys Ops Administrator.
  • Experience in multi-tenant SaaS environments.
  • Experience working in globally distributed teams.
  • Familiarity with Chat Ops practices.
  • Experience improving monitoring quality and reducing alert fatigue.
PERKS AND BENEFITS
  • Professional growth: Mentorship, Tech Talks, and personalized growth roadmaps.
  • Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
  • Exciting projects: Modern solutions with Fortune 500 and top product companies.
  • Flextime: Flexible schedule with remote and office options.
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Recife

Híbrido
BRL 298 000 - 399 000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Full Stack Engineer – Java / React Id55426 Em Recife (pe)
Full Stack Engineer – Java / React Id55426 Em Recife (pe)

AgileEngine • Pernambuco

Híbrido
BRL 294 000 - 393 000
Mentorship
Education budget
Flexible schedule
Site Reliability Engineer ID55632
Site Reliability Engineer ID55632

AgileEngine • São Paulo

Híbrido
Professional growth opportunities
Competitive USD-based compensation
Exciting projects with top companies
+1
Platform Engineer / Devops Engineer Em Olinda (pe)
Platform Engineer / Devops Engineer Em Olinda (pe)

Sky Systems, Inc. • Pernambuco

Presencial
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Riograndina

Híbrido
BRL 385 000 - 551 000
Professional growth
Competitive compensation
Flextime
+1
Platform Engineer / Devops Engineer Em Cabo De Santo Agostinho (pe)
Platform Engineer / Devops Engineer Em Cabo De Santo Agostinho (pe)

Sky Systems, Inc. • Pernambuco

Presencial
Platform Engineer / Devops Engineer Em Caruaru (pe)
Platform Engineer / Devops Engineer Em Caruaru (pe)

Sky Systems, Inc. • Pernambuco

Presencial
Senior Software Engineer, Platform Reliability Em Recife (pe)
Senior Software Engineer, Platform Reliability Em Recife (pe)

Hyqoo • Pernambuco

Presencial
ESPECIALISTA DE TECNOLOGIA DA INFORMAÇÃO (SRE - Site Reliability Engineer)
ESPECIALISTA DE TECNOLOGIA DA INFORMAÇÃO (SRE - Site Reliability Engineer)

Qualicorp • São Paulo

Teletrabalho
BRL 120 000 - 160 000
Vale-refeição ou Vale-alimentação
Assistência médica
Assistência odontológica
+3
SITE RELIABILITY ENGINEER (SRE) (HYBRID / REMOTE)
SITE RELIABILITY ENGINEER (SRE) (HYBRID / REMOTE)

iTRTech Group • São Paulo

Híbrido
BRL 180 000 - 260 000