Senior DevOps Analyst – SRE, Kubernetes, Cloud

Jobtailor

São Paulo

Presencial

BRL 90 000 - 130 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Jobtailor in São Paulo, Brazil seeks a DevOps Engineer/SRE to implement and evolve SRE practices, administer Linux/Unix environments, and build robust CI/CD pipelines.

You will manage Docker and Kubernetes, craft IaC with Terraform or Pulumi, and enhance observability with Prometheus, Grafana, and Alertmanager, aiming for high reliability. Collaboration across development, architecture, and operations is essential for continuous improvement.

Qualificações

  • Solid experience in DevOps, SRE, or similar roles.
  • Strong command of Git and GitHub (branching, PRs, versioning).
  • Advanced Linux/Unix knowledge and troubleshooting skills.
  • Experience with Docker, including image build, optimization, and security.
  • Experience with Kubernetes for deployment, ops, and diagnostics.
  • Terraform, Pulumi, or similar IaC tools knowledge.
  • Observability mindset with metrics, logs, tracing.
  • Experience with Prometheus, Grafana, and Alertmanager.
  • Experience defining and managing SLI/SLO/SLA.
  • CI/CD pipelines experience and automation mindset.
  • Cloud exposure (Azure, AWS, OCI) beneficial.
  • Networking knowledge and security basics (IAM, TLS, secrets).
  • Familiarity with ELK or Loki logging platforms.

Responsabilidades

  • Implement and evolve Site Reliability Engineering practices.
  • Administer and optimize high-criticality Linux/Unix environments.
  • Create, maintain, and improve CI/CD pipelines.
  • Manage containerized environments using Docker and Kubernetes.
  • Develop and maintain IaC using Terraform, Pulumi, or similar tools.
  • Implement observability with metrics, logs, and tracing.
  • Define and track reliability indicators (SLI, SLO, SLA).
  • Lead management and response to critical incidents.
  • Perform advanced troubleshooting in distributed environments.
  • Ensure security best practices related to IAM, TLS, secrets management, access control.
  • Manage monitoring and observability platforms such as Prometheus, Grafana, and Alertmanager.
  • Implement strategies to reduce failures, eliminate toil, increase automation.
  • Conduct RCAs, postmortems, and continuous improvement initiatives.
  • Collaborate with development, architecture, and operations to drive operational excellence.

Ferramentas

Git
GitHub
Linux
Kubernetes
Docker
Terraform
Pulumi
Prometheus
Grafana
Alertmanager
ELK Stack
CI/CD
IAM/TLS
Service Mesh
Multi-Cloud

Descrição da oferta de emprego

Responsibilities
  • Implement and evolve Site Reliability Engineering (SRE) practices
  • Administer and optimize high-criticality Linux/Unix environments
  • Create, maintain, and improve CI/CD pipelines
  • Manage containerized environments using Docker and Kubernetes
  • Develop and maintain Infrastructure as Code (IaC) using Terraform, Pulumi, or similar tools
  • Implement observability solutions using metrics, logs, and tracing
  • Define and track reliability indicators such as SLI, SLO, and SLA
  • Lead management and response to critical incidents
  • Perform advanced troubleshooting in distributed environments
  • Ensure security best practices related to IAM, TLS, secrets management, and access control
  • Manage monitoring and observability platforms such as Prometheus, Grafana, and Alertmanager
  • Implement strategies to reduce failures, eliminate repetitive work (toil), and increase automation
  • Conduct root cause analyses (RCA), postmortems, and continuous improvement initiatives
  • Collaborate with development, architecture, and operations teams to drive operational excellence
Requirements
  • Solid experience working as a DevOps Engineer, SRE, or in similar roles
  • Strong command of Git and GitHub (branching, pull requests, and versioning strategies)
  • Advanced experience with Linux and Unix
  • Deep knowledge of application and infrastructure troubleshooting
  • Experience with Docker, including image build, optimization, and security
  • Experience with Kubernetes for deployment, operations, and diagnostics
  • Knowledge of observability (metrics, logs, and tracing)
  • Experience with Prometheus, Grafana, and Alertmanager
  • Experience defining and managing SLI, SLO, and SLA
  • Experience with Infrastructure as Code using Terraform, Pulumi, or similar tools
  • Experience in Cloud environments (Azure, AWS, or OCI)
  • Knowledge of automation through scripting
  • Experience with CI/CD pipelines
  • Solid knowledge of networking and protocols
  • Experience with infrastructure security, IAM, TLS, and secrets management
  • Familiarity with logging platforms such as the ELK Stack or Loki
  • Experience in incident management, reliability, and continuous improvement
  • Analytical, collaborative profile with a problem-solving orientation
  • Differentials: Experience with Service Mesh; Knowledge of resilient and distributed architectures; Experience with performance and load testing; Experience in multi-cloud environments; Knowledge of FinOps and cloud cost optimization; Experience with Apache Kafka; Knowledge of advanced deployment strategies; Experience as a technical reference or technical lead for teams; Cloud, Kubernetes, or DevOps certifications.
Core Competencies

Demonstrates expertise in Site Reliability Engineering practices, including managing high-criticality Linux/Unix environments and implementing Infrastructure as Code using Terraform or similar tools. Proficient in CI/CD pipeline creation, observability solutions, and incident management to drive operational excellence.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • São Paulo

Presencial
BRL 180 000 - 260 000
Mid-level SRE
Mid-level SRE

Jobtailor • São Paulo

Presencial
BRL 180 000 - 240 000
Senior SRE
Senior SRE

Jobtailor • São Paulo

Presencial
BRL 180 000 - 260 000
Senior Staff Engineer - DevOps
Senior Staff Engineer - DevOps

Nagarro • Rio de Janeiro

Presencial
BRL 120 000 - 180 000
Senior DevOps Analyst
Senior DevOps Analyst

Jobtailor • Uberlândia

Presencial
BRL 180 000 - 320 000
DevOps Engineer I
DevOps Engineer I

Jobtailor • Blumenau

Presencial
BRL 90 000 - 150 000
Senior SRE, Specialist
Senior SRE, Specialist

Jobtailor • São Paulo

Presencial
BRL 180 000 - 320 000
Senior SRE / Infrastructure Engineer
Senior SRE / Infrastructure Engineer

Jobtailor • São Paulo

Presencial
BRL 240 000 - 360 000
Senior Reliability and Platform Engineer
Senior Reliability and Platform Engineer

Jobtailor • São Paulo

Presencial
BRL 180 000 - 300 000
DevOps Coordinator – SRE
DevOps Coordinator – SRE

Jobtailor • São Paulo

Presencial
BRL 180 000 - 320 000