Destaca-te para esta função — gera um currículo e uma carta de apresentação personalizados em cerca de um minuto.
iTRTech Group in Brazil seeks a Senior Site Reliability Engineer to improve reliability, scalability, and operability of enterprise cloud environments. You will automate operations, enhance observability, and ensure high availability across mission-critical systems.
The role requires advanced English, 6+ years of experience, and strong SRE/DevOps skills with AWS/Azure/GCP, Terraform, Docker, Kubernetes, and CI/CD expertise.
SITE RELIABILITY ENGINEER (SRE) (HYBRID / REMOTE BRAZIL)
Brazilian company hires for hybrid or remote position
Location: Brazil (any location)
Only candidates already based in Brazil will be considered
Work Model: Hybrid for candidates living in state capitals and Remote for candidates living in countryside/cities outside the state capitals
Language Requirements: Advanced/Fluent English – Mandatory
Seniority: Senior (6+ years)
Compensation: Please inform your salary expectations when applying.
Build Highly Reliable Cloud Platforms at Enterprise Scale
We are looking for an experienced Site Reliability Engineer (SRE) to improve the reliability, scalability, performance, and operational excellence of enterprise cloud environments.
You will work closely with development, platform, and infrastructure teams to automate operations, improve observability, optimize deployments, and ensure high availability across mission-critical systems.
If you enjoy solving complex infrastructure challenges through engineering, automation, and modern cloud technologies, this opportunity is for you.
We are seeking a highly skilled Site Reliability Engineer with extensive experience in cloud infrastructure, DevOps practices, automation, and distributed systems.
The ideal candidate combines strong infrastructure knowledge with software engineering principles, helping organizations improve operational efficiency through automation, Infrastructure as Code (IaC), observability, and continuous improvement.
You should be comfortable working in high-availability environments, responding to critical incidents, and continuously enhancing platform resilience.
All requirements below are mandatory and eliminatory. Candidates who cannot clearly demonstrate these qualifications in their CV are unlikely to proceed in the recruitment process.
This position is intended for a Senior Site Reliability Engineer (SRE) with extensive experience in cloud infrastructure, automation, observability, DevOps, and platform reliability.
Candidates whose experience is primarily focused on traditional infrastructure administration, system support, or operations without demonstrated expertise in Infrastructure as Code, Kubernetes, cloud platforms, automation, CI/CD, and Site Reliability Engineering practices are unlikely to meet the expectations for this role.
Site Reliability Engineer, SRE, DevOps Engineer, Platform Engineer, Cloud Engineer, Infrastructure Engineer, Cloud Infrastructure, AWS, Amazon Web Services, Microsoft Azure, Google Cloud Platform, GCP, Terraform, Infrastructure as Code, IaC, Docker, Kubernetes, Linux, Unix, Python, Bash, Shell Scripting, CI/CD, Jenkins, GitLab CI, GitHub Actions, Monitoring, Observability, Prometheus, Grafana, ELK Stack, OpenTelemetry, Distributed Tracing, Logging, Metrics, SLI, SLO, SLA, Incident Management, Root Cause Analysis, RCA, Capacity Planning, Auto Scaling, Self-Healing, Disaster Recovery, Chaos Engineering, DevSecOps, Platform Engineering, Apache Kafka, High Availability, Scalability, Enterprise Infrastructure, Financial Services
#EY BR