Site Reliability Engineer (Middle) ID38916

AgileEngine

São José dos Campos

Híbrido

BRL 80 000 - 100 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Professional growth opportunities
Competitive USD-based compensation
Flexible schedule
Budgets for education and fitness

Resumo da oferta

A leading software company is seeking a Mid-Senior AWS Cloud Engineer in São José dos Campos, Brazil. The role involves managing alerts, providing 24×7 support, and deploying to EKS/K8s. Candidates should have over 2 years of experience in AWS Cloud Engineering, proficiency with Datadog, and strong communication skills. The position offers flexible working options and contributions towards professional growth.

Qualificações

  • 2+ years of professional experience.
  • Experience working with Datadog.
  • Hands-on experience with AWS as a Cloud Engineer.
  • Good understanding of AWS IAM roles and policies.
  • Experience logging AWS resources with CloudWatch.
  • Excellent communication skills, both written and verbal.

Responsabilidades

  • Manage alerts and check systems.
  • Provide 24×7 on-call support for critical SaaS events.
  • Document issues and remediation steps.
  • Deploy to EKS/K8s cluster using Terraform and Helm.
  • Collaborate with teams to ensure high-level support.

Conhecimentos

AWS Cloud Engineering
Datadog
EKS/Terraform/Helm
Docker and Docker Swarm
Bash and/or Python scripting
REST APIs
Grafana and Prometheus
Linux Environment
Customer-facing communication

Descrição da oferta de emprego

Overview

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

What you will do
  • Shift: Monday – Thursday 8AM – 7PM PST (11AM – 10PM EST) with rotating on-call
  • On-call shifts: every 6 weeks, for one week as primary responder and the next week as secondary
  • Manage alerts daily, check systems, and escalate issues as needed
  • Be part of a team that provides 24×7 on-call support for critical SaaS events
  • Be available in case of emergencies when team members are not available or need help
  • Document issues and remediation steps
  • Proactively create appropriate monitors in the EKS/K8S ecosystem
  • Deploy to EKS/K8s cluster using Terraform and Helm
  • Learn and maintain existing infrastructure running under Docker Swarm
  • Improve existing infrastructure health by implementing checks and scripts to correct known issues
  • Maintain and develop deployment code
  • Automate manual tasks
  • Implement/integrate new technologies in our Cloud Infrastructure
  • Collaborate with other teams and departments to provide the highest level of support and assistance
  • Apply a real customer focus when planning deployments/updates, having the customer in the forefront of the mind, and considering the impact on them before making changes
  • Work closely with Support, Customer Success, Migration, and Professional Services teams to provide the best in class SaaS service to our customers
  • Perform RCA and take necessary corrective actions to prevent recurrence of issues
  • Create and assign alert-related actions to the appropriate team after the investigation
  • Handle support requests for environment-specific actions
  • Identify and provide automation requirements to improve RCA
MUST HAVES
  • 2+ years of professional experience
  • Experience working with Datadog
  • Hands-on experience as an AWS Cloud Engineer
  • Working knowledge of EKS/Terraform/Helm
  • Working experience with Docker and Docker Swarm
  • Good understanding of AWS IAM roles and policies
  • Experience logging and monitoring AWS resources using CloudWatch logs
  • Experience working in a Linux environment
  • Proficient in Bash and/or Python scripting
  • A strong understanding of web technologies such as REST APIs
  • Working experience with monitoring solutions, such as Grafana and Prometheus
  • Excellent oral and written communication skills
  • Customer-facing communication skills to effectively explain issues and RCAs to them
  • Experience in Product/Application Support for SaaS-based products
  • Understanding of APIs, Databases, Systems Architecture, and Design
  • Designing, implementing, and operating in a DevSecOps environment
  • Excellent communication skills, both written and verbal
  • Ability to work independently as well as within a collaborative environment
  • A technical aptitude with the desire to learn new and evolving technologies
  • Upper-Intermediate English level
NICE TO HAVES
  • Experience with GCP or Azure
  • Certifications: AWS Certified DevOps Engineer – Professional or AWS Certified Advanced Networking Specialty
PERKS AND BENEFITS
  • Professional growth: Accelerate your professional journey with mentorship, TechTalks, and personalized growth roadmaps
  • Competitive compensation: Competitive USD-based compensation and budgets for education, fitness, and team activities
  • A selection of exciting projects: Projects with modern solutions development and top-tier clients including Fortune 500 enterprises
  • Flextime: Flexible schedule with options to work from home or office to suit your productivity
Seniority level
  • Mid-Senior level
Employment type
  • Full-time
Industries
  • IT Services and IT Consulting

Referrals increase your chances of interviewing at AgileEngine by 2x

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • São Bernardo do Campo

Híbrido
BRL 517 000 - 673 000
Professional growth
Competitive compensation
Exciting projects
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Brasília

Híbrido
BRL 414 000 - 622 000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • São Bernardo do Campo

Híbrido
BRL 250 000 - 360 000
Professional growth
Competitive compensation
Exciting projects
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Florianópolis

Híbrido
BRL 414 000 - 622 000
Professional growth
Competitive compensation
Exciting projects
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Recife

Híbrido
BRL 180 000 - 250 000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Recife

Híbrido
BRL 298 000 - 399 000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Site Reliability Engineer ID55632
Site Reliability Engineer ID55632

AgileEngine • São Paulo

Híbrido
Professional growth opportunities
Competitive USD-based compensation
Exciting projects with top companies
+1
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Riograndina

Híbrido
BRL 385 000 - 551 000
Professional growth
Competitive compensation
Flextime
+1
DevOps Engineer ID56470
DevOps Engineer ID56470

AgileEngine • São Paulo

Híbrido
BRL 396 000 - 595 000
Professional growth through mentorship
USD-based competitive compensation
Exciting projects with Fortune 500 companies
+1
Technical Lead ID71009
Technical Lead ID71009

AgileEngine, LLC. • São Paulo

Presencial
BRL 280 000 - 420 000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3