Site Reliability Engineer

Infosys

Richardson (TX)

On-site

USD 80,000 - 120,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

An established industry player is seeking a Mid-Senior level professional to join their dynamic team. This role focuses on implementing end-to-end monitoring solutions and cloud technologies, utilizing industry-leading tools to enhance observability and automation. You will be responsible for developing innovative solutions, supporting critical incident resolutions, and leading a small team in a global delivery model. If you are passionate about cloud technologies and enjoy tackling complex challenges, this opportunity is perfect for you. Join a forward-thinking company that values creativity and technical excellence, and make a significant impact in the IT services sector.

Qualifications

  • Experience in implementing monitoring solutions and SLOs/SLIs.
  • Proficiency in cloud technologies and infrastructure as code.

Responsibilities

  • Develop observability solutions and support incident resolution.
  • Lead a small team and develop Proof of Concepts based on client needs.

Skills

Monitoring Solutions
SLOs and SLIs
Cloud Technologies (AWS, GCP, Azure)
Infrastructure as Code
Scripting Languages
Automation Tools
AI/ML-based Monitoring
Chaos Engineering

Tools

Datadog
Dynatrace
AppDynamics
New Relic
Terraform
Ansible

Job description

Observability: Implementing end-to-end monitoring solutions, implementing SLOs and SLIs for customer journeys, using industry tools like Datadog, Dynatrace, AppDynamics, etc.

DevSecOps: Setting up CD pipelines using tools.

Cloud Technologies: One of the major cloud technologies - AWS, GCP, or Azure – for key services – Compute, Storage, and Networking.

Infrastructure as Code: Solution design and implementation with industry tools like Terraform, Ansible, etc.

Scripting and Automation: Scripting languages and automation tools.

Preferred Skills:

  1. Develop observability solution implementations – monitoring, anomaly detection, alerting, and self-healing using industry tools like Datadog, Dynatrace, AppDynamics, New Relic, etc.
  2. Support critical incident resolution in a complex environment – applications hosted on cloud or datacenters, containerized applications, databases, etc.
  3. Set up SLOs and SLIs using industry-leading tools.
  4. Play the role of an individual contributor and lead a small team in a global delivery model.
  5. Develop Proof of Concepts (PoCs) and perform hands-on technical tasks based on client needs.
  6. Support responding to Requests for Proposal (RFPs) from clients.
  7. Analyze and identify improvement opportunities for automation and automate them.
  8. Experience in Implementing AI/ML-based monitoring and self-healing solutions.
  9. Experience in Implementing Chaos Engineering/testing.
Seniority level

Mid-Senior level

Employment type

Full-time

Job function

Consulting, Analyst, and Engineering

Industries

Information Services, IT Services and IT Consulting, and Software Development

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr Observability Engineer
Sr Observability Engineer

IT Associates • Irvine (CA)

On-site
USD 150,000 - 210,000
Senior Observability Engineer — AI-Driven Reliability
Senior Observability Engineer — AI-Driven Reliability

IT Associates • Irvine (CA)

On-site
Confidential
Site Reliability Engineer
Site Reliability Engineer

BlueSky Resource Solutions • Duluth (GA)

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Brooksource • San Antonio (TX)

On-site
USD 80,000 - 120,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • United States

On-site
USD 140,000 - 190,000
Site Reliability Engineer (Observability)
Site Reliability Engineer (Observability)

Cognizant • Town of Hartford (WI)

Hybrid
USD 132,000 - 152,000
Medical Insurance
PTO & Holidays
401(k) plan
+3
Sr. Site Reliability Engineer(Local to Atlanta GA Only)
Sr. Site Reliability Engineer(Local to Atlanta GA Only)

Trigint Solutions LLC • Atlanta (GA)

Hybrid
USD 124,000 - 220,000
Dynatrace Observability Consultant
Dynatrace Observability Consultant

Infosat IT Services LLC • United States

Remote
USD 150,000 - 190,000
SRE/Observability Engineer
SRE/Observability Engineer

Software Guidance & Assistance, Inc. (SGA, Inc.) • United States

On-site
USD 103,320 - 137,760
Technical Operations Lead
Technical Operations Lead

First Citizens Bank • Phoenix (AZ)

On-site
USD 140,000 - 190,000
Benefits program