Site Reliability Engineer

Qualitest

Riverwoods (IL)

On-site

USD 110,000 - 130,000

Full time

27 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Diversity and inclusion
Internal rotation
Career progression
Casual culture
Tech academy
Referral bonuses

Job summary

QualityAI is seeking a Site Reliability Engineer to join our growing team in Riverwoods, IL. You will partner with development teams to build resiliency, implement SLIs/SLOs, and enhance end-to-end observability across critical payment and data platforms.

The role emphasizes automation, incident response, and capacity management in a fast-paced AI-first quality engineering environment. Ideal candidates have 6–12 years as an SRE, strong Linux and AWS/cloud experience, and proficiency with

Qualifications

  • Experience as SRE with SDLC Application delivery.
  • Ability to translate requirements into NFT Automation tests.
  • Experience with DevOps and CI/CD tools.
  • Good experience of Linux, AWS Cloud and on Prem deployments.

Responsibilities

  • Partner with Development teams to build resiliency for applications.
  • Implement service level objectives and observability.
  • Build out end-to-end observability and dashboards.
  • Implement monitoring, alerting and dashboards for apps.
  • Automate operational processes and capacity planning tools.
  • Define DR plans for critical apps.
  • Participate in on-call rotation and production incident support.
  • Develop chaos testing processes.

Skills

AWS
Python
Java
Go
Linux
Datadog
Kubernetes
Jenkins
OpenShift
Grafana
Kibana
JIRA
SNOW

Tools

Docker
OpenShift
Kubernetes
Grafana
Kibana
Datadog
Jenkins

Job description

Are you interested in working with the World’s leading AI-first Quality Engineering Company? Ready to advance your career, team up with global thought leaders across industries and make a difference every day? Join us at QualityAI!

We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States!

Responsibilities
  • Partner with Application Development teams to build resiliency for Payment application.
  • Partner with our Application Develop teams to implement service level objectives.
  • Partner with our Application Development teams and other SREs to build out end to end observability.
  • Implement monitoring, alerting and dashboards needed for our apps.
  • Automated operational processes.
  • Help to develop our capacity management and performance management tools.
  • Help to define the DR plan needed for our critical apps.
  • Participate in an on-call rotation and support production Incidents.
SRE Skillsets - Expectations from Pricing & Settlements team
  • Good understanding of hybrid infrastructure.
  • Expertise with AWS.
  • Expertise in one or more general purpose programming languages: Python, Go, shell scripting (Unix/Linux), Java.
  • Experience in CI/CD pipelines preferably Jenkins expertise.
  • Experience in container technology (OpenShift, Kubernetes).
  • Expertise in automation tools experience (preferably Ansible).
  • Expertise in observability tools including APM (Datadog), synthetic monitoring and log aggregation (Elk).
  • Experience in dashboarding tools such as Grafana and Kibana.
  • Understating of Agile concepts and experience in JIRA.
  • Basic understating of Release Management.
  • Hands-on experience on SNOW.
SRE Skillsets - Expectations from Data Platform team
  • Expertise in Message Broker (preferably Rabbit MQ, Kafka).
  • Expertise on Hadoop, spark commands JSON formatting.
Qualifications
  • 6-12 years of overall experience.
  • Professional experience as a Site Reliability Engineer (SRE).
  • Software development “hands on” engineer with excellent understanding of SDLC Application delivery.
  • Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience with DevOps, CI/CD tools.
  • Good experience of Linux, AWS Cloud and on Prem deployments.
  • Good experience in Systems Observability and APM tools, preferably Datadog.
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Strong JIRA knowledge.
Skills
  • Expertise in AWS-Lambda Services 4/5.
  • Strong programming skills (Java, Python, Shell Optional Java Script) 4/5.
  • Proficiency in Database concepts, Strong knowledge of SQL and experience with MySQL 4/5.
  • APM tools - Datadog 4/5.
  • Hands-on experience on SNOW.
Must have
  • Professional experience as a Site Reliability Engineer (SRE).
  • Experience in performance testing, Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience of AWS Cloud Application (Must for sure).
  • Experience of Linux, AWS Cloud(Must) and on Prem deployments.
  • Good experience in Systems Observability and APM tools, preferably Datadog.
  • Experience in dashboarding tools such as Grafana and Kibana.
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Expertise in one or more programming languages: Python, shell scripting (Unix/Linux), Java.
Nice to have
  • Hands-on experience on SNOW.
  • Experience in container technology (OpenShift, Kubernetes).
  • Strong JIRA knowledge.
  • Basic understating of Release Management.
  • Experience in CI/CD pipelines preferably Jenkins expertise.
Job description

Are you interested in working with the World’s leading AI-first Quality Engineering Company? Ready to advance your career, team up with global thought leaders across industries and make a difference every day? Join us at QualityAI!

We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States!

Responsibilities
  • Partner with Application Development teams to build resiliency for Payment application.
  • Partner with our Application Develop teams to implement service level objectives.
  • Partner with our Application Development teams and other SREs to build out end to end observability.
  • Implement monitoring, alerting and dashboards needed for our apps.
  • Automated operational processes.
  • Help to develop our capacity management and performance management tools.
  • Help to define the DR plan needed for our critical apps.
  • Help to develop a chaos testing process.
  • Participate in an on-call rotation and support production Incidents.
SRE Skillsets - Expectations from Pricing & Settlements team
  • Good understanding of hybrid infrastructure.
  • Expertise with AWS.
  • Expertise in one or more general purpose programming languages: Python, Go, shell scripting (Unix/Linux), Java.
  • Experience in CI/CD pipelines preferably Jenkins expertise.
  • Experience in container technology (OpenShift, Kubernetes).
  • Expertise in automation tools experience (preferably Ansible).
  • Expertise in observability tools including APM (Datadog), synthetic monitoring and log aggregation (Elk).
  • Experience in dashboarding tools such as Grafana and Kibana.
  • Understating of Agile concepts and experience in JIRA.
  • Basic understating of Release Management.
  • Hands-on experience on SNOW.
SRE Skillsets - Expectations from Data Platform team
  • Expertise in Message Broker (preferably Rabbit MQ, Kafka).
  • Expertise on Hadoop, spark commands JSON formatting.
Qualifications
  • 6-12 years of overall experience.
  • Professional experience as a Site Reliability Engineer (SRE).
  • Software development “hands on” engineer with excellent understanding of SDLC Application delivery.
  • Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience with DevOps, CI/CD tools.
  • Good experience of Linux, AWS Cloud and on Prem deployments.
  • Good experience in Systems Observability and APM tools, preferably Datadog.
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Strong JIRA knowledge.
Skills
  • Expertise in AWS-Lambda Services 4/5.
  • Strong programming skills (Java, Python, Shell Optional Java Script) 4/5.
  • Proficiency in Database concepts, Strong knowledge of SQL and experience with MySQL 4/5.
  • APM tools - Datadog 4/5.
  • Hands-on experience on SNOW.
Must have
  • Professional experience as a Site Reliability Engineer (SRE).
  • Experience in performance testing, Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience of AWS Cloud Application (Must for sure).
  • Experience of Linux, AWS Cloud(Must) and on Prem deployments.
  • Good experience in Systems Observability and APM tools, preferably Datadog.
  • Experience in dashboarding tools such as Grafana and Kibana.
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Expertise in one or more programming languages: Python, shell scripting (Unix/Linux), Java.
Nice to have
  • Hands-on experience on SNOW.
  • Experience in container technology (OpenShift, Kubernetes).
  • Strong JIRA knowledge.
  • Basic understating of Release Management.
  • Experience in CI/CD pipelines preferably Jenkins expertise.
Profile description

We offer:

Benefits
  • Be a part of a company who strives to support for diversity and inclusion in the workplace - we are one, we are many at QualityAI. Celebrate culture, share knowledge with engineers from around the globe, and inspire each other through our differences.
  • Local and global opportunities - we offer you internal rotation and international mobility opportunities to grow your career.
  • Clear view of your career and progression with the company - QualityAI is growing massively (since Jan 2021 - added more than 2000 engineers) and giving you the opportunity to grow with us.
  • Work hard and play harder with our flexible and casual culture. Take a break from work and join an employee event, or enjoy the amenities and games provided from one of our Employees Centers.
  • Never stop experimenting and learning with QualityAI Tech academy: 3000+ training courses, mentorship programs, technical tribes, sponsored certifications, leadership programs and much more.
  • Earn bonuses via our Client Referral and Employee Referral Program’s. Refer and earn - tap your network for net-worth.
  • A Competitive pay, the salary range for the role is $110,000 - $130,000.
Why QualityAI?

QualityAI is an AI-first quality engineering company helping enterprises deploy and scale complex systems with greater confidence. Operating across data, models, platforms, infrastructure, and operational environments, the company provides assurance and engineering expertise that helps organizations ensure systems perform reliably in real-world conditions.

Formerly Qualitest, QualityAI supports global enterprises across regulated and technology-driven industries, combining deep engineering heritage with AI-enabled delivery, operational assurance, and lifecycle expertise to help clients achieve certainty at go-live.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Quality Ai • Northern (KY)

Hybrid
USD 110,000 - 130,000
Competitive pay
Global opportunities
Technical training & certification
Sr IT Project Manager - Agile Delivery & Operations
Sr IT Project Manager - Agile Delivery & Operations

Qualitest • Alameda (CA)

Hybrid
USD 120,000 - 140,000
Support Tech - Site Generalist
Support Tech - Site Generalist

Quality Ai • Burlingame (CA)

On-site
USD 42,000
Quantitative Research Analyst
Quantitative Research Analyst

Quality Ai • Northern (KY)

Hybrid
USD 120,000 - 160,000
Diversity and inclusion programs
Internal rotation and internationalMob
Tech academy and training
Gen AI Architect
Gen AI Architect

Qualitest • United States

On-site
USD 180,000 - 200,000
Diversity & inclusion
International mobility opportunities
Career progression
+4
Verification Engineer
Verification Engineer

Qualitest • Ball Ground (GA)

On-site
USD 80,000 - 100,000
Diversity and inclusion
Internal rotation opportunities
Career progression
+1
Sr QA Engineer (automation)
Sr QA Engineer (automation)

Qualitest • San Diego (CA)

On-site
USD 80,000 - 85,000
Diversity and inclusion focus
Internal rotation opportunities
Global mobility opportunities
+5
Sr QA Engineer (Automation)
Sr QA Engineer (Automation)

Quality Ai • San Diego (CA)

On-site
USD 80,000 - 85,000
Diversity & inclusion
Global opportunities
Career progression
+3
Data Engineer
Data Engineer

Quality Ai • Northern (KY)

Hybrid
USD 110,000 - 120,000
401k plan
Healthcare benefits
Flexible and casual culture
Sr. Quality Engineer- First Dollar (Remote)
Sr. Quality Engineer- First Dollar (Remote)

Inspira Financial • Oak Brook (IL)

On-site
USD 91,000 - 111,000
Healthcare
401K savings plan
Paid time off
+2