Epic Site Reliability Engineer II

Quest Diagnostics Incorporated

Secaucus (NJ)

Hybrid

USD 111,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Day 1 Medical, supplemental health, dental & vision
401(k) with company match
Annual health assessment program
Education assistance
Well-being programs
Vacation and Health/Flex Time
Employee stock purchase plan
Life and disability insurance

Job summary

Quest Diagnostics Incorporated is seeking a Performance II-Epic to provide reliability engineering services through observability and performance engineering techniques. The role supports high availability and system performance, ensuring incidents are resolved effectively within a hybrid work setting.

Ideal candidates should possess a strong background in scripting and cloud platforms, being part of an innovative team working toward operational excellence. The position offers a pay range of $111,000-130,000 plus bonuses.

Qualifications

  • 4+ years of experience with multiple APM tools and extensive experience with Dynatrace.
  • 3+ years SRE experience.
  • Experience in software development, infrastructure, or operations roles.
  • Certifications in relevant technologies (AWS, DevOps, Kubernetes, etc.) required.
  • Working experience building CI/CD pipelines and version control systems.
  • Excellent problem-solving and communication skills.

Responsibilities

  • Implement and maintain robust observability solutions.
  • Proactively identify and address performance bottlenecks.
  • Conduct capacity planning exercises based on patterns and growth projections.
  • Develop and maintain automation scripts for routine tasks.
  • Document system architecture and configurations.
  • Work closely with software engineers to integrate observability tools.

Skills

Dynatrace
Prometheus
Grafana
AWS
Azure
GCP
Python
Bash
Go

Education

Bachelor's Degree in Computer Science, Engineering or related field

Tools

Neoload
Jmeter
Terraform
Ansible
Splunk

Job description

Job Description

As a Performance II-Epic, your role is to provide reliability engineering services through observability and performance engineering techniques. Using monitoring and performance tools to deliver detailed feedback to product owners and development teams. You will partner with Product Owners to define service level objectives and develop service level indicators. Collaborate with cross-functional teams to design, build, automate, and maintain scalable infrastructure. Your responsibilities will include ensuring high availability, monitoring system performance, and aiding support staff with resolving incidents. This role requires a strong background in scripting, cloud platforms, and a passion for optimizing operational efficiency. You will use Site Reliability Engineering practices to deliver a seamless user experience.

Pay Range: $111,000-130,000, plus yearly bonus

Salary offers are based on a wide range of factors including relevant skills, training, experience, education, and, where applicable, certifications obtained. Market and organizational factors are also considered. Successful candidates may be eligible to receive annual performance bonus compensation.

Remote: This position supporting Epic can be 100% remote if not located near a hub within certain criteria.

This position is hybrid and will require 3 days on site at one of the following Quest sites: Secaucus, NJ or Schaumburg, IL.

Benefits Information: We are proud to offer best-in-class benefits and programs to support employees and their families in living healthy, happy lives. Our pay and benefit plans have been designed to promote employee health in all respects physical, financial, and developmental. Depending on whether it is a part-time or full-time position, some of the benefits offered may include:

  • Day 1 Medical, supplemental health, dental & vision for FT employees who work 30+ hours
  • Best-in-class well-being programs
  • Annual, no-cost health assessment program
  • Blueprint for Wellness
  • healthyMINDS mental health program
  • Vacation and Health/Flex Time
  • 6 Holidays plus 1 MyDay off
  • FinFit financial coaching and services
  • 401(k) pre-tax and/or Roth IRA with company match up to 5% after 12 months of service
  • Employee stock purchase plan
  • Life and disability insurance, plus buy-up option
  • Flexible Spending Accounts Annual incentive plans
  • Matching gifts program
  • Education assistance through MyQuest for Education Career advancement opportunities and so much more!
Responsibilities
System Monitoring and Analysis
  • Implement and maintain robust observability solutions to monitor system performance, identifying bottlenecks, and ensuring optimal operation.
  • Utilize tools to gather, analyze, and visualize key performance metrics.
Performance Optimization
  • Proactively identify and address performance bottlenecks through in-depth analysis and optimization strategies.
  • Work closely with development teams to implement performance improvements and enhance overall system efficiency.
Capacity Planning
  • Conduct capacity planning exercises based on observed patterns and future growth projections.
  • Collaborate with infrastructure and development teams to ensure adequate resources are available to meet system demands.
Automation and Scripting
  • Develop and maintain automation scripts for routine tasks, enabling efficient monitoring and response procedures.
  • Implement automated processes for scaling and provisioning resources based on observed workload patterns.
Documentation
  • Document system architecture, configurations, and observability best practices to facilitate knowledge transfer and onboarding for team members.
  • Keep documentation up-to-date to reflect changes in the system and its monitoring setup.
Collaboration with Development Teams
  • Work closely with software engineers to integrate observability tools into the development lifecycle.
  • Provide guidance on building observable systems and assist in instrumenting applications for effective monitoring.
Continuous Improvement
  • Stay informed about industry best practices and emerging technologies related to observability and performance engineering.
  • Drive continuous improvement initiatives to enhance the reliability and performance of systems.
Security and Compliance
  • Collaborate with security teams to implement monitoring and observability measures that align with security requirements and compliance standards.
  • Participate in security incident response activities and contribute to ongoing security assessments.
Training and Knowledge Sharing
  • Conduct training sessions for team members and other stakeholders on observability tools, best practices, and performance engineering concepts.
  • Foster a culture of knowledge sharing within the organization.

And other duties as assigned.

Qualifications
Required Work Experience
  • 4 plus years of experience with multiple APM tools and extensive experience with Dynatrace
  • 3 plus years SRE experience
  • Experience in software development, infrastructure, or operations roles
  • Certifications in relevant technologies (e.g. AWS, DevOps, Kubernetes, Dynatrace, Azure, etc.)
  • Working experience building CI/CD pipelines and version control systems
  • Working experience with scripting languages (e.g. Python, Bash, Go, etc.)
  • Excellent problem-solving and communication skills.
  • Ability to work collaboratively in a fast-paced, agile environment.
Preferred Work Experience
  • Working experience with Neoload, Jmeter or equivalent performance testing tool.
  • Experience executing software load and performance testing in an enterprise environment.
  • Experience testing applications hosted in the cloud.
  • Experience with infrastructure as code tools such as Terraform or CloudFormation.
  • Deep understanding of Linux systems administration and networking principles.
  • Experience with containerization and orchestration technologies such as Docker and Kubernetes.
  • Experience or familiarity with IIS, HTML, Java, Jboss.
  • Experience in Chaos Engineering
  • Programming experience using.NET, C, C++, Java, or other popular programming languages. Perl/Python/JavaScript scripting experience may be considered equivalent.
  • Terraform and Ansible experience.
  • Exposure to Splunk tools.
  • Exposure to microservices.
  • Dynatrace Certifications
  • AWS/Azure/GCP Certifications
  • Chaos Engineering Certifications
  • Agile Certifications
Physical and Mental Requirements
  • Ability to sit/stand for long periods of time.
  • Ability to handle high stress situations.
  • Ability to lift up to 50 lbs.
Knowledge
  • Site Reliability Engineering Principles
  • DevSecOps Principles
  • Agile (SAFe)
  • Healthcare industry
  • ITLT
  • ServiceNow
  • Jira/Confluence
Skills
  • Dynatrace/Prometheus/Grafana
  • Neoload/Jmeter
  • Splunk
  • AWS/Azure/GCP
  • SAFe Agile
  • Strong communication skills (written/verbal)
  • Time management
  • Analytic problem solver
  • Self-starter
  • Result oriented and proven ability in organizing priorities
Education
  • Bachelor's Degree Bachelor's degree in Computer Science, Engineering, or a related field (Required)
Licenses and Certifications
  • Agile Certification (Project Management) (Preferred)

Quest Diagnostics honors our service members and encourages veterans to apply.

While we appreciate and value our staffing partners, we do not accept unsolicited resumes from agencies. Quest will not be responsible for paying agency fees for any individual as to whom an agency has sent an unsolicited resume.

Equal Opportunity Employer: Race/Color/Sex/Sexual Orientation/Gender Identity/Religion/National Origin/Disability/Vets or any other legally protected status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Epic Site Reliability Engineer II
Epic Site Reliability Engineer II

QUEST DIAGNOSTICS INC • Secaucus (NJ)

Hybrid
USD 111,000 - 130,000
Medical, dental & vision insurance
401(k) with company match
Education assistance
+2
Performance Engineer II - Epic
Performance Engineer II - Epic

Quest Diagnostics • Secaucus (NJ)

On-site
USD 111,000 - 130,000
Day 1 Medical insurance
401(k) with company match
Flexible Spending Accounts
+1
Epic Site Reliability Engineer II
Epic Site Reliability Engineer II

Quest Diagnostics • Secaucus (NJ)

Hybrid
USD 85,000 - 110,000
Medical benefits
401(k) match
Education assistance
Epic Site Reliability Engineer II
Epic Site Reliability Engineer II

Quest Diagnostics Incorporated • Schaumburg (IL)

Hybrid
USD 85,000 - 110,000
Medical benefits
401(k) match
Employee stock purchase plan
+1
Epic Cloud Network Engineer
Epic Cloud Network Engineer

QUEST DIAGNOSTICS INC • Secaucus (NJ)

Hybrid
USD 112,000 - 130,000
Day 1 Medical
Well-being programs
Annual incentive plans
+5
Senior SRE (Site Reliability Engineer)
Senior SRE (Site Reliability Engineer)

Vytwo • Dallas (TX)

Hybrid
USD 130,000 - 160,000
Flexible work from home options
Senior SRE: Dynatrace & Azure Observability Expert
Senior SRE: Dynatrace & Azure Observability Expert

RaceTrac • Atlanta (GA)

On-site
DevSecOps Engineer II
DevSecOps Engineer II

Quest Diagnostics Incorporated • Schaumburg (IL)

Hybrid
USD 120,000 - 180,000
401(k) match appreciation
Education assistance
Flexible Spending Accounts
+3
Epic Sr. Certification Engineer
Epic Sr. Certification Engineer

Quest Diagnostics Incorporated • Wood Dale (IL)

Hybrid
USD 90,000 - 130,000
Health insurance
401(k) with company match
Education assistance
Senior Site Reliability Engineer (SRE) – Dynatrace & Azure Observability Expert
Senior Site Reliability Engineer (SRE) – Dynatrace & Azure Observability Expert

RaceTrac • Atlanta (GA)

On-site