Senior Site Reliability Engineer

DigiCert

Lehi (UT)

On-site

USD 125,000 - 145,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation and comprehensive health, dental, and vision coverage
Retirement savings programs with company matching (401(k) or RRSP)
Generous paid time off and holidays
Paid parental leave and family support benefits
Flexible spending and health savings options
Health and wellness support, gym reimbursement
Education assistance and professional development opportunities

Job summary

DigiCert is seeking a Site Reliability Engineer to embed reliability, scalability, and performance best practices. You will collaborate closely with development teams to ensure mission-critical systems remain resilient and highly available.

The position offers an annual base salary range of $125,000 — $145,000 and comprehensive benefits, including health coverage and retirement savings programs.

Qualifications

  • Extensive experience in distributed systems and cloud-native architectures.
  • Strong scripting and automation skills in Python, Go, Bash, or similar languages.
  • Ability to troubleshoot complex production issues efficiently.

Responsibilities

  • Design and build fault-tolerant, high-performing systems.
  • Implement monitoring, alerting, and logging.
  • Act as a first responder for production incidents.
  • Develop self-healing, automated deployments.
  • Improve CI/CD pipelines for reliable feature rollouts.
  • Ensure production environments meet security and compliance requirements.

Skills

Distributed systems
Cloud-native architectures (AWS, GCP, Azure)
DevOps practices
Kubernetes
Terraform
CI/CD pipelines
Infrastructure as Code (IaC)
Scripting skills (Python, Go, Bash)
Observability tools (Prometheus, Grafana)
Troubleshooting complex production issues

Job description

Who we are

DigiCert is a global leader in intelligent trust. We protect the digital world by ensuring the security, privacy, and authenticity of every interaction. Our AI‑powered DigiCert ONE platform unifies PKI, DNS, and certificate lifecycle management, to secure infrastructure, software, devices, messages, AI content and agents. Learn why more than 100,000 organizations, including 90% of the Fortune 500, choose DigiCert to stop today’s threats and prepare for a quantum‑safe future at www.digicert.com

Job summary

The Site Reliability Engineer (SRE) collaborates with development teams to embed reliability, scalability, and performance best practices throughout the software development lifecycle. This role bridges software engineering and cloud operations, ensuring mission‑critical systems remain highly available and resilient. By integrating reliability early, the SRE fosters a culture of shared responsibility while enabling rapid and safe feature delivery.

What you will do
  • Design and build fault‑tolerant, high‑performing systems that meet Service Level Objectives (SLOs) and Service Level Agreements (SLAs).
  • Implement monitoring, alerting, distributed tracing, and logging to ensure real‑time system health visibility and proactive issue resolution.
  • Act as a first responder for production incidents, conduct blameless postmortems, and drive root cause analysis (RCA) and corrective actions.
  • Develop self‑healing, automated deployments, and scaling solutions to minimize toil and improve system efficiency.
  • Improve continuous integration and deployment pipelines to enable safe, rapid, and reliable feature rollouts.
  • Review code, debug issues, and perform quality assurance (QA) on software components to enhance system reliability and performance.
  • Work closely with development teams to ensure best practices in software architecture, coding standards, and operational readiness.
  • Forecast scalability needs and optimize cloud infrastructure costs while balancing performance and efficiency.
  • Ensure production environments meet security and compliance requirements, collaborating with teams to mitigate vulnerabilities and enforce best practices.
  • Work closely with development teams to embed reliability at every stage rather than treating it as an afterthought.
  • Use error budgets to balance feature velocity with system stability.
  • Implement observability and automation‑first principles to measure system health and drive continuous improvement.
  • Leverage game days, chaos engineering, and resilience testing to validate system robustness and refine operational processes.
What you will have
  • Extensive experience in distributed systems, cloud‑native architectures (AWS, GCP, Azure), and DevOps practices.
  • Proficiency in Kubernetes, Terraform, CI/CD pipelines, and Infrastructure as Code (IaC).
  • Strong scripting and automation skills in Python, Go, Bash, or similar languages.
  • Expertise in observability tools such as Prometheus, Grafana, Datadog, Splunk, New Relic, and OpenTelemetry.
  • Ability to troubleshoot complex production issues and drive scalable, resilient solutions.
  • Experience reviewing code, debugging applications, and conducting software testing to ensure high reliability and quality.
Benefits
  • Competitive compensation and comprehensive health, dental, and vision coverage
  • Retirement savings programs with company matching (401(k) or RRSP)
  • Generous paid time off, including holidays, and vacation
  • Paid parental leave and family support benefits
  • Life and disability coverage
  • Flexible spending and health savings options (where applicable)
  • Health and wellness support, including gym reimbursement and wellness programs
  • Employee Assistance Program with 24/7 confidential support for employees and families
  • Education assistance and professional development opportunities
  • Access to LinkedIn Learning and continuous learning resources
  • Employee referral bonus program and additional company perks and discounts
  • Internal rewards and recognition platform (Motivosity) to celebrate and acknowledge project wins, milestone achievements, and the outstanding contributions of our colleagues
  • Business travel insurance and global employee support programs
Equal Opportunity Employer

DigiCert is an Equal Opportunity employer and is committed to diversity in its workforce. In compliance with applicable federal and state laws, DigiCert prohibits discrimination on the basis of race or ethnicity, religion, color, national origin, sex, age, sexual orientation, gender identity/expression, veteran’s status, status as a qualified person with a disability, or genetic information. Individuals from historically underrepresented groups, such as minorities, women, qualified person with disabilities, and protected veterans are strongly encouraged to apply.

Compensation Transparency

The annualized base salary range for this position is outlined below.

Each candidate’s compensation offer will be determined based on factors including experience, skills, qualifications, job duties, business needs, and location. For roles that include additional compensation components, total compensation may include base pay, bonus, equity, or other incentives.

This role may also be eligible for benefits, which will be discussed during the hiring process. We are committed to fair and transparent pay practices and comply with all applicable pay transparency requirements. If you would like more information about compensation or benefits, we are happy to provide additional details during the hiring process.

For more information regarding our comprehensive benefits, see the benefits section.

Base Salary

$125,000 — $145,000 USD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

Hybrid
USD 140,000 - 150,000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1
Senior Devops Engineer
Senior Devops Engineer

DigiCert • United States

On-site
USD 110,000 - 150,000
Competitive compensation and comprehensive health coverage
Retirement savings programs with matching
Generous paid time off
+1
Devops Engineer
Devops Engineer

DigiCert • United States

On-site
USD 130,000 - 150,000
Health insurance
401(k) with company matching
Paid time off
+6
Senior Devops Engineer
Senior Devops Engineer

DigiCert • Lehi (UT)

On-site
USD 110,000 - 140,000
Health, dental, and vision coverage
Retirement savings programs with company matching
Generous paid time off
Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX INC • Chicago (IL)

On-site
USD 124,200 - 230,800
Health Benefits
401(k) with match
Wellness platform
+1
Site Reliability Engineer
Site Reliability Engineer

VantageScore® • San Francisco (CA)

On-site
USD 150,000
Medical insurance
Dental insurance
401(k) plan
+1
Lead, Site Reliability Engineer
Lead, Site Reliability Engineer

CardWorks • Pittsburgh

Hybrid
USD 146,000 - 163,000
Competitive Pay
Medical, Dental, and Vision Benefits
401(k) Plan with Company Match
+1
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Oaks (PA)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare coverage
401(k) matching
Tuition reimbursement
+1