Staff Site Reliability Engineer

Altium

San Diego (CA)

On-site

USD 170,000 - 210,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive benefits package

Job summary

Altium seeks a Senior Site Reliability Engineer to ensure reliability, availability, and performance of the Altium Cloud Platform. You will automate operations, enhance observability, and contribute to incident management while collaborating with development and technology teams to build scalable applications.

Join Altium to help standardize reliability patterns across regions, drive IaC initiatives, and improve deployment efficiency in a fast-paced SaaS environment.

Qualifications

  • 6+ years in SRE, DevOps or related role in a large-scale environment.
  • 3+ years professional software development experience.
  • Experience with .NET development is a plus.
  • Strong understanding of SDLC, microservices and HA.
  • Familiarity with observability tools and incident management.
  • Proficient with Kubernetes in production and AWS.
  • Knowledge of CI/CD tooling and IaC (Terraform, Ansible).
  • Basic networking knowledge; relational DB experience is a plus.
  • US permanent resident (Citizen or Green Card).

Responsibilities

  • Understand how the Altium Cloud Platform works.
  • Improve observability with logging, monitoring and APM.
  • Develop reliability frameworks for SaaS products across regions.
  • Educate engineering teams on reliability best practices.
  • Contribute to incident response and post-incident analysis.
  • Participate in system design, platform management and capacity planning.
  • Collaborate with DevOps to enhance automation and deployment.
  • Promote service ownership principles and participate in on-call rotations.

Skills

SRE/DevOps experience
Software development
Observability
System design/HA
On-call & incident response
Collaboration with engineering teams

Tools

Kubernetes
AWS
NewRelic / ELK / Grafana / PagerDuty / OTEL
CI/CD: Jenkins / GitLab / GitHub / ArgoCD
Terraform
Ansible
Relational databases (MySQL / PostgreSQL)

Job description

  • Compensation: USD 170,000 - USD 210,000 - yearly
Job Description

Senior Site Reliability Engineerensures the reliability, availability, and performance of large-scale software systems through a blend of software engineering and systems administration. Key responsibilities involve automating operational tasks,improving observability, andcontributing to incident management, while also collaborating with developmentand technologyteams to build more reliable and scalable applications.

Join Altium as a Senior Site Reliability Engineer to ensure the reliability and performance of the Altium Cloud Platforms.

Key Responsibilities:

  • Understanding how an Altium Cloud Platform works
  • Pioneer improvements in observability, including logging, monitoring, and application performance management (APM), ensuring system reliability and proactive issue detection.
  • Develop and implement reliability frameworks and patterns that standardize and elevate the resilience of our SaaS products across multiple regions and environments.
  • Cultivate a shared responsibility model where the SRE team collaborates with and educates engineering teams on reliability best practices.
  • Contribute to incident response and management, ensuring rapid resolution, clear stakeholder communication, and post-incident analysis for continuous improvement.
  • Participate in system design consulting, platform management,infrastructureupgradesand capacity planning.
  • Partner closely with engineering and development teams to enhance product stability, observability, and manageability through best practices in reliability engineering.
  • Partner closely with DevOps/Operations, drive automation initiatives, promote Infrastructure as Code (IaC), and streamline deployment processes to improve operational efficiency and scalability.
  • Champion Service-Oriented Organization (SOO) principles to ensure accountability and clarity in service ownership.
  • Participate in on-call rotation and drive operational improvements after incidents.
Qualifications
  • 6+ years in SRE, DevOps or related role in a large-scale environment
  • 3+ years professional experience in software development
  • Software development experience(ideally working with and as a .NET developer)
  • Strong understanding of SDLC, microservice and HA architecture
  • Observability - NewRelic, ELK, Grafana, PagerDuty, OTEL or similar
  • Experience with Kubernetes clusters in production setting, AWS, IOC
  • Experience with operational tasks
  • Knowledge of CI-CD tooling Jenkins, Gitlab, GitHub, ArgoCD or similar
  • Knowledge of IaaC Terraform, Ansible
  • Basic knowledge of networking fundamentals
  • Experience with relational databases (mysql, postgres) as a plus
  • MUST be a US permanent resident (Citizen or Green Card)
Additional Information

Altium Limited, a part of the Renesas Group and headquartered in San Diego, California, is a global software company accelerating the pace of electronics innovation. We are redefining electronic product creation in a software-defined world with our industry-first cloud-based platform that unites every stakeholder and phase ofelectronicsdevelopment.

From startups to world’s technology giants,our digital platforms give more power to PCB designers, supply chain, and manufacturing, letting them collaborate as never before. At Altium, our teams are empowered to innovate, collaborate globally, and help create the future of electronics development.

  • We believe in rewarding our employees with a competitive benefits package alongside their salary. More information will be provided during the hiring process.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer
Staff Site Reliability Engineer

Renesas Electronics • San Diego (CA)

On-site
USD 170,000 - 210,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Renesas Electronics Corporation • San Diego (CA)

On-site
USD 120,000 - 180,000
Senior Product Support Engineer
Senior Product Support Engineer

Altium • Frisco (TX)

On-site
USD 90,000 - 120,000
Bonus opportunities
Benefits package
Senior Product Support Engineer
Senior Product Support Engineer

Renesas Electronics • San Diego (CA)

On-site
USD 100,000 - 130,000
Medical benefits
Health Savings Account
Dental
+4
Senior SRE: Cloud Reliability, Observability & Automation
Senior SRE: Cloud Reliability, Observability & Automation

Renesas Electronics • San Diego (CA)

On-site
USD 170,000 - 210,000
Senior Product Support Engineer
Senior Product Support Engineer

Renesas Electronics • La Jolla Ranch (CA)

On-site
USD 100,000 - 130,000
Medical
Dental
Vision
+5
Senior Product Support Engineer
Senior Product Support Engineer

Renesas Electronics Corporation • Frisco (TX)

On-site
USD 90,000 - 120,000
Medical benefits
Life insurance / AD&D
Paid time off
Senior Implementation Engineer
Senior Implementation Engineer

Renesas Electronics Corp. • Town of Texas (WI)

On-site
USD 102,000 - 138,000
Medical insurance
Health savings account
Dental insurance
+6
Senior Product Support Engineer
Senior Product Support Engineer

Renesas Electronics Corporation • La Jolla Ranch (CA)

On-site
USD 100,000 - 130,000
Bonus opportunities
Medical benefits
Health Savings Account
+7
Altium Enterprise Pre-Sales Field Application Engineer
Altium Enterprise Pre-Sales Field Application Engineer

Renesas Electronics • San Diego (CA)

On-site
USD 135,000 - 180,000
Medical, dental, vision benefits
Paid holidays & vacation