CardWorks, Inc. is seeking a Site Reliability Engineer to lead operational improvements in uptime and scalability in Pittsburgh, PA. The role requires expertise in AI/ML for operational workflows, strong observability skills, and a deep understanding of Infrastructure as Code. Candidates should hold a Master’s degree and have over 7 years of experience in SRE. This position offers a hybrid work model, competitive salary ranging from $146,032 to $162,257, and a comprehensive benefits package including health insurance and retirement plans.
Qualifications
7+ years of experience in Site Reliability Engineering.
Strong observability and telemetry expertise.
Proven ability to define and manage Service Level Indicators.
Responsibilities
Establish the SRE operating model for cross-team adoption.
Define and manage reliability metrics and error budgets.
Oversee incident and problem management processes.
Skills
Site Reliability Engineering
AI/ML Operations
Infrastructure as Code
CI/CD Pipeline Design
Containerization
Education
Master’s degree in Computer Science or equivalent
Tools
Terraform
Ansible
Azure DevOps
Docker
Kubernetes
Job description
CardWorks, Inc. is seeking a Site Reliability Engineer to lead operational improvements in uptime and scalability in Pittsburgh, PA. The role requires expertise in AI/ML for operational workflows, strong observability skills, and a deep understanding of Infrastructure as Code. Candidates should hold a Master’s degree and have over 7 years of experience in SRE. This position offers a hybrid work model, competitive salary ranging from $146,032 to $162,257, and a comprehensive benefits package including health insurance and retirement plans.