Staff Site Reliability Engineer

Carrier

Town of Florida (NY)

On-site

USD 96,000 - 192,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health Care Benefits
Retirement Benefits
Paid time off
Training opportunities

Job summary

Carrier's BAS Cloud organization seeks a Senior Site Reliability Engineer to build, scale, and continuously improve our cloud-native SaaS platform. You will define reliability standards, drive automation, and enable engineering teams to operate services at scale.

You will collaborate with software engineering, architecture, cybersecurity, DevOps, and product teams to improve resilience, accelerate delivery, and reduce toil.

Qualifications

  • Bachelor’s degree in computer science, software engineering, information technology, or a technical field with 7+ years in SRE/Platform/DevOps/Cloud Engineering (or Master’s with 5+ years).
  • 3+ years of hands-on experience designing and supporting large-scale cloud environments.
  • 3+ years of experience with Infrastructure as Code (Terraform, CloudFormation, AWS CDK).
  • 3+ years of experience building and maintaining CI/CD pipelines and deployment automation.
  • 3+ years of extensive AWS experience (EKS, EC2, VPC, RDS, Lambda, CloudWatch, IAM).
  • 3+ years of observability tooling: Prometheus, Grafana, OpenTelemetry, Datadog, Splunk, or New Relic.

Responsibilities

  • Design, implement, and maintain highly available, scalable cloud infrastructure for Carrier's SaaS platforms.
  • Define and drive SLOs, SLIs, and SLAs across critical services.
  • Build self-service platform capabilities for development teams to deploy and operate services.
  • Develop reliability frameworks and best practices to improve platform resiliency and customer experience.
  • Lead incident response and perform root cause analyses for production events.

Skills

Python
Go
Bash
PowerShell
Cloud architecture
Incident response

Education

Bachelor's degree in Computer Science/Engineering/IT
Master's degree in Computer Science/Engineering/IT

Tools

Terraform
AWS
CloudFormation
AWS CDK
Kubernetes
Docker
Prometheus
Grafana
Datadog
New Relic

Job description

About Carrier

Carrier Global Corporation, global leader in intelligent climate and energy solutions, is committed to creating innovations that bring comfort, safety and sustainability to life. Through cutting-edge advancements in climate solutions such as temperature control, air quality and transportation, we improve lives, empower critical industries and ensure safe transport of food, lifesaving medicines and more. Since inventing modern air conditioning in 1902, we lead with purpose: enhancing the lives we live and the world we share. We continue to lead because of our world-class, inclusive workforce that puts the customer at the center of everything we do. For more information, visit corporate.carrier.com or follow on Carrier social media at @Carrier.

About Carrier

Carrier Global Corporation, global leader in intelligent climate and energy solutions, is committed to creating innovations that bring comfort, safety and sustainability to life. Through cutting-edge advancements in climate solutions such as temperature control, air quality and transportation, we improve lives, empower critical industries and ensure safe transport of food, lifesaving medicines and more. Since inventing modern air conditioning in 1902, we lead with purpose: enhancing the lives we live and the world we share. We continue to lead because of our world-class, inclusive workforce that puts the customer at the center of everything we do. For more information, visit corporate.carrier.com or follow on Carrier social media at @Carrier.

About The Role

Carrier's Building Automation Systems (BAS) Cloud organization is seeking a highly skilled Senior Site Reliability Engineer (SRE) to help build, scale, and continuously improve our cloud-native SaaS platform. This role combines software engineering, cloud infrastructure, platform engineering, observability, automation, and operational excellence to ensure world-class reliability, security, and performance for business-critical customer solutions. As a Senior SRE, you will play a key role in defining reliability standards, building self-service platform capabilities, driving automation, and enabling engineering teams to operate highly available services at scale. You will partner closely with software engineering, architecture, cybersecurity, DevOps, cloud operations, and product teams to improve system resilience, accelerate delivery, and reduce operational toil. This is an opportunity to influence the technical direction of a rapidly evolving cloud platform while helping establish engineering best practices across a global organization.

What You'll Do
Platform Engineering & Reliability
  • Design, implement, and maintain highly available, scalable, and secure cloud infrastructure supporting Carrier's SaaS platforms.
  • Define and drive Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Service Level Agreements (SLAs) across critical services.
  • Build self-service platform capabilities that enable development teams to deploy and operate services efficiently and consistently.
  • Develop reliability frameworks, standards, and operational best practices that improve platform resiliency and customer experience.
  • Identify reliability bottlenecks and architect solutions to eliminate single points of failure.
Automation & Infrastructure as Code
  • Develop automation solutions that eliminate repetitive operational tasks and reduce manual intervention.
  • Design and maintain Infrastructure as Code (IaC) using Terraform, AWS CloudFormation, AWS CDK, or similar technologies.
  • Build and enhance CI/CD pipelines to improve deployment velocity, consistency, and security.
  • Implement automated remediation, self-healing capabilities, and operational workflows.
Observability & Incident Response
  • Design and enhance modern observability solutions using metrics, logs, traces, and distributed monitoring technologies.
  • Develop dashboards, monitoring standards, alerting frameworks, and reliability reporting mechanisms.
  • Lead incident response efforts for critical production events, including root cause analysis and long-term corrective actions.
  • Drive post-incident reviews focused on systemic improvements rather than individual fault.
Cloud Operations & Resilience
  • Improve platform performance, availability, scalability, disaster recovery, and business continuity capabilities.
  • Validate backup, restoration, and recovery processes through regular testing and continuous improvement efforts.
  • Collaborate with engineering teams to implement resilient architectures capable of meeting defined reliability and compliance objectives.
  • Support operational readiness reviews for new services and platform capabilities.
Engineering Excellence & Leadership
  • Mentor engineers and provide technical leadership in reliability engineering practices.
  • Partner with development teams to improve application reliability throughout the software development lifecycle.
  • Establish and promote a culture of automation, ownership, continuous learning, and operational excellence.
  • Stay current on emerging technologies and industry best practices within cloud infrastructure, platform engineering, AI-assisted operations, and Site Reliability Engineering.
On-Call & Operational Excellence
  • Participate in a shared on-call rotation supporting production environments.
  • Drive initiatives focused on reducing operational toil through automation, self-healing systems, and platform improvements.
  • Continuously improve incident response processes and platform reliability practices.
Required Qualifications
  • Bachelor's degree in Computer Science, Software Engineering, Information Technology, or a technical field with 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering -OR- Master’s degree in Computer Science, Software Engineering, Information Technology, or a technical field with 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering.
  • 3+ years of hands-on experience designing and supporting large-scale cloud environments.
  • 3+ years of experience with Infrastructure as Code (such as Terraform, CloudFormation, AWS CDK).
  • 3+ years of experience building and maintaining CI/CD pipelines and deployment automation.
  • 3+ years of extensive experience with Amazon Web Services (AWS), including services such as EKS, EC2, VPC, RDS, Lambda, CloudWatch, and IAM.
  • 3+ years of experience implementing observability platforms utilizing technologies such as Prometheus, Grafana, Open Telemetry, Datadog, Splunk, or New Relic.
Preferred Qualifications
  • Strong understanding of modern SRE principles, including reliability engineering, observability, toil reduction, incident management, and error budgets.
  • Strong scripting or programming experience using Python, Go, PowerShell, Bash, or similar languages.
  • Strong knowledge of networking, security, systems architecture, and distributed systems concepts.
  • Proven experience leading production incident response and root cause analysis activities.
  • Proven experience with cloud-native platforms and container technologies, including Kubernetes and Docker.
  • Experience operating SaaS platforms serving large-scale customer environments.
  • Experience supporting compliance frameworks such as SOC 2, ISO 27001, NIST, or similar standards.
  • Experience implementing AI-assisted engineering solutions to improve operational efficiency, troubleshooting, automation, and service reliability.
  • Knowledge of platform engineering concepts including Internal Developer Platforms (IDP), developer self-service capabilities, and engineering enablement practices.
  • AWS certifications or other cloud certifications.
  • Excellent communication, leadership, and collaboration skills.
Pay Range

The annual salary for this position is between $96,000.00 - $192,000.00 annually. Factors which may affect pay within this range include, but are not limited to, skills, education, experience, and other unique qualifications of the successful candidate.

Other Compensation

Thisposition is entitled to short-term cash incentives, subject to plan requirements.

Benefits
  • Health Care Benefits: Medical, Dental, Vision; Wellness incentives
  • Retirement Benefits
  • Time off and Leave: Paid vacation days, up to 15 days; paid sick days, up to 5 days; paid personal leave, up to 5 days; paid holidays, up to 13 days; birth and adoption leave; parental leave; family and medical leave; bereavement leave; jury duty leave; military leave; purchased vacation
  • Disability: Short-term and long-term disability
  • Life Insurance and Accidental Death and Dismemberment
  • Tax-Advantaged Accounts: Health Savings Account; Health Care Spending Account; Dependent Care Spending Account
  • Tuition Assistance

To learn more about our benefits offering, please click here Work with us | Carrier Corporate.

Carrier EEO Statement and
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer
Staff Site Reliability Engineer

Carrier Global Corporation • Town of Florida (NY)

On-site
USD 96,000 - 192,000
Health Care Benefits
Retirement Benefits
Paid time off
Observability Technical Lead
Observability Technical Lead

Carrier • Town of Florida (NY), Northern (KY)

On-site
USD 96,000 - 192,000
Health Care Benefits: Medical, Dental,
Retirement Benefits
Paid vacation days
Associate Director, Business Partnership & Delivery
Associate Director, Business Partnership & Delivery

Carrier Global Corporation • Atlanta (GA)

Hybrid
USD 143,000 - 286,000
Health Care Benefits
Retirement Benefits
Paid Time Off
Associate Director of Start-Up and Commissioning
Associate Director of Start-Up and Commissioning

Carrier • Town of Florida (NY)

On-site
USD 143,000 - 286,000
Health Care Benefits
Retirement Benefits
Time off and Leave
+3
Data Engineer
Data Engineer

Carrier • Town of Florida (NY)

On-site
USD 65,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior HR Business Partner
Senior HR Business Partner

Carrier Global Corporation • Beverly (MA)

Hybrid
USD 117,000 - 235,000
Health Care Benefits
Retirement Benefits
Paid vacation days
+1
Senior Service Engineer
Senior Service Engineer

Carrier Global Corporation • City of Syracuse (NY), Northern (KY)

Hybrid
USD 79,000 - 158,000
Health Care Benefits
Retirement Benefits
Paid time off
Senior Service Engineer
Senior Service Engineer

Carrier • Village of East Syracuse (NY)

On-site
USD 79,000 - 158,000
Health benefits
Retirement benefits
Paid time off
+2
Senior Operations Manager, Building Automation and Controls - Automated Logic
Senior Operations Manager, Building Automation and Controls - Automated Logic

Carrier Global Corporation • Canton (MA), Northern (KY)

On-site
USD 118,000 - 235,000
Health benefits
Retirement plan
Paid time off
+3