Site Reliability Engineer

Cloudbeds

United States

Remote

USD 120,000 - 150,000

Full time

41 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote First
PTO
Home office stipend
Cloudbeds University
Upskilling opportunities

Job summary

Cloudbeds seeks a Site Reliability Engineer to safeguard the reliability and performance of our hospitality platform. You will architect scalable AWS cloud solutions and maintain Kubernetes (EKS) clusters, while supporting CI/CD with ArgoCD and GitHub Actions, and automating deployments with Terraform.

You will develop and improve product observability with Grafana, Prometheus, DataDog, and CloudWatch, and participate in incident management and RCA to minimize downtime.

Qualifications

  • 5+ years as a DevOps or SRE in AWS environments.
  • Strong Kubernetes (EKS) experience and Helm charts.
  • CI/CD design with ArgoCD and GitHub Actions.
  • Infrastructure-as-code with Terraform.
  • Observability with Grafana/Prometheus/DataDog/CloudWatch.

Responsibilities

  • Design and run reliable, scalable AWS architecture.
  • Maintain and support large Kubernetes (EKS) clusters.
  • Support CI/CD with ArgoCD and GitHub Actions.
  • Automate deployments with Terraform IaC.
  • Improve observability and monitoring across platforms.
  • Lead incident management and RCA activities.
  • Collaborate with dev teams on reliability best practices.

Skills

AWS
Kubernetes
CI/CD
Observability
Incident Management
Networking
English communication

Education

Bachelor’s degree in CS or equivalent

Tools

EKS
ArgoCD
GitHub Actions
Terraform
Grafana/Prometheus
Datadog
CloudWatch
Nginx/Ingress

Job description

What Makes Cloudbeds Unique

At Cloudbeds, we're not just building software, we’re transforming hospitality. Our platform powers hotels across 150 countries, processing billions in bookings annually. From boutique properties to large hotel groups, we help hoteliers transform operations and uplevel their commercial strategy through a unified platform that integrates with hundreds of partners. And we do it with a completely remote team. Imagine working alongside global innovators to build AI-powered solutions that solve hoteliers' biggest challenges. Since our founding in 2012, we've continually been recognized for excellence, most recently earning top honors at the 2026 HotelTechAwards for Best Hotel Management Software and landing on Deloitte's Technology Fast 500 again, but we're just getting started.

How You'll Make an Impact:

As a Site Reliability Engineer, you'll be the guardian of our platform's reliability and performance, ensuring millions of hospitality transactions flow seamlessly across the globe. You'll architect and implement scalable AWS cloud solutions that keep the most ambitious hotels running 24/7, while fostering a culture of automation, resilience, and continuous improvement across our engineering teams.

Our SRE Team:

We're a bottom-up, collaborative team that thrives on healthy debate and shared ownership of our infrastructure. You'll have endless opportunities to influence architecture decisions while working with cutting-edge cloud technologies at scale. We believe the best solutions come from engineers who are empowered to innovate, experiment, and challenge the status quo.

What You Bring to the Team:
  • Design and implement a reliable and scalable AWS architecture to meet the needs of the organization.
  • Maintain and support highly loaded Kubernetes (EKS) clusters and infrastructure-related components.
  • Support the CICD process with ArgoCD and GitOps.
  • Automate the platform deployments with Terraform infrastructure-as-code.
  • Develop and continuously improve product Observability and Monitoring systems based on the Grafana, Prometheus, DataDog, and Cloudwatch.
  • Respond and participate with Incident Management and Root Cause Analysis, ensuring minimal impact on services.
  • Optimize system performance and troubleshoot issues as they arise.
  • Collaborate with development teams to establish monitoring best practices and ensure systems meet reliability targets.
  • Collaborate with security teams to implement and maintain security best practices.
  • Infrastructure support rotation providing guidance to other engineering teams.
What Sets You Up for Success:
  • 5+ years of experience as aDevOps or SRE working within the AWS ecosystem.
  • 5+ years of experience with Kubernetes (EKS) and Helm charts.
  • Experience with designing, building, and supporting CI/CD pipelines with ArgoCD and GitHub actions.
  • Experience with infrastructure-as-code methodologies with Terraform.
  • Experience with Observability and Monitoring with Grafana, Prometheus, DataDog, and Cloudwatch.
  • Experience with Incident Management, full stack troubleshooting, performance analysis and root cause analysis (RCA).
  • Experience with Web application systems such as Nginx, Ingress controllers, load balancing and Content Delivery Networks.
  • Experience with Databases (MySQL, PostgreSQL, Aurora) and Middleware technologies (Redis, Memcached and SQS)
  • Good networking skills with VPC, Security Groups and Network ACLs.
  • Ability to work remotely and manage your own time in a global team.
  • Good written and verbal communication in English.
  • Bachelor’s degree in Computer Science or equivalent experience.
Bonus Skills to Stand Out:
  • Advanced experience with Database Administration (Aurora, MySQL, PostgreSQL).
  • Experience working in a PCI-compliant environment.
  • Experience working with Kong API Gateway.
Compensation:

Depending on your skills and experience, you can expect your annual compensation to be between $120,000 - $150,000

#LI-IK1

Work Authorization:

Please note that applicants must be currently authorized to work in the location where the position is located without requiring visa sponsorship. At this time, Cloudbeds is unable to provide sponsorship for work visas.

What to Expect - Your Journey with Us

Behind Cloudbeds' revolutionary technology is a team redefining what's possible in hospitality. We're 650+ team members across 40+ countries, bringing together elite engineers, AI architects, world-class designers, hoteliers, and hospitality veterans to solve challenges others haven't dared to tackle. Our diverse team speaks 30+ languages, but we all share one language: a passion for innovation and travel. From pioneering breakthroughs in machine learning to revolutionizing how hotels operate, we're not just watching the future of hospitality unfold – we're coding it, designing it, writing it, and shipping it. If you're ready to work alongside some of the brightest minds in tech who are obsessed with using AI to transform a trillion-dollar industry, this is your chance to be part of something extraordinary.

Learn more online at cloudbeds.com

Cloudbeds Awards:
  • Winner – Hotel Management Software | HotelTechAwards (2026)
  • Overall 10 Best Places to Work | HotelTechAwards (2026)
  • Overall Top 10 Hotelier's Choice | HotelTechAwards (2026)
  • Finalist – Property Management Systems (PMS) & Channel Managers | HotelTechAwards (2026)
  • Most Loved Workplace® Certified (2024)
  • Deloitte Technology Fast 500 (2024)
Discover our Benefits:
  • Remote First, Remote Always
  • PTO in accordance with local labor requirements
  • Fully Paid Parental Leave
  • Home office stipend based on country of residency
  • Professional development courses in Cloudbeds University
  • Access to professional development, including manager training, upskilling, and knowledge transfer
Everyone is Welcome - A Culture of Inclusion

Cloudbeds is proud to be an Equal Opportunity Employer that celebrates the diversity in our global team! We do not discriminate based upon race, religion, color, national origin, gender (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics.

Cloudbeds is committed to the full inclusion of all qualified individuals. As part of this commitment, Cloudbeds will ensure that persons with disabilities are provided reasonable accommodations in the hiring process. We encourage deaf, hard-of-hearing, deaf-blind, and deaf-disabled individuals to apply. If reasonable accommodation is needed to participate in the job application or interview process or to perform essential job functions, please contact our HR team by phone at (858) 201-7832 or via email at accommodations@cloudbeds.com. Cloudbeds will provide an American Sign Language (ASL) interpreter where needed as a reasonable accommodation for the hiring process.

To all Staffing and Recruiting Agencies: Our Careers Site is only for individuals seeking a job at Cloudbeds. Staffing, recruiting agencies, and individuals being represented by an agency are not authorized to use this site or to submit applications, and any such submissions will be considered unsolicited. Cloudbeds does not accept unsolicited resumes or applications from agencies. Please do not forward resumes to our jobs alias, Cloudbeds employees, or any other company location. Cloudbeds is not responsible for any fees related to unsolicited resumes/applications.

#LI-REMOTE

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

DevOps Engineer
DevOps Engineer

Cloudbeds • Houston (TX)

On-site
USD 90,000 - 120,000
Remote-first culture
Home office stipend
Fully Paid Parental Leave
+3
DevOps Engineer
DevOps Engineer

Cloudbeds • Philadelphia

On-site
USD 90,000 - 120,000
Home office stipend
Parental leave
PTO per local law
DevOps Engineer
DevOps Engineer

Cloudbeds • Orlando (FL)

On-site
USD 90,000 - 120,000
Remote-first company
Home office stipend
Fully Paid Parental Leave
+1
DevOps Engineer
DevOps Engineer

Cloudbeds • Miami (FL)

On-site
USD 90,000 - 120,000
Remote-first team
Home office stipend
Fully Paid Parental Leave
+1
DevOps Engineer
DevOps Engineer

Cloudbeds • Dallas (TX)

On-site
USD 90,000 - 120,000
Home office stipend
Fully paid parental leave
Monthly Wellness Fridays
+1
DevOps Engineer
DevOps Engineer

Cloudbeds • Atlanta (GA)

On-site
USD 90,000 - 120,000
Home office stipend based on country
Fully Paid Parental Leave
Monthly Wellness Fridays
+2
DevOps Engineer
DevOps Engineer

Cloudbeds • Pittsburgh

On-site
USD 90,000 - 120,000
Remote-first team
Home office stipend
DevOps Engineer
DevOps Engineer

Cloudbeds • Northern (KY)

Hybrid
USD 90,000 - 120,000
Remote-first culture
PTO and Wellness Fridays
Home office stipend
+3
Security Engineer
Security Engineer

Third-Party Job Posts • United States

Remote
USD 86,000 - 115,000
Remote work policy
Home office stipend
Wellness Fridays
Engineering Manager
Engineering Manager

Cloudbeds • United States

On-site
GBP 70,000 - 90,000
Remote First, Remote Always
Paid Time Off (PTO)
Corporate Apartment Accommodations
+4