Senior Platform / Site Reliability Engineer (SRE)

Zohorecruit

United States

À distance

USD 140 000 - 180 000

Plein temps

Il y a 2 jours
Soyez parmi les premiers à postuler
Générateur de candidature

Une candidature conçue pour ce poste — un CV et une lettre de motivation personnalisés qui correspondent à l’offre.

Passez les filtres ATS

Résumé du poste

Zohorecruit is seeking a Senior Platform / Site Reliability Engineer to own and improve a cloud-based SaaS environment. The role is remote with Australia/US scope and focuses on reliability, security, and scalable infrastructure on AWS.

You will lead Terraform-driven IaC implementations, maintain CI/CD pipelines, and collaborate with software teams to deliver robust, performant systems. The ideal candidate combines hands-on expertise with proactive problem-solving and ownership.

Qualifications

  • Minimum 5 years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Cloud Infrastructure.
  • At least 3 years of hands-on experience managing AWS production environments for SaaS applications.
  • Strong experience with Terraform and Infrastructure as Code (IaC).
  • Hands-on knowledge of CI/CD pipelines, automated deployments, and rollback procedures.
  • Experience with monitoring, alerting, logging, and observability tools.
  • Strong PostgreSQL and Amazon RDS experience, including performance tuning, backups, restores, and scalability.
  • Experience developing and testing disaster recovery plans, including RTO and RPO.
  • Good understanding of AWS security, networking, access management, and infrastructure management.
  • Experience with capacity planning, performance optimization, and production incident resolution.
  • Ability to independently manage production infrastructure and take full technical ownership.
  • Strong problem-solving and communication skills.

Responsabilités

  • Manage AWS production environments across Australia and the United States using Terraform.
  • Improve monitoring, logging, alerting, and observability to identify and resolve issues early.
  • Maintain CI/CD pipelines and improve deployment processes, including automated health checks and rollbacks.
  • Ensure PostgreSQL and Amazon RDS databases remain reliable, secure, and optimized for performance.
  • Manage database backups, restoration procedures, query optimization, and scaling.
  • Develop, document, and regularly test disaster recovery procedures.
  • Monitor infrastructure capacity and prepare systems for increasing customer demand.
  • Collaborate with software development teams to improve application reliability and performance.
  • Manage infrastructure maintenance, server updates, security patches, and vulnerability fixes.
  • Investigate production incidents, identify root causes, and implement long‑term solutions.
  • Improve infrastructure automation and operational processes without disrupting ongoing development.

Connaissances

Platform Engineering
Site Reliability
AWS Expertise
Terraform
IaC
CI/CD Pipelines
Monitoring & Observability
PostgreSQL / RDS
Disaster Recovery
Security & Networking
Incident Response
Ownership
Communication

Description du poste

Senior Platform / Site Reliability Engineer (SRE)

Job Title: Senior Platform / Site Reliability Engineer (SRE)

Location: Remote – Australia

Employment Type: Full-Time.

ABOUT THE OPPORTUNITY

We are looking for an experienced Senior Platform / Site Reliability Engineer to take ownership of a growing cloud-based SaaS environment.

This is a hands‑on technical role for someone who enjoys solving infrastructure challenges, improving system reliability, and working independently. The selected candidate will be responsible for maintaining a secure, stable, and scalable AWS environment while supporting ongoing product development and business growth.

KEY RESPONSIBILITIES
  • Manage AWS production environments across Australia and the United States using Terraform.
  • Improve monitoring, logging, alerting, and observability to identify and resolve issues early.
  • Maintain CI/CD pipelines and improve deployment processes, including automated health checks and rollbacks.
  • Ensure PostgreSQL and Amazon RDS databases remain reliable, secure, and optimized for performance.
  • Manage database backups, restoration procedures, query optimization, and scaling.
  • Develop, document, and regularly test disaster recovery procedures.
  • Monitor infrastructure capacity and prepare systems for increasing customer demand.
  • Collaborate with software development teams to improve application reliability and performance.
  • Manage infrastructure maintenance, server updates, security patches, and vulnerability fixes.
  • Investigate production incidents, identify root causes, and implement long‑term solutions.
  • Improve infrastructure automation and operational processes without disrupting ongoing development.
REQUIRED SKILLS AND EXPERIENCE
  • Minimum 5 years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Cloud Infrastructure.
  • At least 3 years of hands‑on experience managing AWS production environments for SaaS applications.
  • Strong experience with Terraform and Infrastructure as Code (IaC).
  • Hands‑on knowledge of CI/CD pipelines, automated deployments, and rollback procedures.
  • Experience with monitoring, alerting, logging, and observability tools.
  • Strong PostgreSQL and Amazon RDS experience, including performance tuning, backups, restores, and scalability.
  • Experience developing and testing disaster recovery plans, including RTO and RPO.
  • Good understanding of AWS security, networking, access management, and infrastructure management.
  • Experience with capacity planning, performance optimization, and production incident resolution.
  • Ability to independently manage production infrastructure and take full technical ownership.
  • Strong problem‑solving and communication skills.
PREFERRED QUALIFICATIONS
  • Familiarity with SOC 2, ISO 27001, or similar compliance standards.
  • Experience working with managed-service providers or external support teams.
  • Background supporting enterprise SaaS applications, particularly in industrial or supply chain environments.
  • Previous experience as the primary or sole Platform/SRE Engineer.
  • Experience managing cloud infrastructure across multiple regions.
IDEAL CANDIDATE

We are looking for a proactive, hands‑on engineer who can independently manage and improve a production SaaS environment.

The ideal candidate will have strong AWS, Terraform, PostgreSQL/RDS, and CI/CD experience, along with proven skills in monitoring, disaster recovery, infrastructure security, and platform scalability.

This position is suitable for someone who takes ownership, works independently, and focuses on practical improvements that strengthen reliability and support business growth.

Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

SRE (Site Realiability Engineer)
SRE (Site Realiability Engineer)

STRATIS Cloud Tech Solutions INC • Arkansas

Sur place
USD 110 000 - 150 000
Competitive salary
Growth and learning opportunities
Friendly, collaborative team
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GoGuardian • El Segundo (CA)

Hybride
USD 180 000 - 240 000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • États-Unis

Sur place
USD 140 000 - 210 000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Clera • États-Unis

À distance
USD 150 000 - 210 000
Site Reliability Engineer (SRE) - 175164
Site Reliability Engineer (SRE) - 175164

Piper Companies • États-Unis

À distance
USD 120 000 - 145 000
Health insurance
Vision insurance
Dental insurance
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

Sur place
USD 140 000 - 150 000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The ReWork Group • New York (NY)

Sur place
USD 120 000 - 160 000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Good co India • États-Unis

À distance
USD 120 000 - 160 000
Site Reliability Engineer
Site Reliability Engineer

Harvey Nash • États-Unis

À distance
USD 120 000 - 150 000
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

VITG • Ellicott City (MD)

Sur place
USD 90 000 - 120 000
401(k) with employer contribution
Medical/Dental/Vision insurance
Paid vacation (PTO)