Site Reliability Engineer II

Restaurant365

San Francisco (CA)

Hybrid

USD 98,583 - 138,016

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Comprehensive medical benefits, 100% paid for employee
401k + matching
Equity Option Grant
Unlimited PTO + Company holidays
Wellness initiatives

Job summary

Restaurant365 in San Francisco seeks a Site Reliability Engineer II to support and enhance its cloud infrastructure. Candidates should have 2-4 years of experience in site reliability engineering or DevOps, proficiency with cloud platforms like Azure or AWS, and automation experience with tools such as Terraform and Ansible. This role requires a collaborative approach to incident response and system reliability, offering competitive salary and unlimited PTO as part of a comprehensive benefits package.

Qualifications

  • 2-4 years of experience in site reliability engineering, DevOps, or cloud operations.
  • Experience with cloud platforms (Azure or AWS), including services such as AKS, ECS, Functions/Lambda.
  • Proficiency with infrastructure‑as‑code and automation tools.

Responsibilities

  • Respond to production incidents and perform triage and troubleshooting.
  • Identify and automate manual processes to improve efficiency.
  • Enhance monitoring tools and platforms for better observability.

Skills

Incident response
System monitoring
Automation
Performance troubleshooting
Linux engineering
Scripting (Python, Bash, PowerShell)

Education

BS in Computer Science, Information Systems, or related field

Tools

Terraform
Ansible
Cloud platforms (Azure, AWS)
Monitoring tools (Prometheus, Grafana, ELK)
CI/CD tools (GitLab, Git)

Job description

Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back‑office operations for restaurants. Restaurant365’s culture is focused on empowering team members to produce top‑notch results while elevating their skills. We’re constantly evolving and improving to make sure we are and always will be “Best in Class” … and we want that for you too!

This role requires a hybrid work schedule based out of one of our office locations: Austin, TX; Irvine, CA; or Akron, OH.

Site Reliability Engineer II will be responsible for supporting, enhancing, and maintaining Restaurant365’s cloud infrastructure and applications. Qualified candidates will demonstrate growing expertise in site reliability practices, with skills in incident response, system monitoring, automation, and performance troubleshooting. You will collaborate with DevOps, development, and infrastructure teams to resolve moderately complex issues, propose improvements, and strengthen the reliability, scalability, and security of our SaaS platform.

How you’ll add value:
  • Execution & Collaboration
    • Respond to production incidents, perform triage and troubleshooting, and contribute to post‑incident analysis.
    • Identify and automate manual processes to improve efficiency and reduce risk.
    • Enhance and evolve monitoring tools and platforms to improve observability.
    • Promote and apply best practices for reliability, scalability, and performance across engineering.
    • Implement and support cloud automation using Terraform, Ansible, or CloudFormation.
    • Work within change management protocols to provide maximum uptime for production systems.
    • Participate in on‑call rotation, providing 24x7 support for incidents and contributing to root cause analysis.
    • Partner with developers, architects, vendors, and IT teams to ensure reliable system operations.
    • Research and remediate vulnerabilities in coordination with security teams.
    • Maintain documentation of infrastructure, monitoring, runbooks, and incident response procedures.
  • Standards & Process
    • Apply company policies and procedures when handling operational tasks and incidents.
    • Suggest and implement improvements to operational processes and monitoring practices.
    • Contribute to technical diagrams, documentation, and runbooks for system reliability.
  • Learning & Growth
    • Expand expertise in cloud services (Azure, AWS, or GCP) and container platforms (EKS, ECS, AKS).
    • Build proficiency with observability and monitoring tools (Prometheus, Grafana, ELK, Site24x7, Nagios).
    • Develop scripting and automation skills using Python, Bash, PowerShell, or similar.
    • Participate in planning discussions by contributing technical input on system stability and reliability.
What you’ll need to be successful in this role:
  • BS in Computer Science, Information Systems, or related field (or equivalent experience).
  • 2-4 years of experience in site reliability engineering, DevOps, or cloud operations.
  • Experience with cloud platforms (Azure or AWS), including services such as AKS, ECS/EKS, Functions/Lambda, S3, and Blob storage.
  • Proficiency with infrastructure‑as‑code and automation (Terraform, Ansible, YAML, Python, Bash, PowerShell).
  • Strong Linux engineering skills; working knowledge of Windows administration.
  • Experience supporting production environments and participating in on‑call rotations.
  • Familiarity with web servers and middleware (Nginx, Apache Tomcat).
  • Experience with CI/CD tools (GitLab, Git, or similar).
  • Strong written, oral, and interpersonal communication skills.
Preferred Qualifications
  • Experience with monitoring tools (Prometheus, Grafana, ELK, Site24x7, Nagios).
  • Knowledge of performance analysis and system vulnerability remediation.
  • Cloud certification (AWS or Azure) preferred.
  • Familiarity with restaurant industry SaaS platforms and customer‑facing applications.
R365 Team Member Benefits & Compensation
  • This position has a salary range of $98,583‑$138,016 annually. The above range represents the expected salary range for this position. The actual salary may vary based on several factors, including, but not limited to, relevant skills/experience, time in the role, business line, and geographic location. Restaurant365 focuses on equitable pay for our team and aims for transparency with our pay practices.
  • Comprehensive medical benefits, 100% paid for employee.
  • 401k + matching.
  • Equity Option Grant.
  • Unlimited PTO + Company holidays.
  • Wellness initiatives.

DYN365, Inc d/b/a Restaurant365 is an equal opportunity employer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

Restaurant365 • Denver (CO)

On-site
USD 98,583 - 138,016
100% employee medical benefits
401k + matching
Equity Option Grant
+2
Site Reliability Engineering (SRE)/Dev Ops
Site Reliability Engineering (SRE)/Dev Ops

Spectraforce • Louisville (KY)

Hybrid
USD 80,000 - 90,000
Senior Customer Architect
Senior Customer Architect

Restaurant365 • Irvine (CA)

On-site
USD 130,000 - 174,000
Medical benefits
401k + matching
Equity Option Grant
+3
Site Reliability Engineer
Site Reliability Engineer

Staffworxs • Louisville (KY)

Hybrid
USD 120,000 - 180,000
Customer Architect
Customer Architect

Restaurant365 • San Francisco (CA)

On-site
USD 116,000 - 174,000
Medical benefits
401k + matching
Equity option grant
+2
Senior Implementation Specialist
Senior Implementation Specialist

Restaurant365 • San Francisco (CA)

On-site
USD 120,000 - 170,000
Competitive compensation package
Comprehensive medical benefits
401k + matching
+4
Customer Architect
Customer Architect

Restaurant365 • Austin (TX)

On-site
USD 116,000 - 174,000
401k + matching
Equity option grant
Unlimited PTO + holidays
+2
Senior Implementation Specialist
Senior Implementation Specialist

Restaurant365 • Palo Alto (CA)

On-site
USD 140,000 - 180,000
Competitive compensation package
Comprehensive medical benefits
401k + matching
+4
Customer Architect
Customer Architect

Restaurant365 • Palo Alto (CA)

On-site
USD 116,000 - 174,000
Medical benefits
401k + matching
Equity option grant
+3
Senior Customer Success Manager
Senior Customer Success Manager

Restaurant365 • Irvine (CA)

On-site
USD 96,000 - 145,000
Medical benefits
401k + matching
Equity Option Grant
+3