Cloud SRE – 24x7 Production Uptime (Hybrid)

Skyhigh Security

Frisco (TX)

Hybrid

USD 110,000 - 140,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Retirement Plans
Medical, Dental and Vision Coverage
Paid Time Off
Paid Parental Leave
Support for Community Involvement

Job summary

A leading cloud security company based in Texas seeks a Site Reliability Engineer to monitor and maintain high-availability production environments. The role involves incident management, troubleshooting, and collaboration with engineering teams. Candidates should have a Bachelor's degree in Computer Science and 7+ years of SRE experience, as well as expertise in Linux systems and various monitoring tools like Prometheus and Grafana. This position offers a hybrid work model, enhancing work-life flexibility.

Qualifications

  • Bachelor’s degree in computer science or related area with 7+ years of SRE experience.
  • System admin experience on Linux environments.
  • Experience with Prometheus, Grafana for monitoring setups.

Responsibilities

  • Monitor and troubleshoot operational issues in a high-availability production environment.
  • Perform root cause analysis on major incidents.
  • Implement proactive monitoring and alerting solutions.

Skills

SRE experience in a large enterprise organization
System admin experience on Linux environments
Experience with end-to-end monitoring setup
Experience with Prometheus, Grafana, ELK
Experience with Cloud Technologies like AWS
Experience with containerized workloads tools
Network knowledge (TCP/IP, UDP, DNS)
Ability to script/program with languages like Python
Experience with configuration management tools
Strong communication and analytical skills

Education

Bachelor’s degree in computer science or related

Tools

Prometheus
Grafana
AWS
Kubernetes
Jenkins
GitHub

Job description

A leading cloud security company based in Texas seeks a Site Reliability Engineer to monitor and maintain high-availability production environments. The role involves incident management, troubleshooting, and collaboration with engineering teams. Candidates should have a Bachelor's degree in Computer Science and 7+ years of SRE experience, as well as expertise in Linux systems and various monitoring tools like Prometheus and Grafana. This position offers a hybrid work model, enhancing work-life flexibility.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cloud-Native SRE: Reliability, Automation & Observability
Cloud-Native SRE: Reliability, Automation & Observability

IBM • Lowell (MA)

On-site
USD 85,000 - 110,000
Hybrid SRE: Cloud Infrastructure & Automation
Hybrid SRE: Cloud Infrastructure & Automation

TikTok • Seattle (WA)

Hybrid
USD 112,725 - 177,840
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
Hybrid Cloud SRE Engineer – Automation & Reliability
Hybrid Cloud SRE Engineer – Automation & Reliability

DryvIQ • United States

Hybrid
USD 120,000 - 160,000
Senior Site Reliability Engineer: Cloud, Kubernetes Uptime
Senior Site Reliability Engineer: Cloud, Kubernetes Uptime

Compunnel, Inc. • Greenwood Village (CO)

On-site
USD 120,000 - 150,000
Hybrid Cloud SRE - Automate & Scale Global Infra
Hybrid Cloud SRE - Automate & Scale Global Infra

TikTok • San Jose (CA)

Hybrid
USD 118,000 - 260,000
Hybrid SRE: Infra & Assurance Services
Hybrid SRE: Infra & Assurance Services

TikTok • Seattle (WA)

Hybrid
USD 112,000 - 178,000
Cloud SRE - Database & Distributed Systems
Cloud SRE - Database & Distributed Systems

Ll Oefentherapie • Seattle (WA)

On-site
USD 100,000 - 130,000
Senior Site Reliability Engineer - Build SRE Ops (Hybrid)
Senior Site Reliability Engineer - Build SRE Ops (Hybrid)

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000
Cloud SRE: Scale Microservices Across Multi-Cloud
Cloud SRE: Scale Microservices Across Multi-Cloud

TP-Link Corporation Limited • Irvine (CA)

On-site
USD 100,000 - 140,000
Free snacks and drinks
Fully paid medical insurance
401k contributions
+2
Senior Site Reliability Engineer (SRE) – Automation & Cloud Ops
Senior Site Reliability Engineer (SRE) – Automation & Cloud Ops

Knack Solutions • Reston (VA)

On-site
USD 120,000 - 160,000