Staff Site Reliability Engineer - Paze

Early Warning Services LLC

Scottsdale (AZ)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Healthcare Coverage
401(k) Retirement Plan
Flexible Time Off
Paid Parental Leave
Family Planning Support

Job summary

Early Warning Services LLC is seeking a Staff Site Reliability Engineer to collaborate with development teams on ensuring application availability and resiliency. You will design and implement solutions that enhance performance, troubleshoot technical issues, and mentor team members.

This position requires a Bachelor's Degree and 8+ years of related experience, including expertise in technologies like Python, AWS, and Docker. They offer comprehensive benefits including competitive healthcare coverage, a 401(k) match, and generous paid time off.

Qualifications

  • 8+ years of experience managing complex projects in technical settings.
  • Strong troubleshooting skills in complex environments.
  • Experience leading high-priority incidents.

Responsibilities

  • Build automation and tooling for application management.
  • Design and implement monitoring systems to detect problems.
  • Serve as a technical liaison for application teams.

Skills

Python
Docker
Microservices Architecture
AWS
Java
Jenkins
Linux Administration

Education

Bachelor's Degree in Business or Computer Science

Tools

Kafka
Redis
Docker

Job description

Overall Purpose

The Staff Site Reliability Engineer partners with development teams by defining availability standards and implementing availability and resiliency patterns in applications and infrastructure.

Essential Functions
  • Design and implement software and tools to improve the performance, availability, scalability, and latency, while delivering end products to customers with the highest efficiency and meeting all security standards.
  • Support the company's commitment to risk management and protecting the integrity and confidentiality of systems and data.
  • Build automation and tooling around application management, such as deployments, configuration changes and disaster recovery scenarios.
  • Design, implement and evangelize observability and monitoring systems to proactively detect problems and identify cause.
  • Evaluate capacity of the application on a continuous basis to provide stats to the Product/Business teams and recommend an efficient path to scale for future needs.
  • Identify performance bottlenecks and work with cross-functional teams to troubleshoot and resolve issues.
  • Serve as a technical liaison for the application and provide documents and runbooks to Level 1 and Level 2 teams.
  • Participate in 24X7 on‑call rotation.
  • Be a champion of excellent processes; take the initiative in developing repeatable patterns and standard, reusable work across teams.
  • Work directly with application development teams to provide feedback and technical requirements to the software development lifecycle, implementing best‑practice microservice design patterns and other modern software development approaches.
  • Understand and support the adoption of best‑practice microservice design patterns and other modern software reliability approaches and techniques.
  • Be a thought leader: a senior point of expertise on site reliability engineering issues, industry trends and developing technologies. Be a role model to others on the team. Coach and mentor team members.
Minimum Qualifications
  • Education and experience typically obtained through completion of a Bachelor's Degree in Business and/or Computer Science or related field.
  • Typically 8+ years of related progressive experience managing large complex projects in a technical or software development environment inclusive of post‑graduate degree.
  • Proven ability to lead a team through high‑priority incidents and improve the RCA process.
  • Excellent troubleshooting skills and proven experience resolving technical issues in complex environments.
  • Hands‑on experience in designing and developing using one or more of the following technologies: Python, Go, Java, Docker, Microservices Architecture, Messaging frameworks such as Kafka, SQS or JMS, Database Technologies like Oracle, Dynamo DB, Aurora, Caching layers such as Redis and memcached, Strong understanding of Linux administration, Experience with CI/CD pipeline implementation including GIT, Chef, Maven, Jenkins, Strong understanding and hands‑on experience on TCP/UDP/IP protocols, Experience in leading cross‑functional teams to create technical solutions, Proven track record designing and building complex end‑to‑end systems (full stack developer).
  • Background and drug screen.
Preferred Qualifications
  • Good programming skills in one or more of the following languages: Java, Ruby, Python, JavaScript and Go.
  • Hands‑on experience in supporting applications in a 24X7 customer‑facing production environment.
  • Working knowledge of AWS, Docker, Kubernetes, Swarm.

Employee must be able to perform essential functions and physical requirements of position with or without reasonable accommodation.

Physical Requirements

Working conditions consist of a normal office environment. Work is primarily sedentary and requires extensive use of a computer and involves sitting for periods of approximately four hours. Work may require occasional standing, walking, kneeling and reaching. Must be able to lift 10 pounds occasionally and/or negligible amount of force frequently. Requires visual acuity and dexterity to view, prepare, and manipulate documents and office equipment including personal computers. Requires the ability to communicate with internal and/or external customers.

Benefits
  • Healthcare Coverage—Competitive medical (PPO/HDHP), dental, and vision plans as well as company contributions to your Health Savings Account (HSA) or pre‑tax savings through flexible spending accounts (FSA) for commuting, health & dependent care expenses.
  • 401(k) Retirement Plan—Featuring a 100% company safe harbor match on your first 6% deferral immediately upon eligibility.
  • Paid Time Off—Flexible Time Off for exempt (salaried) employees, as well as generous PTO for non‑exempt (hourly) employees, plus 11 paid company holidays and a paid volunteer day.
  • 12 weeks of Paid Parental Leave.
  • Maven Family Planning—provides support through your parenting journey including egg freezing, fertility, adoption, surrogacy, pregnancy, postpartum, early pediatrics, and returning to work.

Early Warning Services is an affirmative action and equal opportunity employer.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Early Warning Services, LLC ("Early Warning") considers for employment, hires, retains and promotes qualified candidates on the basis of ability, potential, and valid qualifications without regard to race, religious creed, religion, color, sex, sexual orientation, genetic information, gender, gender identity, gender expression, age, national origin, ancestry, citizenship, protected veteran or disability status or any factor prohibited by law, and as such affirms in policy and practice to support and promote equal employment opportunity and affirmative action, in accordance with all applicable federal, state, and municipal laws. The company also prohibits discrimination on other bases such as medical condition, marital status or any other factor that is irrelevant to the performance of our employees.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer - Paze
Staff Site Reliability Engineer - Paze

Early Warning Services LLC • San Francisco (CA)

On-site
USD 130,000 - 160,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Staff Site Reliability Engineer - Paze
Staff Site Reliability Engineer - Paze

Early Warning Services LLC • Chicago (IL)

On-site
USD 120,000 - 150,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+1
Senior Site Reliability Engineer: Observability & Resiliency
Senior Site Reliability Engineer: Observability & Resiliency

Early Warning • Scottsdale (AZ)

Hybrid
USD 116,000 - 145,000
Discretionary incentive plan
Sr. Software Engineer - Java / SpringBoot / AWS - Zelle
Sr. Software Engineer - Java / SpringBoot / AWS - Zelle

Early Warning® • Scottsdale (AZ)

On-site
USD 142,000 - 183,000
Healthcare Coverage
401(k) Matching
Paid Time Off
+1
Senior DevOps Engineer - Cloud, Kubernetes & IaC
Senior DevOps Engineer - Cloud, Kubernetes & IaC

Early Warning Services LLC • San Francisco (CA)

Hybrid
USD 128,000 - 156,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Sr. Staff Software Engineer - Commerce - Java/SpringBoot/AWS
Sr. Staff Software Engineer - Commerce - Java/SpringBoot/AWS

Early Warning Services LLC • San Francisco (CA)

On-site
USD 180,000 - 240,000
Healthcare coverage
401(k) with company match
Paid time off
+1
Senior Java & SpringBoot Engineer — AWS Cloud
Senior Java & SpringBoot Engineer — AWS Cloud

Early Warning • San Francisco (CA)

On-site
Sr. Site Reliability Engineer - Paze
Sr. Site Reliability Engineer - Paze

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Sr. Site Reliability Engineer - Paze
Sr. Site Reliability Engineer - Paze

Early Warning • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Matching
Paid Time Off
+1
Sr. Site Reliability Engineer - Paze
Sr. Site Reliability Engineer - Paze

Early Warning • Scottsdale (AZ)

Hybrid
USD 106,000 - 156,000
Healthcare Coverage
401(k) Company Match
Paid Time Off
+2