Site Reliability Engineer

Pacificacontinental

Pacifica (CA)

Hybrid

USD 140,000 - 190,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Pacificacontinental is seeking a Site Reliability Engineer to strengthen our DevOps-driven infrastructure across hybrid cloud and on-premises environments. You will work with Windows and Linux servers, VMware, and Azure to enable cloud-native applications while meeting security and regulatory requirements.

Responsibilities include improving deployment velocity, incident remediation, and collaboration across teams.

Qualifications

  • 5+ years of hands-on experience with Windows and Linux servers across hybrid environments.
  • Experience with VMware, cloud platforms (Azure preferred), and Active Directory.
  • Secrets management with Consul and Vault or similar systems.
  • Configuration management tools like Salt and Ansible; Terraform for IaC.
  • Firewalls/load balancers (F5); web servers IIS/NGINX; DBs SQL Server & PostgreSQL.
  • APM with New Relic; infrastructure monitoring (Sensu, Nagios, Azure App Insights).
  • CI/CD tools: TeamCity, Octopus Deploy, Concourse, Azure DevOps, GitHub Actions.
  • Log aggregation with SumoLogic or Splunk; network concepts (DNS, DHCP, proxies).
  • Security ops with SAST/DAST/RAST and WAF; Infrastructure as Code is essential.

Responsibilities

  • Improve internal processes to reduce lead time and increase deployment frequency.
  • Enhance security, reliability, and performance of our infrastructure.
  • Increase velocity by leveraging cross-functional expertise.
  • Remediate production incidents quickly and safely, reducing outages.
  • Collaborate with other teams on best practices and implementation strategies.
  • Advocate for IaC, monitoring, high availability, disaster recovery, security, and DevOps methodologies.
  • Create SLIs, SLOs, and SLAs; participate in capacity planning.

Skills

Windows servers
Linux servers
VMware
Azure
Active Directory
Secrets management
Terraform
Salt
Ansible
F5
IIS
NGINX
MS SQL Server
PostgreSQL
New Relic
Sensu
Nagios
Azure DevOps
GitHub Actions
SumoLogic
Splunk
DNS
DHCP
WAF
SAST
DAST
RAST
Infrastructure as Code
PowerShell
BASH

Tools

TeamCity
Octopus Deploy
Concourse
Azure DevOps
GitHub Actions

Job description

Ourengineering team has built the largest private Medicare marketplace in the country. We passionately focus on the continuous improvement of the systems we build.


We have spent many years growing and fostering a DevOps culture by bridging the divide between our Software and Infrastructure Engineering departments. We want the cross-functional teams that we are building to include Site Reliability Engineers. We operate in a complex, multi-tenant, hybrid cloud and on-premises infrastructure that spans both the Windows and Linux OS. We strive for security, reliability, and automation in line with DevOps and Site Reliability Engineering principles. If you are passionate about learning and improvement through metrics and automation, and passionate about engendering that mindset in others, we want to hear from you.


About the role:

Maintains shared cloud resources in use by numerous software engineering teams within our business unit. We aim to enable software engineering teams to build cloud native applications that adhere to security and regulatory requirements with limited handholding by our cloud engineers. We do still have a fair number of applications hosted in on-premise data centers, which we aim to support migrating to the cloud.


Requirements:

Hands-on Engineering


5+ years of hands-on experience with a majority of the following technologies, along with a willingness to become proficient in the remaining areas:



  • Windows and Linux Servers

  • VMware

  • Cloud platforms, preferably with Azure

  • Active Directory

  • Secrets management with Consul and Vault or similar systems

  • Configuration management tools like Salt, Ansible and Terraform

  • Firewalls and load balancers such as F5

  • Web servers, including IIS andNGINX

  • Database Server Infrastructure like Microsoft SQL Server and PostgreSQL

  • Application Performance Monitoring with tools like New Relic

  • Infrastructure monitoring with tools like Sensu, SolarWinds, Nagios, or Azure App Insights

  • CI/CD tools like TeamCity, Octopus Deploy, Concourse, Azure DevOps, or GitHub Actions

  • Log Aggregation tools like SumoLogic or Splunk

  • Network theory and protocols such as DNS, DHCP, proxy servers, and firewalls

  • Security operations with tools for SAST, DAST, RAST, and WAF

  • Infrastructure as Code or automation experience.


Proficiency, high-comfort, and familiarity with:



  • One or more scripting languages, such as PowerShell and BASH

  • Command line tools such as (git, netcat, npm, terraform, etc.)


Responsibilities


  • Make improvements to internal processes to reduce lead time and increase deployment frequency

  • Identify improvements to the quality, security, and performance of our infrastructure

  • Increase the velocity with which teams deliver, leveraging expertise from various functional disciplines

  • Identify how to remediate production incidents more quickly and safely while reducing the frequency of outages

  • Actively engage with other teams and departments to collaborate on best practices and implementation strategy

  • Adhere to and advocate for best practices, including Infrastructure as Code, monitoring, high availability,disaster recovery,security, and DevOps methodologies

  • Create SLIs, SLOs, and SLAs

  • Contribute to capacity planning, advise and consult with teams who will be load/stress testing

  • Keep up with industry innovations, recommending new tools or practices when appropriate

  • Actively mentor peers, developing their expertise and inspiring others to innovate

  • Provide timely assistance and remediation solutions during critical situations and production incident

  • Document and share “lessons learned” from production, including root cause analysis

  • Explore new ways of improving communication between other Site Reliability Engineers and with other teams

  • Write and maintain architectural, stakeholder, and policy documentation


Attach resume as .pdf, .doc, .docx, .odt, .txt, or .rtf (limit 5MB) orpaste resume


References: Please enter names and contact information:*

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Brooksource • San Antonio (TX)

On-site
USD 80,000 - 120,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 140,000 - 210,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

O.C. Tanner • Salt Lake City (UT)

On-site
USD 180,000 - 260,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Site Reliability Engineer -- SINDC5717546
Site Reliability Engineer -- SINDC5717546

Compunnel Inc. • Denton (TX)

Hybrid
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000
Mgr IT Site Reliability Eng
Mgr IT Site Reliability Eng

Kforce, Inc • Town of Florida (NY)

On-site
USD 110,000 - 140,000
Site Reliability Engineer
Site Reliability Engineer

FORT • United States

Hybrid
USD 150,000 - 180,000
Healthcare benefits
Flexible work environment
Large-scale cloud platform project
+1
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior SRE Engineer
Senior SRE Engineer

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 140,000 - 190,000