Senior Engineer – Reliability

Jobtailor

South Carolina

On-site

USD 120,000 - 180,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Jobtailor in the United States is seeking an experienced Site Reliability Engineer focused on platform engineering and production support. You will partner with product owners and engineering teams to ensure operational readiness and stable releases across enterprise initiatives.

The role emphasizes incident management, CI/CD adoption, observability, and collaboration with global delivery teams to reduce recurring issues and improve deployment practices.

Qualifications

  • Bachelore9 degree is required.
  • 6+ years of experience in SRE, platform engineering, production support or related environments.
  • 5+ years supporting enterprise platforms (AWS, Dynatrace, ELK, ServiceNow, SolarWinds).
  • Experience troubleshooting complex production incidents.
  • Experience executing changes via formal change and release management.
  • Experience in financial services or regulated industry.
  • Experience with CI/CD pipelines and deployment processes.
  • Experience with observability platforms and monitoring tools.
  • Experience collaborating with offshore/global delivery teams.

Responsibilities

  • Partner with Product Owners and engineering teams to provide operational and release support for technology initiatives
  • Ensure platform changes meet operational readiness requirements, including rollback procedures and handoffs
  • Maintain production stability during upgrades and enterprise initiatives
  • Support platform ownership transitions and readiness activities across global delivery teams
  • Troubleshoot application and platform issues to restore services
  • Serve as escalation point for complex production incidents and operational challenges
  • Participate in incident triage, root cause analysis, and corrective action planning
  • Collaborate with engineering to implement long-term solutions reducing recurring incidents
  • Conduct health checks, configuration reviews, and performance assessments
  • Validate vendor releases and configuration changes prior to production deployment
  • Enhance monitoring, alerting, and observability with analytics teams
  • Execute platform changes through SDLC, change control, and release processes
  • Improve release and deployment practices with Platform Eng, DevOps, QE, Scrum teams
  • Support source control, environment separation, release automation, and CI/CD adoption
  • Maintain runbooks, support documentation, incident playbooks, and release procedures
  • Document incident findings, lessons learned, and process improvement opportunities
  • Contribute to standardized, repeatable, scalable operational practices

Skills

Platform Engineering
Incident Management
CI/CD Pipeline Implementation
AWS Experience
Change Management

Education

Bachelore2s degree in computer science, information technology, engineering, or a related field

Tools

AWS
Dynatrace
ELK
ServiceNow
SolarWinds
Monitoring Tools
Observability Platforms

Job description

  • Partner with Product Owners and engineering teams to provide operational and release support for technology initiatives
  • Ensure platform changes meet operational readiness requirements, including rollback procedures, runbook documentation, integration standards, and support handoffs
  • Maintain production stability throughout platform upgrades, enhancements, and enterprise initiatives
  • Support platform ownership transitions and operational readiness activities across global delivery teams
  • Troubleshoot application and platform issues to restore services and minimize business impact
  • Serve as an escalation point for complex production incidents and operational challenges
  • Participate in incident triage, root cause analysis, corrective action planning, and resolution activities
  • Collaborate with engineering teams to implement long-term solutions that reduce recurring incidents
  • Conduct health checks, configuration reviews, and performance assessments to identify operational risks
  • Validate vendor releases, hotfixes, and configuration changes prior to production deployment
  • Enhance monitoring, alerting, and issue detection capabilities with observability and analytics teams
  • Execute platform changes through SDLC, change management, and release management processes
  • Improve release and deployment practices with Platform Engineering, DevOps, Quality Engineering, and Scrum teams
  • Support source control, environment separation, release automation, and CI/CD adoption
  • Maintain runbooks, support documentation, configuration records, incident playbooks, and release procedures
  • Document incident findings, lessons learned, and process improvement opportunities
  • Contribute to standardized, repeatable, and scalable operational practices
Requirements
  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field
  • 6+ years of experience in SRE, platform engineering, production support, enterprise application operations, or related technology environments
  • 5+ years of experience supporting enterprise platforms, including AWS, Dynatrace, ELK, ServiceNow, and SolarWinds
  • Experience troubleshooting complex production incidents within enterprise-scale technology environments
  • Experience executing technology changes through formal change management and release management processes
  • Experience within financial services or another regulated industry
  • Experience implementing or supporting CI/CD pipelines, release automation, and deployment processes across development, testing, and production environments
  • Experience with observability platforms, monitoring tools, performance dashboards, or application monitoring solutions
  • Experience collaborating with offshore, nearshore, or global delivery teams
Core Competencies

Demonstrates expertise in platform engineering and production support, with a strong focus on operational readiness, incident management, and CI/CD practices. Proficient in collaborating with cross-functional teams to enhance system stability and implement effective monitoring and release strategies.

Highest-signal resume keywords
  • Platform Engineering
  • Incident Management
  • CI/CD Pipeline Implementation
  • AWS Experience
  • Change Management
ATS Optimization Keywords
Hard Skills
  • SRE
  • Production Support
Soft Skills
  • Collaboration
  • Problem-Solving
  • Communication
Industry Keywords
  • Financial Services
  • Regulated Industry
Tools & Technologies
  • AWS
  • Dynatrace
  • ELK
  • ServiceNow
  • SolarWinds
  • Monitoring Tools
  • Observability Platforms
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Jobtailor • New Jersey

On-site
USD 120,000 - 180,000
Technical Platform Operations Support, Manager
Technical Platform Operations Support, Manager

Jobtailor • Town of Texas (WI)

On-site
USD 120,000 - 180,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Jobtailor • Arizona

On-site
USD 180,000 - 240,000
Senior Platform Engineer
Senior Platform Engineer

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Senior/Staff Infrastructure – Platform Engineer
Senior/Staff Infrastructure – Platform Engineer

Jobtailor • Santa Clara (CA)

On-site
USD 170,000 - 250,000
Senior DevOps Engineer
Senior DevOps Engineer

Jobtailor • Lehi (UT)

On-site
USD 120,000 - 180,000
Site Reliability Engineer – Lead
Site Reliability Engineer – Lead

Jobtailor • Arizona

On-site
USD 140,000 - 230,000
AVP SRE, Cloud Solutions
AVP SRE, Cloud Solutions

Jobtailor • Arlington (TX)

On-site
USD 180,000 - 240,000
Senior Staff Engineer – DevOps
Senior Staff Engineer – DevOps

Jobtailor • California (MO)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer – Digital Assets
Senior Site Reliability Engineer – Digital Assets

Jobtailor • Arizona

On-site
USD 120,000 - 170,000