Senior Platform Support Engineer

Saviynt

Atlanta (GA)

On-site

USD 110,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Saviynt is seeking a Senior Platform Support Engineer to help ensure the 24x7x365 operation of our Enterprise Identity Cloud. You will troubleshoot at the pod level in AKS/EKS, analyze application and DB performance, and drive improvements across monitoring and incident response.

The role collaborates with engineering teams to reduce toil and uphold SLAs. The ideal candidate will manage alerts and service requests end-to-end, develop runbooks, and mentor junior engineers in a dynamic cloud

Responsibilities

  • Strong pod-level troubleshooting skills in AKS/EKS (not just restarting pods).
  • Analyze application and DB (RDS, MySQL) performance issues.
  • Deeply investigate and analyze application performance issues (Java, Grails, Hibernate), identifying root causes and implementing solutions.
  • Oversee the monitoring of our SaaS applications and underlying infrastructure (Kubernetes on AWS and Azure, VPN connections, customer applications, Elastic Search, MySQL) for alerts and performance issues.
  • Strong understanding of basic computing concepts like DNS, IP addressing, Networking, and LDAP.
  • Effectively participate and contribute in on-call escalations with a strong operational mindset and provide technical guidance during critical incidents.
  • Proactively communicate with customers on technical issues when required.
  • Ability to guide junior engineers when needed technically.
  • Manage the full lifecycle of alerts, incidents, and service requests reported through FreshService, ensuring timely and accurate logging, prioritization, resolution, and escalation.
  • Develop, implement, and maintain operational procedures, runbooks, and knowledge base articles to standardize incident resolution and service request fulfillment.
  • Drive continuous improvement initiatives to optimize operational efficiency, reduce incident rates, and improve service request turnaround times via engineering automation to reduce toil and waste.
  • Collaborate with backend engineering and development teams to troubleshoot complex issues, identify root causes, and implement preventative measures.
  • Ensure adherence to defined SLAs (Service Level Agreements) and KPIs (Key Performance Indicators) for operational performance.
  • Maintain operational documentation, including system diagrams, contact lists, and escalation paths.
  • Ensure compliance with relevant security and compliance policies.
  • Plan and coordinate scheduled maintenance activities with minimal impact to service availability.

Job description

Senior Platform Support Engineer

As a Sr. Platform Support Engineer in our SRE Operations team, you will be a key player in ensuring the 24x7x365 smooth operation of Saviynt's Enterprise Identity Cloud. This role focuses on maintaining the stability, performance, and reliability of our platform with a strong emphasis on application layer support and operational ownership. You will be working closely with other operations team members, development, and engineering to resolve issues, implement improvements, and provide exceptional support. This is an opportunity for someone who enjoys operational challenges and problem-solving in a dynamic cloud environment and wants to see their work through to completion.

WHAT YOU WILL BE DOING

  • Strong pod-level troubleshooting skills in AKS/EKS (not just restarting pods).

  • Analyze application and DB (RDS, MySQL) performance issues.

  • Deeply investigate and analyze application performance issues (Java, Grails, Hibernate), identifying root causes and implementing solutions.

  • Oversee the monitoring of our SaaS applications and underlying infrastructure (Kubernetes on AWS and Azure, VPN connections, customer applications, Elastic Search, MySQL) for alerts and performance issues.

  • Strong understanding of basic computing concepts like DNS, IP addressing, Networking, and LDAP.

  • Effectively participate and contribute in on-call escalations with a strong operational mindset and provide technical guidance during critical incidents.

  • Proactively communicate with customers on technical issues when required.

  • Ability to guide junior engineers when needed technically.

  • Manage the full lifecycle of alerts, incidents, and service requests reported through FreshService, ensuring timely and accurate logging, prioritization, resolution, and escalation.

  • Develop, implement, and maintain operational procedures, runbooks, and knowledge base articles to standardize incident resolution and service request fulfillment.

  • Drive continuous improvement initiatives to optimize operational efficiency, reduce incident rates, and improve service request turnaround times via engineering automation to reduce toil and waste.

  • Collaborate with backend engineering and development teams to troubleshoot complex issues, identify root causes, and implement preventative measures.

  • Ensure adherence to defined SLAs (Service Level Agreements) and KPIs (Key Performance Indicators) for operational performance.

  • Maintain operational documentation, including system diagrams, contact lists, and escalation paths.

  • Ensure compliance with relevant security and compliance policies.

  • Plan and coordinate scheduled maintenance activities with minimal impact to service availability.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Reliability Engineer (SaaS, 24/7 Ops)
Senior Platform Reliability Engineer (SaaS, 24/7 Ops)

Saviynt • Atlanta (GA)

On-site
USD 110,000 - 140,000
Senior / Staff Site Reliability, Platform Engineering
Senior / Staff Site Reliability, Platform Engineering

Saviynt • Atlanta (GA)

On-site
USD 120,000 - 150,000
Competitive compensation
Benefits package
Career growth opportunities
Principal Engineer, Cloud Platforms
Principal Engineer, Cloud Platforms

Saviynt • Milpitas (CA)

On-site
USD 235,000 - 250,000
Competitive compensation
Benefits and growth opportunities
Work on a mission-critical platform
Senior Engineer, Federal Cloud Platform
Senior Engineer, Federal Cloud Platform

Rival • Atlanta (OH)

On-site
USD 120,000 - 160,000
Competitive compensation
Growth opportunities
Mission-critical projects
Principal Engineer (Federal Cloud Platform)
Principal Engineer (Federal Cloud Platform)

Saviynt • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
Benefits
Growth opportunities
Senior Platform Engineer
Senior Platform Engineer

Incedo Inc. • San Francisco (CA)

On-site
USD 150,000 - 210,000
Platform Engineer
Platform Engineer

Synergy • Chicago (IL)

On-site
USD 100,000 - 150,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

State of Wisconsin Investment Board • Madison (WI)

On-site
USD 150,000 - 190,000
Senior SRE
Senior SRE

Selby Jennings • Austin (TX)

On-site
USD 140,000 - 190,000
Principal Platform Engineer, Application Services
Principal Platform Engineer, Application Services

Selby Jennings • Austin (TX)

On-site
USD 150,000 - 230,000