System Reliability Engineer

Ascendion

Jacksonville (FL)

On-site

USD 110,000 - 135,000

Full time

28 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) retirement plan
Long-term disability
Short-term disability
Paid time off
Paid vacation
Holidays
Learning management

Job summary

Ascendion is seeking a Systems Reliability Engineer to join the Technical Operations team supporting the Embedded Finance platform. You will own the reliability and resiliency of a large-scale enterprise platform, ensuring services remain highly available, performant, and secure.

Responsibilities include designing monitoring with Splunk, Dynatrace, Grafana, and Datadog; defining SLOs/SLIs; leading incident response; driving RCA; building automation to reduce toil; collaborating with engineering,

Qualifications

  • Hands-on experience with monitoring, observability and alerting tools (Splunk, Dynatrace, Grafana, Datadog).
  • Proven experience operating and supporting a large-scale enterprise platform.
  • Led or contributed to incident response and RCA processes.

Responsibilities

  • Own the reliability, resiliency and availability of the Embedded Finance platform.
  • Define and track SLOs/SLIs and error budgets.
  • Lead and participate in incident response and RCA processes.
  • Build automation to reduce toil and improve MTTD/MTTR.
  • Collaborate with engineering, product and risk teams to embed reliability practices.

Skills

Observability & monitoring
Incident response leadership
Root cause analysis
SRE fundamentals
Cross-team communication

Tools

Splunk
Dynatrace
Grafana
Datadog

Job description

Ascendion is an AI-native software engineering disruptor helping businesses innovate faster, smarter, and with greater impact. We partner with enterprise clients across North America, the UK, Europe, and APAC to solve complex challenges in data, experience design, software product engineering, and workforce transformation. Powered by expert engineers, thousands of AI agents, and our Engineering to the Power of AI (EngineeringAI) method, we deliver measurable outcomes that build trust, unlock value, and accelerate growth. Learn more at https://ascendion.com/.

Engineering to the Power of AI™, AAVA™, Engineering AI , Engineering to Elevate Life™, Enterprise Platforms AI , Data & Insights AI , Experience AI , GCC AI , Operations AI , Platform Engineering AI , Product AI , and Quality Engineering AI are trademarks or service marks of Ascendion ® . AAVA™ is pending registration. Unauthorized use is strictly prohibited.

Ascendion | Engineering to elevate life
About Ascendion

Ascendion is an AI-native software engineering disruptor helping businesses innovate faster, smarter, and with greater impact. We partner with enterprise clients across North America, the UK, Europe, and APAC to solve complex challenges in data, experience design, software product engineering, and workforce transformation. Powered by expert engineers, thousands of AI agents, and our Engineering to the Power of AI (EngineeringAI) method, we deliver measurable outcomes that build trust, unlock value, and accelerate growth. Learn more at https://ascendion.com/.

Engineering to the Power of AI™, AAVA™, Engineering AI , Engineering to Elevate Life™, Enterprise Platforms AI , Data & Insights AI , Experience AI , GCC AI , Operations AI , Platform Engineering AI , Product AI , and Quality Engineering AI are trademarks or service marks of Ascendion ® . AAVA™ is pending registration. Unauthorized use is strictly prohibited.

We have a culture built on opportunity, inclusion, and a spirit of partnership. Come, change the world with us:
  • Build the coolest tech for the world’s leading brands
  • Solve complex problems - and learn new skills
  • Experience the power of transforming digital engineering for Fortune 500 clients
  • Master your craft with leading training programs and hands‑on experience

Experience a community of change makers!

About the Role

We are seeking a Systems Reliability Engineer to join the Technical Operations team supporting our Embedded Finance (EmFi) platform. In this role, you will own the reliability and resiliency of a large‑scale enterprise platform, ensuring that our services remain highly available, performant and secure.

Job Title

SRE

Key Responsibilities
  • Own the reliability, resiliency and availability of the Embedded Finance platform, proactively identifying and mitigating risks to service continuity.
  • Design, implement and maintain comprehensive monitoring and alerting frameworks leveraging Splunk, Dynatrace, Grafana and Datadog to provide end‑to‑end observability across the platform.
  • Define and track service level objectives (SLOs), service level indicators (SLIs) and error budgets to measure and improve platform health.
  • Lead and participate in incident response, serving as a technical driver during remediation calls and coordinating with impacted and impacting technical and product teams.
  • Own and advance the root cause analysis (RCA) process — investigating incidents, documenting the sequence of events and remediating actions, and clearly identifying underlying root causes to prevent recurrence.
  • Ensure timely creation and management of incident tickets (e.g., ServiceNow) and accurate incident tracking, aging and reporting.
  • Build automation and tooling to reduce toil, improve mean time to detection (MTTD) and mean time to resolution (MTTR), and increase operational efficiency.
  • Collaborate with engineering, product and risk stakeholders to embed reliability best practices into the platform lifecycle.
Minimum Qualifications
  • Hands‑on experience with monitoring, observability and alerting tools, specifically Splunk, Dynatrace, Grafana and Datadog.
  • Proven experience operating and supporting a large‑scale enterprise platform environment.
  • Demonstrated experience with incident response and leading or contributing to root cause analysis (RCA) processes.
  • Strong understanding of reliability engineering principles, including availability, resiliency, monitoring and alerting best practices.
  • Experience with ticketing and incident management workflows (e.g., ServiceNow).
  • Excellent communication skills, with the ability to drive remediation efforts and collaborate across technical, product and risk teams.
Desired Qualifications
  • Experience in financial services, payments or embedded finance environments.
  • Proficiency with scripting or programming languages for automation (e.g., Python, Go, Bash).
  • Familiarity with cloud platforms, containerization and CI/CD pipelines.
  • Experience defining and managing SLOs, SLIs and error budgets.
Location

Jacksonville, FL, Berkeley Heights, NJ , Alpharetta, GA , Toronto, ON (Onsite)

Salary and Other Compensation

The annual [salary/hourly rate]for this position is between [$110K- $135K annually]/[$55/hr -65 per hour]. Factors which may affect pay within this range may include geography/market, skills, education, experience and other qualifications of the successful candidate.

Benefits

The Company offers the following benefits for this position, subject to applicable eligibility requirements: [medical insurance] [dentals insurance] [vision insurance] [401(k) retirement plan] [long‑term disability insurance] [short‑term disability insurance] [5 personal days accrued each calendar year. The Paid time off benefits meet the paid sick and safe time laws that pertains to the City/ State] [10-15 days of paid vacation time] [6 paid holidays and 1 floating holiday per calendar year] [Ascendion Learning Management System]

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer
Senior Software Engineer

Ascendion • Plano (TX)

On-site
USD 100,000 - 110,000
Medical insurance
Dental insurance
Vision insurance
+3
Technical Program Manager
Technical Program Manager

Ascendion • Berkeley Heights (NJ)

On-site
USD 72,000 - 103,000
Medical insurance
Dental insurance
Vision insurance
+7
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hobbsnews • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Discretionary incentive plan
Senior Site Reliability Engineer
Senior Site Reliability Engineer

National Black MBA Association • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Annual discretionary plan
Tech Lead Java Developer
Tech Lead Java Developer

Ascendion • Columbus (OH)

On-site
USD 140,000 - 170,000
medical insurance
dental insurance
vision insurance
+5
Senior Site Reliability Engineer
Senior Site Reliability Engineer

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Infrastructure Lead
Infrastructure Lead

Ascendion • Vancouver (WA)

On-site
USD 100,000 - 112,000
Medical insurance
Dental insurance
Vision insurance
+6
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Donnelley Financial Solutions (DFIN) • Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Donnelley Financial Solutions (DFIN) • United States

On-site
USD 150,000 - 190,000
Compliance Analyst
Compliance Analyst

Ascendion • Sacramento (CA)

On-site
USD 103,000 - 107,000
Medical insurance
Dental insurance
Vision insurance
+7