Senior Site Reliability Engineer

Donnelley Financial Solutions (DFIN)

United States

On-site

USD 150,000 - 190,000

Full time

9 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Donnelley Financial Solutions is seeking Senior Site Reliability Engineers to ensure SaaS products are fast, stable, and scalable. You will drive SRE culture, automate runbooks, and own CI/CD pipelines while collaborating across engineering teams.

The role emphasizes AI-enabled observability, proactive incident management, and continuous improvement of system performance, reliability, and scalability in a global SaaS environment.

Qualifications

  • 5+ years designing and maintaining cloud infrastructure in Azure or AWS.
  • Experience applying AI capabilities within CloudOps operations.
  • Certifications or training in AI, Cloud AI services or AIOps are a plus.
  • 5+ years writing software in C#/.NET or Java.
  • 5+ years automating deployments with CI/CD tools.
  • 5+ years implementing production monitoring with tools like New Relic, Dynatrace, DataDog or AppDynamics.
  • 5+ years scripting in PowerShell or Python/Bash.
  • 5+ years supporting public client-facing revenue systems.
  • DevOps focus and IaC with Terraform.
  • Experience with databases (SQL, Cosmos) monitoring.
  • Experience with Kubernetes (AKS or EKS).
  • BS in CS or equivalent.

Responsibilities

  • Champion and implement SRE culture to maintain a high-quality platform.
  • Leverage AI tools for observability, incident prediction, and automated remediation.
  • Evaluate AI-powered ops solutions to improve performance and scalability.
  • Monitor and alert to prevent client-impacting issues, maintain SLOs/SLAs.
  • Automate runbooks and CI/CD pipelines.
  • Collaborate with SRE and software teams, communicate progress.
  • Participate in on-call duties 24/7 and lead incident RCA.
  • Stay ahead of latest tools to improve processes.
  • Learn continuously and share best practices.

Skills

Cloud infra
Azure
AWS
SRE practices
CI/CD
Automation
Python
PowerShell
Monitoring
Kubernetes

Education

BS in CS

Tools

Harness
Azure DevOps
Ansible
Jenkins
Terraform
New Relic
Dynatrace
DataDog
AppDynamics
SQL Monitoring

Job description

Select how often (in days) to receive an alert:

Join a dynamic team at the pulse of global markets, where we deliver innovative software and service solutions for essential financial reporting and capital markets transactions. At DFIN, we are a values-driven organization that empowers you to build a fulfilling career while bringing your authentic self to work every day. Our Win as One mentality ensures that our team’s success is directly linked to Client, Shareholder and Employee Satisfaction.

In 2026, DFIN was named #1 on the 2026 Top 100 Global Most Loved Workplaces® by Best Practice Institute. We have also been recognized as one of America’s Most Loved Workplaces® for five consecutive years and a Built In Best Place to Work for six years, reflecting our continued commitment to supporting employees’ total well-being. Enjoy competitive compensation, a flexible workplace, comprehensive benefits, and opportunities for professional growth. Bring your passion and talents to DFIN – because being YOU thrives here.

Summary:

We are looking for technical team members at all levels who want to push themselves to deliver best in market SaaS solutions. We offer a challenging environment where you will have to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise.The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers. SRE’s at DFIN take on availability, performance, managing change, monitoring, response and are guardians of non-functional requirements. You either have an SaaS infrastructure background with a programmatic, automated mindset or are someone that comes with a software engineering background with SaaS infrastructure experience. The SRE goal is to build automated systems that reduce or eliminate manual work to keep our products up and running and performing optimally. We are looking for someone who thrives on collaboration within the team and across other groups and can operate independently to deliver solutions.

Responsibilities:
  • Champion and implement a culture of SRE to maintain a high-quality platform infrastructure in DFIN SaaS products
  • Leverage AI tools to enhance system reliability, including intelligent observability, incident prediction and automated remediation across cloud infrastructure
  • Evaluate and implement emerging AI powered operations and observability solutions to proactively improve system performance, reliability and scalability
  • Champion and implement application and infrastructure monitoring and alerting to prevent client impacting issues by ensuring system availability, performance and scalability to maintain SLOs and SLAs
  • Optimize application performance at scale
  • Automate everything including system operational runbooks
  • Define and support continuous integration and deployment pipelines (CI/CD) aligned to branchingand quality assurance strategies
  • Dive deep into technology and stay on the forefront of the latest tools, technologies, and strategies; help evaluate, prototype, and integrate them into work processes
  • Perform with broad independence and deliver on project milestones and tasks on schedule while communicating progress regularly
  • Build strong relationships with SRE team members and software engineering teams to hold each other accountable for quality expectations
  • Learn continuously and apply lessons learned
  • Evangelize best practices, eliminate bottlenecks, and improve process
  • Participate in on-call duties 365/24/7 and lead the triage and RCA of production incidents
Qualifications:
  • 5+years experiencedesigning, building, securing, monitoring and maintaining cloud infrastructure in Azure or AWS
  • Experience applying AI capabilities withinCloudOpsoperations
  • Relevant certifications or training in AI, Cloud AI services or AIOps platforms are a plus
  • 5+years experiencewriting software in any modern software language such as C# .NET, Java
  • 5+years experiencecreating automated deployments with tools such as Harness, Azure DevOps, Ansible or Jenkins to manage Infrastructure as Code and software build and deployment in a continuous integration (CI) / continuous delivery (CD) environment
  • 5+years experienceimplementing production performance, availability, and scalability monitoring and alerting using a tool such as New Relic, Dynatrace,DataDogor AppDynamics
  • 5+years experiencewriting scripts in PowerShell or Python/Bash to automate system operations as runbooks for Windows or Linux environments.
  • 5+years experiencesupporting public client facing revenue generating systems
  • Strong DevOps focus and experience building and deploying Infrastructure as Code with Terraform or similar technology
  • Experiencing monitoring and preventing issues with databases and database queries (SQL, Cosmos) using tools likeSolarwindsDatabase Performance Analyzer,IderaSQL Diagnostic Manager, or Redgate SQL Monitor
  • Experience planning, coordinating, developing and executing all stages of post deployment verification test scripts
  • Experience securing Windows or Linux systems in 24x7 production environment
  • Experience with containerization and managing Kubernetes clusters (AKS or EKS)
  • Experience with common cloud networking, firewall and load balancing configuration
  • BS in Computer Science or equivalent work experience

It is the policy of Donnelley Financial Solutions to select, place, and manage all its employees without discrimination based on race, color, national origin, gender, age, religion, actual or perceived disability, veteran status, actual or perceived sexual orientation, genetic information or any other protected status.

If you are a qualified individual with a disability or a disabled veteran, you have the right to request a reasonable accommodation if you are unable or limited in your ability to use or access jobs.dfinsolutions.com as a result of your disability. You can request a reasonable accommodation by sending an email to talentacquisition@dfinsolutions.com .

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Donnelley Financial Solutions (DFIN) • Northern (KY)

Hybrid
USD 120,000 - 180,000
Software Engineer – Remote at Donnelley Financial Solutions (DFIN)
Software Engineer – Remote at Donnelley Financial Solutions (DFIN)

Feedinkoo • United States

Remote
USD 90,000 - 120,000
Flexible workplace
Comprehensive benefits
Opportunities for professional growth
Software Engineer – Remote at ExecutivePlacements.com
Software Engineer – Remote at ExecutivePlacements.com

Feedinkoo • United States

Remote
USD 120,000 - 160,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

T. Rowe Price • Juneau (AK)

Hybrid
USD 140,000 - 190,000
Flexible remote work
Health care benefits
Tuition assistance
+3
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)

Mindlance • Irving (TX)

Hybrid
USD 120,000 - 160,000
Senior Application Support Engineer (SRE)
Senior Application Support Engineer (SRE)

The Depository Trust & Clearing Corporation (DTCC) • Tampa (FL)

Hybrid
USD 120,000 - 170,000
Competitive compensation
Health & life insurance
Pension / Retirement benefits
+2
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Senior Application Support Engineer (SRE)
Senior Application Support Engineer (SRE)

The Depository Trust & Clearing Corporation (DTCC) • Jersey City (NJ)

On-site
USD 120,000 - 180,000
Base pay
Health & life insurance
Pension
+2
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Calabrio • United States

On-site
USD 90,000 - 130,000