AVP SRE and Cloud Solutions

GM Financial

Arlington (TX)

Hybrid

USD 180,000 - 230,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

401K matching
Bonding leave for new parents
Tuition assistance
Training
GM employee auto discount
Community service pay
Nine company holidays

Job summary

GM Financial is seeking an AVP, Platform Site Reliability Engineering (PSRE) to lead reliability engineering vision, strategy, and operating model. Build an AI-enabled, self-healing platform ensuring availability, performance, and security across the software delivery lifecycle.

You will develop high‑performing teams of SREs and SRAs, establish governance, SLOs/SLIs, and partner with Cybersecurity, Architecture, Product, and Engineering to reduce risk and debt.

Qualifications

  • 7-10 years IT industry experience
  • 3-5 years technical leadership in large enterprises
  • 2-4 years early career software development
  • experience in financial services industry preferred
  • Bachelor's Degree in Computer Science or equivalent required
  • Master's Degree preferred

Responsibilities

  • Define and execute the PSRE strategy, roadmap, standards, and operating model aligned with business and technology objectives.
  • Lead the adoption and responsible governance of agentic AI across reliability engineering, incident management, monitoring, observability, technology risk, vulnerability remediation, cloud cost optimization and non-functional requirements.
  • Establish automated controls and intelligent agents that validate reliability, security, operational readiness, and non-functional requirements before software progresses through the delivery lifecycle.
  • Drive resilient, self-healing systems that detect, diagnose, remediate, and recover from known failure conditions with safeguards and auditability.
  • Champions engineering practices improving availability, scalability, recoverability, performance, maintainability, and reducing toil.
  • Establish observability strategies across metrics, logs, traces, app performance and service health with agent-driven correlation and anomaly detection.
  • Drive measurable improvements in incident trends, MTTR, service availability and customer impact.
  • Ensure recurring incidents are addressed through root cause remediation and durable engineering improvements.
  • Establish governance for incident management, runbooks, supportability and service health.
  • Partner with Cybersecurity, Architecture, Product, and Engineering to reduce risk and improve remediation, debt, resiliency and standards.
  • Establish SLOs, SLIs, error budgets, reliability KPIs, and executive reporting.
  • Build and lead high-performing teams through coaching, mentorship and career development.
  • Foster knowledge sharing, engineering excellence, and continuous improvement.
  • Provide executive leadership during major incidents and influence strategy and priorities.

Skills

Leadership
Agentic AI
AIOps
Observability
Incident management
CI/CD
IaC
Cloud cost
Automation
Executive communication

Education

Bachelor's Degree in Computer Science or equivalent
Master's Degree
High School Diploma
Associate Degree

Job description

Why GM Financial Technology?

Innovation isn’t just a talking point at GM Financial, it’s how we operate. From generative AI and cloud-native technologies to peer-led learning and hackathons, our tech teams are building real solutions that make a difference. We’re committed to AI-powered transformation, using advanced machine learning and automation to help us reimagine customer interactions and modernize operations, positioning GM Financial as a leader in digital innovation within a dynamic industry.

Job Description
Why GM Financial Technology?

Innovation isn’t just a talking point at GM Financial, it’s how we operate. From generative AI and cloud-native technologies to peer-led learning and hackathons, our tech teams are building real solutions that make a difference. We’re committed to AI-powered transformation, using advanced machine learning and automation to help us reimagine customer interactions and modernize operations, positioning GM Financial as a leader in digital innovation within a dynamic industry.

Join us and discover a workplace where your ideas matter, your development is prioritized, and you can truly make a global impact.

This position will be posted until filled.

Responsibilities
About The Role:

The AVP, Platform Site Reliability Engineering (PSRE), is a strategic technology and people leader responsible for executing the reliability engineering vision, strategy, and operating model for the organization. This leader advances an AI-enabled reliability organization focused on resilient, self-healing systems, operational excellence, technology risk reduction, and software engineering best practices.

The AVP establishes standards, governance, and agentic AI capabilities that automate reliability, operational readiness, vulnerability and technology debt remediation, cloud cost optimization, and non-functional requirement validation throughout the software delivery lifecycle. The role develops high-performing teams of Site Reliability Engineers (SREs) and Site Reliability Analysts (SRAs) while fostering innovation, continuous learning, accountability, and toil elimination.

  • Define and execute the PSRE strategy, roadmap, standards, and operating model aligned with business and technology objectives.
  • Lead the adoption and responsible governance of agentic AI across reliability engineering, incident management, monitoring, observability, technology risk, vulnerability and technology debt remediation, cloud cost optimization, and non-functional requirements.
  • Establish automated controls and intelligent agents that validate reliability, security, operational readiness, and non-functional requirements before software progresses through the delivery lifecycle, including before pull requests where practical.
  • Drive resilient, self-healing systems that detect, diagnose, remediate, and recover from known failure conditions with appropriate safeguards, human oversight, and auditability.
  • Champion engineering practices that improve availability, scalability, recoverability, performance, maintainability, and operational sustainability while reducing manual intervention and toil.
  • Establish observability strategies across metrics, logs, traces, application performance, service health, and user-experience signals using agent-driven correlation, anomaly detection, diagnostics, and proactive remediation.
  • Drive measurable improvements in incident trends, recurring-issue elimination, MTTR, service availability, alert quality, customer impact, and operational efficiency.
  • Ensure recurring incidents are systematically addressed through root cause remediation, automation, self-healing capabilities, and durable engineering improvements.
  • Establish operational governance for incident and escalation management, operational readiness, service ownership, documentation, runbooks, supportability, and service health.
  • Partner with Cybersecurity, Architecture, Product, and Engineering teams to reduce technology risk and improve vulnerability remediation, technology debt, resiliency, and engineering standards.
  • Establish SLOs, SLIs, error budgets, reliability KPIs, operational metrics, and executive reporting that connect technology performance to customer and business outcomes.
  • Build and lead high-performing teams by developing engineers, analysts, and technical leaders through coaching, mentorship, structured career development, succession planning, and continuous technical upskilling.
  • Establish a culture of knowledge sharing, innovation, engineering excellence, accountability, and continuous improvement that strengthens organizational capability and grows future technical and people leaders.
  • Provide executive leadership during major incidents and influence technology strategy, investment decisions, and engineering priorities across the organization.
Qualifications
What Makes You an Ideal Candidate?
  • Experience leading Site Reliability Engineering, Platform Engineering, Operational Excellence, or comparable technology organizations.
  • Demonstrated leadership in agentic AI, AIOps, intelligent automation, or autonomous operational capabilities, including governance, evaluation, security, human oversight, and auditability.
  • Deep understanding of observability, incident and problem management, resiliency engineering, operational readiness, technology risk, SLOs, SLIs, error budgets, and toil reduction.
  • Experience improving incident trends, recurring incidents, MTTR, service availability, alert quality, customer impact, and operational efficiency.
  • Experience building or enabling resilient, self-healing systems and automation that reduce manual intervention.
  • Experience with CI/CD, Infrastructure as Code (IaC), DevOps automation, policy as code, test automation, and modern software delivery practices.
  • Experience leading vulnerability remediation, technology debt reduction, non-functional requirement automation, and cloud cost optimization initiatives.
  • Strong software engineering foundation and technical judgment related to automation, APIs, debugging, performance, maintainability, and scalable solution design.
  • Demonstrated success developing engineering talent, growing future technical leaders, and building high-performing organizations focused on continuous learning and technical excellence.
  • Strong executive communication, stakeholder management, collaboration, and organizational influence skills.
  • Ability to operate effectively in a fast-paced environment, provide quality service to internal and external customers, work a flexible schedule when business needs require, and travel on a limited basis.
Education And Experience
  • 7-10 years IT industry experience Req
  • 3-5 years technical leadership in large enterprises Req
  • 2-4 years early career experience as a software developer Req
  • experience in financial services industry Pref
  • High School Diploma Required
  • Associate Degree
  • Bachelor’s Degree in Computer Science or equivalent work experience Required
  • Master’s Degree preferred
What We Offer
  • 401K matching
  • bonding leave for new parents (12 weeks, 100% paid)
  • tuition assistance
  • training
  • GM employee auto discount
  • community service pay
  • nine company holidays
Our Culture

Our team members define and shape our culture — an environment that welcomes innovative ideas, fosters integrity, and creates a sense of community and belonging. Here we do more than work — we thrive.

Compensation

Competitive pay and bonus eligibility.

Work Life Balance

Flexible hybrid work environment, 3-days a week in office.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer — Cloud, AI & Leadership
Senior Software Engineer — Cloud, AI & Leadership

GM Financial • Arlington (TX)

Hybrid
USD 100,000 - 150,000
Manager - Data Engineering
Manager - Data Engineering

GM Financial • Arlington (TX)

On-site
USD 120,000 - 170,000
401K matching
Tuition assistance
Training
+2
AVP IT Delivery Transformation
AVP IT Delivery Transformation

GM Financial • Detroit (MI)

On-site
USD 120,000 - 180,000
401K matching
Tuition assistance
GM employee auto discount
+2
AVP IT Delivery Transformation
AVP IT Delivery Transformation

GM Financial • Dallas (TX)

On-site
USD 140,000 - 190,000
401K matching
Tuition assistance
Training
+3
Manager - Data Engineering
Manager - Data Engineering

GM Financial • Irving (TX)

On-site
USD 120,000 - 150,000
401K matching
Parental leave (12 weeks, 100% paid)
Tuition assistance
+2
Engineering Manager, Developer Experience
Engineering Manager, Developer Experience

General Motors • Sunnyvale (CA)

On-site
USD 219,000 - 335,000
Health benefits
Retirement plan
GM vehicle discounts
AVP Enterprise Resilience
AVP Enterprise Resilience

GM Financial • Irving (TX)

On-site
USD 140,000 - 210,000
401K matching
Bonding leave for new parents
Training
+3
Engineering Manager, Developer Experience
Engineering Manager, Developer Experience

General Motors • Warren (MI)

On-site
USD 219,000 - 335,000
Health and wellbeing benefits
GM vehicle discounts
Paid vacation & holidays
+2
AVP Cloud Data Analytics Architecture
AVP Cloud Data Analytics Architecture

GM Financial • Irving (TX)

On-site
USD 180,000 - 240,000
401K matching
Tuition assistance
Training
+2
Software Development Engineer I
Software Development Engineer I

re-zoo-me • Arlington (TX)

Hybrid
USD 90,000 - 130,000
401K matching
Bonding leave for new parents
Tuition assistance
+5