Vice President - Production Support Manager

Morgan Stanley

South Jordan (UT)

On-site

USD 100,000 - 130,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PowerToFly is hiring a Production Support Manager in South Jordan, Utah, to manage critical incidents and ensure operational excellence across platforms. The role requires strong technical skills and a minimum of 8 years of experience in production support or technical leadership.

The successful candidate will oversee a global team, promote automation, and support cloud applications. Excellent communication skills and flexibility for rotational on-call coverage are essential.

Qualifications

  • Minimum 8+ years of hands-on experience in Production Support or similar technical leadership role.
  • Experience troubleshooting large-scale distributed applications and managing critical incidents.
  • Flexibility in working hours, including rotational on-call and weekend coverage.

Responsibilities

  • Provide strategic leadership for complex distributed and cloud platforms.
  • Lead and manage critical incidents, ensuring timely resolution.
  • Mentor and manage a high-performing production support team.

Skills

Production Support
UNIX/Linux
Scripting languages
Relational databases
Cloud platforms
Communication skills
Agile development practices

Education

Bachelor of Computer Science, Engineering, or related field

Tools

Azure Monitor
Kubernetes
Docker
Terraform

Job description

Position Description

We are seeking an experienced Production Support Manager to join our Global Operations Reliability and Production Engineering team in South Jordan, Utah. The successful candidate will represent and manage critical incidents across a diverse portfolio of over 1,000 applications, ensuring stability, performance, and regulatory compliance for internal and external clients. This role requires both technical and business acumen, exceptional communication skills, and flexibility to operate in a dynamic, high‑pressure environment, including rotational on‑call and weekend coverage.

What you’ll do in the role
  • Provide strategic leadership and oversight for complex distributed and cloud platforms, ensuring operational excellence and regulatory compliance.
  • Lead and manage critical incidents, ensuring timely resolution and effective communication with executive management and business stakeholders.
  • Troubleshoot and resolve issues across hardware, software, application, network, and cloud stacks.
  • Build and maintain relationships with senior stakeholders, downstream consumers, IT partners, and development teams globally.
  • Mentor and manage a high‑performing production support team, promoting continuous improvement, learning, and resilience.
  • Drive automation, toil reduction, and enhancements in observability, monitoring, and reliability across platforms.
  • Own and evolve documentation, knowledge sharing, and best practices for global teams.
  • Collaborate with development and infrastructure teams to resolve support issues and implement reliability solutions.
  • Represent production support in executive forums, influencing technology and business decisions.
  • Operate in a 'follow‑the‑sun' support model, with rotational on‑call and weekend coverage.
  • Develop and implement programs to establish and enhance reliability and production management practices in the department.
  • Function as a buffer between support and development teams, reducing escalations and resolving issues within production management.
  • Support end users and business functions in day‑to‑day operations.
What you’ll bring to the role
  • Bachelor of Computer Science, Engineering, or a related field.
  • Minimum 8+ years of hands‑on experience in Production Support, Production Management, or a similar technical leadership role.
  • Proven people management and team leadership experience.
  • Strong working knowledge of UNIX/Linux operating systems, scripting languages (e.g., Shell, Python, Perl, JavaScript), and relational databases (e.g., Sybase, DB2, SQL, Postgres, Snowflake, MongoDB).
  • Experience troubleshooting large‑scale distributed applications and managing critical incidents.
  • Hands‑on experience supporting applications deployed on cloud platforms, particularly Azure.
  • Expertise in analyzing, debugging, and troubleshooting complex applications, infrastructure, and database issues.
  • Excellent and confident communicator, able to manage high‑pressure environments and executive‑level communication.
  • Flexibility in working hours, including rotational on‑call and weekend coverage.
  • Experience supporting financial industry systems is highly valued.
  • Knowledge of ITIL, SDLC, and Agile development practices.
  • Experience with cloud/distributed computing technologies and certifications is a plus.
  • Familiarity with modern observability and monitoring tools (e.g., Azure Monitor, AppInsights, Prometheus, Grafana, Datadog, Kubernetes, Docker, Ansible).
  • Experience using and configuring DevOps tooling (e.g., Terraform), and in instrumenting application endpoints for logging, metrics, and events.
  • Strong documentation and knowledge‑sharing skills.
  • Experience in implementing reliability engineering and production management program.
Equal Employment Opportunity

Morgan Stanley is an equal opportunity employer committed to building and maintaining a workforce that is diverse in experience and background. Our recruiting efforts reflect our strong commitment to a culture of inclusion, where individuals are hired, developed, and advanced based on their skills and talents.

Our workforce reflects a broad cross‑section of the global communities in which we operate, bringing a variety of backgrounds, talents, perspectives, and experiences.

For more information, please visit: https://www.morganstanley.com/people-opportunities/eeo.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

PowerToFly • Alpharetta (GA)

On-site
USD 100,000 - 130,000
SRE/Production Support Lead
SRE/Production Support Lead

Morgan Stanley • Alpharetta (GA)

On-site
USD 125,000 - 175,000
Production Support Engineer
Production Support Engineer

LanceSoft, Inc. • South Jordan (UT)

On-site
USD 100,000 - 150,000
Production Management Support
Production Management Support

Soho Square Solutions • South Jordan (UT)

Hybrid
USD 70,000 - 110,000
Lead Site Reliability Engineer, Vice President
Lead Site Reliability Engineer, Vice President

Morgan Stanley • New York (NY)

On-site
USD 150,000 - 190,000
Compute & Storage Engineer VP
Compute & Storage Engineer VP

Morgan Stanley • United States

On-site
USD 180,000 - 240,000
Application Support
Application Support

PowerToFly • New York (NY)

On-site
USD 120,000 - 165,000
Comprehensive employee benefits
Opportunity for career advancement
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Relha LLC • Alpharetta (GA), Northern (KY)

Hybrid
USD 125,000 - 175,000
Vice President, Production Services Application Support
Vice President, Production Services Application Support

BNY Mellon • New York (NY)

On-site
USD 180,000 - 250,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Morgan-Stanley • Alpharetta (GA)

On-site
USD 125,000 - 175,000
Comprehensive benefits