Lead Site Reliability Engineer, Vice President

Morgan Stanley

New York (NY)

On-site

USD 150,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Morgan Stanley is seeking a Senior Site Reliability Engineer, VP, to join the WM Product Technology team in New York. The role focuses on production support, automation, and reliable systems, collaborating with BU and development partners to minimize outages.

Candidates should have 10+ years in production, strong scripting, cloud, and CI/CD experience, and excellent communication skills. The role emphasizes automating deployments, improving platform reliability, and delivering results in a

Qualifications

  • 10+ years in a production environment with software development background.
  • Strong troubleshooting and end-to-end diagnosis capabilities.
  • Expertise in scripting languages (Shell, Python, Perl) and cloud development.
  • Experience with CI/CD pipelines and automation.
  • Knowledge of databases: DB2, Sybase, Oracle.
  • Experience with batch scheduling: Autosys or similar.
  • Proven ability to work in Agile environments (Scrum).
  • Experience with cloud deployments (Azure, AWS).

Responsibilities

  • Maintain live applications with monitoring of availability, latency and health.
  • Lead lifecycle improvements from design to deployment and operation.
  • Scale systems via automation to improve reliability and velocity.
  • Troubleshoot infrastructure issues and update knowledge bases.
  • Collaborate with development teams to build production-management tools.
  • Interface with data providers and consumers to minimize escalations.
  • Develop scripts and code changes for operational tasks.
  • Own and maintain knowledge bases and documentation.
  • Identify risks and act with urgency in a team or independent setting.

Skills

SRE experience
Scripting (Shell, Python, Perl)
Databases
CI/CD
Cloud (Azure, AWS)
Automation
Troubleshooting
Linux/UNIX
Agile/Scrum

Tools

Autosys
Jenkins
Train
Windeploy

Job description

Senior Site Reliability Engineer, VP

At Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always do so with a standard of excellence. We are a leading global financial services firm that conducts its business through three principal business segments—Institutional Securities, Wealth Management (WM), and Investment Management. The Firm's employees serve clients worldwide from more than 1,200 offices in 43 countries.

As a market leader, the talent and passion of our people is critical to our success. Together, we share a common set of values rooted in integrity, excellence, and strong team ethic. Morgan Stanley can provide a superior foundation for building a professional career - a place for people to learn, to achieve and grow. A philosophy that baances personal lifestyles, perspectives and needs is an important part of our culture.

Position Description

We are seeking for a Site Reliability Engineer with a minimum of 10 years of industry experience, preferably working in the financial IT community. The position in the WM Product Technology team is focused on delivering exceptional services to both BU and Dev partners to minimize/avoid any production outages. The role will focus on production support within the WM Product Technology automating deployments and working with the agile teams to build and support stable and reliable production systems. The ideal candidate will be passionate about automation and skilled in one of the programming language Python/PERL/SHELL, Ruby, JAVA, C#, or the like. Candidate should possess a strong understanding of database concepts, job scheduler, MQ, Web services, UNIX/LINUX/Windows OS as well as experience with debugging applications. We are looking for a strong leader with excellent communications skills who is committed to continuously improving and delivering results. Candidate should be organized, disciplined, detail-oriented, self-motivated, and delivery-focused.

What You’ll Do In The Role
  • Maintain applications once they are live by measuring and monitoring availability, latency and overall system health with a focus on business activities and continuously evaluate cost and TOIL.
  • Engage in and improve the whole lifecycle of services from inception and design, through deployment, operation, capacity planning and launch reviews.
  • Scale systems sustainably through mechanisms like automation and evolve systems by pushing for changes that improve reliability and velocity; includes automation for other various operational needs.
  • Troubleshoot infrastructure issues, reviewing log files, updating documentation, and having knowledge base with resolutions
  • Work closely with the application Development team to understand the platform and create tools/utilities to help with production management
  • Work with upstream data providers and upstream consumers, and reducing the amount of escalation to development teams
  • Develop scripts and assist with code changes along with operational tasks/activities.
  • Work closely with Application Development to ensure that the support team has excellent knowledge of the application set, own and maintain support knowledgebase and documents.
  • Use analytical skills to find trends in the environment and drive out problems.
  • Lead effort to determine improvement areas to stabilize the plant.
  • Identify risks and work with a sense of urgency, working within a team or independently.
  • Test and tune network, hardware, and software configurations to maximize performance needs.
  • Troubleshoot infrastructure issues, reviewing log files, updating documentation, and having knowledge base with resolutions
  • Work closely with the application Development team to understand the platform and create tools/utilities to help with production management
  • Work with upstream data providers and upstream consumers, and reducing the amount of escalation to development teams
  • Develop scripts and assist with code changes along with operational tasks/activities.
  • Work closely with Application Development to ensure that the support team has excellent knowledge of the application set, own and maintain support knowledgebase and documents.
  • Use analytical skills to find trends in the environment and drive out problems.
  • Lead effort to determine improvement areas to stabilize the plant.
  • Identify risks and work with a sense of urgency, working within a team or independently.
  • Test and tune network, hardware, and software configurations to maximize performance * Interface with different teams like IT Dev managers, Infrastructure teams and lead as a Subject Matter Expert (SME) for the application(s) supported.
  • Understand the overall business flow of supported application systems and its interface with clients
  • Take ownership and managing production requests, questions, issues and perform Root Cause Analysis for outages/incident
  • Understand the overall business flow of supported application systems and its interface with clients
  • Be flexible to provide weekend on call rotation and available for offshore time lead
  • Be accountable for the Production Environments as well as the non-Production Environments for the existing GBOT team and be part of 24/7 production support coverage
What You’ll Bring To The Role
  • 10+ years of experience in a production environment with a solid software development background and understanding of performance tuning, end-to-end troubleshooting, networking fundamentals and appropriate attention to detail
  • Ability to focus, provide resolutions for production issues in a high demanding and pressured environment
  • 10+ years hands-on experience in designing, developing, and implementing technical solutions, or significant experience in deep technical support
  • Strong experience in scripting language (Shell scripting, Python, Perl, etc.) and cloud driven development
  • Strong database skills with DB2, Sybase or Oracle
  • Hands-on experience with Autosys or other batch scheduling software
  • Strong experience in Continuous Integration and Continuous Deployment
  • Strong experience in environment on demand for both Virtual Machines and containers
  • Knowledge and hands-on experience on with monitoring tools like Splunk, IP Soft, Sockeye
  • Practical experience on Agile Methodology (e.g. Scrum)
  • Knowledge or experience with automating deployments using Jenkins, Train or Windeploy
  • Ability to diagnose technical problems, debug, optimize code, and automate routine tasks
  • Hands-on experience in application and database troubleshooting/issue resolution in a fast-paced environment
  • Knowledge of Cloud based deployment, security, networking concepts in Azure and AWS
What You Can Expect From Morgan Stanley

At Morgan Stanley, we raise, manage and allocate capital for our clients – helping them reach their goals. We do it in a way that’s differentiated – and we’ve done that for 90 years. Our values - putting clients first, doing the right thing, leading with exceptional ideas, committing to diversity and inclusion, and giving back - aren’t just beliefs, they guide the decisions we make every day to do what's best for our clients, communities and more than 80,000 employees in 1,200 offices across 42 countries. At Morgan Stanley, you’ll find an opportunity to work alongside the best and the brightest, in an environment where you are supported and empowered. Our teams are relentless collaborators and creative thinkers, fueled by their diverse backgrounds and experiences. We are proud to support our employees and their families at every point along their work-life journey, offering some of the most attractive and comprehensive employee benefits and perks in the industry. There’s also ample opportunity to move about the business for those who show passion and grit in their work.

To learn more about our offices across the globe, please copy and paste https://www.morganstanley.com/about-us/global-offices into your browser.

Expected base pay rates for the role will be between $150,000 to $190,000 per year at the commencement of employment. However, base pay if hired will be determined on an individualized basis and is only part of the total compensation package, which, depending on the position, may also include commission earnings, incentive compensation, discretionary bonuses, other short and long-term incentive packages, and other Morgan Stanley sponsored benefit programs.

Morgan Stanley is an equal opportunity employer committed to building and maintaining a workforce that is diverse in experience and background. Our recruiting efforts reflect our strong commitment to a culture of inclusion, where individuals are hired, developed, and advanced based on their skills and talents.

Our workforce reflects a broad cross-section of the global communities in which we operate, bringing a variety of backgrounds, talents, perspectives, and experiences.

For more information, please visit: https://www.morganstanley.com/people-opportunities/eeo.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

Relha LLC • Alpharetta (GA), Northern (KY)

Hybrid
USD 125,000 - 175,000
Application Support
Application Support

PowerToFly • New York (NY)

On-site
USD 120,000 - 165,000
Comprehensive employee benefits
Opportunity for career advancement
SRE/Production Support Lead
SRE/Production Support Lead

Morgan Stanley • Alpharetta (GA)

On-site
USD 125,000 - 175,000
Executive Director – Site Reliability Engineering – WM Technology
Executive Director – Site Reliability Engineering – WM Technology

PowerToFly • New York (NY)

On-site
USD 195,000 - 215,000
Comprehensive employee benefits
Opportunity for career advancement
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Morgan-Stanley • Alpharetta (GA)

On-site
USD 125,000 - 175,000
Comprehensive benefits
Site Reliability Engineer
Site Reliability Engineer

PowerToFly • Alpharetta (GA)

On-site
USD 100,000 - 130,000
Software Engineer - Backend
Software Engineer - Backend

Socket.dev • New York (NY)

On-site
USD 120,000 - 170,000
Full Stack Java/Scala/DB2 Engineer - Vice President
Full Stack Java/Scala/DB2 Engineer - Vice President

PowerToFly • New York (NY)

On-site
USD 155,000 - 215,000
Linux Systems, Hardware and Infrastructure Services Engineer - Vice President
Linux Systems, Hardware and Infrastructure Services Engineer - Vice President

Socket.dev • New York (NY)

On-site
USD 155,000 - 215,000
Senior Network Engineer - Critical Infrastructure & Patch Deployment - Director
Senior Network Engineer - Critical Infrastructure & Patch Deployment - Director

PowerToFly • New York (NY)

On-site
USD 120,000 - 165,000