Site Reliability Engineer

mthree

Charlotte (NC)

On-site

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Strong culture of equality
Inclusion initiatives
Diversity promotion

Job summary

A financial technology consultancy is looking for a Production Support Analyst to join their Charlotte team. Ideal candidates will possess excellent communication skills, a problem-solving mindset, and the ability to thrive in a fast-paced environment. The role demands experience ranging from 3-10 years, with a background in service management and payments preferred. Applicants must be authorized to work in the US on a full-time basis, with no sponsorship available.

Qualifications

  • 3-5 years experience for juniors, 10+ years for mid-level.
  • Hands-on experience with UNIX and SQL for troubleshooting.
  • Knowledge in scripting languages like Python, Bash, Perl, Ruby.

Responsibilities

  • Embed SRE and production engineering principles into Payments Modernization.
  • Define and validate non-functional requirements (NFRs).
  • Lead performance testing for event-driven payment flows.

Skills

Service management experience
Payments knowledge
Analytical skills
Communication skills
Problem solving mindset

Tools

Kubernetes
SpringBoot
MongoDB
Kafka
SQL
UNIX

Job description

Want to work in technology in the financial industry?

We are looking for a Production Support Analyst to join our client's team in Charlotte, NC. We are looking for someone with excellent written and verbal communication skills, energetic, good follow-up that has a curious nature that drives them to solve problems. Ability to work in a fast paced and high demand environment. Financial background is desired.

About mthree:

Since 2010, mthree has been helping clients solve their business and technological challenges. We are a technology and business consultancy with a global workforce delivering significant business and IT projects in some of the largest financial services organizations worldwide.

Core Services:

  • Consulting and Advisory
  • Managed Services
  • Alumni Graduate Program

We have a global presence and are experts in delivering exceptional quality to our client base, providing consulting services across Risk, Regulation & Compliance; Vendor Products; Application Support; Application Development; Cyber & Information Security; Data Science and DevOps areas.

Our Expert program offers experienced professionals access to top roles in tech, finance, aviation and insurance. Join us to work on groundbreaking technology projects, from international trading platforms to critical applications for leading airlines. We recruit professionals who are eager to fast-track their careers in technology or operations within prestigious global organizations.

Responsibilities:
Platform & Reliability Engineering
  • Embed SRE and production engineering principles into Payments Modernization from design through early life support
  • Define and validate non-functional requirements (NFRs) covering resilience, scalability, observability, recovery, and operability
  • Drive replay, retry, and exception-handling validation for event-driven payment flows
  • Lead capacity and performance testing, including volume growth and peak event scenarios (e.g. FedNow, CHIPS, SWIFT)
Service Transition & Operational Readiness
  • Own Permit-to-Operate readiness across environments (NFR Testing)
  • Define cutover, shadow support, and early life support models
  • Ensure runbooks, support procedures, on-call readiness, and escalation paths are production-grade before go-live
  • Partner with Change Assurance to apply risk-based release controls, canary/blue-green strategies, and rollback automation
Observability & Stability
  • Implement end-to-end observability across Kafka, MongoDB, API layers, and downstream payment components
  • Define and monitor SLOs, error budgets, and golden signals
  • Reduce alert noise through signal design, correlation, and automation
  • Analyze early defects and exception patterns (ACK/NACKs, business errors) to drive stabilization
  • Design and execute controlled failure testing (chaos engineering) to validate recovery patterns and blast radius
  • Lead blameless RCAs, ensuring corrective actions are owned and recurrence is prevented
  • Drive continuous service improvement (CSI) initiatives, including automation, resilience uplift, and technical debt reduction
Required Experience
  • Range from juniors with 3-5 years experience to mid range, 10+ years.
  • Service management experience, payments knowledge and tech wise knowledge on framework such as springboot, mongodb, kakfa, Kubernetes/ CI/CD pipelines
  • Hands on experience with UNIX, SQL to assist with troubleshooting
  • Knowledge of Automation Related activities using scripting languages such as Python, Bash, Perl, Ruby
  • Excellent analytical and communication skills
  • Ability to prioritize and willingness to take ownership
  • Problem solving mindset and solution enabler

At mthree, our values support courageous teammates, needle movers, and learning champions all while striving to support the health and well-being of all employees. We take great pride in celebrating the diversity of each individual who contributes to making mthree the company it is today and will be in the future. We value diversity both within mthree and with our partner companies, and we're proud to provide an environment where all our colleagues can flourish. That means promoting a strong culture of equality but, most importantly, inclusion.

Applicants must be currently authorized to work in the United States on a full-time basis. The Company will not sponsor applicants for work visas.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Reliability & Production Engineer (RPE)
Reliability & Production Engineer (RPE)

mthree • New York (NY)

On-site
USD 85,000 - 110,000
Reliability & Production Engineer
Reliability & Production Engineer

mthree Recruiting Portal • New York (NY)

On-site
USD 85,000 - 110,000
Production Support Analyst
Production Support Analyst

mthree • New York (NY)

On-site
USD 71,000 - 119,000
Full Stack Developer
Full Stack Developer

mthree Recruiting Portal • Jersey City (NJ)

On-site
USD 140,000 - 170,000
Site Reliability Engineer / Production Support Analyst
Site Reliability Engineer / Production Support Analyst

mthree Recruiting Portal • United States

On-site
USD 70,000 - 90,000
Flexible benefits package
Ongoing training and support
Valuable industry experience
Site Reliability Engineer / Production Support Analyst
Site Reliability Engineer / Production Support Analyst

jobr.pro • Salt Lake City (UT)

On-site
USD 56,000 - 58,000
Generous salary and annual increases
Flexible benefits package
Ongoing training and support
+1
Site Reliability Engineer / Production Support Analyst
Site Reliability Engineer / Production Support Analyst

Dormont Manufacturing Co • South Jordan (UT)

On-site
USD 56,000 - 58,000
Annual salary increases
Flexible benefits package
In-depth training from industry experts
+1
Production Support/SRE Analyst at mthree
Production Support/SRE Analyst at mthree

University of Delaware • United States

Hybrid
USD 50,000 - 70,000
Flexible benefits package
Ongoing training and support
Annual salary increase
Full Stack Developer
Full Stack Developer

mthree • New Jersey

On-site
USD 140,000 - 170,000
Production Support/SRE Analyst - Houston, New York
Production Support/SRE Analyst - Houston, New York

mthree Recruiting Portal • New York (NY), Houston (TX)

On-site
USD 58,000 - 70,000
Training via mthree Academy
Annual pay rise
Competitive compensation package
+2