SRE Observability Technical Lead - Vice President

Citigroup Inc.

Belfast City District

On-site

GBP 70,000 - 110,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

27 days annual leave
Discretionary annual bonus
Private Medical Care & Life Insurance
Employee Assistance Program
Pension Plan
Paid Parental Leave
Employee discounts
Learning and development resources

Job summary

Citigroup Inc. in Belfast is seeking an Observability/SRE specialist to engineer the future of finance through scalable telemetry, SLOs and dashboards for critical payment flows.

You will work with SREs, developers and platform teams to embed telemetry, define SLOs, and build real‑time visualizations across Payments. The role offers hybrid work options and a path for growth within Citi Tech.

Qualifications

  • Hands-on experience in SRE, Observability Engineering, or platform infrastructure roles.
  • Deep knowledge of observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, and Splunk.
  • Strong understanding of SLIs, SLOs, and telemetry best practices in high-availability environments.
  • Experience troubleshooting integration issues across hybrid platforms (on-prem, cloud, containers).
  • Experience building dashboards aligned to business outcomes and incident workflows, especially in payments.
  • Familiarity with AI/ML capabilities in observability tooling; alert tuning.
  • Excellent collaboration across federated teams and central infrastructure groups.
  • Experience enabling platform teams and scaling best practices.

Responsibilities

  • Define the roadmap for engineering enablers for Project Orion aligned with enterprise reliability goals.
  • Translate strategy into actionable delivery plans with Services Products, Operations & Engineering.
  • Develop End-to-End monitoring solutions for critical business services.
  • Build scalable telemetry dashboards and visualizations for key journeys (Payments).
  • Review monitoring toil and remediate with stakeholders to reduce MTTR.
  • Guide LOB teams in SLIs/SLOs, golden signals and alerting practices.
  • Support integration and adoption of observability tooling across on‑prem, cloud and containers.

Skills

SRE experience
Observability
Hybrid/cloud environments

Education

Bachelor’s degree in Computer Science or related field

Tools

Grafana
Prometheus
OpenTelemetry
ELK
Splunk

Job description

Engineer the future of global finance. At Citi, our Tech team doesn’t just support finance – we are helping to redefine it. Every day, $5 trillion crosses through our network. We do business in 180+ countries operating at a scale few can match. From deploying advanced AI to helping shape global markets, we build systems that matter. Look to join a team where your work helps influence economies, your ideas can drive innovation and outcomes, and your growth is backed by mentorship, continuous learning and flexibility with potential hybrid work opportunities. Help solve real-world challenges that touch millions and get the opportunity to build the future of finance with Citi Tech.

The SRE Observability Specialist is a hands‑on expert, delivering the future of Observability across Services Technology. This role is a part of a central SRE enablement team within Services Production, working closely with SREs, developers, and platform teams to embed telemetry, implement SLOs, and build meaningful visualizations for key production flows — particularly in critical Payments Business.

The ideal candidate will have deep technical knowledge, a collaborative mindset, and the ability to translate strategy into scalable engineering outcomes. You will also act as a bridge between Services Technology teams and central infrastructure/CTO teams, prioritising observability needs from line‑of‑business teams and driving improvements. A strong understanding of observability tooling, evolving AI/ML capabilities, and enterprise tooling ecosystems will be essential.

This role requires providing technological Support solution for Function called Project Orion which provides End‑to‑End payment monitoring like Building an End‑to‑End payments Dashboard, Toil Reduction, Transformation of legacy monitoring into observability based monitoring solution, requires good understanding of different Payments Taxonomy (ACH, Wires, Instant Payments, etc.). Strong commercial awareness, technical credibility, and excellent communication skills are essential to negotiate internally, influence peers, and drive change. Some external communication may be necessary.

Key Responsibilities
  • Define the roadmap for Engineering enablers for Project Orion team aligned with enterprise reliability and SRE Services organization goals.

  • Translate Organization strategy into an actionable delivery plan in partnership with Services Products, Operations & Engineering function, delivering incremental, high‑value milestones.

  • Understand Critical Business Services functional scope and translate into End‑to‑End monitoring solutions.

  • Deliver against the observability roadmap for Services Technology by building scalable, reusable telemetry solutions.

  • Periodic review and analyze application monitoring TOIL and collaborate with stakeholders and remediate them as per organization goal.

  • Create and maintain dashboards and visualizations for critical client journeys, including real‑time flows across Payments.

  • Guide line‑of‑business teams in implementing SLIs/SLOs, golden signals, and effective alerting to support operational excellence.

  • Support integration and adoption of observability tooling across on‑prem, public cloud (AWS/GCP), and containerized environments (ECS, Kubernetes).

  • Customize shared dashboards and observability components in partnership with CTI and other central Engineering functions, ensuring usability and flexibility.

  • Provide technical support and implementation guidance to SREs and developers facing integration or tooling challenges.

  • Effectively manage the observability book of work for Services Technology and drive initiatives to reduce MTTD and improve recovery outcomes.

  • Serve as a key connection point between line‑of‑business SREs and central infrastructure functions by gathering tooling feedback, surfacing systemic issues, and influencing platform enhancements via the Services Observability Forum.

  • Stay current with observability trends, including AI/ML‑driven insights, anomaly detection, and emerging OSS practices, and assess their applicability.

  • Maintain strong knowledge of observability platform features and vendor offerings to advise teams and maximize the value of tooling investments.

  • Foster AI adoption by building use cases performed by Orion L1 Functions and remediation using Citi AI tech stack.

Qualifications
  • Experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry.

  • Hands‑on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.

  • Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high‑availability environments.

  • Proven ability to troubleshoot integration issues and support observability across hybrid platforms (on‑prem, cloud, containers).

  • Experience building dashboards aligned to business outcomes and incident workflows, especially in critical flows like payments.

  • Familiarity with modern observability tooling ecosystems, including AI/ML capabilities, trace correlation, baselining, and alert tuning.

  • Strong interpersonal and collaboration skills — able to operate across federated engineering teams and central infrastructure groups.

  • Experience in enablement or platform teams with a track record of scaling best practices across diverse business units.

Education
  • Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.

What we’ll provide you

By joining Citi, you will not only be part of a business casual workplace with a hybrid working model (up to 2 days working at home per week), but also receive a competitive base salary (which is annually reviewed), and enjoy a whole host of additional benefits such as:

  • 27 days annual leave (plus bank holidays)

  • A discreational annual performance related bonus

  • Private Medical Care & Life Insurance

  • Employee Assistance Program

  • Pension Plan

  • Paid Parental Leave

  • Special discounts for employees, family, and friends

  • Access to an array of learning and development resources

Alongside these benefits Citi is committed to ensuring our workplace is where everyone feels comfortable coming to work as their whole self, every day. We want the best talent around the world to be energized to join us, motivated to stay and empowered to thrive.

#LI-BH3

Job Family Group:

Technology

Job Family:

Applications Support

Time Type:

Full time

Most Relevant Skills

Please see the requirements listed above.

Other Relevant Skills

For complementary skills, please see above and/or contact the recruiter.

Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.

If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi.

View Citi’s EEO Policy Statement and the Know Your Rights poster.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Observability Technical Lead - Vice President
SRE Observability Technical Lead - Vice President

Citi • Belfast City District

Hybrid
GBP 90,000 - 110,000
Hybrid work options
Competitive base salary
27 days annual leave
+5
Observability Engineer - Assistant Vice President
Observability Engineer - Assistant Vice President

Citibank (Switzerland) AG • Greater London

Hybrid
Confidential
Annual leave 27d
Discretionary bonus
Medical & life insurance
+5
Application Support Technical Lead Analyst - Vice President
Application Support Technical Lead Analyst - Vice President

Citi • Belfast City District

Hybrid
GBP 90,000 - 120,000
27 days annual leave
Discretional annual bonus
Private medical & life insurance
+3
Observability Engineer - Assistant Vice President
Observability Engineer - Assistant Vice President

Citigroup Inc. • Greater London

On-site
GBP 120,000 - 180,000
Observability Engineer - Assistant Vice President
Observability Engineer - Assistant Vice President

Citi • Greater London

On-site
GBP 90,000 - 120,000
Custody Support Applications Support - Assistant Vice President
Custody Support Applications Support - Assistant Vice President

Citigroup Inc. • Belfast City District

Hybrid
GBP 70,000 - 110,000
27 days annual leave
Performance related bonus
Private Medical Care & Life Insurance
+5
Full Stack Developer - Assistant Vice President
Full Stack Developer - Assistant Vice President

Citigroup Inc. • Belfast City District

Hybrid
GBP 85,000 - 115,000
27 days annual leave
Discretionary bonus
Private Medical Care & Life Insurance
+5
Custody Support - Application Support Technical Lead - Vice President
Custody Support - Application Support Technical Lead - Vice President

Citi • Belfast City District

Hybrid
GBP 70,000 - 95,000
27 days annual leave (plus bankhols)
Discretional annual bonus
Private Medical Care & Life Insurance
+5
Principal Site Reliability Engineer, Infrastructure Observability
Principal Site Reliability Engineer, Infrastructure Observability

United States Digital Space LLC • Greater London

Hybrid
GBP 120,000 - 170,000
Hybrid work up to 3 days per week
Senior Scala Engineer (SolstiCE) – Equity Derivatives Tech – VP
Senior Scala Engineer (SolstiCE) – Equity Derivatives Tech – VP

Citi • Greater London

Hybrid
GBP 90,000 - 120,000
Paid Parental Leave
27 days annual leave
Private Medical Care & Life Insurance
+2