Observability Engineer

Vitality

Johannesburg

Hybrid

ZAR 900,000 - 1,300,000

Full time

41 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Bonus scheme
Retirement match
Health plan
Life & disability insurance

Job summary

Vitality invites an experienced Observability Engineer to join the Platform Operations team in a hybrid role based in Sandton. You will design, implement and maintain observability solutions to boost health, performance and reliability of applications and infrastructure across cloud environments.

You will work with CI/CD pipelines, SRE and DevOps teams to embed observability-first practices, create dashboards, automate logging and tracing, and support rapid incident response in a dynamic,

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering or related field.
  • 5+ years’ experience in DevOps, Site Reliability Engineering, Platform Engineering or similar technical role.
  • 5+ years’ hands-on experience with observability tools and platforms.
  • Strong knowledge of cloud platforms including AWS, Azure or GCP.
  • Experience with Kubernetes and containerised environments.
  • Understanding of microservices architectures and distributed systems.
  • Experience with AIOps and predictive monitoring approaches.
  • Knowledge of FinOps principles and cost observability practices.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Dynatrace or similar platforms.
  • Knowledge of CI/CD tools and Infrastructure-as-Code technologies including Terraform, ARM or CloudFormation

Responsibilities

  • Design and implement observability frameworks across distributed systems and cloud-based environments
  • Establish standards for metrics, logging, tracing and monitoring to improve visibility across applications and infrastructure
  • Build and maintain dashboards that provide real-time insight into system performance, availability and reliability
  • Implement intelligent alerting solutions that reduce noise and improve issue detection and response times
  • Develop and enhance centralised logging and distributed tracing capabilities to support root cause analysis
  • Support incident response activities by providing actionable diagnostics and operational insights
  • Identify observability gaps through post-incident reviews and drive continuous improvement initiatives
  • Automate observability processes using scripting, Infrastructure-as-Code and platform tooling
  • Integrate monitoring, logging and observability practices into CI/CD pipelines and engineering workflows
  • Collaborate with DevOps, SRE and development teams to embed observability-first practices across the organisation

Skills

OpenTelemetry
CI/CD
APM

Education

Bachelor's degree in CS/IT/Engineering or related field

Tools

Prometheus
Grafana
Dynatrace

Job description

Team - V1 Technology Operations
Working Pattern

Hybrid - 2 days per week in the Vitality Sandton Office. Full time, 37.5 hours per week.

We are happy to discuss flexible working!

Top 3 skills needed for this role
  • Proven OpenTelemetry Implementation Expertise
  • Strong CI/CD Pipeline Integration Capability
  • Advanced APM Platform Mastery
What this role is all about

We are seeking an Observability Engineer to join the Platform Operations function, reporting to the Head of Platform Operations. This role is responsible for designing, implementing and maintaining observability solutions that provide deep visibility into the health, performance and reliability of applications and infrastructure. The position plays a critical role in strengthening platform resilience, improving incident response capabilities and enabling data-driven operational decision making across complex cloud and distributed environments.

Key Actions
  • Design and implement observability frameworks across distributed systems and cloud-based environments
  • Establish standards for metrics, logging, tracing and monitoring to improve visibility across applications and infrastructure
  • Build and maintain dashboards that provide real-time insight into system performance, availability and reliability
  • Implement intelligent alerting solutions that reduce noise and improve issue detection and response times
  • Develop and enhance centralised logging and distributed tracing capabilities to support root cause analysis
  • Support incident response activities by providing actionable diagnostics and operational insights
  • Identify observability gaps through post-incident reviews and drive continuous improvement initiatives
  • Automate observability processes using scripting, Infrastructure-as-Code and platform tooling
  • Integrate monitoring, logging and observability practices into CI/CD pipelines and engineering workflows
  • Collaborate with DevOps, SRE and development teams to embed observability-first practices across the organisation
What do you need to thrive
  • Bachelor's degree in Computer Science, Information Technology, Engineering or a related field, or equivalent industry experience
  • 5+ years’ experience in DevOps, Site Reliability Engineering, Platform Engineering or a similar technical role
  • 5+ years’ hands-on experience working with observability tools and platforms
  • Strong knowledge of cloud platforms including AWS, Azure or GCP
  • Experience working with Kubernetes and containerised environments
  • Understanding of microservices architectures and distributed systems
  • Experience with AIOps and predictive monitoring approaches
  • Knowledge of FinOps principles and cost observability practices
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Dynatrace or similar platforms
  • Knowledge of CI/CD tools and Infrastructure-as-Code technologies including Terraform, ARM or CloudFormation
So, what’s in it for you
  • Bonus Schemes - A bonus that regularly rewards you for your performance
  • Retirement support - We will match your contributions up to 5% of your salary
  • Health & wellbeing - Discovery Medical Health Scheme
  • Financial protection - Life assurance, income protection and short and long-term disability insurance
Fantastic Benefits. Exciting rewards. Great career opportunities

These are just some of the many perks that we offer!

If you are successful in your application and join us at Vitality, this is our promise to you, we will:

  • Help you to be the healthiest you’ve ever been.
  • Create an environment that embraces you as you are and enables you to be your best self.
  • Give you flexibility on how, where and when you work.
  • Help you advance your career by playing you to your strengths.
  • Give you a voice to help our business grow and make Vitality a great place to be.
  • Give you the space to try, fail and learn.
  • Provide a healthy balance of challenge and support.
  • Recognise and reward you with a competitive salary and amazing benefits.
  • Be there for you when you need us.
  • Provide opportunities for you to be a force for good in society.

We commit to all these things because we want you to feel that you belong, and are supported to be happy and healthy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Vitality • Johannesburg

Hybrid
ZAR 700,000 - 1,000,000
Bonus schemes
Retirement support
Health & wellbeing
+1
Observability Engineer: Cloud, CI/CD & SRE
Observability Engineer: Cloud, CI/CD & SRE

Vitality • Johannesburg

Hybrid
ZAR 900,000 - 1,300,000
Bonus scheme
Retirement match
Health plan
+1
Senior Observability Specialist - SOS
Senior Observability Specialist - SOS

Nedbank • Johannesburg

On-site
ZAR 1,200,000 - 2,000,000
Senior Administrator Application Support
Senior Administrator Application Support

Vodacom • Johannesburg

On-site
ZAR 720,000 - 960,000
Incentives
Medical aid benefits
Staff discounts
Platform Engineering Lead
Platform Engineering Lead

Vodacom • Johannesburg

On-site
ZAR 1,200,000 - 1,900,000
Medical aid benefits
Retirement funds
Staff discounts
+1
DevOps Engineer Corporate intelligence Cape Town
DevOps Engineer Corporate intelligence Cape Town

S-RM Intelligence and Risk Consulting • Cape Town

On-site
ZAR 700,000 - 1,000,000
23 days holiday per year
Pension contribution up to 7%
Hybrid working
+3
Senior Administrator Application Support
Senior Administrator Application Support

Vodafone • Midrand

On-site
ZAR 420,000 - 560,000
Enticing incentive programs
Retirement funds
Medical aid benefits
+3
Business Intelligence Analyst
Business Intelligence Analyst

LekkeSlaap • Cape Town

On-site
ZAR 600,000 - 800,000
Hybrid work model & flexible start times
Free lunch when in office
Travel vouchers and discounts
+2
Site Reliability Engineer
Site Reliability Engineer

Impact.Com • Cape Town

Hybrid
ZAR 800,000 - 1,400,000
Flexible Working
Health and Wellness
RSUs with 3-year vesting
+3
Senior Full-Stack Software Developer
Senior Full-Stack Software Developer

ATS Client • Sandton

Hybrid
ZAR 900,000 - 1,500,000
Hybrid work model
Equipment provided
Data & connectivity allowance
+1