Network Operations Centre Engineer

DigiOutsource

Cape Town

On-site

ZAR 400,000 - 600,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Learning and development programmes
Performance feedback tools
Group Life Cover
Retirement Annuity Subsidy
Medical Aid Subsidy

Job summary

DigiOutsource in Cape Town is seeking a skilled Operations Analyst to monitor system performance and manage incidents effectively. Candidates should have a strong understanding of various monitoring tools and incident management practices, with relevant experience in operational environments.

The ideal candidate will possess excellent communication skills, IT certifications, and experience using tools like Jira for incident tracking.

Join us for development programs, a supportive feedback culture, and comprehensive benefits!

Qualifications

  • 3-5 years’ experience in Operations, NOC, or Incident Management.
  • Experience using major monitoring platforms like Nagios, SolarWinds.
  • ITIL Incident Management knowledge, ideally with ITIL v4 certification.

Responsibilities

  • Monitor system health and performance using designated tools.
  • Manage incidents using structured escalation and tracking.
  • Automate operational tasks to enhance efficiency.

Skills

Monitoring tools (Grafana, Datadog)
TCP/IP, DNS, HTTP, TLS
Incident Management (ITIL)
Scripting (Bash, Python)
Cloud technologies (AWS, Azure, GCP)

Education

CompTIA A+ / Network+ or Cisco CCNA
Relevant IT qualification

Tools

Jira
PagerDuty
Docker
Kubernetes

Job description

What You’ll Be Doing
  • Monitoring & Observability
    • Using monitoring tools such as Grafana, Datadog, SolarWinds, and Nagios to interpret dashboards, review alerts, and identify abnormal performance patterns or traffic deviations.
    • Correlating real‑time metrics, logs, and telemetry to detect system health concerns and escalating appropriately.
  • Networking & Platform Operations
    • Applying solid understanding of TCP/IP, DNS, HTTP, TLS, load balancing, and CDNs to support troubleshooting of platform issues.
    • Using working knowledge of distributed systems, caching, and messaging components to assist with fault isolation and impact assessment during incidents.
  • Incident Management Tooling
    • Using Jira for structured incident tracking, escalation, and resolution workflows.
    • Operating on‑call platforms such as PagerDuty and maintaining knowledge base/runbook documentation for consistent incident response.
  • Diagnostics & Troubleshooting
    • Performing first‑pass triage on server health, application performance, API latency, and database connectivity.
    • Analysing logs, metrics, and system indicators to narrow down root‑cause direction during high‑pressure incidents.
  • Scripting & Automation
    • Using basic Bash, Python, or PowerShell scripts for log extraction, parsing, or system checks.
    • Assisting with automating recurring operational tasks to reduce manual effort and improve consistency.
  • Cloud & Container Technologies
    • Understanding cloud fundamentals in AWS, Azure, or GCP to support cloud‑based troubleshooting or triage.
    • Basic exposure to Docker, Kubernetes, or log stacks such as ELK/Opensearch, Splunk, or Loki/Promtail to aid with diagnosing distributed workloads.
What You’ll Bring
  • Clear, confident communication (written and verbal), and the ability to break down complex ideas.
  • A collaborative mindset, working smoothly with cross‑functional teams to hit shared goals.
  • Strong organisational skills and the ability to manage multiple projects without dropping the ball.
  • Exceptional attention to detail and a commitment to high‑quality work.
  • Adaptability – staying sharp, productive and positive in fast‑moving environments.
  • A relevant IT qualification or industry certification, such as CompTIA A+ / Network+ or Cisco CCNA, or equivalent intermediate‑level technical certification.
  • 3 – 5 years’ experience in an Operations, NOC, or Incident Management environment with a focus on real‑time monitoring, incident detection, and structured escalation.
  • 3 – 5 years’ hands‑on experience using at least one major monitoring platform (e.g., Nagios, SolarWinds, Datadog, Grafana, Zabbix), including alert interpretation and basic correlation.
  • 3 – 5 years’ experience using enterprise ITSM tools such as Jira, ServiceNow, or Freshservice.
  • Practical exposure to ITIL Incident Management, demonstrated by ITIL v4 Foundation certification, or documented participation in a structured incident‑response workflow.
  • 3 – 5 years’ experience collaborating with engineering, support, platform, or SRE teams in a 24/7 operational or incident‑driven environment.
  • Prior experience in environments with rapid changes, live‑system dependencies, or peak‑traffic events (e.g., sports, e‑commerce, gaming, streaming, financial trading).
Desirable Skills You’ve Got Up Your Sleeve
  • A relevant Information Technology tertiary qualification.
  • Knowledge of sports betting markets, including odds calculation, betting types and market trends.
  • Previous experience in the online gaming or casino industry, with a strong understanding of player behaviour and industry regulations.
  • Familiarity with gambling regulations and compliance requirements in various jurisdictions, ensuring adherence to legal standards.
  • Experience with real‑time systems or high‑traffic platforms, ensuring performance and stability during peak betting or gaming events.
  • Strong verbal and written communication skills, with the ability to convey complex ideas clearly and effectively.
  • Experience working collaboratively in cross‑functional teams, with a focus on achieving shared goals.
Benefits
  • Learning and development programmes to level up fast.
  • Performance tool ensuring meaningful feedback and career support.
  • Employee Assistance Programme offering resources for you and your family.
  • Group Life Cover.
  • Funeral Fund Benefit.
  • Income Continuation Benefit.
  • Medical Aid Subsidy.
  • Retirement Annuity Subsidy.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Operations Centre Engineer
Network Operations Centre Engineer

Digital Outsource Services • Cape Town

On-site
Learning and development programmes
Employee Assistance Programme
Group Life Cover and Medical Aid Subsidy
Network Administrator
Network Administrator

Digital Outsource Services • Cape Town

On-site
ZAR 400,000 - 600,000
Comprehensive learning and development programmes
Employee assistance programme
Health and insurance benefits
+1
IT Support Technician
IT Support Technician

Digital Outsource Services • Cape Town

On-site
ZAR 200,000 - 300,000
Comprehensive learning and development programmes
Free daily meals
On-site gym
+2
Software Engineer
Software Engineer

Digital Outsource Services • Cape Town

On-site
ZAR 500,000 - 700,000
Comprehensive learning and development programmes
Free daily meals
On-site gym
+3
Operations Engineer (TTD)
Operations Engineer (TTD)

Sabenza IT & Recruitment • Pretoria

On-site
ZAR 420,000 - 660,000
Infrastructure Engineer
Infrastructure Engineer

ATS Client • Cape Town

Hybrid
ZAR 600,000 - 800,000
Discovery medical aid
21 days leave
Discretionary company performance bonu
NOCdesk Operator
NOCdesk Operator

AtripleA recruitment & temps • Pretoria

On-site
ZAR 300,000 - 450,000
Software Engineer (Full Stack)
Software Engineer (Full Stack)

Digital Outsource Services • Cape Town

On-site
Comprehensive learning and development programmes
Free daily meal
On-site gym
+4
Security Operations Analyst
Security Operations Analyst

DigiOutsource • Cape Town

On-site
ZAR 480,000 - 720,000
Learning & development programs
Performance management tools
Employee Assistance Programme
+5
IT Specialist
IT Specialist

OneDayOnly.co.za • Cape Town

On-site
ZAR 300,000 - 420,000