IT Observability Engineer

Home Credit Philippines

Taguig

On-site

PHP 1,200,000 - 2,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

HMO on Day 1
Dependent coverage
Mental health support
Wellness leaves
Birthday leave
Internal career mobility
Learning opportunities
Up to 20% performance bonus

Job summary

Home Credit Philippines is seeking an IT Operations Observability Engineer to ensure smooth operation of IT infrastructure and applications through proactive monitoring, dashboards, and data-driven insights. You will prioritize incidents and collaborate across teams to maintain uptime.

The role requires 4+ years in IT operations focused on monitoring, observability, and incident response, with experience in Prometheus, Grafana, Splunk, Jaeger, and cloud monitoring (AWS/Azure).

Qualifications

  • 4+ years in IT operations or similar role with monitoring/observability focus.
  • Strong analytical and problem-solving skills with data-driven insights.
  • Excellent communication and collaboration abilities across diverse teams.
  • Passion for continuous learning and observability expertise development.
  • Knowledge of ITIL v3 processes.
  • Proficient in scripting (Python, Bash, or Go) for automation.
  • Experience with distributed tracing and containerized environments.
  • Familiarity with incident response methodologies and best practices.
  • Hands-on experience with Prometheus, Grafana, Splunk, Jaeger, or similar tools.

Responsibilities

  • Design and implement monitoring systems with metrics, logs, and traces for network, apps, and servers.
  • Build dashboards to convert data into actionable insights for stakeholders.
  • Create intelligent alerts to enable proactive issue prevention.
  • Investigate incidents, diagnose root causes, and drive resolution with teams.
  • Stay current with observability tools and technologies; continuously improve monitoring landscape.
  • Collaborate with engineers to optimize performance and uptime across systems.
  • Prepare and maintain detailed RCA reports and knowledge-sharing materials.
  • Maintain monitoring dashboards and ensure ongoing effectiveness through reviews.
  • Perform related duties as assigned.

Skills

Monitoring and observability
Incident prioritization
Analytical thinking
Scripting/automation
Communication
Collaboration across teams
ITIL familiarity
Cloud monitoring

Tools

Prometheus
Grafana
Splunk
Jaeger
Zipkin
Docker
Kubernetes
AWS CloudWatch
Azure Monitor

Job description

The IT Operations Observability Engineer primary purpose is to ensure the smooth and efficient operation of IT infrastructure and applications through proactive and data-driven monitoring and analysis. Ensure system stability and enable proactive issue resolution through comprehensive visibility and monitoring. This involves a combination of technical skills and analytical thinking to provide deep insights into system performance and behavior. Must be able to prioritize incidents effectively and elevate to the correct teams.

What You’ll Do
  • Design and implement monitoring systems: Craft a robust toolkit of metrics, logs, and traces to keep tabs on the health and performance of our network, applications, and servers.
  • Unleash the power of visualization: Build intuitive dashboards that transform complex data into actionable insights, informing stakeholders and guiding decision‑making.
  • Become the alert whisperer: Set up intelligent alerts that shout "Hold on!" before an issue snowballs, ensuring proactive problem prevention.
  • Dive into the details: Investigate and diagnose incidents with a surgeon's precision, unearthing the root cause and resolving challenges efficiently.
  • Embrace the evolution: Stay current with the latest observability tools and technologies, constantly learning and adapting to keep our monitoring landscape cutting‑edge.
  • Collaborate with the best: Partner with engineers across teams to share knowledge, optimize system performance, and build a culture of proactive observability.
  • Participates in the creation and maintenance of detailed reports summarizing root cause investigation results and key findings for knowledge sharing within the team.
  • Partner with engineers across teams to safeguard system performance and deliver exceptional uptime.
  • Proactively ensures the ongoing effectiveness and accuracy of group-built monitoring dashboards through regular maintenance and performance reviews.
  • Performs all other related duties as assigned.
What You Need To Have
  • Brings 4+ years of experience in IT operations or a similar role, with a strong foundation in monitoring and observability principles.
  • Demonstrates excellent analytical and problem‑solving skills, adept at uncovering insights within complex data sets.
  • Has strong communication and collaboration skills, able to effectively communicate technical information to diverse audiences.
  • Exhibits a passion for continuous learning and personal growth, eager to refine their observability expertise.
  • Knowledgeable in ITIL v3 Processes.
  • Enjoys scripting and automation, with proficiency in languages like Python, Bash, or Go for data manipulation and automation tasks.
  • Understanding of distributed tracing systems and containerized environments.
  • Familiarity with incident response methodologies and best practices.
  • Knowledge or background in Splunk or other monitoring system.
  • Expertise in Prometheus, Grafana, Splunk, Jaeger, or similar industry-leading tools.
  • Experience with cloud-based observability platforms like AWS CloudWatch or Azure Monitor.
  • Knowledge of security monitoring tools and incident response best practices.
  • Has experience with incident response methodologies and best practices.
What Can Set You Apart
  • Experience with distributed tracing systems like Jaeger or Zipkin.
  • Familiarity with containerized environments and orchestrators like Docker and Kubernetes.
  • Up to 20% variable performance-based bonus.
  • HMO on Day 1 and HMO dependents coverage including same-sex partners.
  • Access to mental health and wellness partners.
  • Wellness Leaves and Birthday Leave.
  • Internal career mobility options.
  • Local and international learning opportunities.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

IT Observability Engineer
IT Observability Engineer

Home Credit Philippines • Philippines

On-site
PHP 900,000 - 1,200,000
Permanent dayshift schedule
Up to 20% variable bonus
HMO day 1 and dependents coverage
+4
IT Infrastructure Engineer (Monitoring and Observability Team)
IT Infrastructure Engineer (Monitoring and Observability Team)

ACCION LABS PHILIPPINES, INC. • Quezon City

On-site
PHP 600,000 - 1,000,000
Observability Engineer
Observability Engineer

Avaloq AG • Makati

On-site
PHP 900,000 - 1,500,000
IT Specialist, IT Service Intelligence
IT Specialist, IT Service Intelligence

DSV - Global Transport and Logistics • Parañaque

On-site
PHP 1,000,000 - 1,500,000
Modern observability & cloud tech
Exposure to large-scale enterprise env
Collaborative, engineering-driven kult
+1
IT Specialist, IT Service Intelligence
IT Specialist, IT Service Intelligence

Dsv Air & Sea SAU • Manila

On-site
PHP 800,000 - 1,200,000
Senior ITSMA Observability Engineer
Senior ITSMA Observability Engineer

HedgeServ (Philippines) LLC – Manila Branch • Manila

On-site
PHP 1,200,000 - 2,000,000
Competitive salary
Benefits packages
Ongoing learning and development opportunities
IT Operations Engineer
IT Operations Engineer

ING Hubs Philippines • Manila

On-site
PHP 600,000 - 1,200,000
Collaborative work environment
Continuous learning and development
Exposure to global systems and large‑m
+1
Observability Engineer
Observability Engineer

Avaloq • Makati

Hybrid
PHP 900,000 - 1,500,000
Hybrid work model
Flexible working
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Omilia • Philippines

On-site
PHP 1,000,000 - 1,800,000
Fixed compensation
Vacation leaves
Professional development opportunities
+3
Monitoring, Observability and Event Management Architect
Monitoring, Observability and Event Management Architect

Gratitude Philippines • Quezon

On-site
PHP 800,000 - 1,200,000
Competitive compensation
Learning and development support
Meaningful products and problems