Reliability Engineer: SRE, Observability & Cloud Ops

E4 Software Services Pvt Ltd.

Hinoba-an

On-site

PHP 893,000 - 1,674,000

Full time

9 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

E4 Software Services Pvt Ltd. is seeking an experienced SRE/DevOps professional to monitor, manage and improve production environments. You will focus on reliability, observability and automation to reduce toil.

Responsibilities include incident handling, RCA, root cause actions, and collaboration with development and infrastructure teams to boost resilience. Knowledge of cloud platforms (AWS/Azure/GCP), Kubernetes, Docker, and CI/CD is essential.

Qualifications

  • 4+ years in SRE, Production Support, DevOps or Cloud Operations.
  • Strong Linux/Unix administration and troubleshooting skills.
  • Experience with monitoring and observability platforms.
  • Hands-on experience with Splunk, ELK, Dynatrace, Datadog, Prometheus or Grafana.
  • Strong scripting using Python, Bash or Shell.
  • Experience with incident management, problem management and RCA.
  • Good understanding of cloud platforms such as AWS, Azure or GCP.
  • Knowledge of networking concepts including TCP/IP, DNS and load balancing.
  • Experience with CI/CD and DevOps practices.
  • Strong production troubleshooting and operational support experience.

Responsibilities

  • Monitor and maintain the availability, performance and reliability of production environments.
  • Design and manage monitoring, logging, alerting and observability solutions.
  • Analyze application and infrastructure logs to identify performance and reliability issues.
  • Handle production incidents, service requests and operational escalations.
  • Perform root cause analysis and implement permanent corrective actions.
  • Develop automation to reduce repetitive operational activities and manual intervention.
  • Define and monitor reliability indicators, SLIs, SLOs and operational KPIs.
  • Support cloud environments across AWS, Azure and/or GCP.
  • Troubleshoot Linux/Unix, networking, application and infrastructure issues.
  • Support Kubernetes, Docker and CI/CD environments where required.
  • Maintain operational runbooks and improve incident response processes.
  • Work with development and infrastructure teams to improve system resilience.
  • Participate in on-call and production support activities.

Skills

SRE Experience
Linux Administration
Monitoring/Observability
Monitoring Tools
Scripting
Incident RCA
Cloud Platforms
Networking
CI/CD
Operational Support

Tools

Kubernetes
Docker
Apache/Tomcat
ITSM tools (ServiceNow/Jira)
Ansible/Chef/Puppet
SAP support
OpenTelemetry/Moogsoft/Rundeck
Prometheus/Grafana

Job description

E4 Software Services Pvt Ltd. is seeking an experienced SRE/DevOps professional to monitor, manage and improve production environments. You will focus on reliability, observability and automation to reduce toil.

Responsibilities include incident handling, RCA, root cause actions, and collaboration with development and infrastructure teams to boost resilience. Knowledge of cloud platforms (AWS/Azure/GCP), Kubernetes, Docker, and CI/CD is essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. • Hinoba-an

On-site
PHP 893,000 - 1,674,000
Senior SRE Operations Engineer - Cloud & Observability
Senior SRE Operations Engineer - Cloud & Observability

Teciem • Hinoba-an

Hybrid
PHP 2,772,000 - 4,291,000
Cloud Platform & SRE Engineer — Kubernetes & Automation
Cloud Platform & SRE Engineer — Kubernetes & Automation

Global Recruitment and Consultancy OPC • Cebu City

On-site
PHP 1,200,000 - 2,400,000
Associate SRE — Platform Reliability & Automation
Associate SRE — Platform Reliability & Automation

Railway Corp • Mexico

On-site
PHP 5,474,000 - 7,908,000
Senior SRE: Cloud Reliability, Automation & Observability
Senior SRE: Cloud Reliability, Automation & Observability

Broadridge Financial Solutions • Manila, Hinoba-an

On-site
PHP 1,800,000 - 2,600,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Senior Site Reliability Engineer: Cloud & Observability
Senior Site Reliability Engineer: Cloud & Observability

Omilia • Philippines

On-site
PHP 1,000,000 - 1,800,000
Fixed compensation
Vacation leaves
Professional development opportunities
+3
Senior Cloud SRE — AWS, Automation & Observability
Senior Cloud SRE — AWS, Automation & Observability

Broadridge • Metro Manila

On-site
PHP 1,200,000 - 2,400,000
Senior Cloud SRE: AWS & Kubernetes (Remote, On-Call Ownership)
Senior Cloud SRE: AWS & Kubernetes (Remote, On-Call Ownership)

Salve.Inno Consulting • Manila

On-site
PHP 1,200,000 - 1,800,000
Fully remote position
Full-time B2B cooperation
Work on complex cloud environments
+1
Senior SRE: Scale, Automate & Self-Healing Systems
Senior SRE: Scale, Automate & Self-Healing Systems

Replit • España

On-site
PHP 2,461,000 - 4,308,000
Competitive Salary & Equity
Health, Dental, Vision and Life Insurance
Flexible Time Off (FTO) + Holidays
+2