Senior Site Reliability Engineer - Cloud & Automation
OpsWerks
Mandaluyong
On-site
PHP 892,800 - 1,116,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
A technology solutions provider is seeking an experienced professional to serve as a Subject Matter Expert for hybrid cloud applications. Responsibilities include leading incident management, ensuring system performance, and driving operational improvements. Candidates need a Bachelor's degree and at least 5 years of experience in critical systems support, with strong skills in Linux administration and automation tools. This role is crucial for maintaining high-availability production systems in a collaborative DevOps environment.
Qualifications
Minimum 5 years of experience supporting critical, high-availability production systems.
Proven success in cross-functional collaboration within modern DevOps environments.
Experience supporting and maintaining PaaS environments and scalable, resilient architectures.
Responsibilities
Serve as Subject Matter Expert (SME) for distributed applications on hybrid cloud platforms.
Lead incident management and troubleshooting.
Proactively monitor, troubleshoot, and optimize application performance.
Skills
Linux Administration & Troubleshooting
Distributed Applications
Logging & Monitoring
Version Control
Automation using Bash/Python
Education
Bachelor’s degree in Information Technology, Engineering, or a related technical field
Tools
RHEL
CentOS
Ubuntu
Microservices
Splunk
Grafana
Prometheus
Git
GitHub
GitLab
Job description
A technology solutions provider is seeking an experienced professional to serve as a Subject Matter Expert for hybrid cloud applications. Responsibilities include leading incident management, ensuring system performance, and driving operational improvements. Candidates need a Bachelor's degree and at least 5 years of experience in critical systems support, with strong skills in Linux administration and automation tools. This role is crucial for maintaining high-availability production systems in a collaborative DevOps environment.