JD IC4 - Sr Infra Engineer - SRE

Spin Careers

Philippines

On-site

PHP 1,200,000 - 2,100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Spin Careers is seeking a Senior SRE Engineer in the Philippines to lead the reliability strategy across IT infrastructure and applications. You will champion advanced monitoring, incident response, automation, and performance analysis while mentoring junior engineers and shaping IT operations.

A strong focus on security, documentation, and cross-functional collaboration will be essential. You will drive continuous improvement, capacity planning, and disaster recovery planning, partnering with

Qualifications

  • Bachelor's degree in computer science, information technology, or related field, or equivalent work experience.
  • 7+ years of experience in site reliability engineering or related fields.
  • Deep understanding of system reliability concepts, including monitoring, automation and incident response.
  • Proficiency with multiple scripting languages and automation tools.
  • Strong problem-solving and troubleshooting skills.
  • Experience with cloud platforms and containerization technologies.
  • Excellent communication and teamwork skills.
  • Proven leadership and mentorship abilities.
  • Strong strategic thinking and decision-making skills.
  • Adaptability and willingness to learn new technologies.

Responsibilities

  • Advanced System Monitoring: design and maintain sophisticated monitoring solutions for infrastructure and applications.
  • Incident Response: lead complex incident response activities and conduct post-incident reviews.
  • Automation and Scripting: develop automation scripts to enhance reliability and efficiency.
  • Performance Analysis: collect and interpret performance data to identify trends and issues.
  • Documentation: maintain up-to-date system configurations and procedures.
  • Collaboration: work with cross-functional teams on reliability projects.
  • Mentorship: guide junior and mid-level engineers.
  • Security Compliance: implement security measures and uphold policies.
  • Continuous Improvement: drive adoption of new technologies and methods.
  • Strategic Leadership: influence IT strategy and operations direction.
  • Capacity Planning: plan for future growth and demand.
  • Disaster Recovery: develop and maintain DR plans.
  • Change Management: participate in change processes.
  • Root Cause Analysis: lead RCA for major incidents.
  • Autonomous Work Culture: promote initiative and self-motivation.
  • Spin Culture Ambassador: embody Spin's values and foster inclusive environment.

Skills

System Monitoring
Incident Response
Automation & Scripting
Performance Analysis
Documentation
Collaboration
Mentorship
Security Compliance
Continuous Improvement
Strategic Leadership
Capacity Planning
Disaster Recovery
Change Management
Root Cause Analysis
Autonomous Work Culture
Spin Culture Ambassador

Education

Bachelor's degree in Computer Science or Information Technology

Tools

Cloud platforms
Containerization

Job description

Objective of the Role

The Senior SRE Engineer is a highly experienced role responsible for leading the enhancement and maintenance of the reliability, availability, and performance of the company's IT infrastructure and applications. This role focuses on advanced system stability, efficiency, and scalability through sophisticated monitoring, automation, and incident response. The Senior SRE Engineer plays a critical role in driving strategic initiatives and mentoring junior engineers, contributing significantly to the success ofIT operations.

Main Responsibilities
  • Advanced System Monitoring: Design, implement, and maintain sophisticated monitoring solutions to ensure optimal health and performance of infrastructure and applications.
  • Incident Response: Lead complex incident response activities, diagnose and resolve critical system reliability issues, and conduct thorough post-incident reviews to prevent recurrence.
  • Automation and Scripting: Develop and implement advanced automation scripts and tools to enhance system reliability, operational efficiency, and scalability.
  • Performance Analysis: Collect, analyze, and interpret complex performance data to identify trends, anomalies, and potential issues, providing strategic insights and recommendations.
  • Documentation: Ensure comprehensive and up-to-date documentation of system configurations, processes, and procedures, and contribute to knowledge sharing within the team.
  • Collaboration: Collaborate closely with cross-functional teams and departments to support and lead reliability engineering projects and initiatives.
  • Mentorship: Provide advanced guidance and support to junior and mid-level engineers, fostering their technical growth and development.
  • Security Compliance: Implement and enforce robust security measures to protect systems and ensure compliance with security policies and industry standards.
  • Continuous Improvement: Drive continuous improvement initiatives, exploring and integrating new technologies and methodologies to enhance system reliability and performance.
  • Strategic Leadership: Actively contribute to strategic planning and decision-making processes, leveraging expertise to influence the direction of IT infrastructure and operations.
  • Capacity Planning: Conduct capacity planning to ensure systems can handle future growth and demand.
  • Disaster Recovery: Develop and maintain disaster recovery plans to ensure business continuity in case of system failures.
  • Change Management: Participate in change management processes to ensure smooth implementation of system updates and changes.
  • Root Cause Analysis: Lead root cause analysis for major incidents, identifying underlying issues and implementing long-term solutions to prevent recurrence.
  • Autonomous Work Culture: Promote and embody an autonomous work culture by taking initiative, being self-motivated, and collaborating effectively in an agile and lean environment.
  • Spin Culture Ambassador: Embody and promote Spin's values in every action, fostering a positive
    and inclusive work environment
Required Knowledge and Experience
  • Education: Bachelor's degree in computer science, Information Technology, or a related field, or equivalent work experience.
  • Experience: Minimum of 7+ years of experience in site reliability engineering or related fields.
  • Deep understanding of system reliability concepts, including advanced monitoring, automation, and incident response.
  • Proficiency with multiple scripting languages and automation tools.
  • Strong problem-solving and troubleshooting skills.
  • Experience with cloud platforms and containerization technologies.
  • Excellent communication and teamwork skills.
  • Proven leadership and mentorship abilities.
  • Strong strategic thinking and decision-making skills.
  • Adaptability: Willingness to learn and adapt to new technologies and processes
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Infojini Inc • Mexico

Hybrid
MXN 900,000 - 1,300,000
Site Reliability Engineer
Site Reliability Engineer

Alsons/AWS Information Systems Inc. • Cebu City

Hybrid
PHP 600,000 - 1,000,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Senior SRE Engineer: Reliability, Automation & Leadership
Senior SRE Engineer: Reliability, Automation & Leadership

Spin Careers • Philippines

Remote
PHP 1,200,000 - 2,100,000
Incident Response Lead
Incident Response Lead

Permhunt • Cebu City

On-site
PHP 900,000 - 1,700,000
SRE (Site Reliability Engineer)
SRE (Site Reliability Engineer)

GCash • Manila

On-site
Opportunity for career growth and development
Dynamic collaborative team environment
Highly competitive compensation and benefits package
Senior Engineer - Site Reliability
Senior Engineer - Site Reliability

Dencom Consultancy and Manpower Services • Parañaque

On-site
SRE (Site Reliability Engineer)
SRE (Site Reliability Engineer)

GCash • Metro Manila

On-site
PHP 1,800,000 - 2,400,000
Career growth
Collaborative team
Competitive compensation
SERVICE RELIABILITY ENGINEER
SERVICE RELIABILITY ENGINEER

Metrobank • Taguig

On-site
PHP 600,000 - 900,000
Senior SRE: Scale, Automate & Self-Healing Systems
Senior SRE: Scale, Automate & Self-Healing Systems

Replit • España

On-site
PHP 2,461,000 - 4,308,000
Competitive Salary & Equity
Health, Dental, Vision and Life Insurance
Flexible Time Off (FTO) + Holidays
+2