Site Reliability Engineer

Philtech Inc.

Taguig

On-site

PHP 1,200,000 - 1,800,000

Full time

39 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Philtech Inc. is seeking an experienced Site Reliability Engineer to monitor and maintain our production environment, ensuring availability, performance, and reliability.

The role emphasizes automation, collaboration with development teams, and incident management to minimize downtime and improve efficiency.

Candidates should have 4+ years of SRE or related experience, a CS/Engineering degree, and hands-on skills with Linux, containers, and cloud platforms.

Qualifications

  • 4+ years in site reliability, operations, or software engineering.
  • Strong Linux/Unix, networking, and infrastructure knowledge.
  • Proficiency with Python or Ruby scripting.
  • Experience with Docker and Kubernetes in production.
  • Excellent troubleshooting and cross-functional communication.

Responsibilities

  • Monitor production systems for availability and performance, responding to incidents.
  • Develop automation tools to reduce manual work and improve efficiency.
  • Collaborate with development teams to design scalable, reliable systems.
  • Analyze metrics to identify bottlenecks and optimize performance.
  • Lead incident response, perform root-cause analysis, and implement safeguards.
  • Create and maintain documentation for architecture, processes, and procedures.
  • Plan capacity to ensure future growth.

Skills

Scripting (Python Ruby)
Linux/Unix troubleshooting
Communication
Collaboration
Problem-solving

Education

Bachelor's degree in Computer Science or Engineering

Tools

Docker
Kubernetes
AWS
GCP
Azure
Ansible
Chef
Puppet
Jenkins
GitLab CI
Prometheus
Grafana
ELK Stack

Job description

  • Monitor and Maintain Systems: Ensure the availability, performance, and reliability of our production environment by monitoring system health and responding to incidents.
  • Automation: Develop and implement automation tools to reduce manual intervention and improve system efficiency.
  • Collaboration: Work closely with development teams to design and implement scalable and reliable systems.
  • Performance Tuning: Analyze system metrics to identify performance bottlenecks and optimize system performance.
  • Incident Management: Lead incident response efforts, conduct root cause analysis, and implement preventive measures.
  • Documentation: Create and maintain comprehensive documentation for system architecture, processes, and procedures.
  • Capacity Planning: Conduct capacity planning and ensure systems can handle future growth.
Qualifications:
  • Experience: 4+ years of experience in site reliability engineering, operations, or software engineering.
  • Education: Bachelor's degree in Computer Science, Engineering, or a related field.
  • Technical Skills: Proficiency in scripting languages (e.g., Python, Ruby), experience with containerization (Docker, Kubernetes), and familiarity with cloud platforms (AWS, GCP, Azure).
  • System Knowledge: Strong understanding of Linux/Unix systems, networking, and infrastructure components.
  • Problem-Solving: Excellent troubleshooting and problem-solving skills.
  • Communication: Strong communication and collaboration skills to work effectively with cross-functional teams.
  • Certifications: Relevant certifications (e.g., AWS Certified Solutions Architect, Certified Kubernetes Administrator) are a plus.
Preferred Skills:
  • Experience with configuration management tools (e.g., Ansible, Chef, Puppet).
  • Knowledge of CI/CD pipelines and tools (e.g., Jenkins, GitLab CI).
  • Familiarity with monitoring and logging tools (e.g., Prometheus, Grafana, ELK stack).
Why Join Us:
  • Innovative Environment: Work on cutting-edge technologies and projects.
  • Growth Opportunities: Opportunities for professional development and career advancement.
  • Collaborative Culture: Join a team that values collaboration, diversity, and inclusion.
  • Competitive Benefits: Comprehensive benefits package including health insurance, retirement plans, and more.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

IDEMIA • Philippines

On-site
PHP 900,000 - 1,500,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Site Reliability Engineer
Site Reliability Engineer

EROAD Limited • Manila

On-site
PHP 900,000 - 1,500,000
Site Reliability Engineer
Site Reliability Engineer

Hammerjack Pty Ltd • Metro Manila

On-site
PHP 900,000 - 1,300,000
Site Reliability Engineer, VP
Site Reliability Engineer, VP

JobCubby • Hinoba-an

On-site
PHP 1,200,000 - 1,600,000
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Accenture • Cebu City

On-site
PHP 700,000 - 1,100,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Omilia • Philippines

On-site
PHP 1,000,000 - 1,800,000
Fixed compensation
Vacation leaves
Professional development opportunities
+3
Site Reliability Engineer
Site Reliability Engineer

Alsons/AWS Information Systems Inc. • Cebu City

Hybrid
PHP 600,000 - 1,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Acquire Intelligence • Taguig

On-site
PHP 900,000 - 1,500,000
Site Reliability Engineers
Site Reliability Engineers

Trinity Workforce Solutions, Inc. • Makati

On-site