Site Reliability Engineer

Philtech Inc.

Taguig

On-site

PHP 670,000 - 1,339,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
Retirement plans
Career growth opportunities

Job summary

Philtech Inc. is seeking a Site Reliability Engineer to monitor and maintain our production systems, implement automation, and collaborate with development teams on scalable, reliable architectures. The role emphasizes incident leadership, performance tuning, and comprehensive documentation.

The ideal candidate brings 4+ years in SRE/ops with strong Linux, networking, scripting (Python/Ruby), and containerization (Docker/Kubernetes), plus AWS/GCP/Azure experience.

Qualifications

  • 4+ years of experience in site reliability engineering, operations, or software engineering.
  • Bachelor's degree in Computer Science, Engineering, or related field.
  • Proficiency in scripting (Python/Ruby), containerization (Docker/Kubernetes), and cloud platforms (AWS/GCP/Azure).
  • Strong Linux/Unix, networking, and infrastructure knowledge; excellent troubleshooting skills.

Responsibilities

  • Monitor and maintain production systems to ensure availability and performance.
  • Develop automation tools to reduce manual work and improve efficiency.
  • Collaborate with development teams to build scalable, reliable systems.
  • Tune performance by analyzing metrics and addressing bottlenecks.
  • Lead incident response, perform root cause analysis, and implement preventive measures.
  • Create and maintain comprehensive system documentation.

Skills

Python
Ruby
Linux/Unix
Networking
Troubleshooting
Communication
AWS

Education

Bachelor's degree in Computer Science or Engineering

Tools

Docker
Kubernetes
AWS
GCP
Azure
Jenkins
GitLab CI
Prometheus
Grafana
ELK Stack
Ansible

Job description

  • Monitor and Maintain Systems: Ensure the availability, performance, and reliability of our production environment by monitoring system health and responding to incidents.
  • Automation: Develop and implement automation tools to reduce manual intervention and improve system efficiency.
  • Collaboration: Work closely with development teams to design and implement scalable and reliable systems.
  • Performance Tuning: Analyze system metrics to identify performance bottlenecks and optimize system performance.
  • Incident Management: Lead incident response efforts, conduct root cause analysis, and implement preventive measures.
  • Documentation: Create and maintain comprehensive documentation for system architecture, processes, and procedures.
  • Capacity Planning: Conduct capacity planning and ensure systems can handle future growth.
Qualifications
  • Experience: 4+ years of experience in site reliability engineering, operations, or software engineering.
  • Education: Bachelor's degree in Computer Science, Engineering, or a related field.
  • Technical Skills: Proficiency in scripting languages (e.g., Python, Ruby), experience with containerization (Docker, Kubernetes), and familiarity with cloud platforms (AWS, GCP, Azure).
  • System Knowledge: Strong understanding of Linux/Unix systems, networking, and infrastructure components.
  • Problem-Solving: Excellent troubleshooting and problem-solving skills.
  • Communication: Strong communication and collaboration skills to work effectively with cross-functional teams.
  • Certifications: Relevant certifications (e.g., AWS Certified Solutions Architect, Certified Kubernetes Administrator) are a plus.
Preferred Skills
  • Experience with configuration management tools (e.g., Ansible, Chef, Puppet).
  • Knowledge of CI/CD pipelines and tools (e.g., Jenkins, GitLab CI).
  • Familiarity with monitoring and logging tools (e.g., Prometheus, Grafana, ELK stack).
Why Join Us
  • Innovative Environment: Work on cutting-edge technologies and projects.
  • Growth Opportunities: Opportunities for professional development and career advancement.
  • Collaborative Culture: Join a team that values collaboration, diversity, and inclusion.
  • Competitive Benefits: Comprehensive benefits package including health insurance, retirement plans, and more.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

NCS Philippines • Taguig

Hybrid
PHP 900,000 - 1,300,000
Site Reliability Engineer
Site Reliability Engineer

Private Advertiser • Makati

On-site
PHP 1,200,000 - 1,800,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

LTM • Mexico

On-site
PHP 900,000 - 1,300,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

EPAM Systems • Mexico

On-site
PHP 5,846,000 - 8,616,000
Site Reliability Engineering Onsite
Site Reliability Engineering Onsite

Gratitude Philippines • Manila

On-site
PHP 900,000 - 1,800,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Senior Platform Engineer (DevOps / Site Reliability Engineer) RTO 1x In A Month
Senior Platform Engineer (DevOps / Site Reliability Engineer) RTO 1x In A Month

Avensys Consulting • Philippines

On-site
PHP 1,339,000 - 2,009,000
Site Reliability Engineer
Site Reliability Engineer

Alsons/AWS Information Systems Inc. • Cebu City

Hybrid
PHP 600,000 - 1,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Omilia • Philippines

On-site
PHP 1,000,000 - 1,800,000
Fixed compensation
Vacation leaves
Professional development opportunities
+3
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Acquire Intelligence • Taguig

On-site
PHP 900,000 - 1,500,000