Senior SRE: Kubernetes, Java & Observability Expert

V2 Solutions

Hinoba-an

On-site

PHP 900,000 - 1,500,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

V2 Solutions is seeking an experienced Site Reliability Engineer to own the reliability of production systems. You will work with Kubernetes, Java microservices, and SQL databases to ensure availability, performance, and scalability.

You will implement monitoring with Datadog and Prometheus, define SLIs/SLOs, and participate in on-call rotations. Collaboration across Dev, Infra, Security, and Business teams is essential.

Qualifications

  • 5+ years in IT with SRE/DevOps or production support.
  • Hands-on Kubernetes and containerized apps experience.
  • Datadog and/or Prometheus monitoring experience.
  • Java/J2EE microservices in production; strong SQL knowledge.
  • REST APIs, Linux, CI/CD, and scripting skills.

Responsibilities

  • Own reliability and health of production applications and services.
  • Provide L2/L3 production support, incident triage, and RCA communication.
  • Monitor apps on Kubernetes, including pods, deployments, services, and scaling.
  • Implement and maintain monitoring dashboards, metrics, logs, and alerts.
  • Define SLIs/SLOs/SLAs and runbooks; drive MTTR improvements.
  • Troubleshoot across Java apps, APIs, and databases; data issue analysis.
  • Collaborate with Development, Infra, DevOps, Security, and Business teams.

Skills

Kubernetes
Datadog/Prometheus monitoring
Java/J2EE microservices
SQL relational databases
Incident management
CI/CD pipelines
Shell scripting
Linux/Unix
Observability

Tools

Datadog
Prometheus
Kubernetes
ELK/OpenSearch/Splunk
Docker
GitLab CI/CD

Job description

V2 Solutions is seeking an experienced Site Reliability Engineer to own the reliability of production systems. You will work with Kubernetes, Java microservices, and SQL databases to ensure availability, performance, and scalability.

You will implement monitoring with Datadog and Prometheus, define SLIs/SLOs, and participate in on-call rotations. Collaboration across Dev, Infra, Security, and Business teams is essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

V2 Solutions • Hinoba-an

On-site
PHP 900,000 - 1,500,000
Senior DevOps & SRE | Hybrid Kubernetes Architect
Senior DevOps & SRE | Hybrid Kubernetes Architect

V2 Solutions • Hinoba-an

Hybrid
PHP 1,200,000 - 2,400,000
Senior SRE - Kubernetes & Cloud Reliability (Mexico)
Senior SRE - Kubernetes & Cloud Reliability (Mexico)

Sezzle • Mexico

Hybrid
Senior SRE – Cloud SaaS & Observability
Senior SRE – Cloud SaaS & Observability

Cloud Blue • Philippines

On-site
PHP 5,058,000 - 7,948,000
Remote work
Competitive salary
Career development
+1
Senior SRE Operations Engineer - Cloud & Observability
Senior SRE Operations Engineer - Cloud & Observability

Teciem • Hinoba-an

Hybrid
PHP 2,772,000 - 4,291,000
Senior SRE: Scale, Automate & Self-Healing Systems
Senior SRE: Scale, Automate & Self-Healing Systems

Replit • España

On-site
PHP 2,461,000 - 4,308,000
Competitive Salary & Equity
Health, Dental, Vision and Life Insurance
Flexible Time Off (FTO) + Holidays
+2
Senior SRE - Kubernetes Reliability Lead (Remote)
Senior SRE - Kubernetes Reliability Lead (Remote)

BairesDev • Mexico

On-site
PHP 7,299,000 - 10,949,000
100% remote work
Competitive USD compensation
Hardware and software setup
+3
Senior SRE: Production Reliability on AWS & EKS (Remote)
Senior SRE: Production Reliability on AWS & EKS (Remote)

Salve.Inno Consulting • Philippines

On-site
PHP 1,000,000 - 1,800,000
Fully remote
B2B cooperation
Ownership & impact
Cloud-Native Platform & Reliability Engineer
Cloud-Native Platform & Reliability Engineer

V2 Solutions • Hinoba-an

On-site
PHP 1,004,000 - 1,451,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000