DevOps & Site Reliability Engineer (GCP)

Urban Ridge Supplies

Islamabad

On-site

PKR 24,965,000 - 38,835,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Urban Ridge Supplies is seeking a mid‑level DevOps Engineer to own production reliability. You will design and maintain cloud infrastructure on Google Cloud, automate deployments with Docker and Kubernetes, and drive observability with Prometheus, Elasticsearch, and New Relic.

You will lead incident response, define SLOs/SLIs, and work with cross‑functional teams to optimize performance and cost as we scale globally.

Qualifications

  • 3–5+ years in a DevOps or similar role.
  • Strong Docker and Kubernetes experience.
  • Hands-on with GCP services (Compute Engine, Cloud Run, GKE).
  • Experience with Elasticsearch for monitoring/logging/search.
  • Proficient with MongoDB administration and optimization.
  • Proven track record operating production systems and incident response.

Responsibilities

  • Design, deploy, and manage cloud infrastructure on GCP.
  • Implement and maintain container orchestration with Docker and Kubernetes.
  • Develop and maintain CI/CD pipelines (GitHub and Google Cloud Run).
  • Set up monitoring, logging, dashboards, and alerts (Prometheus, Kibana, New Relic).
  • Lead incident response and blameless postmortems; drive to resolution.
  • Manage MongoDB backups, recovery, and DR strategies.

Skills

Docker
Kubernetes
GCP
Elasticsearch
MongoDB
Python/Bash scripting
Incident response / SRE practices

Education

Bachelor's degree in CS/Engineering

Tools

Terraform
Ansible
GitHub Actions
CircleCI
Compute Engine / Cloud Run / GKE

Job description

As a DevOps Engineer, you will design, implement, and maintain the infrastructure that supports our applications and services — and, just as importantly, you will keep that infrastructure reliable in production. You will work closely with development, QA, and IT teams to automate and streamline operations, build the observability that lets us catch problems before customers do, and lead the response when incidents happen. This is a hands‑on role where you own the health of production, not only its build‑out. We are hiring at a mid level (3–5 years) for someone with strong production instincts and the judgment to grow into a senior reliability owner as we scale globally.

Key Responsibilities
Infrastructure Management
  • Design, deploy, and manage cloud infrastructure on Google Cloud Platform (GCP).

  • Implement and maintain scalable container orchestration using Docker.

CI/CD Pipeline
  • Develop and maintain continuous integration/continuous deployment (CI/CD) pipelines to automate testing, building, and deployment processes through GitHub and Google Cloud Run.

  • Collaborate with development teams to integrate new features into the CI/CD process.

Monitoring & Observability
  • Set up and manage monitoring, logging, and alerting systems using tools like Elasticsearch, Kibana, Prometheus, and New Relic.

  • Build dashboards, metrics, and actionable alerts so engineering detects degradations before customers do — and so no critical signal (e.g. a database running hot for hours) ever goes unnoticed.

Reliability & Incident Response (SRE)
  • Define, measure, and own service‑level objectives (SLOs/SLIs) and error budgets for critical services.

  • Lead incident response end to end: triage by severity, form a hypothesis and confirm it with metrics before taking action, mitigate, and drive to resolution. Act as incident commander on major incidents and coordinate communication across stakeholders.

  • Run blameless postmortems, identify true root causes, and track corrective actions to closure.

  • Participate in an on‑call rotation; continuously reduce toil and mean‑time‑to‑recovery (MTTR) through automation.

  • Plan for capacity, performance, and cost as we scale toward multi‑region, global traffic.

Database Management
  • Administer and optimize MongoDB instances, ensuring data integrity, performance, and security.

  • Implement backup, recovery, and disaster recovery strategies for MongoDB and other databases.

Security & Compliance
  • Implement security best practices across infrastructure, applications, and data.

  • Ensure compliance with industry standards and internal policies.

Automation & Scripting
  • Automate infrastructure provisioning, configuration management, and system operations using GCP services.

  • Develop custom scripts as needed to enhance automation and operational efficiency.

Collaboration & Support
  • Work closely with development and QA teams to support the software development lifecycle.

  • Provide technical guidance and support to resolve infrastructure‑related issues.

Qualifications
Education
  • Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent experience).

Experience
  • 3–5+ years of experience as a DevOps Engineer or in a similar role.

  • Strong experience with Docker, including container orchestration using Kubernetes.

  • Hands‑on experience with Google Cloud Platform (GCP) services, including Compute Engine, Cloud Storage, Cloud Run, and GKE.

  • Experience with Elasticsearch for monitoring, logging, and search.

  • Proficiency in administering and optimizing MongoDB databases.

  • Demonstrated experience operating production systems and responding to incidents — not only building infrastructure. You can point to real production incidents you owned, quantify their impact concretely (users affected, duration, consequence), and describe what you changed afterward.

Skills
  • Strong scripting skills in Python, Bash, or similar languages.

  • Proficiency in infrastructure as code (IaC) tools such as Terraform or Ansible.

  • Experience with CI/CD tools such as GitHub Actions or CircleCI.

  • Sound production‑debugging methodology: observe and form a hypothesis before acting, reason about how components fail together (e.g. how a queue backlog interacts with the database), and update your approach when new information appears.

  • Familiarity with defining alerts, dashboards, and SLOs; comfort reasoning about availability, latency, saturation, and error budgets.

  • Excellent problem‑solving and troubleshooting skills.

  • Strong communication skills and ability to work collaboratively across teams.

  • Strong ownership and a blameless, collaborative posture under pressure — focused on diagnosing and fixing problems rather than assigning blame.

Nice to Have
  • Experience with additional cloud platforms (e.g. AWS or Azure).

  • Familiarity with Agile/Scrum methodologies.

  • Certifications in Docker, GCP, or related technologies.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (Remote)
Site Reliability Engineer (Remote)

Urban Ridge Supplies • Pakistan

On-site
PKR 892,800 - 1,339,200
Senior Site Reliability Engineer - (GCP) - Afternoon
Senior Site Reliability Engineer - (GCP) - Afternoon

10Pearls • Islamabad

On-site
PKR 1,800,000 - 2,600,000
Senior Site Reliability Engineer - (GCP) - Afternoon
Senior Site Reliability Engineer - (GCP) - Afternoon

10Pearls • Islamabad

On-site
PKR 3,500,000 - 6,500,000
Senior Site Reliability Engineer - (GCP) - Afternoon
Senior Site Reliability Engineer - (GCP) - Afternoon

10Pearls, LLC • Pakistan

On-site
PKR 19,434,000 - 33,315,000
GCP DevOps & SRE: Production Reliability Lead
GCP DevOps & SRE: Production Reliability Lead

Urban Ridge Supplies • Islamabad

On-site
PKR 24,965,000 - 38,835,000
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)

10Pearls, LLC • Lahore, Karachi Division

On-site
PKR 2,500,000 - 5,000,000
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)

10Pearls • Karachi Division

On-site
PKR 2,400,000 - 4,200,000
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)
Staff/Senior DevOps Consultant - SRE (BI & Data Ecosystems)

10Pearls, LLC • Pakistan

On-site
PKR 2,400,000 - 4,800,000
DevOps Engineer
DevOps Engineer

DevVibe • Multan

On-site
PKR 450,000 - 650,000
Senior GCP SRE: Cloud Reliability & DevOps Leader
Senior GCP SRE: Cloud Reliability & DevOps Leader

10Pearls • Islamabad

On-site
PKR 3,500,000 - 6,500,000