Site Reliability Engineer (SRE) - Google Cloud Platform

Aziro

Hyderabad

Hybrid

INR 1,500,000 - 3,200,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Aziro in Hyderabad is seeking a Site Reliability Engineer (SRE) with solid GCP experience to optimize, automate, and manage our cloud infrastructure. You will focus on reliability and performance of services on Google Cloud, including Cloud Run, Pub/Sub, VMs, Kubernetes, and Dataflow.

You will own monitoring, logging, and deployment pipelines, and work with developers to deliver resilient solutions. Practical Terraform/Python/Bash and CI/CD proficiency required.

Qualifications

  • Minimum proven experience as an SRE or similar role.
  • GCP expertise including Pub/Sub, VMs, Kubernetes Engine, App Engine and Dataflow.
  • Strong automation skills with Terraform, Python, Bash.
  • Experience with monitoring/logging tools (Sentry, Grafana) and dashboards.
  • CI/CD pipelines knowledge with Bitbucket or similar tools.
  • MySQL performance tuning and query optimization.
  • Familiarity with Docker and Kubernetes orchestration.

Responsibilities

  • Design, build, and maintain scalable infrastructure on Google Cloud Platform.
  • Automate deployments with IaC (Terraform) and CI/CD pipelines.
  • Enhance observability by logging and monitoring solutions (Sentry, Grafana).
  • Oversee deployment pipelines using Bitbucket for reliable releases.
  • Monitor and optimize MySQL databases, indexes, and queries.
  • Collaborate with developers, QA, and product managers to ensure reliability.

Skills

SRE Experience
GCP Expertise
Automation Scripting
CI/CD Pipelines
MySQL Tuning
Docker & Kubernetes
Monitoring & Logging
IaC (Terraform)
Bitbucket

Tools

Terraform
Python
Bash
Cloud Run
Pub/Sub
Kubernetes
Dataflow
Grafana
Sentry
Bitbucket
Docker
App Engine

Job description

Job Title:

Site Reliability Engineer (SRE) - Google Cloud Platform

About Us:

We are a forward-thinking organization looking to enhance our infrastructure and operational efficiency. As we grow, we are seeking a talented Site Reliability Engineer (SRE) with expertise in Google Cloud Platform (GCP) to optimize, automate, and manage the deployment of our cloud-based infrastructure.

Role Overview:

As an SRE on our team, you will play a key role in ensuring the reliability, performance, and scalability of our systems. You will monitor, optimize, and automate the deployment of our infrastructure using GCP services, including Cloud Run, Pub/Sub, Virtual Machines, Kubernetes, and Dataflow. You will also take ownership of our logging and monitoring tools, streamline our deployment pipelines, and work to enhance database performance.

Key Responsibilities:
  • Infrastructure Management:
  • Design, build, and maintain highly available and scalable infrastructure on Google Cloud Platform (GCP).
  • Work with GCP services like Cloud Run, Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Automate and optimize deployment processes using Infrastructure as Code (IaC) preferably Terraform and CI/CD pipelines.
  • Monitoring and Optimization:
  • Enhance system observability by setting up and maintaining logging and monitoring solutions, including Sentry.
  • Proactively monitor the health and performance of our systems, troubleshoot incidents, and implement resolutions.
  • Continuously identify opportunities to improve system efficiency and reduce latency.
  • Deployment Pipeline Management:
  • Oversee the deployment pipeline using Bitbucket to ensure smooth, reliable, and repeatable deployment processes.
  • Collaborate with development teams to ensure the integration of scalable and maintainable deployment practices.
  • Database Optimization:
  • Monitor and maintain MySQL databases, focusing on performance and reliability.
  • Analyze and optimize database indexes and address poorly performing queries.
  • Collaboration and Best Practices:
  • Work closely with software developers, QA engineers, and product managers to deliver high-quality, reliable solutions.
  • Advocate for best practices in SRE, including monitoring, alerting, and incident management.
Qualifications:
  • Proven experience as a Site Reliability Engineer or similar role.
  • Hands-on expertise with Google Cloud Platform (GCP), including Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Strong scripting and automation skills using tools like Terraform, Python, bash, or equivalent.
  • Proficiency in setting up and managing alerting, monitoring and logging tools (e.g., Sentry, Stackdriver, or similar) as well as dashboarding (e.g. Grafana).
  • Experience managing CI/CD pipelines with Bitbucket or similar tools.
  • Solid understanding of MySQL database performance tuning and query optimization.
  • Familiarity with containerization and orchestration tools, such as Docker and Kubernetes.
Preferred Skills:
  • Experience with Infrastructure as Code (IaC) using Terraform or similar tools.
  • Knowledge of incident response and on‑call best practices.
  • Familiarity with modern software delivery practices like DevOps and Agile.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) – GCP Platform
Site Reliability Engineer (SRE) – GCP Platform

ITC Infotech • Bengaluru

On-site
INR 900,000 - 1,300,000
Intermediate Applications Developer
Intermediate Applications Developer

UPS • Chennai District

On-site
INR 1,500,000 - 2,000,000
Site Reliability Manager
Site Reliability Manager

Google • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Site Reliability Manager
Site Reliability Manager

Google Inc. • Bengaluru

On-site
INR 4,000,000 - 8,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UST • Pune District

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineers - Google Cloud Platform GCP RedHat OpenShift Administration
Site Reliability Engineers - Google Cloud Platform GCP RedHat OpenShift Administration

UPS • Thiruvallur District

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps INC. • Pune District

On-site
INR 1,400,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Mumbai

On-site
INR 800,000 - 1,200,000