Site Reliability Engineer- GCP

Aziro

Hyderabad

Hybrid

INR 1,400,000 - 2,100,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Aziro is seeking a Site Reliability Engineer (SRE) with strong GCP expertise to optimize and automate our cloud-based infrastructure in Hyderabad. You will own deployment pipelines, observability, and performance tuning across services such as Cloud Run, Dataflow, and VM instances.

You will design scalable infrastructure, implement IaC with Terraform, manage CI/CD with Bitbucket, and ensure reliability through monitoring tools like Sentry and Grafana.

Qualifications

  • Proven experience as a Site Reliability Engineer or similar role.
  • Hands-on expertise with Google Cloud Platform (GCP), including Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Strong scripting and automation skills using tools like Terraform, Python, bash, or equivalent.
  • Proficiency in setting up and managing alerting, monitoring and logging tools (e.g., Sentry, Stackdriver, or similar) as well as dashboarding (e.g. Grafana).
  • Experience managing CI/CD pipelines with Bitbucket or similar tools.
  • Solid understanding of MySQL database performance tuning and query optimization.
  • Familiarity with containerization and orchestration tools, such as Docker and Kubernetes.

Responsibilities

  • Design, build, and maintain highly available and scalable infrastructure on Google Cloud Platform (GCP).
  • Work with GCP services like Cloud Run, Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Automate and optimize deployment processes using Infrastructure as Code (IaC) preferably Terraform and CI/CD pipelines.
  • Enhance system observability by setting up and maintaining logging and monitoring solutions, including Sentry.
  • Proactively monitor the health and performance of our systems, troubleshoot incidents, and implement resolutions.
  • Continuously identify opportunities to improve system efficiency and reduce latency.
  • Oversee the deployment pipeline using Bitbucket to ensure smooth, reliable, and repeatable deployment processes.
  • Collaborate with development teams to ensure the integration of scalable and maintainable deployment practices.
  • Monitor and maintain MySQL databases, focusing on performance and reliability.
  • Analyze and optimize database indexes and address poorly performing queries.
  • Work closely with software developers, QA engineers, and product managers to deliver high-quality, reliable solutions.
  • Advocate for best practices in SRE, including monitoring, alerting, and incident management.

Skills

GCP expertise
Terraform automation
CI/CD pipelines
MySQL performance tuning
Kubernetes
Python scripting

Education

Bachelor's degree in Computer Science

Tools

Terraform
Bitbucket
Grafana
Sentry
Stackdriver

Job description

We are a forward-thinking organization looking to enhance our infrastructure and operational efficiency. As we grow, we are seeking a talented Site Reliability Engineer (SRE) with expertise in Google Cloud Platform (GCP) to optimize, automate, and manage the deployment of our cloud-based infrastructure.

Role Overview

As an SRE on our team, you will play a key role in ensuring the reliability, performance, and scalability of our systems. You will monitor, optimize, and automate the deployment of our infrastructure using GCP services, including Cloud Run, Pub/Sub, Virtual Machines, Kubernetes, and Dataflow. You will also take ownership of our logging and monitoring tools, streamline our deployment pipelines, and work to enhance database performance.

Key Responsibilities
Infrastructure Management
  • Design, build, and maintain highly available and scalable infrastructure on Google Cloud Platform (GCP).
  • Work with GCP services like Cloud Run, Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Automate and optimize deployment processes using Infrastructure as Code (IaC) preferably Terraform and CI/CD pipelines.
Monitoring and Optimization
  • Enhance system observability by setting up and maintaining logging and monitoring solutions, including Sentry.
  • Proactively monitor the health and performance of our systems, troubleshoot incidents, and implement resolutions.
  • Continuously identify opportunities to improve system efficiency and reduce latency.
Deployment Pipeline Management
  • Oversee the deployment pipeline using Bitbucket to ensure smooth, reliable, and repeatable deployment processes.
  • Collaborate with development teams to ensure the integration of scalable and maintainable deployment practices.
Database Optimization
  • Monitor and maintain MySQL databases, focusing on performance and reliability.
  • Analyze and optimize database indexes and address poorly performing queries.
Collaboration and Best Practices
  • Work closely with software developers, QA engineers, and product managers to deliver high-quality, reliable solutions.
  • Advocate for best practices in SRE, including monitoring, alerting, and incident management.
Qualifications
  • Proven experience as a Site Reliability Engineer or similar role.
  • Hands-on expertise with Google Cloud Platform (GCP), including Pub/Sub, Virtual Machines, Kubernetes Engine, App Engine, and Dataflow.
  • Strong scripting and automation skills using tools like Terraform, Python, bash, or equivalent.
  • Proficiency in setting up and managing alerting, monitoring and logging tools (e.g., Sentry, Stackdriver, or similar) as well as dashboarding (e.g. Grafana).
  • Experience managing CI/CD pipelines with Bitbucket or similar tools.
  • Solid understanding of MySQL database performance tuning and query optimization.
  • Familiarity with containerization and orchestration tools, such as Docker and Kubernetes.
Preferred Skills
  • Experience with Infrastructure as Code (IaC) using Terraform or similar tools.
  • Knowledge of incident response and on-call best practices.
  • Familiarity with modern software delivery practices like DevOps and Agile.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) - Google Cloud Platform
Site Reliability Engineer (SRE) - Google Cloud Platform

Aziro • Hyderabad

Hybrid
INR 1,500,000 - 3,200,000
Site Reliability Engineer
Site Reliability Engineer

Advance Career Solutions • Pune District

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer (SRE) – GCP Platform
Site Reliability Engineer (SRE) – GCP Platform

ITC Infotech • Bengaluru

On-site
INR 900,000 - 1,300,000
Intermediate Applications Developer
Intermediate Applications Developer

UPS • Chennai District

On-site
INR 1,500,000 - 2,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Infobell It Solutions • Hyderabad

On-site
INR 2,800,000 - 4,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UST • Pune District

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer with GCP
Site Reliability Engineer with GCP

Appsierra Group • Pune District

On-site
INR 1,080,000 - 1,320,000
Senior Data Site Reliability Engineer | GCP is mandatory
Senior Data Site Reliability Engineer | GCP is mandatory

Anlage Infotech • Chennai District

On-site
INR 3,500,000 - 6,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000