Site Reliability Engineer - SaaSOps

ValGenesis

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

A leading digital validation platform provider is seeking a Site Reliability Engineer based in Hyderabad. The role involves ensuring reliability for a cloud-native software platform by embedding SRE best practices, automating processes, and leading incident response efforts. Candidates should have over 3 years of experience in SRE, strong scripting skills, and deep Azure experience. The company values innovation and collaboration, working onsite five days a week.

Qualifications

  • Minimum 3 years of experience in Site Reliability Engineering, cloud-native applications.
  • Prior experience as a DevOps engineer or cloud system administrator.
  • Deep hands-on experience with Microsoft Azure in production environments.

Responsibilities

  • Define and embed SRE best practices across the SaaS platform.
  • Establish and maintain SLAs, SLIs, SLOs, and error budgets.
  • Lead incident response efforts with structured troubleshooting.

Skills

Site Reliability Engineering (SRE)
Scripting (Python, PowerShell)
Microsoft Azure
Terraform
Ansible
Kubernetes
PostgreSQL performance tuning
CI/CD pipelines
GitOps principles

Job description

About ValGenesis

ValGenesis is a leading digital validation platform provider for life sciences companies. ValGenesis suite of products are used by 30 of the top 50 global pharmaceutical and biotech companies to achieve digital transformation, total compliance and manufacturing excellence/intelligence across their product lifecycle.

Learn more about working for ValGenesis, the de facto standard for paperless validation in Life Sciences: https://www.valgenesis.com/about

About the Role:
Responsibilities:
  • Define and embed SRE best practices across the SaaS platform, ensuring reliability is built into the system from the ground up.
  • Establish and maintain meaningful SLA, SLIs, SLOs, and error budgets to protect customer experience and guide engineering priorities.
  • Design and continuously improve high-availability and disaster recovery strategies.
  • Automate manual processes, manage incident response, optimize performance (SLI/SL0).
  • Bridge the gap between development IT operations.
  • Ensure strong tenant isolation and consistent performance within a DB-per-tenant architecture.
  • Strengthen system resiliency across both Azure and on-prem deployments in our hybrid environment.
  • Lead incident response efforts with structured troubleshooting and clear communication.
  • Drive thorough root cause analysis (RCA) and conduct blameless postmortems focused on long-term improvements.
  • Translate incidents into systemic fixes rather than temporary patches.
  • Develop and maintain operational runbooks to standardize responses.
  • Design and maintain a comprehensive observability framework for both cloud and on-prem environments.
Requirements:
  • Must have a minimum of 3+ years of hands‑on experience in Site Reliability Engineering (SRE), supporting production‑grade, cloud‑native enterprise software platform/applications.
  • Prior experience as a DevOps engineer, cloud system administrator or software developer.
  • Strong proficiency in scripting languages such as Python, PowerShell etc.
  • Deep hands‑on experience working with Microsoft Azure in production environments.
  • Possess a solid understanding of Terraform, Ansible, Kubernetes internals, including networking, scheduling, scaling, and resource management.
  • Have proven experience in PostgreSQL performance tuning and optimization in production systems.
  • Demonstrate hands‑on experience with Azure Monitor, Application Insights, and Log Analytics for cloud‑based observability.
  • Implement and manage Prometheus and Grafana for Kubernetes and on‑prem monitoring.
  • Understand how to turn metrics, logs, and traces into actionable insights that improve reliability and performance.
  • Troubleshoot and improve CI/CD pipelines to ensure stable and predictable releases.
  • Apply GitOps principles to manage deployments and infrastructure changes in a controlled and auditable manner.

In 2005, we disrupted the life sciences industry by introducing the world’s first digital validation lifecycle management system. ValGenesis VLMS® revolutionized compliance‑based corporate validation activities and has remained the industry standard.

Today, we continue to push the boundaries of innovation ― enhancing and expanding our portfolio beyond validation with an end‑to‑end digital transformation platform. We combine our purpose‑built systems with world‑class consulting services to help every facet of GxP meet evolving regulations and quality expectations.

Our customers’ success is our success. We keep the customer experience centered in our decisions, from product to marketing to sales to services to support. Life sciences companies exist to improve humanity’s quality of life, and we honor that mission.

We work together. We communicate openly, support each other without reservation, and never hesitate to wear multiple hats to get the job done.

We think big. Innovation is the heart of ValGenesis. That spirit drives product development as well as personal growth. We never stop aiming upward.

We’re in it to win it. We’re on a path to becoming the number one intelligent validation platform in the market, and we won’t settle for anything less than being a market leader.

How We Work

Our Chennai, Hyderabad and Bangalore offices are onsite, 5 days per week. We believe that in‑person interaction and collaboration fosters creativity, and a sense of community, and is critical to our future success as a company.

ValGenesis is an equal‑opportunity employer that makes employment decisions on the basis of merit. Our goal is to have the best‑qualified people in every job. All qualified applicants will receive consideration for employment without regard to race, religion, sex, sexual orientation, gender identity, national origin, disability, or any other characteristics protected by local law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Engineer - DevOps
Lead Engineer - DevOps

Valgenesis • Chennai District

On-site
INR 4,000,000 - 7,500,000
Senior Software Engineer, Fullstack
Senior Software Engineer, Fullstack

ValGenesis, Inc. • Chennai District

On-site
INR 1,800,000 - 3,200,000
Lead Engineer - DevOps
Lead Engineer - DevOps

ValGenesis, Inc. • Chennai District

On-site
INR 2,500,000 - 4,000,000
Lead Engineer - DevOps
Lead Engineer - DevOps

TymblHub • Chennai District

On-site
INR 4,200,000 - 7,000,000
Senior Software Engineer, Full-stack
Senior Software Engineer, Full-stack

ValGenesis, Inc. • Hyderabad

On-site
INR 2,000,000 - 3,500,000
Senior Software Engineer, Full-stack
Senior Software Engineer, Full-stack

ValGenesis • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Senior Software Engineering Manager, Fullstack
Senior Software Engineering Manager, Fullstack

ValGenesis • Chennai District

On-site
INR 3,500,000 - 7,000,000
Lead Software Engineer, Data Engineering
Lead Software Engineer, Data Engineering

ValGenesis • Chennai District

On-site
INR 1,000,000 - 2,000,000
Lead Software Engineer, Data Engineering
Lead Software Engineer, Data Engineering

ValGenesis, Inc. • Chennai District

On-site
INR 2,800,000 - 4,000,000
Senior Manager, Engineering
Senior Manager, Engineering

ValGenesis • Chennai District

On-site
INR 2,000,000 - 3,500,000
Collaborative work environment
Career growth opportunities
Health and wellness benefits