Site Reliability Engineer

Gemini Solutions

Gurugram District

On-site

INR 2,500,000 - 4,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Gemini Solutions is seeking a Site Reliability Engineer (SRE) with 6–7 years of experience in techops and application support to drive production reliability. The role emphasizes ownership of incidents, secrets management, and operational hygiene across critical apps hosted on AWS.

Responsibilities include incident lifecycle management, monitoring setup, and collaboration with development, platform, and infra teams. SRE-first with DevOps elements introduced as needed to meet delivery timelines.

Qualifications

  • Minimum 6 to 7 years of experience in SRE, Application Support, or TechOps roles.
  • Experience owning end-to-end incident lifecycle and post-incident documentation.
  • Hands-on monitoring & observability using industry tools and alerting policies.
  • Secrets and credentials lifecycle management using AWS Secrets Manager, HashiCorp Vault, Beacon Vault.

Responsibilities

  • Own incident resolution lifecycle including triage, resolution, RCA, and post-incident reports.
  • Configure monitoring, alerting, health checks, and escalation policies (Datadog, PagerDuty).
  • Manage secrets lifecycle, rotation, expiry monitoring, and cross-team coordination.
  • Support AWS operations including S3, IAM, key management, and access provisioning.
  • Develop operational automation scripts (Bash or similar) and contribute to CI/CD onboarding.

Job description

  • We are seeking a Site Reliability Engineer (SRE) with 6 to 7 years of experience in techoperations and application support, with a strong focus on production reliability, monitoring,and operational ownership. The engineer will own incident resolution, maintain secrets andcredentials lifecycle, and drive tech hygiene and operational improvements across businesscritical applications hosted on AWS.
POSITION RESPONSIBILITIES
  • Minimum of 6 to 7 years of experience in SRE, Application Support, or TechOps roles.
  • Incident Management: Experience owning end-to-end incident lifecycle including triage, resolution, root cause analysis, and post-incident documentation following ITIL practices.
  • Monitoring and Observability: Hands-on experience configuring monitoring, alerting, health checks, and escalation policies using Datadog and PagerDuty. Ability to proactively identify monitoring gaps, reduce alert noise, and ensure correct alert routing and ownership across applications and environments.
  • Secrets and credentials management: Experience managing end-to-end secrets lifecycle including rotation scheduling, expiry monitoring, and coordination with application and DBA teams using AWS Secrets Manager, HashiCorp Vault, and Beacon Vault.
  • Aws and cloud operations: Working knowledge of AWS S3 and IAM for operational tasks such as
  • key management, access provisioning, and permission-related requests.
  • Scripting for operational automation: Ability to write and maintain scripts using Bash or equivalent for operational tasks such as alert automation, health checks, and monitoring scripts. AI-assisted scripting tools may be used to support delivery.
  • Version control and pipeline onboarding: Basic working knowledge of GitHub and GitHub Actions
  • to onboard repositories onto standardised CI/CD workflows provided by the platform team by importing templates and configuring repository-specific variables.
  • Documentation and coordination: Ability to create and maintain runbooks, wikis, and operational procedures, and to coordinate effectively with development, platform, infrastructure, and business teams to drive deliverables to closure.
EXPERIENCE AND REQUIRED SKILL SETS
  • Minimum of 6 to 7 years of experience in SRE, Application Support, or TechOps roles.
  • The engineer will coordinate closely with development, platform, and infrastructure teams on
  • operational deliverables, where tasks require skills or capacity beyond the SRE's scope, the
  • expectation is to actively coordinate and follow through to closure.
  • The role is SRE-first, with DevOps responsibilities introduced progressively based on bandwidth and team need.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Innodata Inc. • India

On-site
INR 2,400,000 - 4,000,000
Site Reliability Engineer (SRE) – Core IT Infrastructure
Site Reliability Engineer (SRE) – Core IT Infrastructure

TECEZE • Chennai District

On-site
INR 1,000,000 - 2,000,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Lonvec Technologies Private Limited • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Site Reliability Engineer
Site Reliability Engineer

Lloyds Technology Centre • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Unified Consultancy Services • Karnataka

Hybrid
INR 1,500,000 - 3,000,000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Lyzr AI • Bengaluru

Hybrid
INR 1,000,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Tekskills • Pune District

On-site
INR 1,200,000 - 1,800,000
Restaurant d'entreprise
Indemnités de stage/alternance
Site Reliability Engineer
Site Reliability Engineer

Tekskills • Chennai District

On-site
INR 1,800,000 - 3,000,000
Lead SRE
Lead SRE

JobItUs • Secunderabad, Hyderabad

On-site
INR 1,200,000 - 1,600,000
SRE Lead
SRE Lead

Acldigital • Ahmedabad District

On-site
INR 1,500,000 - 2,000,000