Senior SRE

TecQubes Technologies

Bengaluru

Hybrid

INR 2,250,000 - 2,750,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TecQubes Technologies is looking for a Site Reliability Engineer (SRE) with over 10 years of experience to join our team in Bengaluru. This role involves designing and maintaining highly available systems, with a focus on AWS, Kubernetes, and observability tools like Grafana and Elasticsearch.

The ideal candidate will work in a hybrid model and will be responsible for driving reliability and operational excellence across our infrastructure.

Qualifications

  • 10+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering.
  • Strong hands-on experience with AWS services.
  • Advanced expertise in Kubernetes (EKS preferred), Helm, and container orchestration.

Responsibilities

  • Design, build, and operate highly reliable, scalable, and fault-tolerant systems in AWS.
  • Implement and manage Kubernetes (EKS) clusters.
  • Own and improve SLIs, SLOs, and SLAs.

Skills

Site Reliability Engineering
AWS (EC2, EKS, S3, RDS, IAM, VPC, CloudWatch, Auto Scaling)
Kubernetes (EKS preferred)
Elasticsearch (cluster management, performance tuning)
Grafana
Linux system administration
Infrastructure as Code (Terraform, CloudFormation)
Scripting (Python, Bash, Go)

Job description

SRE with AWS, Elastic Search, Kubernetes, Graphana for Bangalore

Experience: 10+ years. Hybrid: 3 days. Location: Kormangla, Bangalore. Salary: 25LPA.

Job Description

We are seeking a highly experienced Site Reliability Engineer (SRE) with 10+ years of experience in designing, implementing, and maintaining highly available, scalable, and resilient systems. The ideal candidate will have deep expertise in AWS, Kubernetes, Elasticsearch, Grafana, and modern SRE practices, with a strong focus on automation, observability, and operational excellence.

Key Responsibilities
  • Design, build, and operate highly reliable, scalable, and fault-tolerant systems in AWS cloud environments.
  • Implement and manage Kubernetes (EKS) clusters, including deployment strategies, scaling, upgrades, and security hardening.
  • Own and improve SLIs, SLOs, and SLAs, driving reliability through data-driven decisions.
  • Architect and maintain observability platforms using Grafana, Prometheus, and Elasticsearch.
  • Manage and optimize Elasticsearch clusters, including indexing strategies, performance tuning, scaling, and backup/restore.
  • Develop and maintain monitoring, alerting, and logging solutions to ensure proactive incident detection and response.
  • Lead incident management, root cause analysis (RCA), postmortems, and continuous improvement initiatives.
  • Automate infrastructure and operations using Infrastructure as Code (IaC) and scripting.
  • Collaborate with development teams to improve system reliability, deployment pipelines, and release processes.
  • Implement CI/CD best practices and reduce deployment risk through canary, blue-green, and rolling deployments.
  • Ensure security, compliance, and cost optimization across cloud infrastructure.
  • Mentor junior SREs and drive adoption of SRE best practices across teams.
Required Skills & Qualifications
  • 10+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering.
  • Strong hands-on experience with AWS services (EC2, EKS, S3, RDS, IAM, VPC, CloudWatch, Auto Scaling).
  • Advanced expertise in Kubernetes (EKS preferred), Helm, and container orchestration.
  • Deep knowledge of Elasticsearch (cluster management, indexing, search optimization, performance tuning).
  • Strong experience with Grafana and observability stacks (Prometheus, Loki, ELK).
  • Proficiency in Linux system administration and networking fundamentals.
  • Experience with Infrastructure as Code tools (Terraform, CloudFormation).
  • Strong scripting skills in Python, Bash, or Go.
Core Technical Skills
  • 10+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering.
  • Strong hands-on experience with AWS services (EC2, EKS, S3, RDS, IAM, VPC, CloudWatch, Auto Scaling).
  • Advanced expertise in Kubernetes (EKS preferred), Helm, and container orchestration.
  • Deep knowledge of Elasticsearch (cluster management, indexing, search optimization, performance tuning).
  • Strong experience with Grafana and observability stacks (Prometheus, Loki, ELK).
  • Proficiency in Linux system administration and networking fundamentals.
  • Experience with Infrastructure as Code tools (Terraform, CloudFormation).
  • Strong scripting skills in Python, Bash, or Go.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE Technical Specialist
Senior SRE Technical Specialist

United States Digital Space LLC • Karnataka

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

F-Prime Capital • Pune District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Hilabs • Pune District

On-site
INR 1,500,000 - 2,500,000
SRE - AWS DevOPS Engineer
SRE - AWS DevOPS Engineer

Prowess Publishing • Hyderabad

On-site
INR 900,000 - 1,500,000
Assistant Manager - Azure Site Reliability Engineer
Assistant Manager - Azure Site Reliability Engineer

Promaynov Advisory Services Pvt. Ltd • Bengaluru

On-site
INR 1,400,000 - 2,100,000
Site Reliability Engineer (SRE) / Observability Engineer
Site Reliability Engineer (SRE) / Observability Engineer

N Human Resources & Management Systems • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work
Certification reimbursement
Structured learning
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Site Reliability Engineer
Site Reliability Engineer

Talentrouters • Chennai District

On-site
INR 2,500,000 - 4,200,000