Site Reliability Engineer

EvoluteIQ

Bengaluru

On-site

INR 3,500,000 - 6,000,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Growth opportunities
Innovative culture

Job summary

EvoluteIQ is seeking an experienced Site Reliability Engineer to strengthen platform reliability and operational excellence in Bengaluru. The role requires hands-on expertise across cloud infrastructure, Kubernetes, and observability, with collaboration across Engineering and Product teams.

You'll work on incident management, automation, and production operations across AWS/Azure/GCP, ensuring reliability and performance of our platform.

Qualifications

  • 9+ years in Site Reliability/Platform/DevOps roles.
  • Experience troubleshooting Java/Spring Boot apps and SQL databases.
  • Strong skills with Kubernetes, Docker, Linux.
  • Experience with REST APIs and enterprise integrations.
  • Scripting with Python, Bash; cloud-native tooling.

Responsibilities

  • Own incident management and post-incident reviews.
  • Monitor production environments and preempt issues.
  • Develop runbooks and troubleshooting procedures.
  • Define reliability improvements and self-healing mechanisms.
  • Troubleshoot Java/Spring Boot apps, DB connectivity, and APIs.
  • Build monitoring, logging, tracing, and observability.

Skills

Kubernetes
Docker
Linux
Java/Spring Boot
APIs & Integrations
Python
Bash
SQL Databases
Cloud Platforms
Observability
Automation
Networking

Tools

AI tools
Cloud Native Tools

Job description

We at EvoluteIQ believe in the power of transformation. We are committed to building an industry leading technology that will revolutionize the way enterprises conduct business. To make that happen, we need people who are generous, genuine, self-driven, and collaborative.

People who not only want to be a part of a fast-growing and radical thinking company, but who are kind and caring—about each other. We at EvoluteIQ thrive in the company of each other and make each other a better version of ourselves every day.

Could that be you?

We are looking for an experienced Site Reliability Engineer to strengthen our platform reliability and operational excellence. The role requires strong hands‑on expertise across cloud infrastructure, Kubernetes, application troubleshooting, observability, automation, and production operations.

What you'll do at EvoluteIQ:

As an SRE you will be responsible for ensuring the reliability, availability, scalability, performance, and operational excellence of EvoluteIQ's platform across AWS, Azure, and Google Cloud.

You will work closely with Engineering, Product, and Customer‑facing teams to troubleshoot complex production issues, improve platform resilience, automate operational processes, and drive continuous improvements in the reliability of the platform.

The role requires strong hands‑on experience with Kubernetes, Docker, Linux, cloud platforms, Java/Spring Boot applications, databases, networking, APIs, and enterprise integrations.

Platform Reliability & Production Operations
  • Own and drive in incident management, troubleshooting, root‑cause analysis, and post‑incident reviews.
  • Monitor production environments and proactively identify potential reliability and performance issues.
  • Establish and maintain operational runbooks and troubleshooting procedures.
  • Define and implement reliability improvements, operational best practices, and self‑healing mechanisms.
Application & Database Troubleshooting
  • Debug and troubleshoot Java/Spring Boot applications running on Docker and Kubernetes.
  • Analyze application logs, stack traces, resource utilization, and performance issues.
  • Troubleshoot database connectivity, queries, connection pools, and performance issues.
  • Troubleshoot JDBC/ODBC connectivity, REST APIs, enterprise connectors, and integration workflows.
  • Work with Engineering teams to identify application‑level reliability and performance improvements.
Infra & Automation
  • Contribute to self‑healing and automated remediation capabilities by using AI tools, Python, Bash and cloud native tools for quick proof of concepts.
  • Manage and troubleshoot production workloads across AWS, Microsoft Azure, Google Cloud Platform (GCP) and on‑prem.
  • Build and enhance monitoring, alerting, logging, tracing, and observability capabilities.
What will you bring to the team
  • 9+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, or a similar role.
  • Experience troubleshooting Java/Spring Boot applications and good understanding of SQL and databases.
  • Experience with REST APIs, enterprise integrations, JDBC/ODBC, and application connectivity.
  • Strong hands‑on experience with Kubernetes and Docker.
  • Strong Linux administration and troubleshooting skills.
  • Strong scripting/programming skills using AI tools, Python, Bash and cloud native tools.
  • Strong knowledge of networking concepts, DNS, HTTP/HTTPS, TCP/IP, load balancing, and connectivity troubleshooting.
What We Offer
  • Opportunity to shape the strategy of a next‑gen hyper‑automation platform.
  • Work with a cross‑disciplinary team in a fast‑growing, innovation‑driven environment. Competitive compensation and growth opportunities.
  • A culture of innovation, ownership, and continuous learning.

We value a range of diverse backgrounds, experiences, and ideas. We pride ourselves on our diversity and inclusive workplace that provides equal opportunities to all persons regardless of age, race, color, religion, sex, sexual orientation, gender identity and expression, national origin, disability, neurodiversity, military and/or veteran status, or any other protected classes.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - Python
Senior Software Engineer - Python

EvoluteIQ • Bengaluru

On-site
INR 2,500,000 - 3,500,000
Site Reliability Engineer
Site Reliability Engineer

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,200,000
Significant equity in a venture-backed company
Opportunity to work with modern tech stack
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Technical Lead Python
Technical Lead Python

EvoluteIQ • Bengaluru

On-site
INR 4,500,000 - 7,000,000
Site Reliability Engineer
Site Reliability Engineer

Enterpret • Bengaluru

On-site
INR 1,500,000 - 2,700,000
Equity
Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Pivotree Inc. • Bengaluru, Mumbai

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

DeepIQ • Hyderabad

On-site
INR 1,200,000 - 2,100,000
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

United States Digital Space LLC • Bengaluru

Hybrid
INR 6,000,000 - 12,000,000