Lead Software Engineer

Solera Holdings, LLC.

Hyderabad

On-site

INR 7,662,835 - 11,494,252

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Solera Holdings, LLC. is looking for a Lead SRE Engineer in Hyderabad, India. This role is responsible for ensuring the reliability of extensive infrastructure for .NET-based fleet management applications. Candidates should have 8-12 years in Site Reliability Engineering, experience with Kubernetes and CI/CD tools, and strong skills in monitoring, incident management, and security compliance. This position emphasizes automation in operations to enhance deployment reliability and operational efficiency.

Qualifications

  • 8-12 years of relevant experience in DevOps or Site Reliability Engineering.
  • Hands-on expertise in operating production systems and CI/CD pipelines.
  • Deep observability and monitoring skills across distributed application platforms.

Responsibilities

  • Ensure high availability and resilience of production services.
  • Monitor system health and performance proactively.
  • Automate management of platform infrastructure and deployments.

Skills

On-prem infrastructure expertise
Container orchestration (Kubernetes, Rancher)
CI/CD automation experience
Monitoring and observability (Prometheus, Grafana)
Incident management
Security compliance knowledge

Education

Bachelor’s degree in Computer Science

Tools

Terraform
Azure DevOps
Jenkins
Datadog
ELK Stack

Job description

Who We Are

Solera is a leading global technology and data solutions provider, specializing in the automotive industry that strives to transform every touchpoint of the vehicle lifecycle into a connected digital experience. In addition, we provide products and services to protect life’s other most important assets: our homes and digital identities. Today, Solera processes over 300 million digital transactions annually for approximately 235,000 partners and customers in more than 90 countries. Our 6,500 team members foster an uncommon, innovative culture and are dedicated to successfully bringing the future to bear today through cognitive answers, insights, algorithms and automation. For more information, please visit solera.com.



Job Summary

A Lead SRE Engineer responsible for ensuring the reliability, availability, performance, and security of on‑prem infrastructure and .NET‑based fleet management applications. This role blends operational excellence with strong automation, observability, and incident‑response capabilities across high‑scale telemetry and real‑time data systems. As a core member of the development team, you will work to build and maintain robust, reliable infrastructure and automate operational tasks to reduce toil and improve efficiency. We’re seeking an experienced SRE to deliver insights from massive‑scale data in real time.



Essential Responsibilities and Duties


  • Ensure high availability, scalability, and resilience of production services, including APIs, .NET applications, telemetry ingestion pipelines, and on‑prem infrastructure.

  • Run and maintain production environments by continuously monitoring system health, availability, error rates, resource saturation, and end‑to‑end performance.

  • Define, implement, and monitor SLIs/SLOs/SLAs for uptime, latency, throughput, error budgets, and system reliability.

  • Build and maintain software systems to automate the management of platform infrastructure, deployments, and application operations.

  • Measure, analyze, and optimize system performance, proactively identifying bottlenecks and driving architectural improvements.

  • Own incident management, including detection, triaging, mitigation, communication, root cause analysis (RCA), and post‑mortems.

  • Design and maintain monitoring, logging, and observability frameworks (Prometheus, Grafana, Datadog, ELK, APM tools) for distributed services, microservices, telemetry workloads, and on‑prem infrastructure.

  • Develop and enhance automation, CI/CD pipelines to reduce manual toil and improve deployment reliability.

  • Ensure reliability, performance, and best practices are integrated into the SDLC.

  • Manage and operate on‑prem infrastructure, including Rancher, OpenShift, Kubernetes, virtualization, storage, networking, and security controls.

  • Provision, configure, and maintain infrastructure resources using IaC tooling, automation scripts, and configuration management tools.

  • Implement security and compliance best practices, especially around fleet data, driver information, telemetry, GPS, and regulatory requirements.

  • Perform capacity planning and performance tuning for backend services, telemetry systems, and high‑load ingestion pipelines.

  • Provide primary operational support for large‑scale distributed .NET applications and fleet‑critical systems.

  • Maintain detailed documentation on architecture, operational processes, incident playbooks, and system runbooks.



Fleet/Telematics‑Specific Responsibilities


  • Support real‑time data ingestion pipelines (vehicle telemetry, IoT/edge devices, GPS/GNSS streams), ensuring low‑latency and reliable data delivery.

  • Optimize backend systems for load spikes typical in fleet operations (e.g., start‑of‑day vehicle activations, peak trip windows).

  • Monitor the health of vehicle‑facing and driver‑facing data flows, including connectivity, message delivery, and ingestion reliability.

  • Enhance observability for mobile/embedded systems, considering intermittent connectivity, offline sync, and edge constraints.



Qualifications


Education

Bachelor’s degree in Computer Science or equivalent.



Experience

8‑12 years of relevant experience in DevOps or Site Reliability Engineering, with hands‑on expertise in operating production systems, CI/CD pipelines, and distributed application platforms.



Knowledge/Skills/Abilities


  • Strong expertise in on‑prem infrastructure & container orchestration — Rancher, Kubernetes/OpenShift, Docker, virtualization, networking, storage, IP routing, firewalls, and security controls.

  • Deep observability and monitoring skills using Prometheus, Grafana, Datadog, ELK, APMs, log pipelines, distributed tracing, and alerting systems like PagerDuty, with the ability to build end‑to‑end monitoring for APIs, .NET apps, and Java apps, and telemetry pipelines.

  • Advanced reliability engineering capabilities — defining/operationalizing SLIs, SLOs, SLAs, error budgets, availability models, and capacity/performance planning for large‑scale distributed systems.

  • Strong automation and CI/CD experience with GitHub, Octopus, Jenkins/Azure DevOps, IaC (Terraform/Helm/Kustomize), and scripting (PowerShell, Bash, Python) to reduce manual toil and improve deployment reliability.

  • Production operations mastery — incident management (detection → triage → mitigation → RCA/post‑mortem), system health monitoring, performance analysis, scalability improvements, and maintaining high uptime SLAs.

  • Backend performance & systems engineering skills — thread/memory profiling for .NET apps, SQL/No‑SQL Server/Redis tuning, telemetry ingestion optimization, and handling high‑load fleet/telematics workloads.

  • Experience supporting real‑time data flows & IoT/telemetry systems, including GPS/GNSS streams, vehicle connectivity, ingestion reliability, offline/edge constraints, and mobility‑driven scaling patterns.

  • Security and compliance knowledge — secrets management, least‑privilege access, vulnerability scanning, data protection practices for fleet data, driver information, and regulated telemetry workloads.

  • Experience with cloud platforms such as AWS (EKS, EC2, RDS, S3, VPC, IAM) is a plus, especially in hybrid on‑prem + cloud environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Cybage Software • Pune District

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer
Site Reliability Engineer

NOV • Ernakulam

On-site
INR 1,200,000 - 2,000,000
Lead Engineer
Lead Engineer

Solera Holdings, LLC. • Bengaluru Urban

On-site
INR 1,000,000 - 2,000,000
Technical Support Engineer/SRE
Technical Support Engineer/SRE

Boldtek • India

On-site
INR 900,000 - 1,500,000
Hybrid work model
Growth opportunities
SRE Lead
SRE Lead

Hdfc Securities • Mumbai

On-site
INR 3,500,000 - 5,500,000
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Senior Site Reliability Lead
Senior Site Reliability Lead

Generac • Pune District

On-site
INR 3,000,000 - 6,500,000
Senior Manager System Reliability Engineering
Senior Manager System Reliability Engineering

GE Vernova, Inc. • Hyderabad

On-site
INR 4,000,000 - 6,000,000
Relocation Assistance Provided: Yes
Senior Cloud Site Reliability Engineer
Senior Cloud Site Reliability Engineer

Augusta Infotech • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
Senior Platform SRE
Senior Platform SRE

IG Infotech • Bengaluru

On-site
INR 1,200,000 - 1,600,000