Project Lead

HCL Technologies Limited

Bengaluru

On-site

INR 1,400,000 - 2,200,000

Full time

8 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

HCLTech in Bengaluru is seeking a highly skilled Site Reliability Engineer to join the Verizon Project, focusing on reliability, scalability and performance of mission-critical systems.

You will own incident and release management, drive automation and monitoring, and collaborate with cross-functional teams across development, operations and support to deliver robust operational excellence in a fast-paced, customer-focused environment.

Qualifications

  • Proven experience as a Site Reliability Engineer in high-availability environments.
  • Strong knowledge of cloud platforms (AWS/Azure) and Linux-based systems.
  • Hands-on with monitoring/observability tools and incident management.

Responsibilities

  • Incident Management: act as a key resource to minimize impact and drive resolution.
  • Release Management: plan, coordinate, and execute releases across environments.
  • Automation: identify opportunities and develop tooling for resilience and efficiency.
  • Infrastructure Monitoring: establish and maintain comprehensive monitoring for availability.
  • Collaboration: work with cross-functional teams to address requirements and improvements.
  • Post-Mortems: conduct root cause analyses and implement preventive measures.

Skills

Site Reliability Engineer
Incident management
Cloud platforms (AWS/Azure)
Linux/Unix administration
Monitoring/observability (Prometheus,/
ELK/Splunk
CI/CD and Git
OpenShift / Kubernetes

Tools

Prometheus
Grafana
Elasticsearch
Splunk
Git
CI/CD pipelines
OpenShift
Kubernetes

Job description

Looking for a highly skilled and motivated Site Reliability Engineer to join Verizon Project. As a Site Reliability Engineer, you will play a crucial role in ensuring the reliability, scalability, and performance of our systems. You will be responsible for incident management, release management, automation, infrastructure monitoring, and collaborating with cross-functional teams.

Key Responsibilities
  • Incident Management: Act as a key resource in incident management, responding promptly and effectively to incidents to minimize impact. Lead incident resolution efforts, working closely with stakeholders and subject matter experts.
  • Release Management: Manage the planning, coordination, and execution of releases across multiple environments. Ensure smooth release processes, including risk assessment, communication, and rollback strategies.
  • Automation: Identify opportunities for automation and drive the development of tools and frameworks to improve system resiliency, efficiency, and performance. Collaborate with the development and operations teams to implement automation solutions.
  • Infrastructure Monitoring: Establish and maintain comprehensive monitoring systems to ensure high availability and performance of applications and services. Proactively identify potential issues and bottlenecks, and work towards their resolution.
  • Collaboration: Work closely with cross-functional teams, including development, operations, and support, to understand requirements, address issues, and drive continuous improvement. Foster a collaborative and proactive culture within the organization.
  • Incident Post-Mortems: Conduct post-incident analysis and root cause investigations. Identify opportunities for process improvements and work with stakeholders to implement preventive measures.
  • Documentation: Maintain accurate documentation of system configurations, processes, and procedures. Contribute to the knowledge base and provide training and support to team members.
Skill Requirements
  • Proven experience as a Site Reliability Engineer or in a similar role, with a focus on high-availability production environments.
  • Strong understanding of cloud computing platforms, such as Amazon Web Services (AWS) or Microsoft Azure.
  • Solid understanding of Linux/Unix systems administration and troubleshooting.
  • Familiarity with monitoring and observability tools like Prometheus, Grafana, Elasticsearch, or Splunk.
  • Splunk: Expert-level — dashboards, complex queries, production log analysis under pressure
  • Troubleshooting under pressure: Non-negotiable core competency — diagnosing multi-service production failures in real time
  • Strong analytical and problem-solving skills, with the ability to diagnose and resolve complex technical issues.
  • Excellent communication and collaboration skills, with the ability to work effectively in cross-functional teams.
  • Knowledge of DevOps principles and practices, including CI/CD pipelines and version control systems (e.g., Git).
  • Non-negotiable core competency — diagnosing multi-service production failures in real time
Other Requirements

Soft Skills — Equally Important

  • Leadership: Confidently leads triage calls with engineers, DBAs, and senior stakeholders — not a passive participant
  • Communication: Clear, concise verbal and written communication; comfortable on recorded bridge calls
  • Ownership & Accountability: Follows through to completion — open items do not sit unattended, incidents do not close until truly resolved
  • Analytical Thinking: Strong problem-solving skills; drives root cause analysis, not surface-level fixes
  • Composure: Calm, focused, and decisive during high-pressure P1 incidents

Preferred Skills:

  • Experience with Openshift in addition to Kubernetes
  • Familiarity with configuration management tools like Stash or GitHub.
  • Familiarity with networking concepts, REST APIs and security best practices.
  • Prior experience in IoT, telecom, or large-scale connected device platforms
  • Confluence and JIRA proficiency for documentation and ticket management
  • Certification in relevant technologies is a plus.

At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.

HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026totaled $14.8billion.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Track Lead
Senior Track Lead

HCL Technologies Limited • Chennai District

On-site
INR 2,600,000 - 4,800,000
Lead Site Reliability Engineer, Incident Management
Lead Site Reliability Engineer, Incident Management

Qualys • Maharashtra

Hybrid
INR 3,000,000 - 4,500,000
Hybrid/Remote work
On-call leadership
Senior Technical Lead
Senior Technical Lead

HCL Technologies Limited • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior Technical Lead
Senior Technical Lead

HCL Technologies Limited • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Client Delivery Director
Client Delivery Director

HCL Technologies Limited • Chennai District

On-site
INR 4,000,000 - 7,000,000
Track Lead - AWS IAC, Terraform,Python
Track Lead - AWS IAC, Terraform,Python

HCL Technologies Limited • Dadri

On-site
INR 1,200,000 - 1,800,000
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000
Technical Lead
Technical Lead

HCL Technologies Limited • Dadri

On-site
INR 900,000 - 1,500,000
Azure DevOps Senior Technical Specialist
Azure DevOps Senior Technical Specialist

HCL Technologies Limited • Nagpur District

On-site
INR 2,400,000 - 4,000,000
Lead Site Reliability Engineer, Incident Management
Lead Site Reliability Engineer, Incident Management

Qualys • Pune District

Hybrid
INR 4,000,000 - 7,000,000