CloudOps Engineer (L2)

Larsen & Toubro

Mumbai

On-site

INR 1,400,000 - 2,100,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Larsen & Toubro is seeking an experienced Frontline Infrastructure Support Engineer (L2) to provide advanced technical support and incident resolution for enterprise IT infrastructure in Mumbai. The role covers Windows/Linux administration, Kubernetes, virtualization, storage, backup, GPU infrastructure, and data center operations.

You will diagnose complex issues, perform changes, coordinate with L3 teams for permanent resolutions, and mentor L1 engineers to uphold high availability and SLA

Qualifications

  • BE/BTech or equivalent with Computer Science or Electronics & Communication.
  • 4–6 years of hands-on experience in Windows/Linux Administration, Active Directory, VMware/Hyper-V, Kubernetes, Storage & Backup Administration.
  • Experience supporting mission-critical production environments with 24 x 7 operations.
  • 4–6 years in Hardware Support, ITIL processes, Vendor Coordination, Automation, and OPS reporting.

Responsibilities

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.
  • Investigate, troubleshoot, and resolve Level 2 infrastructure incidents within defined SLAs.
  • Perform root cause analysis (RCA) for recurring incidents and coordinate with L3 teams for permanent resolution.
  • Review escalations from L1 teams and provide technical guidance during incident resolution.
  • Participate in Major Incident Management and support service restoration activities.
  • Administer Windows Server environments, including Active Directory, Group Policy, DNS, DHCP, File Services, and Windows services.
  • Manage user accounts, organizational units, security groups, permissions, and authentication-related activities.
  • Troubleshoot server performance, replication, authentication, and domain-related issues.
  • Perform Linux server administration, including user management, package management, service configuration, performance tuning, and security patching.
  • Analyze system logs, troubleshoot kernel and OS issues, and optimize server performance.
  • Administer Kubernetes clusters, including node management, workload deployment, pod lifecycle management, namespace administration, and cluster health monitoring.
  • Troubleshoot cluster failures, resource constraints, networking, and orchestration issues.
  • Administer enterprise storage systems, LUN provisioning, storage allocation, capacity management, and performance optimization.
  • Manage enterprise backup solutions, perform backup validation, restore operations, and troubleshoot backup failures.
  • Administer GPU-based infrastructure supporting AI/ML workloads.
  • Monitor GPU utilization, troubleshoot hardware/software issues, perform firmware validation, and coordinate hardware replacement activities.
  • Diagnose server hardware failures involving CPU, memory, storage, RAID, power supplies, and network interfaces.
  • Coordinate hardware replacements with vendors and perform post-maintenance validation.
  • Execute approved infrastructure changes following organizational change management processes.
  • Participate in Problem Management by identifying recurring issues and implementing preventive actions.
  • Develop and maintain technical documentation, SOPs, runbooks, and knowledge articles.
  • Mentor L1 engineers and contribute to operational excellence and automation initiatives.
  • Support audit, compliance, and governance activities.

Skills

Windows/Linux Administration
Active Directory
Kubernetes
VMware/Hyper-V
Storage & Backup Administration
Networking
Incident Management
Problem Management
Automation (PowerShell/Bash/Python)

Education

BE/BTech or equivalent in Computer Science or ECE

Tools

VMware
Hyper-V
Kubernetes
Storage systems
Backup solutions
Zabbix/Monitoring tools

Job description

Job PurposeThe Frontline Infrastructure Support Engineer (L2) is responsible for providing advanced technical support, incident resolution, and operational management across enterprise IT infrastructure. The role involves diagnosing and resolving complex infrastructure issues, performing administration activities, implementing changes, supporting problem management, and mentoring L1 engineers while ensuring high service availability and SLA compliance.

Roles & Responsibilities

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.

Infrastructure Administration & Operations

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.

Advanced Incident Management

  • Investigate, troubleshoot, and resolve Level 2 infrastructure incidents within defined SLAs.
  • Perform root cause analysis (RCA) for recurring incidents and coordinate with L3 teams for permanent resolution.
  • Review escalations from L1 teams and provide technical guidance during incident resolution.
  • Participate in Major Incident Management and support service restoration activities.

Windows Server & Active Directory Administration

  • Administer Windows Server environments, including Active Directory, Group Policy, DNS, DHCP, File Services, and Windows services.
  • Manage user accounts, organizational units, security groups, permissions, and authentication-related activities.
  • Troubleshoot server performance, replication, authentication, and domain-related issues.

Linux Administration

  • Perform Linux server administration, including user management, package management, service configuration, performance tuning, and security patching.
  • Analyze system logs, troubleshoot kernel and OS issues, and optimize server performance.

Virtualization Administration

  • Administer Cloud native/VMware/Hyper-V environments, including VM provisioning, snapshots, cloning, migration, and resource optimization.
  • Monitor hypervisor performance and troubleshoot virtualization platform issues.

Kubernetes Administration

  • Administer Kubernetes clusters, including node management, workload deployment, pod lifecycle management, namespace administration, and cluster health monitoring.
  • Troubleshoot cluster failures, resource constraints, networking, and orchestration issues.

Storage & Backup Administration

  • Administer enterprise storage systems, LUN provisioning, storage allocation, capacity management, and performance optimization.
  • Manage enterprise backup solutions, perform backup validation, restore operations, and troubleshoot backup failures.

GPU Infrastructure Administration

  • Administer GPU-based infrastructure supporting AI/ML workloads.
  • Monitor GPU utilization, troubleshoot hardware/software issues, perform firmware validation, and coordinate hardware replacement activities.

Hardware & Data Center Support

  • Diagnose server hardware failures involving CPU, memory, storage, RAID, power supplies, and network interfaces.
  • Coordinate hardware replacements with vendors and perform post-maintenance validation.

Change & Problem Management

  • Execute approved infrastructure changes following organizational change management processes.
  • Participate in Problem Management by identifying recurring issues and implementing preventive actions.

Documentation & Continuous Improvement

  • Develop and maintain technical documentation, SOPs, runbooks, and knowledge articles.
  • Mentor L1 engineers and contribute to operational excellence and automation initiatives.
  • Support audit, compliance, and governance activities.

Experience & Educational Requirement

BE/B-Tech or equivalent with Computer Science or Electronics & Communication

Relevant Experience

  • 4–6 years of hands-on experience in Windows/Linux Administration, Active Directory, VMware/Hyper-V, Kubernetes, Storage & Backup Administration, Network Operations, GPU Infrastructure Administration, Troubleshooting, Incident, Problem Management and Infrastructure Troubleshooting.
  • Experience supporting mission – critical production environments with 24 x 7 operations
  • 4–6 years of experience in Hardware Support, Change Management, Documentation, ITIL Processes, Vendor Coordination, Automation (PowerShell/Bash/Python), Operational Reporting, and Team Collaboration.
  • Preferred Exposure: Experience working with cloud platforms including Microsoft Azure, AWS or Google Cloud Platform, cloud native virtualization platforms, VMware, Nutanix, Kubernetes/Openshift environments, cloud native storage, Enterprise storage like NetApp & HPE etc. Backup Solution like Commvault & Cloud native backup tools, enterprise monitoring tools such as Zabbix, Zoho etc. & ITSM Platform including Manage Engine & Service Now etc. and GPU infrastructure supporting AI/ML workload is preferred
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

CloudOps Engineer (L3)
CloudOps Engineer (L3)

Larsen & Toubro-Vyoma • Chennai District

On-site
INR 2,500,000 - 4,500,000
NOC Technical Lead (L3)
NOC Technical Lead (L3)

Larsen & Toubro • Chennai District

On-site
INR 4,500,000 - 7,500,000
L1/L2 NOC
L1/L2 NOC

PeopleStrong • Mumbai

On-site
INR 900,000 - 1,500,000
Senior Engineer
Senior Engineer

Neurealm • Chennai District

On-site
INR 900,000 - 1,300,000
Cloud System Engineer (Server, Storage, Hypervisor and Backup)
Cloud System Engineer (Server, Storage, Hypervisor and Backup)

Larsen & Toubro • Chennai District

On-site
INR 1,000,000 - 1,500,000
SeniorAdministrator - Linux, Red Hat Cluster, Red Hat Satellite
SeniorAdministrator - Linux, Red Hat Cluster, Red Hat Satellite

Gohyred • Uttar Pradesh

On-site
INR 1,800,000 - 2,600,000
Cloud Engineer- L3
Cloud Engineer- L3

Base8 • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Linux & Kubernetes Administrator
Linux & Kubernetes Administrator

Larsen & Toubro • Mumbai

On-site
INR 1,800,000 - 2,400,000
Technical Support Engineer - Data & Cloud Platforms
Technical Support Engineer - Data & Cloud Platforms

Luxoft • Gurugram District

On-site
INR 900,000 - 1,500,000
L2 Azure & Wintel Support Engineer
L2 Azure & Wintel Support Engineer

US Software Group Inc • Bengaluru

On-site
INR 600,000 - 1,200,000