CloudOps Engineer (L2)

Larsen & Toubro

Mumbai

On-site

INR 1,400,000 - 2,100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Larsen & Toubro is seeking an experienced Frontline Infrastructure Support Engineer (L2) to provide advanced technical support and incident resolution for enterprise IT infrastructure in Mumbai. The role covers Windows/Linux administration, Kubernetes, virtualization, storage, backup, GPU infrastructure, and data center operations.

You will diagnose complex issues, perform changes, coordinate with L3 teams for permanent resolutions, and mentor L1 engineers to uphold high availability and SLA

Qualifications

  • BE/BTech or equivalent with Computer Science or Electronics & Communication.
  • 4–6 years of hands-on experience in Windows/Linux Administration, Active Directory, VMware/Hyper-V, Kubernetes, Storage & Backup Administration.
  • Experience supporting mission-critical production environments with 24 x 7 operations.
  • 4–6 years in Hardware Support, ITIL processes, Vendor Coordination, Automation, and OPS reporting.

Responsibilities

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.
  • Investigate, troubleshoot, and resolve Level 2 infrastructure incidents within defined SLAs.
  • Perform root cause analysis (RCA) for recurring incidents and coordinate with L3 teams for permanent resolution.
  • Review escalations from L1 teams and provide technical guidance during incident resolution.
  • Participate in Major Incident Management and support service restoration activities.
  • Administer Windows Server environments, including Active Directory, Group Policy, DNS, DHCP, File Services, and Windows services.
  • Manage user accounts, organizational units, security groups, permissions, and authentication-related activities.
  • Troubleshoot server performance, replication, authentication, and domain-related issues.
  • Perform Linux server administration, including user management, package management, service configuration, performance tuning, and security patching.
  • Analyze system logs, troubleshoot kernel and OS issues, and optimize server performance.
  • Administer Kubernetes clusters, including node management, workload deployment, pod lifecycle management, namespace administration, and cluster health monitoring.
  • Troubleshoot cluster failures, resource constraints, networking, and orchestration issues.
  • Administer enterprise storage systems, LUN provisioning, storage allocation, capacity management, and performance optimization.
  • Manage enterprise backup solutions, perform backup validation, restore operations, and troubleshoot backup failures.
  • Administer GPU-based infrastructure supporting AI/ML workloads.
  • Monitor GPU utilization, troubleshoot hardware/software issues, perform firmware validation, and coordinate hardware replacement activities.
  • Diagnose server hardware failures involving CPU, memory, storage, RAID, power supplies, and network interfaces.
  • Coordinate hardware replacements with vendors and perform post-maintenance validation.
  • Execute approved infrastructure changes following organizational change management processes.
  • Participate in Problem Management by identifying recurring issues and implementing preventive actions.
  • Develop and maintain technical documentation, SOPs, runbooks, and knowledge articles.
  • Mentor L1 engineers and contribute to operational excellence and automation initiatives.
  • Support audit, compliance, and governance activities.

Skills

Windows/Linux Administration
Active Directory
Kubernetes
VMware/Hyper-V
Storage & Backup Administration
Networking
Incident Management
Problem Management
Automation (PowerShell/Bash/Python)

Education

BE/BTech or equivalent in Computer Science or ECE

Tools

VMware
Hyper-V
Kubernetes
Storage systems
Backup solutions
Zabbix/Monitoring tools

Job description

Job PurposeThe Frontline Infrastructure Support Engineer (L2) is responsible for providing advanced technical support, incident resolution, and operational management across enterprise IT infrastructure. The role involves diagnosing and resolving complex infrastructure issues, performing administration activities, implementing changes, supporting problem management, and mentoring L1 engineers while ensuring high service availability and SLA compliance.

Roles & Responsibilities

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.

Infrastructure Administration & Operations

  • Administer and maintain Windows, Linux, virtualization, Kubernetes, storage, backup and GPU infrastructure to ensure high availability and operational stability.
  • Perform proactive health checks, capacity monitoring, performance tuning, and preventive maintenance activities.

Advanced Incident Management

  • Investigate, troubleshoot, and resolve Level 2 infrastructure incidents within defined SLAs.
  • Perform root cause analysis (RCA) for recurring incidents and coordinate with L3 teams for permanent resolution.
  • Review escalations from L1 teams and provide technical guidance during incident resolution.
  • Participate in Major Incident Management and support service restoration activities.

Windows Server & Active Directory Administration

  • Administer Windows Server environments, including Active Directory, Group Policy, DNS, DHCP, File Services, and Windows services.
  • Manage user accounts, organizational units, security groups, permissions, and authentication-related activities.
  • Troubleshoot server performance, replication, authentication, and domain-related issues.

Linux Administration

  • Perform Linux server administration, including user management, package management, service configuration, performance tuning, and security patching.
  • Analyze system logs, troubleshoot kernel and OS issues, and optimize server performance.

Virtualization Administration

  • Administer Cloud native/VMware/Hyper-V environments, including VM provisioning, snapshots, cloning, migration, and resource optimization.
  • Monitor hypervisor performance and troubleshoot virtualization platform issues.

Kubernetes Administration

  • Administer Kubernetes clusters, including node management, workload deployment, pod lifecycle management, namespace administration, and cluster health monitoring.
  • Troubleshoot cluster failures, resource constraints, networking, and orchestration issues.

Storage & Backup Administration

  • Administer enterprise storage systems, LUN provisioning, storage allocation, capacity management, and performance optimization.
  • Manage enterprise backup solutions, perform backup validation, restore operations, and troubleshoot backup failures.

GPU Infrastructure Administration

  • Administer GPU-based infrastructure supporting AI/ML workloads.
  • Monitor GPU utilization, troubleshoot hardware/software issues, perform firmware validation, and coordinate hardware replacement activities.

Hardware & Data Center Support

  • Diagnose server hardware failures involving CPU, memory, storage, RAID, power supplies, and network interfaces.
  • Coordinate hardware replacements with vendors and perform post-maintenance validation.

Change & Problem Management

  • Execute approved infrastructure changes following organizational change management processes.
  • Participate in Problem Management by identifying recurring issues and implementing preventive actions.

Documentation & Continuous Improvement

  • Develop and maintain technical documentation, SOPs, runbooks, and knowledge articles.
  • Mentor L1 engineers and contribute to operational excellence and automation initiatives.
  • Support audit, compliance, and governance activities.

Experience & Educational Requirement

BE/B-Tech or equivalent with Computer Science or Electronics & Communication

Relevant Experience

  • 4–6 years of hands-on experience in Windows/Linux Administration, Active Directory, VMware/Hyper-V, Kubernetes, Storage & Backup Administration, Network Operations, GPU Infrastructure Administration, Troubleshooting, Incident, Problem Management and Infrastructure Troubleshooting.
  • Experience supporting mission – critical production environments with 24 x 7 operations
  • 4–6 years of experience in Hardware Support, Change Management, Documentation, ITIL Processes, Vendor Coordination, Automation (PowerShell/Bash/Python), Operational Reporting, and Team Collaboration.
  • Preferred Exposure: Experience working with cloud platforms including Microsoft Azure, AWS or Google Cloud Platform, cloud native virtualization platforms, VMware, Nutanix, Kubernetes/Openshift environments, cloud native storage, Enterprise storage like NetApp & HPE etc. Backup Solution like Commvault & Cloud native backup tools, enterprise monitoring tools such as Zabbix, Zoho etc. & ITSM Platform including Manage Engine & Service Now etc. and GPU infrastructure supporting AI/ML workload is preferred
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

NOC Technical Lead (L3)
NOC Technical Lead (L3)

Larsen & Toubro • Chennai District

On-site
INR 4,500,000 - 7,500,000
Linux & Cloud Infrastructure Engineer (L2)
Linux & Cloud Infrastructure Engineer (L2)

Lunarays Technologies • Dadri

On-site
INR 900,000 - 1,000,000
NOC Technical Lead L3
NOC Technical Lead L3

Larsen & Toubro • Chennai District

On-site
INR 4,000,000 - 7,000,000
Cloud System Engineer (Server, Storage, Hypervisor and Backup)
Cloud System Engineer (Server, Storage, Hypervisor and Backup)

Larsen & Toubro • Chennai District

On-site
INR 1,000,000 - 1,500,000
CloudOps Engineer – L2
CloudOps Engineer – L2

ESP Engineered • Pune District

On-site
INR 800,000 - 1,200,000
Competitive salary and performance‑based incentives
Professional development and certification opportunities
Career advancement pathways
+2
Infrastructure and Platform Engineer Architect
Infrastructure and Platform Engineer Architect

PwC India • Pune District

On-site
INR 2,000,000 - 4,000,000
Linux & Kubernetes Administrator
Linux & Kubernetes Administrator

Larsen & Toubro • Mumbai

On-site
INR 1,800,000 - 2,400,000
Senior Cloud Infrastructure Engineer
Senior Cloud Infrastructure Engineer

BETSOL • India

On-site
INR 1,200,000 - 2,400,000
Cloud System Engineer
Cloud System Engineer

United States Digital Space LLC • Maharashtra

On-site
INR 2,500,000 - 4,000,000
Cloud Support Engineer
Cloud Support Engineer

Talento HR • Bengaluru

On-site
INR 900,000 - 1,200,000