L1 Operations Engineer

Zoho

Bengaluru

On-site

INR 700,000 - 900,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

ZYBISYS CONSULTING SERVICES LLP is seeking an L1 Operations Engineer to monitor Azure cloud and hybrid on‑prem infrastructure. The role requires strong incident logging, initial triage, and clear communication during incidents in a fast-paced managed services environment.

The ideal candidate has 2–5 years in IT operations, hands-on Azure, Windows Server, and Linux familiarity, plus experience with monitoring tools and ITSM platforms. This is a shift-based full-time position in a 24x7 operation.

Qualifications

  • 2–5 years of experience in IT operations, infrastructure support, or cloud managed services.
  • Working knowledge of Microsoft Azure and basic troubleshooting.
  • Experience with Windows Server and Linux systems.
  • Understanding of hybrid on‑prem and Azure integration.
  • Familiarity with core storage concepts and monitoring tools.

Responsibilities

  • Monitor Azure cloud and hybrid datacenter infrastructure 24x7; triage alerts and escalate per SLA.
  • Execute runbooks and SOPs; document actions with timestamps.
  • Log incidents and tickets with details; update status through lifecycle.
  • Coordinate with L2 engineers during escalations and provide handover notes.
  • Follow priority matrix: P1–P4 response and escalation windows.

Skills

IT operations
Azure
Windows Server
Linux
Hybrid infrastructure
Storage concepts

Tools

Prometheus
Grafana
Azure Monitor
Log Analytics
ITSM platforms

Job description

ZYBISYS CONSULTING SERVICES LLP | Full time

L1 Operations Engineer

EmploymentType: Full-Time|Shift-Based (24x7 Rotational)

About Zybisys

Zybisys is a technology companythat helps banks, financial institutions, and FinTech businesses build and runsecure, reliable, and high-performance technology platforms. We work closelywith some of India's leading stock brokers to manage their cloudinfrastructure, cybersecurity, platform operations, and observability. Withdeep expertise in the Capital Markets domain, we focus on simplifying complextechnology, improving operational resilience, and helping our customers innovatewith confidence.

Job Description

The L1 Operations Engineer isthe first line of defence in Zybisys's 24x7 managed cloud operations. This roleis responsible for continuous monitoring of multi-region Azure cloud and hybridon-premises infrastructure, timely triage of alerts, first-line incidentresponse, and ensuring all issues are accurately logged, prioritised, andescalated within SLA thresholds.

This is a shift-basedoperational role requiring strong attention to detail, disciplined runbookexecution, and clear communication during incidents. The ideal candidate istechnically curious, process-oriented, and comfortable working across cloudmonitoring tools, ITSM platforms, and infrastructure dashboards in a fast-pacedmanaged services environment.

Key Responsibilities
Infrastructure Monitoring &Alert Management
  • Perform continuous 24x7monitoring of Azure cloud and hybrid datacenter infrastructure usingPrometheus, Grafana, Azure Monitor, and related observability dashboards.
  • Triage incoming alerts — assessseverity, validate against known patterns, and determine whether to resolve atL1 or escalation to L2 within defined SLA thresholds.
  • Execute approved runbooks andSOPs for all known alert categories; document actions taken for every incidentwith accurate timestamps and observations.
  • Monitor health and availabilityof compute (VMs), storage, network links, VPN tunnels, and platform servicesacross cloud and on-premises environments.
  • Track sFlow and NetFlowdashboards for network traffic anomalies; flag unusual patterns to the L2 teamfor deeper investigation.
  • Log all incidents, servicerequests, and alerts in the ITSM platform with complete and accurate details —symptoms, affected components, priority, and initial actions taken.
  • Update ticket status throughoutthe incident lifecycle; ensure no incident is left without a current statusupdate beyond the defined response window.
  • Coordinate with L2 engineersduring escalations — provide clear handover notes including timeline, alertcontext, initial diagnostics, and business impact assessment.
  • Follow the priority matrixstrictly: P1 (15 min), P2 (30 min), P3 (4 hr), P4 (8 hr) response andescalation thresholds.
Platform & Service HealthChecks
  • Execute scheduled shift healthchecks across all managed platforms — Azure resources, on-premises servers,network devices, security appliances, and employee services.
  • Verify availability andperformance of core services: DNS, DHCP, NTP, Active Directory, and M365platform components.
  • Monitor security platformdashboards (firewalls, EDR, proxy services) for health status and alert flags;escalate anomalies per defined procedures.
  • Review Azure Cost Managementdashboards for unusual consumption spikes and flag to the lead for review.
Routine Operations &Maintenance Support
  • Execute scheduled batch jobs,backup verifications, replication checks, and housekeeping tasks as per theoperational calendar.
  • Support L2 and Specialistengineers during planned maintenance windows, patch cycles, and changeactivities — providing monitoring coverage and rollback readiness.
  • Validate post-changeinfrastructure health after every approved change; raise a flag immediately ifanomalies are detected post-implementation.
Documentation & KnowledgeManagement
  • Maintain precise shift handoverreports — open tickets, ongoing incidents, recent changes, and watch-items forthe next shift.
  • Contribute to the knowledge baseby documenting recurring alert patterns, resolution steps, and workarounds forL1-resolvable issues.
  • Flag gaps in runbooks or SOPs tothe operations lead so that documentation is continuously improved.
Required Skills & Experience
  • 2 – 5 years of experience in IToperations, infrastructure support, or cloud managed services.
  • Working knowledge of MicrosoftAzure: Azure Portal navigation, VM status checks, resource monitoring, andbasic troubleshooting using Azure Monitor and Log Analytics.
  • Hands-on familiarity withWindows Server (2016/2019/2022) and Linux (RHEL/Ubuntu) — service management,log file reading, process monitoring, and basic fault diagnosis.
  • Understanding of hybridinfrastructure models — on-premises datacenter integrated with Azure cloud viaExpressRoute or VPN.
  • Basic familiarity with storageconcepts: disk performance thresholds, capacity monitoring, backup job status,and replication health checks.
Monitoring & ObservabilityTools
  • Experience reading andinterpreting Prometheus metrics and Grafana dashboards — understanding panelthresholds, alert states, and trend data.
  • Familiarity with Azure Monitoralerts and Log Analytics at a basic level — understanding alert rules, severitylevels, and affected resources.
  • Ability to read sFlow or NetFlowtraffic dashboards for high-level network health assessment.
  • APM dashboard familiarity —understanding response time trends, error rates, and service health indicatorsfrom tools such as Dynatrace, AppDynamics, or Azure Application Insights.
  • Good understanding of corenetworking concepts: TCP/IP, DNS, DHCP, NTP, VLANs, and basic routing.
  • Practical diagnostic skills:ping, traceroute, nslookup, netstat — to perform first-level connectivitychecks and provide meaningful diagnostics to L2.
  • Familiarity with VPN tunnelhealth monitoring and firewall status dashboards at an operational level.
ITSM & Process Discipline
  • Experience working with ITSMplatforms (ServiceNow, Zoho Desk, Freshservice, or equivalent) for incidentlogging, ticket updates, and escalation workflows.
  • Understanding of ITIL incidentmanagement concepts — priority, impact, urgency, escalation paths, and SLAtracking.
  • Ability to follow runbooks andSOPs precisely and consistently, including under pressure during major incidentscenarios.
  • Clear and concise writtencommunication for incident tickets, shift handover notes, and escalationsummaries.
Soft Skills & Work Style
  • Comfortable working in a 24x7rotational shift environment including night shifts, weekends, and publicholidays.
  • High attention to detail —accurate logging, precise documentation, and consistent process adherence.
  • Team-oriented with a proactiveattitude toward learning and improving operational knowledge.
  • Ability to stay calm andsystematic during high-pressure P1/P2 incident scenarios.
Certification (Preferred)
Domain
Certification

AZ-900 (AzureFundamentals)|AZ-104 (Administrator — working towards)

Service Management

ITIL v4 Foundation

Networking

CompTIA Network+|Cisco CCNA (advantageous)

Monitoring

Grafana Certified Associate(advantageous)

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

L1 Operations Engineer
L1 Operations Engineer

Zybisys • Bengaluru

On-site
INR 600,000 - 1,100,000
L2 Engineer
L2 Engineer

Zoho • Bengaluru Urban

On-site
INR 1,200,000 - 1,800,000
L2 Engineer
L2 Engineer

Zybisys • Bengaluru

On-site
INR 1,000,000 - 1,800,000
Systems Engineer
Systems Engineer

Altera Digital Health Inc. • Maharashtra

Remote
INR 1,800,000 - 2,400,000
Systems Engineer
Systems Engineer

Harris Computer • Maharashtra

Remote
INR 1,500,000 - 1,900,000
Azure Engineer (Tier 1)
Azure Engineer (Tier 1)

AlifCloud IT Consulting Pvt. Ltd. • Maharashtra

On-site
INR 300,000 - 500,000
L1 Azure Cloud Engineer
L1 Azure Cloud Engineer

AlifCloud IT Consulting Pvt. Ltd. • Maharashtra

On-site
INR 300,000 - 420,000
Associate Data & Ops Engineer
Associate Data & Ops Engineer

EXL • Maharashtra

On-site
INR 350,000 - 520,000
Azure Engineer (Tier 1)
Azure Engineer (Tier 1)

Zoho • Pune District

On-site
INR 400,000 - 600,000
Network Engineer 2
Network Engineer 2

Moder Solutions • Chennai District

On-site
INR 1,800,000 - 2,400,000