Network Consultant I

Gruve

Maharashtra

On-site

INR 1,200,000 - 1,800,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Gruve, an innovative software services startup, is seeking a shift engineer to own monitoring and first‑fix for the AI Fabrik network estate, including EVPN‑VXLAN data‑center fabric and NGFW platforms. You will support PulseAI infrastructure layers with first‑level diagnostics.

You will work with a dynamic team to maintain telemetry, perform basic diagnostics, and open vendor cases within SLA, ensuring system health and performance.

Qualifications

  • 2–4 years NOC/network operations experience.
  • Hands-on experience on enterprise/data-center switching, routing and next-generation firewall platforms.
  • Understanding of BGP and EVPN‑VXLAN fabric concepts.
  • Awareness of Kubernetes networking constructs (Services, Ingress, CNI/Cilium) for GKE connectivity alerts.
  • Awareness of Kubernetes/OpenShift cluster operations and Linux/GPU health basics.
  • Familiarity with SNMP, syslog and streaming telemetry; change-management discipline.
  • Disciplined ITSM practice; ability to open and follow a vendor support case.

Responsibilities

  • Monitor and first‑fix fabric, edge and firewall alerts; run structured troubleshooting before escalation.
  • Execute standard changes under change control: port turn‑ups, ACL updates, code upgrades in maintenance windows.
  • Own device configuration hygiene: backup verification, drift checks, golden‑config compliance reporting.
  • Monitor PulseAI customer infrastructure via the collector: GPU‑server and node health, OpenShift status, network reachability, storage health.
  • Perform first‑level diagnostics to separate network‑side, hardware‑side, storage‑side and platform‑side faults; open vendor cases within SLA.
  • Verify telemetry reachability per device and monitoring coverage; raise gaps as tickets.
  • Maintain asset, inventory and topology records for the network estate and PulseAI environment (firmware, BIOS, GPU drivers).
  • Track link and capacity utilisation; flag threshold breaches into capacity review.

Skills

NOC operations
BGP & EVPN-VXLAN
Kubernetes networking
OpenShift operations
ITSM practices

Tools

Kubernetes
OpenShift
SNMP
Syslog

Job description

About Gruve

Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in their business strategies utilizing their data to make more intelligent decisions. As a well‑funded early‑stage startup, Gruve offers a dynamic environment with strong customer and partner networks.

About Gruve

Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in their business strategies utilizing their data to make more intelligent decisions. As a well‑funded early‑stage startup, Gruve offers a dynamic environment with strong customer and partner networks.

Position Summary

Shift engineer owning monitoring and first‑fix for the AI Fabrik network estate - EVPN‑VXLAN data‑center fabric, edge routers/firewalls, next‑generation firewall HA pairs and out‑of‑band console management - and L1 for the PulseAI infrastructure layers: monitoring GPU servers, control‑plane/infrastructure nodes, the front‑end network and optional RoCEv2 back‑end fabric, supported customer switches and storage, with first‑level diagnostics and vendor case creation.

Key Responsibilities
  • Monitor and first‑fix fabric, edge and firewall alerts; run structured troubleshooting before escalation.
  • Execute standard changes under change control: port turn‑ups, ACL updates, code upgrades in maintenance windows.
  • Own device configuration hygiene: backup verification, drift checks, golden‑config compliance reporting.
  • Monitor PulseAI customer infrastructure via the collector: GPU‑server and node health (availability, GPU utilisation/thermal, NIC and link errors, out‑of‑band management reachability), OpenShift node and cluster‑network status, front‑end network reachability of every cluster node, back‑end RoCEv2 fabric health, switch telemetry (SNMP/syslog/streaming) and storage capacity/health; acknowledge within the tier SLA.
  • Perform first‑level diagnostics to separate network‑side, hardware‑side, storage‑side and platform‑side faults; confirm hardware faults and open the vendor case within the tier window (60/30 minutes), record the case reference and track to closure.
  • Verify telemetry reachability per device and monitoring coverage of newly onboarded equipment (Covered Environment); raise gaps as tickets.
  • Maintain accurate asset, inventory and topology records for the network estate and the PulseAI covered environment (firmware, BIOS, GPU driver levels).
  • Track link and capacity utilisation; flag threshold breaches into capacity review.
Mandatory Qualifications
  • 2-4 years NOC/network operations experience.
  • Hands‑on operational experience on enterprise/data‑center switching, routing and next‑generation firewall platforms (any major vendor).
  • Working understanding of BGP and EVPN‑VXLAN fabric concepts.
  • Working awareness of Kubernetes networking constructs (Services, Ingress, CNI/Cilium) to monitor fabric‑to‑GKE connectivity alerts and correctly route cluster‑side vs network‑side issues.
  • Working awareness of Kubernetes/OpenShift cluster operations - node health, pod scheduling, oc/kubectl read‑only commands - and of Linux server and GPU‑node health basics, to monitor the PulseAI platform and its infrastructure.
  • Familiarity with SNMP, syslog and streaming‑telemetry based monitoring and with server out‑of‑band management; change‑management discipline.
  • Disciplined ITSM practice; ability to open and follow a vendor support case.
Preferred Qualifications
  • Associate/professional‑level networking certification from a major vendor.
  • Intent‑based networking / fabric automation exposure.
  • GKE VPC‑native networking exposure; basic NetworkPolicy familiarity.
  • Red Hat OpenShift exposure (DO180‑level or equivalent); Linux (RHEL) administration fundamentals.
  • Exposure to RoCEv2 / lossless‑Ethernet GPU fabrics (PFC/ECN), NVLink/NVSwitch topologies, and storage health monitoring (NFS / CSI arrays).
  • Next‑generation firewall operations exposure; DCIM familiarity.
Why Gruve

At Gruve, we foster a culture of innovation, collaboration, and continuous learning. We are committed to building a diverse and inclusive workplace where everyone can thrive and contribute their best work. If you’re passionate about technology and eager to make an impact, we’d love to hear from you.

Gruve is an equal opportunity employer. We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Consultant II
Network Consultant II

Gruve • Maharashtra

On-site
INR 1,800,000 - 3,000,000
Network Consultant I
Network Consultant I

Gruve • Pune District

On-site
INR 600,000 - 1,000,000
Network Consultant II
Network Consultant II

gruve • Pune District

On-site
INR 600,000 - 1,200,000
Network Consultant - L2
Network Consultant - L2

Gruve • Pune District

On-site
INR 800,000 - 1,200,000
Security Operations Consultant
Security Operations Consultant

Gruve • Maharashtra

On-site
INR 1,200,000 - 2,000,000
Network Consultant - L2
Network Consultant - L2

Gruve • Pune District

On-site
INR 1,200,000 - 1,800,000
Security Analyst I
Security Analyst I

Gruve • Maharashtra

On-site
INR 900,000 - 1,300,000
Security Operations Manager
Security Operations Manager

Gruve • Maharashtra

On-site
INR 3,500,000 - 7,000,000
Technical Support Engineer
Technical Support Engineer

Gruve • India

On-site
INR 400,000 - 800,000
Solution Architect – (AI Infrastructure & Hybrid Cloud)
Solution Architect – (AI Infrastructure & Hybrid Cloud)

Gruve • Pune District

On-site
INR 2,800,000 - 5,200,000