Skip to content

Network Consultant I

Gruve

Pune, Maharashtra, India

About Gruve

Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in their business strategies utilizing their data to make more intelligent decisions. As a well-funded early-stage startup, Gruve offers a dynamic environment with strong customer and partner networks.

Position summary:

Shift engineer owning monitoring and first-fix for the AI Fabrik network estate — EVPN-VXLAN data-center fabric, edge routers/firewalls, next-generation firewall HA pairs and out-of-band console management — and L1 for the PulseAI infrastructure layers: monitoring GPU servers, control-plane/infrastructure nodes, the front-end network and optional RoCEv2 back-end fabric, supported customer switches and storage, with first-level diagnostics and vendor case creation.

Key responsibilities:

  • Monitor and first-fix fabric, edge and firewall alerts; run structured troubleshooting before escalation.
  • Execute standard changes under change control: port turn-ups, ACL updates, code upgrades in maintenance windows.
  • Own device configuration hygiene: backup verification, drift checks, golden-config compliance reporting.
  • Monitor PulseAI customer infrastructure via the collector: GPU-server and node health (availability, GPU utilisation/thermal, NIC and link errors, out-of-band management reachability), OpenShift node and cluster-network status, front-end network reachability of every cluster node, back-end RoCEv2 fabric health, switch telemetry (SNMP/syslog/streaming) and storage capacity/health; acknowledge within the tier SLA.
  • Perform first-level diagnostics to separate network-side, hardware-side, storage-side and platform-side faults; confirm hardware faults and open the vendor case within the tier window (60/30 minutes), record the case reference and track to closure.
  • Verify telemetry reachability per device and monitoring coverage of newly onboarded equipment (Covered Environment); raise gaps as tickets.
  • Maintain accurate asset, inventory and topology records for the network estate and the PulseAI covered environment (firmware, BIOS, GPU driver levels).
  • Track link and capacity utilisation; flag threshold breaches into capacity review.

Mandatory Qualifications:

  • 2–4 years NOC/network operations experience.
  • Hands-on operational experience on enterprise/data-center switching, routing and next-generation firewall platforms (any major vendor).
  • Working understanding of BGP and EVPN-VXLAN fabric concepts.
  • Working awareness of Kubernetes networking constructs (Services, Ingress, CNI/Cilium) to monitor fabric-to-GKE connectivity alerts and correctly route cluster-side vs network-side issues.
  • Working awareness of Kubernetes/OpenShift cluster operations — node health, pod scheduling, oc/kubectl read-only commands — and of Linux server and GPU-node health basics, to monitor the PulseAI platform and its infrastructure.
  • Familiarity with SNMP, syslog and streaming-telemetry based monitoring and with server out-of-band management; change-management discipline.
  • Disciplined ITSM practice; ability to open and follow a vendor support case.

Preferred Qualifications:

  • Associate/professional-level networking certification from a major vendor.
  • Intent-based networking / fabric automation exposure.
  • GKE VPC-native networking exposure; basic NetworkPolicy familiarity.
  • Red Hat OpenShift exposure (DO180-level or equivalent); Linux (RHEL) administration fundamentals.
  • Exposure to RoCEv2 / lossless-Ethernet GPU fabrics (PFC/ECN), NVLink/NVSwitch topologies, and storage health monitoring (NFS / CSI arrays).
  • Next-generation firewall operations exposure; DCIM familiarity.

Why Gruve

At Gruve, we foster a culture of innovation, collaboration, and continuous learning. We are committed to building a diverse and inclusive workplace where everyone can thrive and contribute their best work. If you’re passionate about technology and eager to make an impact, we’d love to hear from you.

Gruve is an equal opportunity employer. We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.

Seen 10 days ago · Gruve postings close after a median of 5 days.

Original posting on Gruve's site ↗

Posting text belongs to the employer. Removal requests: contact us.

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

One job at a time

One posting. One CV. $25.

Pick the job you actually want and we write for it.

Get my CV for this job