Skip to content

Monitoring & observability

Streaming telemetry

gNMI and model-driven telemetry for high-frequency metrics — complementing SNMP, not replacing it.

Architecture

Architecture

  1. 01CollectPolling + streaming + flow
    • SNMP
    • gNMI
    • NetFlow/IPFIX
  2. 02Process
    • Telegraf
    • Akvorado
    • Vector
  3. 03StoreTime series and flow
    • Prometheus
    • InfluxDB
    • ClickHouse
  4. 04Visualise
    • Grafana
    • Zabbix
  5. 05AlertRouting and escalation
    • Alertmanager
    • Telegram
    • PagerDuty
SNMP, telemetry and flow together — each at the layer where it works

Capabilities

Capabilities

Every item is marked: verified production experience, or engineering capability.

Site power and batteries

Proven

Battery voltage, the mains feed and temperature on the same map as the traffic. A sagging voltage is visible before the outage — and it explains the share of night-time incidents that otherwise stays unexplained.

  • SNMP
  • Battery voltage
  • Temperature

SNMP polling at scale

Proven

Polling stays the baseline mechanism across most of the fleet — counters, availability and interface state.

  • SNMP
  • Zabbix
  • LibreNMS

gNMI streaming telemetry

Capability

Streaming telemetry replaces polling only where the sampling rate demands it — elsewhere the two complement each other.

  • gNMI
  • gnmic
  • Telegraf

Flow analytics

Capability

NetFlow/IPFIX/sFlow on ClickHouse for peering and traffic analysis.

  • Akvorado
  • ClickHouse
  • IPFIX

Grafana dashboards

Proven
  • Grafana
  • Prometheus
  • InfluxDB

Zabbix infrastructure monitoring

Proven
  • Zabbix
  • SNMP traps
  • Agents

Device inventory and port monitoring

Proven

The whole device fleet and its ports in one system — with autodiscovery, an availability map and ranked error counters, down to the GPON/EPON access layer.

  • LibreNMS
  • SNMP
  • GPON/EPON

SLO and error budgets

Capability

Capacity forecasting

Capability

Growth trends and saturation forecasting for links and storage.

Technology stack

Technology stack

Metrics
PrometheusInfluxDBZabbixLibreNMSVictoriaMetrics
Telemetry
gNMIIOS XR telemetryTelegraf
Flow
AkvoradoClickHouseNetFlow v9IPFIXsFlow
Visualisation
GrafanaZabbix UI
Alerting
AlertmanagerGrafana AlertsTelegram

Engagement model

Engagement model

Project

A one-off scope: audit, migration or implementation with a fixed outcome.

Retainer

Monthly engineering hours — specialist access on demand.

Co-managed

NetWizard and your in-house team together, with split responsibility.

FAQ

FAQ

Zabbix or Prometheus?

Both, for different jobs. Zabbix is strong for appliance and agent monitoring; Prometheus for dynamic, containerised environments. They frequently coexist under one Grafana.

How long is data retained?

Retention is designed against budget and requirement — high resolution short-term, aggregated data long-term.

Tell us about your infrastructure