Monitoring Device
Per-device monitoring drill-down: live CPU, memory, temperature and fan/PSU gauges, traffic and health charts, a per-port table, device thresholds, and alert history.
What the device monitoring page shows
The device monitoring page is a per-device drill-down for a single monitored asset — a switch, router, firewall, PDU, access point, load balancer, or server. It gathers everything about that one device on one screen: live health gauges, time-series charts for CPU, memory, temperature and traffic, a per-port table, and the alert history scoped to that device.
How to reach it
Open Monitoring from the sidebar, switch to the Devices tab, and click any device row in the table. You can also jump here from the Overview tab by clicking a top-talker port or an alert that references a device. The page opens with the device name shown as a large heading, and a Back to Monitoring link sits above it to return to the list.
The header
The top of the page identifies the device and offers quick actions. Under the device name you will see a compact summary line containing:
- Management IP — the address used to reach the device.
- Device type — switch, router, firewall, PDU, and so on.
- Model — shown when the device reports it.
- Uptime — how long the device has been running (for example, “Up 12d 4h”), shown when available.
- Last polled — the timestamp of the most recent successful poll, or “Never” if the device has not been polled yet.
Header action buttons
Three buttons sit on the right side of the header:
- Poll Now
- Triggers an immediate poll of the device instead of waiting for the next scheduled interval. The button shows a spinning icon while the poll runs, then refreshes the gauges, charts, traffic, and port table. Useful right after you change something on the device or want the freshest numbers.
- Thresholds
- Opens the threshold settings window for this specific device, where you set the warning and critical limits that drive alerting and the colour of the gauges and bars. See the table below.
- View Device
- Opens the full inventory record for the device in a window, so you can review or edit its details without leaving this page.
Health gauges
Below the header sits a row of four cards giving an at-a-glance snapshot of current health:
- CPU — current processor utilisation as a percentage.
- Memory — current memory utilisation as a percentage.
- Temperature — current temperature in degrees Celsius.
- Status — the hardware Fan and PSU status, each shown as a coloured label (green for OK, amber for warning, red for critical).
Each of the first three cards has a coloured bar underneath the number. The bar is green while the value is healthy, turns amber once it reaches the warning level, and turns red once it crosses the critical level. A value of N/A means the device did not report that metric on its last poll.
The gauges and the Fan/PSU card only appear once the device has reported health data. A device that was just added, or one where polling has only recently begun, will not show this row until the first successful poll completes. Fan and PSU values appear only when the device reports those hardware sensors; otherwise they read “N/A”.
Choosing a time range
A small toggle above the charts lets you choose the window for every time-series chart on the page. The options are 1 Hour, 6 Hours, 24 Hours, and 7 Days. The default is 24 Hours. Switch to 1 Hour to zoom into a recent spike, or 7 Days to review longer-term trends.
The charts
Aggregate Traffic
This area chart combines the inbound and outbound rate across all of the device’s ports into a single view over the chosen time range. Inbound and Outbound are drawn as separate filled lines, and hovering a point shows the rate in bits per second.
CPU & Memory
A line chart plotting CPU and memory utilisation (both as a percentage, on a 0–100 scale) over the selected range, so you can spot sustained load or correlated spikes.
Temperature
A line chart of temperature in degrees Celsius over the selected range.
Charts only render once enough history exists for the chosen range. If a chart is missing, the device either has not collected data for that period yet or does not report that metric. Pick a shorter range, or wait for more polling intervals to pass, and the chart will appear.
Ports table
When the device exposes ports, a table lists each one. A heading above the table shows how many ports are currently up out of the total. The columns are:
| Column | What it shows |
|---|---|
| Port | The port name, or a numbered label when no name is reported. |
| Status | A coloured dot and label — green for up, red for down. |
| Speed | The link speed in Mbps. (Hidden on narrow screens.) |
| In | Current inbound rate. |
| Out | Current outbound rate. |
| Utilization | A bar plus percentage showing how full the busier direction is. The bar is green when light, amber as it climbs, and red when heavily loaded. (Hidden on smaller screens.) |
Alerts
The Alerts panel lists the active and recent alerts that belong to this device. A heading shows how many are currently active. Each row carries:
- A severity badge — critical, warning, or info.
- The alert type, with the relevant port name beside it when the alert is port-specific.
- A short message describing the condition.
- The timestamp of when the alert was raised.
For an alert that is still active and has not yet been acknowledged, an Acknowledge button appears on the right. Clicking it marks the alert as seen; it stays in the list but no longer shows as needing attention. Alerts that have cleared show a green Resolved badge instead. When the device has no alerts, the panel shows a green tick and an “all clear” message.
Setting thresholds for this device
Click Thresholds in the header to open the Device Thresholds window. Each row has an enable checkbox, a W (warning) value, and a C (critical) value. The settings are grouped into two sections:
| Section | Threshold | Unit |
|---|---|---|
| Traffic | Inbound Utilization | % |
| Outbound Utilization | % | |
| Error Rate | per minute | |
| Discard Rate | per minute | |
| Health | CPU Usage | % |
| Memory Usage | % | |
| Temperature | °C |
To configure a limit, tick its checkbox, enter the warning and critical numbers, and click Save. Untick a row to stop alerting on that metric for this device — its value fields grey out. Click Cancel or the close icon to discard changes.
Thresholds set here apply only to this one device and take priority over the platform-wide defaults. A device with no overrides falls back to the general thresholds configured on the main Monitoring screen, so set device thresholds only where this asset needs to behave differently from the rest.
Common tasks
- Get fresh numbers: click Poll Now and wait for the gauges and charts to refresh.
- Investigate a spike: set the time range to 1 Hour and read the Aggregate Traffic and CPU & Memory charts together.
- Quiet a known issue: click Acknowledge on the alert so it no longer flags as needing attention.
- Tune sensitivity: open Thresholds and adjust the warning and critical values for this device.
- Check a single port: read the Ports table to see which links are up, how fast they are, and how loaded each one is.
