Hardware Health Sensor (SNMP)
Monitors physical hardware sensors via SNMP, including temperature, fan speeds, power supply status, and voltage levels. Enables proactive detection of hardware issues and supports long-term tracking of system health trends.
Overview
The Hardware Health Sensor (SNMP) reads data from physical monitoring components exposed through the ENTITY-SENSOR-MIB. It provides real-time visibility into hardware performance and reliability, allowing administrators to identify issues such as overheating, fan failure, or unstable power conditions before they lead to system downtime.
This sensor is particularly useful for monitoring servers, switches, and other network appliances equipped with onboard environmental sensors.
Requirements
- The monitored device must implement the ENTITY-SENSOR-MIB.
- SNMP must be enabled on the device.
- Valid SNMP credentials (community string or SNMPv3 credentials) are required.
Options
- SNMP Credentials — specify the correct SNMP profile (v1, v2c, or v3) for the device.
- Polling Interval — defines how often sensor data is collected (defaults to the device profile).
Metrics Collected
The sensor automatically retrieves all available hardware readings exposed through SNMP, including:
- Temperature sensors (e.g., CPU, chassis, power supply)
- Fan speeds (RPM)
- Voltage levels
- Power supply status and health
- Other environmental readings defined by the ENTITY-SENSOR-MIB
Collected values are automatically classified and displayed in the hardware health summary, enabling quick identification of abnormal conditions.
Alerts
By default, the sensor raises alerts when:
- Any monitored hardware component reports a failure or critical status.
- Measured values exceed thresholds defined in the device’s MIB or in custom alerting rules.
Administrators can customize alert conditions, severity, and response actions based on operational needs.
Reports and Trend Data
All collected values are stored in the trend database, allowing long-term visualization and reporting through the Trend Viewer.
Reports include:
- Temperature trends per component
- Fan speed variations
- Power and voltage stability reports
- Sensor availability and failure statistics
These insights help detect gradual degradation and predict potential hardware failures.
Use Cases
- Monitoring server room hardware for early signs of overheating.
- Tracking fan degradation across blade chassis or network switches.
- Ensuring consistent power delivery and voltage stability in routers and UPS devices.
- Establishing environmental baselines for predictive maintenance.
Related
- Cisco Hardware Health (SNMP) — monitoring pack for Cisco devices
- Generic UPS (SNMP) — monitoring pack for uninterruptible power supply units