Сегодня покупают
← Вернуться назад
Нужна помощь с конфигурацией?
Пришлите SKU, спецификацию или список требований. Специалист Apltech проверит совместимость компонентов и найдёт актуальные артикулы
Получить консультацию →

HPE Slingshot Monitoring Software: технические характеристики и документация

HPE Slingshot Monitoring Software

Have feedback on QuickSpecs? We're listening

HPE Slingshot Monitoring Software QuickSpecs

 

The HPE Slingshot Monitoring Software (SMS) is designed to deliver real-time insights into the operational health of the HPE Slingshot interconnect fabric, providing a comprehensive suite of tools for monitoring, analysis, and diagnostics.

 

By leveraging standardized descriptive and diagnostic analytics, SMS ensures a consistent and accurate view of the fabric health, enabling users to better understand and address potential issues within their environment. 

 

 

SMS offers robust monitoring capabilities that provide quick and in-depth visibility into critical issues and events over the fabric that might be time consuming for the user to identify. These features make it suitable for systems of any size, from small-scale deployments to large-scale, high-performance computing environments. The software equips users with the tools necessary to rapidly identify and resolve performance bottlenecks, minimizing the time spent on root cause analysis and enhancing operational efficiency.

 

The primary value of SMS lies in its ability to optimize system utilization and availability. By delivering actionable, data-driven insights, it empowers administrators to make informed decisions that improve overall system utilization. Additionally, SMS plays a vital role in reducing machine downtime by proactively identifying issues and facilitating timely intervention through recommended actions and troubleshooting steps. These capabilities collectively help organizations maximize their return on investment in HPE Slingshot interconnect technology while ensuring consistent and reliable system performance. 


What's New

- HPE Slingshot Monitoring Software new version.

- HPE Slingshot Monitoring Software new version with Adapter Kit. 


 

Standard Features for v1.0.1

The HPE Slingshot Monitoring Software(SMS) 1.0.1 release includes various UI/UX improvements and several enhancements to deploy SMS in different High Performance Computing (HPC) cluster configurations. SMS always relies on HPE Performance Cluster Manager (HPCM) to collect Slingshot telemetry, regardless of whether HPCM is actively used for cluster management.

 

Starting with SMS 1.0.1, customers can deploy SMS in the following environments: 

- On a node other than the HPCM admin node within an HPCM-managed cluster 

- In a cluster managed by CSM 

- In a cluster not managed by HPCM nor CSM, including those using a third-party HPC cluster manager 

 

Features in SMS v1.0.1 are: 

- Support for SMS deployments in clusters managed by CSM 1.5 and 1.6 (including all dot releases)

- Support for SMS deployments on a node other than admin node in a HPCM managed cluster

- Support for SMS deployments on a cluster neither managed by HPCM or CSM

- Support for HPCM 1.12+, 1.13 and 1.14

- Support for FM 2.2+ (same as last release of SMS)

- Support for Light and dark theme in Grafana

- Added "Model" panel in Switch Detailed View dashboard that displays the model's name for the selected switch

- Added informative tooltips to majority of panels in application to explain the meaning of each panel and its data

- Adjusted transformations in the Alerts Overview table, so that alerts for switches with a port xname now link to the Switch Detailed View for that switch

- Switches in Switches Overview tabular display are now ordered by health status, such that unhealthy switches are at the top of the table

- The use of red to indicate critical states has been updated across panels to ensure compliance with contrast accessibility standards

- Overhauled theming of SMS configuration page and AIOps Clickthrough panel to use HPE theme

 

UI/UX Changes

- The SMS software is now in the form of a tar file. The tar file contains 2 rpm's. There is a newly introduced rpm called adapter kit. The SMS adapter kit is only be needed if SMS is not being installed on the HPCM admin node 


Standard Features for v1.0

Fabric Overview

Offers a comprehensive overview of the HPE Slingshot interconnect within HPC systems, delivering detailed insights into its operational health

Fabric Overview displays key highlights regarding:

- Fabric Manager (FM) Status(es)

- Switch health

- Link health

- Alerts

 

Fabric Manager Detailed View

- Displays the cluster role of the FM node

    • Displays the CPU utilization, memory utilization, and disk utilization of the selected FM.
    • Displays a table with timestamp, location, alert with severity, cause and recommended action.

 

Switch Overview

- Every switch organized by dragonfly topology group with its health status such as healthy or unhealthy

- Displays a table with switch, status, ASIC temperature, maximum event severity, firmware version, uptime, group ID, and switch ID.

 

Switch Detailed View

- Selected switch detailed view with a time-series visualization of errors or alerts and real-time hardware metrics

- Displays tables of Fabric Health alerts, HSN Link events, and Redfish alerts of the selected switch, including relevant data such as location of the event (for example, port xname), timestamp, cause, and recommended actions

 

Link Overview

- Visualizes all unhealthy links in the system by type and by severity

- Displays a table with Link Health details including both ports connected by the link, the link type (local, edge, or global), status (indicating the severity, which can be warning, critical, or None), the most recent link-related health event, and its corresponding recommended action

 

Link Detailed View

- Displays the error along with its description and recommended action of the selected link

- Provides troubleshooting procedures with links to the HPE Slingshot Troubleshooting Guide

- Displays a time-series visualization of all LinkDown events grouped by cause, HSN link events, location of the event (that is port xname), timestamp, severity, cause, and recommended actions

 

Alerts Overview

- Categorized alerts by severity and by component type

- Displays a table listing all alerts, including their type, severity, component type, location, timestamp, and message

 

The key features of HPE Slingshot Monitoring Software are:

- Designed for simplicity and clarity, the dashboards offer an intuitive interface for monitoring fabric health. This reduces the time needed to pinpoint the root cause of issues and minimizes the reliance on manual troubleshooting efforts.

- SMS gathers telemetry data from multiple interconnect device endpoints, delivering comprehensive, real-time visibility into the fabric's status and operation metrics.

- SMS displays detailed device status, errors, events, and alerts categorized by severity, ensuring that administrators can promptly address critical issues, thereby reducing downtime.

- SMS provides detailed event logs, allowing administrators to trace when and where issues occurred. This facilitates faster root-cause analysis and resolution across both HPE Slingshot hardware and software.

- The software suggests actionable steps and links directly to relevant troubleshooting and operations guides, such as the HPE Slingshot Troubleshooting Guide and the HPE Slingshot Operations Guide, streamlining problem resolution.

 

Hardware Requirements

HPE Performance Cluster Manager software is supported on the following Gen9, Gen10, Gen10+ and Gen11 platforms:

- SGI 8600

- HPE Apollo 2000, 4000, 6000, 6500 and 9000 systems

- HPE Apollo 20 (including CLX-AP) and 40 systems

- HPE ProLiant DL 325 / 345 / 360 / 380 / 385 / 580 servers

- HPE ProLiant Compute DL 384

- HPE ProLiand Compute XD 230

- HPE Apollo 70 system

- HPE Apollo 80 system

- HPE Apollo 35 server

- HPE Cray XD2000, XD6500 systems

    • XD220v, XD224, XD225v, XD295v
    • XD665, XD670

- HPE Cray EX Supercomputers

    • HPE Cray EX235a, EX235n, EX254n, EX255a, EX420, EX425, EX4252
    • HPE Cray EX2500 (chassis, compute blades, switch chassis, and CDU)
    • HPE Cray EX3000 (chassis, compute blades, switch chassis, and CDU)
    • HPE Cray EX4000 (chassis, compute blades, switch chassis, and CDU)

- Superdome Flex Family

 

Please work with the HPE Slingshot Product Management team to develop system specific hardware requirements.

 

Supported Operating Systems

HPE Performance Cluster Manager supports the SUSE Linux Enterprise Server (SLES), Red Hat Enterprise Linux (RHEL), HPE Cray Operating System (COS), Rocky Linux, and Tri-Lab Operating System Stack (TOSS), and CentOS Linux releases noted below. HPCM can manage clusters in which all nodes run the same operating system release, a multi-distro cluster in which compute nodes run a different operating system release than the system management nodes (i.e., admin and leader), or a multi-distro cluster in which compute nodes run a variety of different operating system releases. Review the following details to see which specific operating system releases are tested and supported on the various node types and architectures::

- x86_64

    • admin and leader: RHEL/Rocky 8.10, SLES15 SP6, RHEL/Rocky 9.5*
    • compute/service : RHEL/Rocky 8.9, RHEL/Rocky 8.10,

o RHEL/Rocky 9.4, RHEL/Rocky 9.5

o SLES15 SP5, SLES15 SP6,

o SLES15 SP5-based COS, SLES15 SP6-based COS

o TOSS 4.7, TOSS 4.8

o Ubuntu 22.04.5

o Ubuntu 24.04.1

- Aarch64

    • compute/service : RHEL/Rocky 8.10, RHEL/Rocky 9.5

o SLES15 SP5, SLES15 SP6,

o SLES15 SP5-based COS, SLES15 SP6-based COS

o TOSS 4.7, TOSS 4.8

For more information refer to the latest HPCM version release notes.


For the most up-to-date information on HPE Services, please refer to the HPE Services - Supplemental QuickSpecs, which provides a comprehensive and regularly updated overview of available services.

HPE Slingshot Monitoring Software Products

HPE Slingshot Monitoring Software is not a licensed product. However, SMS has a dependency on HPE Performance Cluster Manager (HPCM). A single node of HPCM for monitoring is required to install and configure SMS.

 

Models

Licensing and Media Options

HPE Performance Cluster Manager 1 Node 3yr 24x7 Support Perpetual E-LTU

Q9V60AAE

Notes:

- One license per node.

- Includes three years of support.

- This is an electronic license.

- This is a perpetual license. The software will continue working even when the support term ends.

HPE Performance Cluster Manager 1 Node 3yr 24x7 Support Perpetual LTU

Q9V60A

Notes:

- One license per node.

- Includes three years of support.

- This is a perpetual license. The software will continue working even when the support term ends.


Date

Version History

Action

Description of Change

20-Jul-2026

Version 4

Changed

- Updated GreenLake references to align with current branding and terminology standards.

- Updated Supplemental Services QuickSpecs content for consistency and accuracy.

02-Mar-2026

Version 3

Changed

Rebranding update applied to QuickSpecs

25-Aug-2025

Version 2

Changed

Overview and Standard Features sections were updated.

Added

Specs for v1.0.1 and v1.0

16-Dec-2024

Version 1

New

New QuickSpecs

 

 

Have feedback on QuickSpecs? We're listening

 

 

 

 

 

Chat now

 

 

© Copyright 2026 Hewlett Packard Enterprise Development LP. The information contained herein is subject to change without notice. The only warranties for Hewlett Packard Enterprise products and services are set forth in the express warranty statements accompanying such products and services. Nothing herein should be construed as constituting an additional warranty. Hewlett Packard Enterprise shall not be liable for technical or editorial errors or omissions contained herein.

 

For hard drives, 1 GB = 1 billion bytes. Actual formatted capacity is less.

 

a50004262enw - 16866 - WorldWide - V4 - 20-July-2026

 

 

HEWLETT PACKARD ENTERPRISE

HPE.com