Cloud Monitoring
Performance Analysis
Performance Analysis displays the performance metrics of key resources monitored externally or internally in the Cloud. You can view the performance analysis or export the analysis report as needed to improve the O&M efficiency.
View Performance Analysis
On the main menu of ZStack Cube Ultimate, choose . Then, the Performance Analysis page is displayed.

- Filters: Supports basic filtering and advanced filtering.
- Basic filtering: Allows you to filter by resource, monitoring
methods, and time spans.
- Resource: Supports VM instances, VPC vRouters, hosts, backup storages, L3 networks, and virtual IPs.
- Monitoring method: Supports external monitoring and internal
monitoring.
- External monitoring: Obtains the VM performance data, such as the CPU, memory, disk I/O, and NIC data from the host by using libvirt.
- Internal monitoring: Obtains the VM performance data, such as the CPU, memory, and disk capacity directly by using the agent and pushes the data to the host. To using internal monitoring, install the agent first.
Note: The memory data obtained by using internal
monitoring is more accurate than that obtained by
external monitoring. Therefore, we recommend that you
use internal monitoring to monitor the memory
data. - Time span: You can select a time span to view the monitoring data. Available time spans: 15 minutes, 1 hour, 1 week, and custom.
- Advanced filtering: Allows you to further filter by filter item,
resource scope, and owner scope.
- Filter item: Allows you to sort and view resources based on monitoring metrics and metric values (for example, CPU utilization >= 75%).
- Resource range: Allows you to view the monitoring information of all resources on the Cloud or specify a resource to view its own monitoring information.
- Owner range: Allows you to view the monitoring information of all owners in the Cloud or specify an owner to view the monitoring information.
- Basic filtering: Allows you to filter by resource, monitoring
methods, and time spans.
- Analysis report: Generates an analysis report based on the filter
conditions.
- Allows you to sort the items by resource name or monitoring metric.
- Allows you to export all the report information or export the information on the current page in CSV format.
- Allows you to customize the number of items to be displayed on each page. By default, 10 items are displayed per page.
Note:
- The VM analysis report page allows you to stop a VM instance.
- The VM analysis report page allows you to filter VM instances based on the VM state.
- The VM/VPC vRouter analysis report allows you to customize the columns to be displayed.
- When you export a VM or VPC vRouter analysis report, you can choose to export the average, maximum, or minimum values of the metrics as needed.
| Resource Type | Monitoring Method | Monitoring Metric | Destription |
|---|---|---|---|
| VM Instance/VPC vRouter | External Monitoring | Default IPv4 | Displays the default IPv4 address of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
| Volume Actual Size | Displays the volume actual size of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
||
| Total Volume Capacity | Displays the total volume capacity of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
||
| CPU Utilization | Displays the average CPU utilization of all
the VM instances/VPC vRouters in the current zone by
default. Note: If a VM instance or VPC vRouter has more than
one CPU, the CPU utilization might be greater than
100%. |
||
| Memory Utilization | Displays the average memory utilization of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Read Rate | Displays the average disk read speed of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Write Rate | Displays the average disk write speed of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC In Rate | Displays the average NIC in rate of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC Out Rate | Displays the average NIC out rate of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Read IOPS | Displays the average disk read IOPS of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Write IOPS | Displays the average disk write IOPS of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC In Packets | Displays the average number of received NIC packets of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC Out Packets | Displays the average number of sent NIC packets of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC In Errors Rate | Displays the average rate of received NIC errors of all the VM instances/VPC vRouters in the current zone by default. | ||
| NIC Out Errors Rate | Displays the average rate of sent NIC errors of all the VM instances/VPC vRouters in the current zone by default. | ||
| Internal Monitoring | Default IPv4 | Displays the default IPv4 address of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
|
| Volume Actual Size | Displays the volume actual size of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
||
| Volume Actual Size | Displays the volume actual size of each VM
instance in the current zone by default. Note: This metric does
not apply to VPC vRouters currently. |
||
| CPU Utilization | Displays the average CPU utilization of all the VM instances/VPC vRouters in the current zone by default. | ||
| CPU Occupancy Rate (System Process) | Displays the average CPU occupancy rate (system process) of all the VM instances/VPC vRouters in the current zone by default. | ||
| CPU Occupancy Rate (User Process) | Displays the average CPU occupancy rate (user process) of all the VM instances/VPC vRouters in the current zone by default. | ||
| CPU Occupancy Rate (Waiting) | Displays the average CPU occupancy rate (waiting) of all the VM instances/VPC vRouters in the current zone by default. | ||
| CPU Idle Rate | Displays the average CPU idle rate of all the VM instances/VPC vRouters in the current zone by default. | ||
| Memory Utilization | Displays the average memory utilization of all the VM instances/VPC vRouters in the current zone by default. | ||
| Memory Idle Rate | Displays the average memory idle rate of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Utilization | Displays the average disk utilization of all the VM instances/VPC vRouters in the current zone by default. | ||
| Disk Idle Rate | Displays the average disk idle rate of all the VM instances/VPC vRouters in the current zone by default. | ||
| Host | / | Disk Read IOPS | Displays the disk read IOPS of all the hosts in the current zone by default. |
| / | Disk Write IOPS | SDisplays the write read IOPS of all the hosts in the current zone by default. | |
| / | Used Disk Storage Percentage | Displays the used disk storage percentage of all the hosts in the current zone by default. | |
| / | Disk Usage | Displays the disk usage of all the hosts in the current zone by default. | |
| / | NIC In Rate | Displays the NIC in rate of all the hosts in the current zone by default. | |
| / | NIC Out Rate | Displays the NIC out rate of all the hosts in the current zone by default. | |
| / | NIC In Errors Rate | Displays the NIC in erros rate of all the hosts in the current zone by default. | |
| / | NIC Out Errors Rate | Displays the NIC out errors rate of all the hosts in the current zone by default. | |
| / | CPU Utilization Average | Displays the average CPU utilization of all the hosts in the current zone by default. | |
| / | Memory Utilization | Displays the memory utilization of all the hosts in the current zone by default. | |
| / | Disk Read Speed | Displays the disk read speed of all the hosts in the current zone by default. | |
| / | Disk Write Speed | Displays the disk write speed of all the hosts in the current zone by default. | |
| / | NIC In Speed | Displays the NIC in speed of all the hosts in the current zone by default. | |
| / | NIC Out Speed | Displays the NIC out speed of all the hosts in the current zone by default. | |
| Backup Storage | / | Backup Storage Capacity Available Percent | Displays the percentage of available capacity of all the backup storages in the current zone by default. |
| L3 Network | / | Used IPs (IPv4) | Displays the number of used IPv4 IPs of all L3 networks in the current zone by default. |
| / | Used IP Percentage (IPv4) | Displays the percentage of used IPv4 IPs of all L3 networks in the current zone by default. | |
| / | Available IPs (IPv4) | Displays the number of available IPv4 IPs of all L3 networks in the current zone by default. | |
| / | Available IP Percentage (IPv4) | Displays the percentage of available IPv4 IPs of all L3 networks in the current zone by default. | |
| Virtual IP | / | Inbound Traffic | Displays the inbound traffic of all virtual IP addresses in the current zone by default. |
| / | Inbound Traffic Rate | Displays the inbound traffic rate of all virtual IP addresses in the current zone by default. | |
| / | Outbound Traffic | Displays the outbound traffic of all virtual IP addresses in the current zone by default. | |
| / | Outbound Traffic Rate | Displays the outbound traffic rate of all virtual IP addresses in the current zone by default. |
Export Analysis Report
You can export all the analysis information of a resource or only information on the current page based on the filter conditions. For VM instance and VPC vRouters, you can customize the monitoring metrics to be exported and choose to export their average values, maximum values, or minimum values as needed.
On the main menu of ZStack Cube Ultimate, choose . Then, the Performance Analysis page is displayed. Taking VM instance as an example, on the analysis report area, click Export CSV and choose Current Page or All. Then, the export page is displayed.

- The export page displays the selected resource, monitoring method, time span, and all the monitoring metrics of the resource.
- By default, metrics displayed in the VM/VPC vRouter list are automatically selected. You can deselect them or select the average, maximum, or minimum value of other metrics.
- You can select all the average, maximum, and minimum values or empty your selections with one click.
Management Node Monitoring
Management Node (MN) monitoring allows you to view the health status of each management node when you use multiple management nodes to achieve high availability.
Monitoring and Alarm
The Monitoring and Alarm feature monitors time-series data and events and sends alarm messages to specified endpoints by using SNS. Resource alarms, event alarms, and extended alarms are supported. The supported endpoints include the system, emails, DingTalk, HTTP applications, text messages, and Microsoft Teams. For some resource alarms, you need to install the agent before they can work as expected.
Concepts
- Monitoring System:A monitoring system provides the following features:
- Monitor the following two types of time-series data:
- Resource utilizations such as CPU utilization of VM instances and memory utilization of hosts
- Resource capacities such as available number of IP addresses and total number of running VM instances
- Event collection: collects events predefined on the Cloud, such as host disconnection and VM HA enabling.
- Alarm: triggers alarms on time-series data or events.
- Audit: records all operations and allows queries.
- Customization: allows you to customize alarms and message templates
and use predefined alarm templates and resource groups.
- The following three types of alarms are supported:
- Resource alarm: triggers alarms on time-series data. For example, you can configure an alarm for VM instances. If the CPU utilization of a VM instance exceeds 80% by five consecutive minutes, send an alarm message to an email address.
- Event alarm: triggers alarms on events, also called event subscription. For example, you can configure an alarm for host disconnection. If a host is disconnected, an alarm message is sent to DingTalk.
- Extended alarm: receives alarm messages from message sources. For example, if a Ceph Enterprise storage pool is downgraded, an alarm message is sent to the system of the Cloud.
-
A message template specifies the text template of a
resource alarm message or event alarm message sent to an SNS system.
- A message template and message recovery template are provided by the system. If you do not create a template, the system uses the predefined templates.
- You can create multiple message templates and can set only one template as the default template. Messages are formatted by using the default template.
- You can use
${}in a template to quote variables configured in an alarm or event. - You can configure email, DingTalk, text message, and Microsoft Teams as an endpoint in a message template. Messages sent by using email, DingTalk, text message, or Microsoft Teams are sent in the specified format.
- A message source is used to take over extended alarm messages. If you configure alarms for message sources, extended alarm messages can be sent to various endpoints. This enables centralized management of alarm messages and improves O&M efficiencies. You can configure a message source to take over alarm messages of Ceph Enterprise.
- An alarm template is a template of alarm rules. If you associate an alarm template with a resource group, an alarm is created to monitor the resources in the group.
- A resource group consists of resources grouped based on your business needs. If you associate an alarm template with a resource group, the alarm rules specified by the template take effect on all the resources in the group.
- The following three types of alarms are supported:
- Monitor the following two types of time-series data:
- SNS:
SNS sends alarm messages to the specified endpoints. The supported types of endpoints include the system, emails, DingTalk, HTTP applications, text messages, and Microsoft Teams.
Endpoints:- The system provides the system-type endpoint. If you associate an alarm with this endpoint, alarm messages will be displayed below the Recent Message button in the top right corner of the UI.
- You can also create an endpoint of the email, DingTalk, HTTP application, text message, or Microsoft Teams type.
Characteristics
- Provides rich metric items to comprehensively monitor and alarm the core resources as well as events of the Cloud platform.
- Supports types of endpoints including system, emails, DingTalk, HTTP application, text messages, and Microsoft Teams for subscription topics. You can choose an appropriate endpoint to receive alarm messages according to the actual situation.
- One alarm can monitor multiple resources at the same time.
- Emails, DingTalk, text messages, and Microsoft Teams support customized alarm message templates. You can set alarm message templates on demand and quickly locate key information from alarm messages.
- Supports for creating a template of alarm rules. If you associate an alarm template with a resource group, an alarm is created to monitor the resources in the group.
Scenarios
The function of monitoring and alarm monitors the core resources and events of the Cloud platform and sets up an alarm receiving mechanism. When core resources are abnormal, the monitoring and alarm will make real-time responses according to the alarm level to help O&M personnel quickly locate and solve the problem.
Global Setting
- Monitoring data is retained locally for 6 months by default, and you can
customize the monitoring data retention period in the basic settings as
follows:
On the main menu of ZStack Cube Ultimate, choose . Then, the Basic tab is displayed. You can set Monitoring Data Retention Period. Enter an integer between 1 and 12. Default: 6. Unit: month.
- Monitoring data is retained locally in a size of 50GB by default, and you can
customize the monitoring data retention size in the basic settings as
follows:
On the main menu of ZStack Cube Ultimate, choose . Then, the Basic tab is displayed. You can set Monitoring Data Retention Size based on your needs. Default: 50 GB.
- ZStack Cube Ultimate supports receiving extended alarm messages. On the main menu, choose . Then, the Advanced tab is displayed. You need to turn on the Extended Alarm Notification switch to use the extended alarm function.
Alarm
Create an Alarm
On the main menu of ZStack Cube Ultimate, choose . Then, the Alarm page is displayed.
- Create a resource alarm
- Create an event alarm
- Create an extended alarm
Create Resource Alarm
The system provides default resource alarms. In addition, you can customize resource alarms based on your needs. On the Resource Alarm tab, click Create Resource Alarm. Then, the Create Resource Alarm page is displayed.
- Name: Enter a name for the resource alarm.
- Description: Optional. Enter a description for the resource alarm.
- Resource Type: Select a resource
type. Valid values: VM Instance, Baremetal
Instance, Elastic Baremetal Instance, VPC
vRouter, Image, Backup Storage, System Data Directory, Host, L3 Network,
Volume, VIP, Primary Storage, Listener, Management Node, Project Resource,
and CDP Task.
Note that you can select Project Resource and CDP Task only after you have purchased Tenant Management License and Continuous Data Protection (CDP) License, respectively.
- Metric Item: Select a metric item
of the selected resource type.
Note:
- Multiple metric items are available for every type of resources. You can select a metric item based on your business needs.
- Some metric items are associated with additional parameter settings. If you select such metric items, you need to configure the additional parameter settings.
- Some metrics items are available only after the agent is installed on the related resource. .
- If you need to monitor memory data, we recommend that you use internal monitoring for this purpose. This is because internal monitoring yields a more accurate memory data than external monitoring.
- You can create resource alarms for key cloud resources such as VM instances, hosts, and primary storage on the details page of the resources.
- Alarm Coverage: Select one or
more resources on which the alarm takes effect.
- If you select multiple resources of the specified type, all the resources for which the alarm is configured are monitored. If the trigger condition of a resource is met, the alarm is triggered.
- If you select only one resource of the specified type, the resource
for which the alarm is configured is monitored. If the trigger
condition of the resource is met, the alarm is triggered.
Note:
- You can configure fine-grained alarms for a resource.
- For example, you can configure an alarm to monitor the utilization of a CPU of a VM instance.
- Alarm Trigger Rule: Select a comparison symbol and specify a threshold and duration.
- Alarm Interval: Select an alarm
interval.
- Only 1 time:
- Alarm is triggered only once for a resource.For example,
- Assume that you configure an alarm for multiple
resources. If the trigger condition of a resource
is met, the alarm is triggered. After that, even
if the trigger condition of this resource is met
again, the alarm is no longer triggered.
Assume that you configure an alarm for a resource. If the trigger condition of the resource is met, the alarm is triggered. After that, even if the trigger condition of the resource is met again, the alarm is no longer triggered.
- Assume that you configure an alarm for multiple
resources. If the trigger condition of a resource
is met, the alarm is triggered. After that, even
if the trigger condition of this resource is met
again, the alarm is no longer triggered.
- The alarm message is sent to the endpoint (if specified) for only once. In addition, the alarm message is displayed on Message Center only once.
- If the resource recovers but later meets the trigger condition again, the one-time alarm is triggered again.
- Alarm is triggered only once for a resource.
- Repetitive Alarm:
- Alarm is triggered multiple times for a resource.For example,
- Assume that you configure an alarm for multiple
resources. If the trigger condition of a resource
is met, the alarm is triggered. If the resource
keeps meeting the trigger condition, the alarm is
repetitively triggered based on the alarm
interval.
Assume that you configure an alarm for a resource. If the trigger condition of the resource is met, the alarm is triggered. If the resource keeps meeting the trigger condition, the alarm is repetitively triggered based on the alarm interval.
- Assume that you configure an alarm for multiple
resources. If the trigger condition of a resource
is met, the alarm is triggered. If the resource
keeps meeting the trigger condition, the alarm is
repetitively triggered based on the alarm
interval.
- Each time the alarm is triggered, the alarm message is sent to the endpoint (if specified). In addition, every alarm message is displayed on Message Center.
- Alarm is triggered multiple times for a resource.
- Only 1 time:
- Emergency Level: Set an emergency level. Valid values: Emergent, Major, and Info. Alarms of different emergency levels correspond to alarm messages of different emergency levels.
- Alarm Recovery Notification: Optional. If enabled, when a resource monitored by a resource alarm recovers from alarmed status, the system receives a notification. The recovery notification is sent according to the default recovery message template. You can customize the message content on the Message Template page.
- Endpoint: Optional. If specified,
alarm messages are sent to the specified endpoint.
Note:
- You can specify multiple endpoints.
- You can specify the system endpoint or customize an endpoint.

Create Event Alarm
The system provides default event alarms. In addition, you can customize event alarms based on your needs. On the Event Alarm tab, click Create Event Alarm. Then, the Create Event Alarm page is displayed.
- Resource Type: Select a resource
type. Valid values: VM Instance, VPC vRouter, Backup Storage, Management
Node, Host, Primary Storage, vCenter, Backup Job, Project Resource, and CDP
Task.
Note that you can select Project Resource and CDP Task only after you have purchased Tenant Management License and Continuous Data Protection (CDP) License, respectively.
- Metric Item: Select a metric item of the selected resource type.
- Emergency Level: Set an emergency level. Valid values: Emergent, Major, and Info. Alarms of different emergency levels correspond to alarm messages of different emergency levels.
- Endpoint: Optional. If specified,
alarm messages are sent to the specified endpoint.
Note:
- You can specify multiple endpoints.
- You can specify the system endpoint or customize an endpoint.

Note:
- An event alarm is triggered only when the configured event occurs. The repetitive alarm mechanism is not available for event alarms.
- When a resource monitored by an event alarm recovers from alarmed status, the system receives a notification. The recovery notification is sent according to the default recovery message template. You can customize the message content on the Message Template page.
- If the configured event occurs again, the event alarm is triggered again.
Create Extended Alarm
Before you can create an extended alarm, you need to enable the extended alarm feature. To do this, choose and turn on the Extended Alarm Notification switch.
- Name: Enter a name for the extended alarm.
- Message Source: Select a source where you need to receive alarm messages.
- Endpoint: Optional. If specified,
alarm messages are sent to the specified endpoint.
Note:
- You can specify multiple endpoints.
- You can specify the system endpoint or customize an endpoint.

One-click Alarm
A one-click alarm integrates multiple metrics of a resource. You can create one-click alarms for multiple resources to monitor these resources.
Alarm Template
Create an Alarm Template
On the main menu of ZStack Cube Ultimate, choose . On the Alarm Template page, click Create Alarm Template. Then, the Create Alarm Template page is displayed.
- Name: Enter a name for the alarm template.
- Description: Optional. Enter a description for the alarm template.
- Resource Type: add an alarm rule to the template.
- Alarm Type: Select resource alarm or event alarm.
- Resource Type: Select a resource type.
- If you create a resource alarm rule, you can select the following resource types: VM Instance, Baremetal Instance, Elastic Baremetal Instance, VPC vRouter, Backup Storage, Host, L3 Network, VIP, Primary Storage, Listener, and License.
- If you create an event alarm rule, you can select the following resource types: VM Instance, VPC vRouter, Backup Storage, Host, and Primary Storage.
- Add Rule: Add an alarm rule for the selected resource.

Resource Group
Create a Resource Group
On the main menu of ZStack Cube Ultimate, choose . On the Resource Group page, click Create Resource Group. Then, the Create Resource Group page appears.
- Name: Enter a name for the resource group.
- Description: Optional. Enter a description for the resource group.
- Resource: Add a resource to the resource group.
- Alarm Template: Optional. Associate a resource group with
an alarm template. Then the alarm template applies to all resources in the
group. You can also associate an alarm template after the resource group is
created.
Note: You can associate a resource group with only one alarm
template. - Endpoint: Optional. If specified,
alarm messages are sent to the specified endpoint.
Note:
- You can specify multiple endpoints.
- You can specify the system endpoint or customize an endpoint.

Message Template
Create a Message Template
On the main menu of ZStack Cube Ultimate, choose . On the Message Template page, click Create Message Template. Then, the Create Message Template page appears.
- Name: Enter a name for the message template.
- Description: Optional. Enter a description for the message template.
- Type: Select the platform type of an endpoint. Valid values: Email, DingTalk, Microsoft Teams, and SMS.
- Alarm Type: Select an alarm type. Valid values: Resource Alarm and Event Alarm.
- Alarm Message Title: Set the title of alarm messages. You can
customize a title or use the sytstem template. Currently, you cannot set a message title
for SMS messages.
The following is a message title template:
Resource Alarm:Alarm ${ALARM_METRIC} ${ALARM_COMPARISON_OPERATOR} ${ALARM_THRESHOLD} ${ALARM_CURRENT_STATUS}Event Alarm:${EVENT_NAME} alarm occurs. - Alarm Message Text: Customize an alarm message text template or
use the system template.
The following is a message template for endpoints of the email and DingTalk types:
Resource Alarm:Alarm ${ALARM_NAME} State Changes To ${ALARM_CURRENT_STATUS} Alarm Details UUID: ${ALARM_UUID} Resource Namespace: ${ALARM_NAMESPACE} Trigger Condition: ${ALARM_METRIC} ${ALARM_COMPARISON_OPERATOR} ${ALARM_THRESHOLD} Trigger Condition Duration: ${ALARM_DURATION} seconds Previous Status: ${ALARM_PREVIOUS_STATUS} Current Value: ${ALARM_CURRENT_VALUE} Tag: ${ALARM_LABELS.join(",")}Event Alarm:Event Details: Name:${EVENT_NAME} Resource Type: ${EVENT_NAMESPACE} Emergency Level:${EVENT_EMERGENCY_LEVEL} Resource UUID:${EVENT_RESOURCE_ID} Name:${EVENT_RESOURCE_NAME} Alarm Trigger Time:${EVENT_TIME} Event Subscription UUID:${EVENT_SUBSCRIPTION_UUID} Error:${EVENT_ERROR}
Note: Message
templates for DingTalk-type endpoints must follow the Markdown syntax. DingTalk
supports the subset of Markdown syntax. For more information, see DingTalk official
website.The following is a message template for endpoints of the Microsoft Teams type:- Resource
alarm:
{ "activityTitle": "Alarm ${ALARM_NAME} ${TITLE_ALARM_RESOURCE_NAME} State Changes To ${ALARM_CURRENT_STATUS}", "facts": [ { "name": "Alarm Details", "value": null }, { "name": "UUID", "value": "${ALARM_UUID}" }, { "name": "Resource Type", "value": "${ALARM_NAMESPACE}" }, { "name": "Trigger Condition", "value": "${ALARM_METRIC} ${ALARM_COMPARISON_OPERATOR} ${ALARM_THRESHOLD}" }, { "name": "Trigger Condition Duration", "value": "${ALARM_DURATION} seconds" }, { "name": "Previous Status", "value": "${ALARM_PREVIOUS_STATUS}" }, { "name": "Current Value", "value": "${ALARM_CURRENT_VALUE}" }, { "name": "Alarm Resource UUID", "value": "${ALARM_RESOURCE_ID}" }, { "name": "Alarm Trigger Time", "value": "${ALARM_TIME}" }, { "name": "Alarm Resource Name", "value": "${ALARM_RESOURCE_NAME}" }, { "name": "Emergency Level", "value": "${ALARM_EMERGENCY_LEVEL}" }, { "name": "Tag", "value": "${ALARM_LABELS.join(\",\")}" } ] } - Even
alarm:
{ "activityTitle": "Event ${EVENT_NAME} Happened", "facts": [ { "name": "Event Details", "value": null }, { "name": "Name", "value": "${EVENT_NAME}" }, { "name": "Resource Type", "value": "${EVENT_NAMESPACE}" }, { "name": "Emergency Level", "value": "${EVENT_EMERGENCY_LEVEL}" }, { "name": "Alarm Resource UUID", "value": "${EVENT_RESOURCE_ID}" }, { "name": "Alarm Resource Name", "value": "${EVENT_RESOURCE_NAME}" }, { "name": "Alarm Trigger Time", "value": "${EVENT_TIME}" }, { "name": "Event Subscription UUID", "value": "${EVENT_SUBSCRIPTION_UUID}" }, { "name": "Error", "value": "${EVENT_ERROR}" } ] }
Note: Message template for endpoints of the Microsoft Teams type must follow the
Webhook syntax of Microsoft Teams. For more information, see Microsoft Teams official
website.The following is a message template for endpoints of the SMS type:- Resource
alarm:
Alarm: ${ALARM_NAME}, Name: ${ALARM_RESOURCE_NAME}, Trigger Condition: ${ALARM_CONDITION}, Emergency Level: ${ALARM_EMERGENCY_LEVEL}, Current Value: ${ALARM_CURRENT_VALUE} - Event
alarm:
Event Name: ${EVENT_NAME}, Name: ${EVENT_RESOURCE_NAME}, Emergency Level: ${EVENT_EMERGENCY_LEVEL}, Error: ${EVENT_ERROR}
Note: Before you can customize a message template for SMS endpoint, you need to
apply for and obtain a third-party SMS signature and SMS template. Currently, only
Alibaba Cloud SMS is supported. If you need to modify a template, you need to modify
on the third-party side. In addition, you need to apply for the signature and template
again. - Resource
alarm:
- Recovery Message Title: When a monitored resource recovers from an alarm status, the Cloud sends alarm recovery messages to the selected endpoint. You can customize the title of the recovery messages as needed. SMS-type endpoints do not support recovery messages.
- Recovery Message Text: Set the text content of the recovery meassage.
- Make Default: Optional. Set the current message template as the default template.


Message Source
Create a Message Source
Before you can create a message source, you need to enable the extended alarm feature. To do this, choose and turn on the Extended Alarm Notification switch.
On the main menu of ZStack Cube Ultimate, choose . On the Message Source page, click Create Message Source. Then, the Create Message Source page appears.
- Name: Enter a name for the message source.
- Description: Optional. Enter a description for the message source.
- Message Source Type: Select the type of message source. Currently, only Ceph Enterprise is supported.
- Login Address and Token: Enter the login address of the
message source and the access token obtained on the message source page, in the
format of
http://{Message Source IP Address}:{Port}/v1/alerts/?token={Access Token} - Alarm Message Conversion Template: Convert third-party
alarm messages into alarm messages that conform to the format required on the
Cloud. The system provides a conversion template that you can use to customize
the parameters.Example:
{ "product":"Ceph Enterprise", "service":"Ceph Enterprise", "message":"${resource_type + '[' + resource_name+'] ' + group + ' ' + alert_value}", "metric":"${resource_type + '::' + group}", "alertLevel":"${level == 'info' ? 'Normal' : level == 'warning' ? 'Important' : 'Emergent'}", "alertTime":"${create}", "dimensions":"{'resource_name':'${resource_name}'}", "dataSource":"Ceph Enterprise" }
Endpoint
Create an Endpoint
On the main menu of ZStack Cube Ultimate, choose . On the Endpoint page, click Create Endpoint. Then, the Create Endpoint page appears.
- Create an endpoint of the email type
- Create an endpoint of the DingTalk type
- Create an endpoint of the HTTP application type
- Create an endpoint of the SMS type
- Create an endpoint of the Microsoft Teams type
Create Endpoint of Email Type
- Messages sent to topics are sent to the specified email address via an email server.
- You can customize a message template or use the system message template to send email messages in a unified format.
- You need to add an email server to the Cloud and test the availability of the server before you can use the email server to send messages.
- Name: Enter a name for the endpoint.
- Description: Optional. Enter a description for the endpoint.
- Type: Select Email.
- Email Address: Enter one or more email addresses. You can enter a maximum of 100 email addresses.
- Email Server: Enter an email server that is added to
the Cloud.
- Test: Test the availability of the email server.
- Message Language: Select a language for alarm messages. Valid values: Simplified Chinese and English.

Create Endpoint of DingTalk Type
- Messages sent to topics are sent to the specified DingTalk robot address via DingTalk. If you specify a contact, the DingTalk user that owns the phone number will be notified of the messages.
- You can customize a message template or use the system message template to send email messages in a unified format.
- You need to create an alarm message template of DingTalk type that follows the Markdown syntax. DingTalk supports the subset of Markdown syntax. For more information, see DingTalk official website.
- Name: Enter a name for the endpoint.
- Description: Optional. Enter a description for the endpoint.
- Type: Select DingTalk.
- Address: Enter a DingTalk robot address.
- Contact: Optional. Specify all members or specific
members in the group.
Note: If you need to specify a member, enter the phone
number of the member, for example, +86-13800000000. - Message Language: Select a language for alarm messages. Valid values: Simplified Chinese and English.

Create Endpoint of HTTP Application Type
- Messages sent to topics are sent to the specified HTTP address by using the HTTP POST method.
- If you set a username and password for the specified HTTP application, enter the username and password for the endpoint.
- Name: Enter a name for the endpoint.
- Description: Optional. Enter a description for the endpoint.
- Type: Select HTTP Application.
- Address: Enter the address of an HTTP application.
Note:
- If the HTTP application is managed by a single management node, the IP address is 127.0.0.1 by default.
- If the HTTP application is managed in a dual-MN environment, the IP address is a VIP.
- User Name: Optional. Enter the username of the HTTP application.
- Password: Optional. Enter the password of the HTTP application.

Create Endpoint of SMS Type
- Messages sent to topics are sent to the specified phone numbers via text messages.
- You need to create a message template and set it as the default template. Then text alarm messages are sent according to the template.
- Name: Enter a name for the endpoint.
- Description: Optional. Enter a description for the endpoint.
- Type: Select SMS.
- AccessKey: Enter a third-party AccessKey pair.
- Phone Number: Enter the phone numbers that receive text messages.

Create Endpoint of Microsoft Teams Type
- Messages sent to topics are sent to the specified Microsoft Teams via Webhook.
- You can create a message template or use the system template for alarm messages to be sent in a unified format.
- Name: Enter a name for the endpoint.
- Description: Optional. Enter a description for the endpoint.
- Type: Select Microsoft Teams.
- Address: Enter the Webhook address obtained in the Microsoft Teams.
- Message Language: Select a language for alarm messages. Valid values: Simplified Chinese and English.

Alarm Item Overview
Resource Alarm Items
Default Alarm Items
| resource type | default alarm | alarm item | description |
|---|---|---|---|
| VM instance | VM instancememoryused percentage | VM instancememoryused percentage≥80% |
|
| VM instanceaverage CPUusage | VM instanceaverage CPUusage≥80% |
|
|
| host | hostroot diskusagealarm | hostroot diskusage≥80% |
|
| CPUtemperaturealarm | CPUtemperature≥80℃ |
|
|
| SSDtemperaturealarm | SSDtemperature≥80℃ |
|
|
| SSDremaining lifealarm | SSDremaining life≤10% |
|
|
| hostaverage CPUusage | hostaverage CPUusage≥80% |
|
|
| hostmemoryused percentage | the hostused memorycapacitypercentage≥80% |
|
|
| image server | image serverstoragecan use capacityalarm | image storagecan use capacitypercentage<20% |
|
| primary storage | primary storagecan use capacityalarm | the primary storagecan use capacitypercentage<20% |
|
| primary storagecan use physical capacityalarm | the primary storagecan use physical capacitypercentage<20% |
|
|
| management node | dual management nodesdatabase is not synchronizedalarm | dual management nodesdatabase is not synchronized |
|
| arbiterIPnot can reachalarm | arbiterIPnot can reach |
|
|
| systemdata directory | systemdata directorydiskcapacityalarm | management nodedata directorydiskoccupy use rate≥70% |
|
| license | licenseexpirationtimealarm | defaultlicenseexpirationtime≤15days |
|
| CDPtask (requiresfor data protectionCDPmodulelicense) | CDPtaskused capacityalarm | CDPtaskused capacityoccupyplannedcapacitypercentage>80% |
|
| CDPtaskRPOoffsetalarm | RPOoffsettime>5minutes |
|
customalarm item
| resource type | subtype | alarm item | description |
|---|---|---|---|
| VM instance | CPU | CPUusage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
| CPUidle rate | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| average CPUusage | batch-monitors multipleVM instance of average CPUusage, any VM instance of average CPUusagemeets the alarm condition triggers an alarm . | ||
| allCPUusage | batch-monitors multipleVM instance of CPUusage, any VM instance of allCPUusagesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| allCPUidle rate | batch-monitors multipleVM instance of CPUidle rate, any VM instance of allCPUidle ratesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| CPUusage(requires installation ofagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| disk | diskreadIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| alldiskreadIOPS | batch-monitors multipleVM instance of diskreadIOPS, any VM instanceall disk of readIOPSsummeets the alarm condition triggers an alarm . | ||
| diskwriteIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwriteIOPS | batch-monitors multipleVM instance of diskwriteIOPS, any VM instanceall disk of writeIOPSsummeets the alarm condition triggers an alarm . | ||
| diskread speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskread speed | batch-monitors multipleVM instance of diskread speed, any VM instanceall disk of read speedsummeets the alarm condition triggers an alarm . | ||
| diskwrite speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwrite speed | batch-monitors multipleVM instance of diskwrite speed, any VM instanceall disk of write speedsummeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacity(requires installation ofagent) | batch-monitors multipleVM instance of diskremainingcapacity, any VM instanceall disk of remainingcapacitysummeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacitypercentage(requires installation ofagent) | batch-monitors multipleVM instance of diskremainingcapacitypercentage, any VM instancealldiskremainingcapacitypercentage (percentage=all diskremainingcapacitysum/all diskcapacitysum)meets the alarm condition triggers an alarm . | ||
| alldiskusedcapacity(requires installation ofagent) | batch-monitors multipleVM instance of diskusedcapacity, any VM instanceall disk of usedcapacitysummeets the alarm condition triggers an alarm . | ||
| alldiskusedcapacitypercentage(requires installation ofagent) | batch-monitors multipleVM instance of diskusedcapacitypercentage, any VM instancealldiskusedcapacitypercentage (percentage=all diskusedcapacitysum/all diskcapacitysum)meets the alarm condition triggers an alarm . | ||
| diskusedcapacity(requires installation ofagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacitypercentage(requires installation ofagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacitypercentage(requires installation ofagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacity(requires installation ofagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NIC | NICinbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| allNICinbound speed | batch-monitors multipleVM instance of NICinbound speed, any VM instanceall NICinbound speedsummeets the alarm condition triggers an alarm . | ||
| NICinbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound packet count | batch-monitors multipleVM instance of NICinbound packet count, any VM instanceall NICinbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICinbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound error count | batch-monitors multipleVM instance of NICinbound error count, any VM instanceall NICinbound error countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound speed | batch-monitors multipleVM instance of NICoutbound speed, any VM instanceall NICoutbound speedsummeets the alarm condition triggers an alarm . | ||
| NICoutbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound packet count | batch-monitors multipleVM instance of NICoutbound packet count, any VM instanceall NICoutbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound error count | batch-monitors multipleVM instance of NICoutbound error count, any VM instanceall NICoutbound error countsummeets the alarm condition triggers an alarm . | ||
| memory | memoryidlecapacity | batch-monitors multipleVM instance of memoryidlecapacity, any VM instance of memoryidlecapacitymeets the alarm condition triggers an alarm . | |
| memoryidlepercentage | batch-monitors multipleVM instance of memoryidlepercentage, any VM instance of memoryidlepercentagemeets the alarm condition triggers an alarm . | ||
| memoryused capacity | batch-monitors multipleVM instance of memoryused capacity, any VM instance of memoryused capacitymeets the alarm condition triggers an alarm . | ||
| memoryused percentage | batch-monitors multipleVM instance of memoryused percentage, any VM instance of memoryused percentagemeets the alarm condition triggers an alarm . | ||
| memoryused percentage(requires installation ofagent) | batch-monitors multipleVM instance of memoryused percentage, any VM instance of memoryused percentagemeets the alarm condition triggers an alarm . | ||
| other | VM instancequantity | monitorcloud platformin all KVMVM instancequantity, meets the alarm condition triggers an alarm . | |
| runningVM instancequantity | monitorcloud platformin all runningstatus of KVMVM instancequantity, meets the alarm condition triggers an alarm . | ||
| runningVM instancepercentage | monitorcloud platformin all runningstatus of KVMVM instancepercentage, meets the alarm condition triggers an alarm . | ||
| stopVM instancequantity | monitorcloud platformin all stopstatus of KVMVM instancequantity, meets the alarm condition triggers an alarm . | ||
| stopVM instancepercentage | monitorcloud platformin all stopstatus of KVMVM instancepercentage, meets the alarm condition triggers an alarm . | ||
| otherstatusVM instancequantity | monitorcloud platformin all otherstatus (not includes: stop/running) of KVMVM instancequantity, meets the alarm condition triggers an alarm . | ||
| otherstatusVM instancepercentage | monitorcloud platformin all otherstatus (not includes: stop/running) of KVMVM instancepercentage, meets the alarm condition triggers an alarm . | ||
| bare metal host (installedbare metalmanagemodulelicense) | CPU | CPUusage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
| disk | diskreadIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| diskwriteIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskread speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskwrite speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NIC | NICinbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| NICinbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICinbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| memory | memorytotal amount | batch-monitors multiplebare metal host of memorytotal amount, any bare metal host of memorytotal amountmeets the alarm condition triggers an alarm . | |
| remainingmemoryamount | batch-monitors multiplebare metal host of remainingmemoryamount, any bare metal host of remainingmemoryamountmeets the alarm condition triggers an alarm . | ||
| used memoryamount | batch-monitors multiplebare metal host of used memoryamount, any bare metal host of used memoryamountmeets the alarm condition triggers an alarm . | ||
| can use memoryamount | batch-monitors multiplebare metal host of can use memoryamount, any bare metal host of can use memoryamountmeets the alarm condition triggers an alarm . | ||
| remainingmemorypercentage | batch-monitors multiplebare metal host of memoryused percentage, any bare metal host of memoryused percentagemeets the alarm condition triggers an alarm . | ||
| memoryusage | batch-monitors multiplebare metal host of memoryusage, any bare metal host of memoryusagemeets the alarm condition triggers an alarm . | ||
| elastic bare metal instance (installedbare metalmanagemodulelicense) | CPU | CPUusage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
| CPUaverage usage | batch-monitors multipleelastic bare metal instance of average CPUusage, any elastic bare metal instance of average CPUusagemeets the alarm condition triggers an alarm . | ||
| disk | diskreadIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| diskwriteIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskread speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskwrite speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NIC | NICinbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| NICinbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICinbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NICoutbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| memory | memorytotal amount | batch-monitors multipleelastic bare metal instance of memorytotal amount, any elastic bare metal instance of memorytotal amountmeets the alarm condition triggers an alarm . | |
| remainingmemoryamount | batch-monitors multipleelastic bare metal instance of remainingmemoryamount, any elastic bare metal instance of remainingmemoryamountmeets the alarm condition triggers an alarm . | ||
| used memoryamount | batch-monitors multipleelastic bare metal instance of used memoryamount, any elastic bare metal instance of used memoryamountmeets the alarm condition triggers an alarm . | ||
| can use memoryamount | batch-monitors multipleelastic bare metal instance of can use memoryamount, any elastic bare metal instance of can use memoryamountmeets the alarm condition triggers an alarm . | ||
| remainingmemorypercentage | batch-monitors multipleelastic bare metal instance of memoryused percentage, any elastic bare metal instance of memoryused percentagemeets the alarm condition triggers an alarm . | ||
| memoryusage | batch-monitors multipleelastic bare metal instance of memoryusage, any elastic bare metal instance of memoryusagemeets the alarm condition triggers an alarm . | ||
| VPCrouter | CPU | CPUusage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
| CPUidle rate | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allCPUusage | batch-monitors multipleVPCrouter CPUusage, any VPCrouter allCPUusagesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| allCPUidle rate | batch-monitors multipleVPCrouter CPUidle rate, any VPCrouter allCPUidle ratesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| disk | diskreadIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| alldiskreadIOPS | batch-monitors multipleVPCrouter diskreadIOPS, any VPCrouterall disk of readIOPSsummeets the alarm condition triggers an alarm . | ||
| diskwriteIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwriteIOPS | batch-monitors multipleVPCrouter diskwriteIOPS, any VPCrouterall disk of writeIOPSsummeets the alarm condition triggers an alarm . | ||
| diskread speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskread speed | batch-monitors multipleVPCrouter diskread speed, any VPCrouterall disk of read speedsummeets the alarm condition triggers an alarm . | ||
| diskwrite speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwrite speed | batch-monitors multipleVPCrouter diskwrite speed, any VPCrouterall disk of write speedsummeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacity(preinstalledagent) | batch-monitors multipleVPCrouter alldiskremainingcapacity, any VPCrouteralldiskremainingcapacitymeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacitypercentage(preinstalledagent) | batch-monitors multipleVPCrouter alldiskremainingcapacitypercentage, any VPCrouteralldiskremainingcapacitypercentagemeets the alarm condition triggers an alarm . | ||
| alldiskusedcapacity(preinstalledagent) | batch-monitors multipleVPCrouter alldiskusedcapacity, any VPCrouteralldiskusedcapacitymeets the alarm condition triggers an alarm . | ||
| alldiskusedcapacitypercentage(preinstalledagent) | batch-monitors multipleVPCrouter alldiskusedcapacitypercentage, any VPCrouteralldiskusedcapacitypercentagemeets the alarm condition triggers an alarm . | ||
| diskremainingcapacity(preinstalledagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacitypercentage(preinstalledagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacity(preinstalledagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacitypercentage(preinstalledagent) | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NIC | NICinbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| allNICinbound speed | batch-monitors multipleVPCrouter NICinbound speed, any VPCrouterall NICinbound speedsummeets the alarm condition triggers an alarm . | ||
| NICinbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound packet count | batch-monitors multipleVPCrouter NICinbound packet count, any VPCrouterall NICinbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICinbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound error count | batch-monitors multipleVPCrouter NICinbound error count, any VPCrouterall NICinbound error countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound speed | batch-monitors multipleVPCrouter NICoutbound speed, any VPCrouterall NICoutbound speedsummeets the alarm condition triggers an alarm . | ||
| NICoutbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound packet count | batch-monitors multipleVPCrouter NICoutbound packet count, any VPCrouterall NICoutbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound error count | batch-monitors multipleVPCrouter NICoutbound error count, any VPCrouterall NICoutbound error countsummeets the alarm condition triggers an alarm . | ||
| memory | memoryidlecapacity | batch-monitors multipleVPCrouter memoryidlecapacity, any VPCrouter memoryidlecapacitymeets the alarm condition triggers an alarm . | |
| memoryidlepercentage | batch-monitors multipleVPCrouter memoryidlepercentage, any VPCrouter memoryidlepercentagemeets the alarm condition triggers an alarm . | ||
| memoryused capacity | batch-monitors multipleVPCrouter memoryused capacity, any VPCrouter memoryused capacitymeets the alarm condition triggers an alarm . | ||
| memoryused percentage | batch-monitors multipleVPCrouter memoryused percentage, any VPCrouter memoryused percentagemeets the alarm condition triggers an alarm . | ||
| image | imagetotal count | monitorcloud platformin imagetotal count, meets the alarm condition triggers an alarm . | |
| can use imagetotal count | monitorcloud platformin can use imagetotal count, meets the alarm condition triggers an alarm . | ||
| can use imagepercentage | monitorcloud platformin can use imagepercentage, meets the alarm condition triggers an alarm . | ||
| root volumeimagequantity | monitorcloud platformin root volumeimagequantity, meets the alarm condition triggers an alarm . | ||
| root volumeimagepercentage | monitorcloud platformin root volumeimagepercentage, meets the alarm condition triggers an alarm . | ||
| data volumeimagequantity | monitorcloud platformin data volumeimagequantity, meets the alarm condition triggers an alarm . | ||
| data volumeimagepercentage | monitorcloud platformin data volumeimagepercentage, meets the alarm condition triggers an alarm . | ||
| ISOimagequantity | monitorcloud platformin ISOimagequantity, meets the alarm condition triggers an alarm . | ||
| ISOimagepercentage | monitorcloud platformin ISOimagepercentage, meets the alarm condition triggers an alarm . | ||
| image server | allimage storagecan use capacity | monitorcloud platformin all image server of can use capacity, all image server of can use capacitysummeets the alarm condition triggers an alarm . | |
| allimage storagecan use capacitypercentage | monitorcloud platformin all image server of can use capacitypercentage (percentage=all image servercan use capacitysum/all image servercapacitysum), meets the alarm condition triggers an alarm . | ||
| image storagecan use capacity | batch-monitors multipleimage server of can use capacity, any image server of can use capacitymeets the alarm condition triggers an alarm . | ||
| image storagecan use capacitypercentage | batch-monitors multipleimage server of can use capacitypercentage, any image server of can use capacitypercentagemeets the alarm condition triggers an alarm . | ||
| allimage storageused capacity | monitorcloud platformin all image server of used capacity, all image server of used capacitysummeets the alarm condition triggers an alarm . | ||
| allimage storageused capacitypercentage | monitorcloud platformin all image server of used capacitypercentage (percentage=all image serverused capacitysum/all image servercapacitysum), meets the alarm condition triggers an alarm . | ||
| image storageused capacity | batch-monitors multipleimage server of used capacity, any image server of used capacitymeets the alarm condition triggers an alarm . | ||
| image storageused capacitypercentage | batch-monitors multipleimage server of used capacitypercentage, any image server of used capacitypercentagemeets the alarm condition triggers an alarm . | ||
| image storagedisabled use capacity | monitorcloud platformin image server of reservedcapacityconfiguration, reservedcapacitymeets the alarm condition triggers an alarm . | ||
| image storagedisabled use capacitypercentage | monitorcloud platformin image server of reservedcapacityconfiguration, any image serverdisabled use capacitypercentage (percentage=reservedcapacity/image servertotalcapacity)meets the alarm condition triggers an alarm . | ||
| systemdata directory | management nodedata directorydiskidlecapacity | monitormanagement nodedata directorydisk of idlecapacity, meets the alarm condition triggers an alarm . | |
| management nodedata directorydiskidle rate | monitormanagement nodedata directorydisk of idle rate, meets the alarm condition triggers an alarm . | ||
| management nodedata directorydiskused capacity | monitormanagement nodedata directorydisk of used capacity, meets the alarm condition triggers an alarm . | ||
| management nodedata directorydiskoccupy use rate | monitormanagement nodedata directorydisk of occupy use rate, meets the alarm condition triggers an alarm . | ||
| host | CPU | CPUidle rate | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
| allCPUidle rate | batch-monitors multiplehost of CPUidle rate, any host of allCPUidle ratesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| CPUusage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| average CPUusage | batch-monitors multiplehost of CPUusage, any host of average CPUusagemeets the alarm condition triggers an alarm . | ||
| allCPUusage | batch-monitors multiplehost of CPUusage, any host of allCPUusagesummeets the alarm condition triggers an alarm . Note: percentagesum, can abilityis greater than 100% . |
||
| allCPUquantity | monitorcloud platformin all hostallCPUquantity, CPUquantitysummeets the alarm condition triggers an alarm . | ||
| usedCPUquantity | monitorcloud platformin all hostused of CPUquantity, usedCPUquantitysummeets the alarm condition triggers an alarm . | ||
| disabled use CPUquantity | monitorcloud platformin all hostdisabled use CPUquantity, disabled use CPUquantitysummeets the alarm condition triggers an alarm . | ||
| usedCPUpercentage | monitorcloud platformin all hosttotalbodyusedCPUpercentage (percentage=all hostusedCPUsum/all hostCPUsum), meets the alarm condition triggers an alarm . | ||
| disabled use CPUpercentage | monitorcloud platformin all hosttotalbodydisabled use CPUpercentage (percentage=all hostdisabled use CPUsum/all hostCPUsum), meets the alarm condition triggers an alarm . | ||
| can use CPUquantity | monitorcloud platformin all hostcan use CPUquantity, can use CPUquantitysummeets the alarm condition triggers an alarm . | ||
| can use CPUpercentage | monitorcloud platformin all hosttotalbodycan use CPUpercentage (percentage=all hostcan use CPUsum/all hostCPUsum), meets the alarm condition triggers an alarm . | ||
| the hostusedCPUquantity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| the hostusedCPUpercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| the hostcan use CPUquantity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| the hostcan use CPUpercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| CPUtemperature | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| memory | memorynotusagecapacity | batch-monitors multiplehost of memorynotusagecapacity, any host of memorynotusagecapacitymeets the alarm condition triggers an alarm . | |
| memorynotusagepercentage | batch-monitors multiplehost of memorynotusagepercentage, any host of memorynotusagepercentagemeets the alarm condition triggers an alarm . | ||
| memoryusagecapacity | batch-monitors multiplehost of memoryusagecapacity, any host of memoryusagecapacitymeets the alarm condition triggers an alarm . | ||
| memoryusagepercentage | batch-monitors multiplehost of memoryusagepercentage, any host of memoryusagepercentagemeets the alarm condition triggers an alarm . | ||
| memorycapacity | monitorcloud platformin all host of memorycapacity, all host of memorycapacitysummeets the alarm condition triggers an alarm . | ||
| usedmemorycapacity | monitorcloud platformin all host of usedmemorycapacity, all host of usedmemorycapacitysummeets the alarm condition triggers an alarm . | ||
| usedmemorypercentage | monitorcloud platformin all host of usedmemorycapacity, all host of totalbodyusedmemorypercentage (percentage=all hostusedmemorycapacitysum/all hostmemorycapacitysum)meets the alarm condition triggers an alarm . | ||
| disabled use memorycapacity | monitorcloud platformin hostreservedmemoryconfiguration, reservedmemorycapacitymeets the alarm condition triggers an alarm . | ||
| disabled use memorycapacitypercentage | monitorcloud platformin hostreservedmemoryconfiguration, any hostreservedmemorycapacitypercentage (percentage=hostreservedmemory/hosttotalmemory)meets the alarm condition triggers an alarm . | ||
| remainingmemorycapacity | monitorcloud platformin all host of remainingmemorycapacity, all host of remainingmemorycapacitysummeets the alarm condition triggers an alarm . | ||
| remainingmemorycapacitypercentage | monitorcloud platformin all host of remainingmemorycapacity, all host of totalbodyremainingmemorycapacitypercentage (percentage=totalremainingmemorycapacity/totalmemorycapacity)meets the alarm condition triggers an alarm . | ||
| the hostused memorycapacity | batch-monitors multiplehost of used memorycapacity, any host of used memorycapacitymeets the alarm condition triggers an alarm . | ||
| the hostused memorycapacitypercentage | batch-monitors multiplehost of used memorycapacitypercentage, any host of used memorycapacitypercentagemeets the alarm condition triggers an alarm . | ||
| the hostcan use memorycapacity | batch-monitors multiplehost of can use memorycapacity, any host of can use memorycapacitymeets the alarm condition triggers an alarm . | ||
| the hostcan use memorycapacitypercentage | batch-monitors multiplehost of can use memorycapacitypercentage, any host of can use memorycapacitypercentagemeets the alarm condition triggers an alarm . | ||
| disk | diskreadIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| alldiskreadIOPS | batch-monitors multiplehost of diskreadIOPS, any host of all diskreadIOPSsummeets the alarm condition triggers an alarm . | ||
| diskwriteIOPS | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwriteIOPS | batch-monitors multiplehost of diskwriteIOPS, any hostall disk of writeIOPSsummeets the alarm condition triggers an alarm . | ||
| diskread speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskread speed | batch-monitors multiplehost of diskread speed, any hostall disk of read speedsummeets the alarm condition triggers an alarm . | ||
| diskwrite speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| alldiskwrite speed | batch-monitors multiplehost of diskwrite speed, any hostall disk of write speedsummeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacity | batch-monitors multiplehost of diskremainingcapacity, any hostall disk of remainingcapacitysummeets the alarm condition triggers an alarm . | ||
| alldiskremainingcapacitypercentage | batch-monitors multiplehost of diskremainingcapacitypercentage, any hostall diskremainingcapacitypercentage (percentage=all diskremainingcapacitysum/all diskcapacitysum)meets the alarm condition triggers an alarm . | ||
| alldiskusedcapacity | batch-monitors multiplehost of diskusedcapacity, any hostall disk of usedcapacitysummeets the alarm condition triggers an alarm . | ||
| alldiskusedcapacitypercentage | batch-monitors multiplehost of diskusedcapacitypercentage, any hostall diskusedcapacitypercentage (percentage=all diskusedcapacitysum/all diskcapacitysum)meets the alarm condition triggers an alarm . | ||
| diskcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskremainingcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacity | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| diskusedcapacitypercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| root diskusage | batch-monitors multiplehost of root diskusage, any host of root diskusagemeets the alarm condition triggers an alarm . | ||
| root diskusagecapacity | batch-monitors multiplehost of root diskusagecapacity, any host of root diskusagecapacitymeets the alarm condition triggers an alarm . | ||
| XFSfile system fragmentation levelpercentage | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| SSDtemperature | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| SSDremaining life | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| NIC | NICinbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
|
| allNICinbound speed | batch-monitors multiplehost of NICinbound speed, any hostall NICinbound speedsummeets the alarm condition triggers an alarm . | ||
| NICinbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound packet count | batch-monitors multiplehost of NICinbound packet count, any hostall NICinbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICinbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICinbound error count | batch-monitors multiplehost of NICinbound error count, any hostall NICinbound error countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound speed | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound speed | batch-monitors multiplehost of NICoutbound speed, any hostall NICoutbound speedsummeets the alarm condition triggers an alarm . | ||
| NICoutbound packet count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound packet count | batch-monitors multiplehost of NICoutbound packet count, any hostall NICoutbound packet countsummeets the alarm condition triggers an alarm . | ||
| NICoutbound error count | whether a configuration is specified, uses different resource monitoring granularities .can as neededselect :
|
||
| allNICoutbound error count | batch-monitors multiplehost of NICoutbound error count, any hostall NICoutbound error countsummeets the alarm condition triggers an alarm . | ||
| hostConntrackconnectioncount | batch-monitors multiplehost of Conntrackconnectioncount, any host of Conntrackconnectioncountmeets the alarm condition triggers an alarm . | ||
| hostConntrackusedpercentage | batch-monitors multiplehost of Conntrackusedamount, any host of Conntrackusedpercentagemeets the alarm condition triggers an alarm . | ||
| other | hostquantity | monitorcloud platformin hostquantity, hosttotal countmeets the alarm condition triggers an alarm . | |
| connectedhostquantity | monitorcloud platformin hostquantity, connectedhostquantitymeets the alarm condition triggers an alarm . | ||
| connectedhostpercentage | monitorcloud platformin hostquantity, connectedhostpercentage (percentage=connectedhostquantity/hosttotal count)meets the alarm condition triggers an alarm . | ||
| disconnectedhostquantity | monitorcloud platformin hostquantity, disconnectedhostquantitymeets the alarm condition triggers an alarm . | ||
| disconnectedhostpercentage | monitorcloud platformin hostquantity, disconnectedhostpercentage (percentage=disconnectedhostquantity/hosttotal count)meets the alarm condition triggers an alarm . | ||
| elastic bare metal gateway node | elastic bare metal gateway nodequantity | monitorcloud platformin elastic bare metal gateway nodequantity, elastic bare metal gateway nodetotal countmeets the alarm condition triggers an alarm . | |
| connectedelastic bare metal gateway nodequantity | monitorcloud platformin elastic bare metal gateway nodequantity, connectedelastic bare metal gateway nodequantitymeets the alarm condition triggers an alarm . | ||
| connectedelastic bare metal gateway nodepercentage | monitorcloud platformin elastic bare metal gateway nodequantity, connectedelastic bare metal gateway nodepercentage (percentage=connectedelastic bare metal gateway nodequantity/elastic bare metal gateway nodetotal count)meets the alarm condition triggers an alarm . | ||
| disconnectedelastic bare metal gateway nodequantity | monitorcloud platformin elastic bare metal gateway nodequantity, disconnectedelastic bare metal gateway nodequantitymeets the alarm condition triggers an alarm . | ||
| disconnectedelastic bare metal gateway nodepercentage | monitorcloud platformin elastic bare metal gateway nodequantity, disconnectedelastic bare metal gateway nodepercentage (percentage=disconnectedelastic bare metal gateway nodequantity/elastic bare metal gateway nodetotal count)meets the alarm condition triggers an alarm . | ||
| L3network | allcan use IPcount (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, all L3network of can use IPcountsummeets the alarm condition triggers an alarm . | |
| allcan use IPpercentage (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, allcan use IPpercentage (allcan use IPpercentage=all L3networkcan use IPtotal count/all L3networkIPtotal count)meets the alarm condition triggers an alarm . | ||
| allused IPcount (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, all L3network of used IPcountsummeets the alarm condition triggers an alarm . | ||
| allused IPpercentage (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, allused IPpercentage (allused IPpercentage=all L3networkused IPtotal count/all L3networkIPtotal count)meets the alarm condition triggers an alarm . | ||
| alldisabled use IPcount (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, all L3network of disabled use IPcountsummeets the alarm condition triggers an alarm . | ||
| alldisabled use IPpercentage (IPv4) | monitorcloud platformin all IPv4typeL3network of IPaddress count, alldisabled use IPpercentage (alldisabled use IPpercentage=all L3networkdisabled use IPtotal count/all L3networkIPtotal count)meets the alarm condition triggers an alarm . | ||
| can use IPcount (IPv4) | batch-monitors multipleIPv4typeL3network of IPaddress count, any L3network of can use IPcountmeets the alarm condition triggers an alarm . | ||
| can use IPpercentage (IPv4) | batch-monitors multipleIPv4typeL3network of IPaddress count, any L3network of can use IPpercentagemeets the alarm condition triggers an alarm . | ||
| used IPcount (IPv4) | batch-monitors multipleIPv4typeL3network of IPaddress count, any L3network of used IPcountmeets the alarm condition triggers an alarm . | ||
| used IPpercentage (IPv4) | batch-monitors multipleIPv4typeL3network of IPaddress count, any L3network of used IPpercentagemeets the alarm condition triggers an alarm . | ||
| volume | volumetotal count | monitorcloud platformin root volume and data volumequantity, root volume and data volumequantitysummeets the alarm condition triggers an alarm . | |
| root volumetotal count | monitorcloud platformin root volumequantity, root volumetotal countmeets the alarm condition triggers an alarm . | ||
| root volumepercentage | monitorcloud platformin root volumepercentage (percentage=root volumequantity/root volume and data volumequantitysum), meets the alarm condition triggers an alarm . | ||
| data volumetotal count | monitorcloud platformin data volumequantity, data volumetotal countmeets the alarm condition triggers an alarm . | ||
| data volumepercentage | monitorcloud platformin data volumepercentage (percentage=data volumequantity/root volume and data volumequantitysum), meets the alarm condition triggers an alarm . | ||
| can use data volumetotal count | monitorcloud platformin data volumestatus, can use data volumequantitymeets the alarm condition triggers an alarm . | ||
| can use data volumepercentage | monitorcloud platformin data volumestatus, can use data volumepercentage (percentage=can use data volumequantity/data volumetotal amount)meets the alarm condition triggers an alarm . | ||
| volume snapshottotal count | monitorcloud platformin volume snapshot (data volume snapshot+root volume snapshot)quantity, volume snapshottotal countmeets the alarm condition triggers an alarm . | ||
| root volume snapshottotal count | monitorcloud platformin root volume snapshotquantity, root volume snapshottotal countmeets the alarm condition triggers an alarm . | ||
| root volume snapshotpercentage | monitorcloud platformin root volume snapshotpercentage (percentage=root volume snapshottotal count/root volume snapshot and data volume snapshotsum), meets the alarm condition triggers an alarm . | ||
| data volume snapshottotal count | monitorcloud platformin data volume snapshotquantity, data volume snapshottotal countmeets the alarm condition triggers an alarm . | ||
| data volume snapshotpercentage | monitorcloud platformin all data volume snapshotpercentage (percentage=data volume snapshottotal count/root volume snapshot and data volume snapshotsum), any data volume snapshotpercentagemeets the alarm condition triggers an alarm . | ||
| volumeusagecapacitypercentage | monitorcloud platformin all volumeusagecapacitypercentage (percentage=usagecapacity/volumetotalcapacity), any volumeusagecapacitypercentagemeets the alarm condition triggers an alarm . Note: thick provisioningtypevolume (exampleAs shown in : Shared
Blockprimary storage on thick provisioningtype of volume) of usagecapacitypercentageas100%, settingsthis alarmwillstandalonethen trigger . |
||
| (100GBandby on )fragmentation level (Extenttotal count) | monitorcloud platformin all capacityexceeds100GBvolume of fragmentation level, any volumefragmentation level of fragmentation level (Extenttotal count)meets the alarm condition triggers an alarm . Note: usagethisalarm itemNote thatby under status:
|
||
| virtualIP | under rownetworktraffic | batch-monitors multiplevirtualIP of under rownetworktraffic, any virtualIP of under rownetworktrafficmeets the alarm condition triggers an alarm . | |
| under rownetworkpacket count | batch-monitors multiplevirtualIP of under rownetworkpacket count, any virtualIP of under rownetworkpacket countmeets the alarm condition triggers an alarm . | ||
| on rownetworktraffic | batch-monitors multiplevirtualIP of on rownetworktraffic, any virtualIP of on rownetworktrafficmeets the alarm condition triggers an alarm . | ||
| on rownetworkpacket count | batch-monitors multiplevirtualIP of on rownetworkpacket count, any virtualIP of on rownetworkpacket countmeets the alarm condition triggers an alarm . | ||
| primary storage | allcapacity | monitorcloud platformin all primary storagecapacity, all primary storagecapacitysummeets the alarm condition triggers an alarm . | |
| allcan use capacity | monitorcloud platformin all primary storagecan use capacity, all primary storagecan use capacitysummeets the alarm condition triggers an alarm . | ||
| allcan use capacitypercentage | monitorcloud platformin all primary storagecan use capacitypercentage (percentage=all primary storagecan use capacitysum/all primary storagecapacitysum), meets the alarm condition triggers an alarm . | ||
| allused capacity | monitorcloud platformin all primary storageused capacity, all primary storageused capacitysummeets the alarm condition triggers an alarm . | ||
| allused capacitypercentage | monitorcloud platformin all primary storageused capacitypercentage (percentage=all primary storageused capacitysum/all primary storagecapacitysum), meets the alarm condition triggers an alarm . | ||
| alldisabled use capacity | monitorcloud platformin primary storagereservedcapacityconfiguration, all primary storagereservedcapacitysummeets the alarm condition triggers an alarm . | ||
| alldisabled use capacitypercentage | monitorcloud platformin all primary storagereservedcapacityconfiguration, primary storagealldisabled use capacitypercentage (percentage=all primary storagereservedcapacitysum/all primary storagecapacitysum)meets the alarm condition triggers an alarm . | ||
| the primary storagecan use capacity | batch-monitors multipleprimary storage of can use capacity, any primary storage of can use capacitymeets the alarm condition triggers an alarm . | ||
| the primary storagecan use capacitypercentage | batch-monitors multipleprimary storage of can use capacitypercentage (percentage=can use capacity/primary storagetotalcapacity), any primary storage of can use capacitypercentagemeets the alarm condition triggers an alarm . | ||
| the primary storageused capacity | batch-monitors multipleprimary storage of used capacity, any primary storage of used capacitymeets the alarm condition triggers an alarm . | ||
| the primary storageused capacitypercentage | batch-monitors multipleprimary storage of used capacitypercentage (percentage=used capacity/primary storagetotalcapacity), any primary storage of used capacitypercentagemeets the alarm condition triggers an alarm . | ||
| the primary storagecan use physical capacity | batch-monitors multipleprimary storage of can use physical capacity, any primary storage of can use physical capacitymeets the alarm condition triggers an alarm . | ||
| the primary storagecan use physical capacitypercentage | batch-monitors multipleprimary storage of can use physical capacitypercentage (percentage=can use physical capacity/primary storagetotalphysical capacity), any primary storage of can use physical capacitypercentagemeets the alarm condition triggers an alarm . | ||
| the primary storageused physical capacity | batch-monitors multipleprimary storage of used physical capacity, any primary storage of used physical capacitymeets the alarm condition triggers an alarm . | ||
| the primary storageused physical capacitypercentage | batch-monitors multipleprimary storage of used physical capacitypercentage (percentage=used physical capacity/primary storagetotalphysical capacity), any primary storage of used physical capacitypercentagemeets the alarm condition triggers an alarm . | ||
| the primary storageroot volumequantity | batch-monitors multipleprimary storagein of root volumequantity, any primary storagein of root volumequantitymeets the alarm condition triggers an alarm . | ||
| the primary storagedata volumequantity | batch-monitors multipleprimary storagein of data volumequantity, any primary storagein of data volumequantitymeets the alarm condition triggers an alarm . | ||
| the primary storagesnapshotquantity | batch-monitors multipleprimary storagein of snapshotquantity, any primary storagein of snapshotquantitymeets the alarm condition triggers an alarm . | ||
| Cephstorage poolcan use capacitypercentage | batch-monitors multipleprimary storagein of Cephstorage poolcan use capacitypercentage, any primary storagein of Cephstorage poolcan use capacitymeets the alarm condition triggers an alarm . | ||
| Cephstorage poolused capacitypercentage | batch-monitors multipleprimary storagein of Cephstorage poolused capacitypercentage, any primary storagein of Cephstorage poolused capacitymeets the alarm condition triggers an alarm . | ||
| listener | sessionused quantity | batch-monitors multiplelistener of sessionquantity, any listener of sessionused quantitymeets the alarm condition triggers an alarm . | |
| sessionused percentage | batch-monitors multiplelistener of sessionused percentage (percentage=sessionused quantity/sessiontotal amount), any listener of sessionused percentagemeets the alarm condition triggers an alarm . | ||
| checknot to healthy of backendservicedevice | batch-monitors multiplelistener of backendservicedevicegroup in of backendservicedevicestatus, any listener of backendservicedevicegroup in has not healthy of backendservicedevice triggers an alarm . | ||
| management node | arbiterIPnot can reach | monitormultiple management nodesarbiterIPwhethercan reach, whenarbiterIPnot can reachwhen triggers an alarm . | |
| dual management nodesdatabase is not synchronized | monitordual management nodescountdatabasestatus, whencountdatabaseabnormal or dual management nodesdatabase is not synchronizedwhen triggers an alarm . | ||
| projectresource (requiresenterprisemanagemodulelicense) | computeresource | VM instancequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
| running in VM instancequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| CPUquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| memoryquotausagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| GPUdevicequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| affinity and groupquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| storageresource | volume snapshotquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
|
| data volumequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| can use storagecapacityquotausagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| imagequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| all imagecapacityquotausagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| can use backupcapacityquotausagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| backupquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| networkresource | VXLANnetworkquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
|
| L3networkquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| securegroupquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| virtualIPquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| elastic performanceIPquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| portforwardquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| load balancingdevicequantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| listenerquantity quota usagepercentage | alarm scopeSupports All Resources and Multiple Resources, can as neededselect :
|
||
| CDPtask (requiresfor data protectionCDPmodulelicense) | CDPtaskused capacityoccupyplannedcapacitypercentage | monitorcloud platformin all CDPtaskused capacityoccupyplannedcapacity of percentage, any CDPtask of used capacity and plannedcapacity of occupyratioreaches the threshold triggers an alarm . | |
| RPOoffsettime | monitorcloud platformin all CDPtask of RPOoffsettime, any CDPtask of RPOoffsettimereaches the threshold triggers an alarm . | ||
Event Alarm Items
Default Alarm Items
| resource type | alarm item | description |
|---|---|---|
| VM instance | VM instancefault | Monitors cloud platformin all running in of VM instancestatus, any running in of VM instancehas fault triggers an alarm . Note: VM instancerequires installation oflatestNewly version of GuestToolstool, andthe toolneeds to is inrunningstatus . |
| VM instancestays in in shutdownstatus |
|
|
| VM instanceOn hostHAstarts |
|
|
| host | hostconnected |
|
| hostgo tomaintenance modetriggerVM instancemigration fails |
|
|
| hostdisconnected |
|
|
| hostNICdisconnected |
|
|
| hostNICconnected |
|
|
| hostmount path error |
|
|
| host on is found with not managed by the system of VM instance |
|
|
| CPUstatusabnormal |
|
|
| memorystatusabnormal |
|
|
| memoryECCalarm |
|
|
| diskstatusabnormal |
|
|
| diskremoved |
|
|
| diskinsert |
|
|
| fanstatusabnormal |
|
|
| image server | image serverconnected |
|
| image serverdisconnected |
|
|
| primary storage | primary storageconnected |
|
| primary storagedisconnected |
|
|
| primary storage to hostconnectionstatuscheck failed |
|
|
| management node | management nodeconnected |
|
| management nodedisconnected |
|
|
| VPCrouter | VPCrouterconnected |
|
| VPCrouterdisconnected |
|
|
| routeractive/standbyswitch |
|
|
| VPCrouterdiskspace occupied byabnormalfileoccupy use |
|
|
| router enabledstatuschanges toaspausedstop |
|
|
| notificationforobject | SMS sending failed |
|
| CDPtask (requiresfor data protectionCDPmodulelicense) | CDPtaskstatusabnormalswitch |
Note: triggerCDPtaskstatusabnormalswitch of reason:
|
| load balancer instance | load balancer instancedisconnected |
|
| load balancer instanceconnected |
|
customalarm item
| resource type | alarm item | description |
|---|---|---|
| VM instance | VM instanceOn hostHAstarts | monitorcloud platformin all NeverStopVM instance, any NeverStopVM instanceOn hostHAstarts triggers an alarm . |
| VM instanceOn host on of statuschanges | monitorcloud platformin all VM instanceOn host on of status, any VM instance of statuschanges triggers an alarm . Note: VM instanceabnormal of statuschanges to change can willtriggeralarm, normal of start, stop, and power operationsoperationnot willtriggeralarm . |
|
| VM instancestays in in shutdownstatus | monitorcloud platformin all VM instancestatus, any VM instancelongtime (about10minutes)is inin shutdownstatus triggers an alarm . | |
| VM instancefault | monitorcloud platformin all running in of VM instancestatus, any running in of VM instancehas fault triggers an alarm . Note: VM instancerequires installation oflatestNewly version of GuestToolstool, andthe toolneeds to is inrunningstatus . |
|
| VPCrouter | VPCrouterdisconnected | monitorcloud platformin all VPCrouter readystatus, any routerdisconnected triggers an alarm . |
| VPCrouterconnected | monitorcloud platformin all VPCrouter readystatus, any VPCrouterconnectionafter triggers an alarm . | |
| routeractive/standbyswitch | monitorcloud platformhighcan use groupin of VPCrouterstatus, any highcan use groupin of VPCrouteractive/standby switchover occursswitch triggers an alarm . | |
| VPCrouterdiskspace occupied byabnormalfileoccupy use | monitorcloud platformall VPCrouterin of diskoccupy use status, any VPCrouterhas exceeds100 MB of single file triggers an alarm . | |
| load balancing | load balancer instancedisconnected | monitorcloud platformin all load balancer instance of readystatus, any load balancer instancedisconnected triggers an alarm . |
| load balancer instanceconnected | monitorcloud platformin all load balancer instance of readystatus, any load balancer instanceconnectionafter triggers an alarm . | |
| image server | image serverdisconnected | monitorcloud platformin all image server of readystatus, any image serverdisconnected triggers an alarm . |
| image serverconnected | monitorcloud platformin all image server of readystatus, any image serverconnectionafter triggers an alarm . | |
| management node | management nodedisconnected | monitorcloud platformin all management node of readystatus, any management nodedisconnected triggers an alarm . |
| management nodeconnected | monitorcloud platformin all management node of readystatus, any management nodeconnectionafter triggers an alarm . | |
| host | host on is found with not managed by the system of VM instance | monitorcloud platformin all host on of VM instancestatus, any actualexampleVM instancenotiscountdatabaserecord triggers an alarm . |
| hostdisconnected | monitorcloud platformin all host of readystatus, any hostdisconnected triggers an alarm . | |
| hostconnected | monitorcloud platformin all host of readystatus, any hostconnectionafter triggers an alarm . | |
| hostNICdisconnected | monitorcloud platformin all is inconnectedstatus of host of NICreadystatus, any hostNICdisconnected triggers an alarm . | |
| hostNICconnected | monitorcloud platformin all is inconnectedstatus of host of NICreadystatus, any hostNICrecoverednormalconnection triggers an alarm . | |
| CPUstatusabnormal | monitorcloud platformin all is inconnectedstatus of host of CPUstatus, any hostCPUstatusabnormal triggers an alarm . | |
| memorystatusabnormal | monitorcloud platformin all is inconnectedstatus of host of memorystatus, any hostmemorystatusabnormal triggers an alarm . | |
| memoryECCalarm | monitorcloud platformin all is inconnectedstatus of hostECCalarm, any hostfoundmemoryECCalarm triggers an alarm . | |
| diskstatusabnormal | monitorcloud platformin all is inconnectedstatus of host of diskreadystatus, any hostdiskconfigurationstatuscheckabnormal triggers an alarm . | |
| diskremoved | monitorcloud platformin all is inconnectedstatus of host of diskconnectionstatus, any hostdiskremoved triggers an alarm . | |
| diskinsert | monitorcloud platformin all is inconnectedstatus of host of diskconnectionstatus, any hostdiskinsert triggers an alarm . | |
| fanstatusabnormal | monitorcloud platformin all is inconnectedstatus of host of fanstatus, any hostfanstatuscheckabnormal triggers an alarm . | |
| primary storage | primary storage to hostconnectionstatuscheck failed | monitorcloud platformin primary storage and host of connectionstatus, whencloud platformnotobtain to primary storage and host of connectionstatuswhen, triggers an alarm . |
| primary storagedisconnected | monitorcloud platformin all primary storage of readystatus, any primary storagedisconnected triggers an alarm . | |
| primary storageconnected | monitorcloud platformin all primary storage of readystatus, any primary storageconnectionafter triggers an alarm . | |
| hostmount path error | monitorcloud platformin specifiedprimary storage (NFS/SharedMountPoint/AliyunNAS) of URL (hostmount path), whenany primary storage of URLcannot be obtained, triggers an alarm . | |
| vCenter | vCenterhosttimeabnormal | monitormanagedvCenterenvironmentin all hosttime and cloud platformsystemtimewhether they are consistent, any vCenterhosttimeabnormal triggers an alarm . |
| vCenterevent message | displaymanagedvCenter of event message, any vCentergenerated of event messageallcan through alarms of displayed . | |
| backuptask | task resultfailed | monitorcloud platformin all backuptask of executestatus, any backuptaskafter failure triggers an alarm . |
| projectresource (requiresenterprisemanagemodulelicense) | projectreclaimed | monitorcloud platformin all projectstatus, any projectreclaimed triggers an alarm . |
| HA | hostgo tomaintenance modetriggerVM instancemigration fails | monitorcloud platformin VM instance of migrationstatus, any non-localstorage of hostgo tomaintenance modetriggerVM instancemigration fails triggers an alarm . |
| CDPtask (requiresfor data protectionCDPmodulelicense) | CDPtaskfailed | monitorcloud platformin all CDPtask of executestatus, any CDPtaskafter failure triggers an alarm . |
| CDPtaskstatusabnormalswitch | monitorcloud platformin all CDPtask of taskstatus, any CDPtaskstatusoccursabnormalswitch triggers an alarm . Note: triggerCDPtaskstatusabnormalswitch of reason:
|
