Platform O&M
Network Topology
A network topology visualizes the network architecture of the Cloud. It allows for efficient planning, management, and improvement of network architecture. Network topologies can be categorized into global topologies and custom topologies.
Cloud Monitoring
Management Node Monitoring
Management Node (MN) monitoring allows you to view the health status of each management node when you use multiple management nodes to achieve high availability.
Performance Analysis
Performance Analysis displays the performance metrics of key resources monitored externally or internally in the Cloud. You can view the performance analysis or export the analysis report as needed to improve the O&M efficiency.
- VM instance: Displays the information such as the name, state, CPU utilization, memory utilization, disk read/write speed, NIC in/out speed, owner, and supported operation (stop VM instance) of a VM instance. You can customize the columns and choose to export the average, maximum, or minimum values of the metrics as needed.
- VPC vRouter: Displays the information such as the name, CPU utilization, memory utilization, disk read speed, disk write IOPS, NIC in rate, NIC in packets, NIC in errors rate, and owner of a VPC vRouter. You can customize the columns and choose to export the average, maximum, or minimum values of the metrics as needed.
- Host: Displays the information such as the CPU utilization, memory utilization, disk read/write speed, disk size, and NIC in/out speed of a host.
- Backup storage: Displays the information such as the name and available capacity percent of a backup storage.
- L3 network: Displays the information such as the name, used IPs, used IP percent, available IPs, and available IP percent of an L3 network.
- VIP: Displays the name, inbound/outbound traffic, inbound/outbound packet rate, and owner of a virtual IP address.
Capacity Management
Capacity Management visualizes the capacities and usages of key resources in the Cloud. You can use this feature to improve O&S efficiency.
Capacity Management displays the physical capacities and usages of key resources by using cards, and displays Top 10 resource capacities and usages, providing you a commanding view of your resource usages and greatly improving O&S efficiency.
Monitoring and Alarm
The Monitoring and Alarm feature monitors time-series data and events and sends alarm messages to specified endpoints by using SNS. Resource alarms, event alarms, and extended alarms are supported. The supported endpoints include the system, emails, DingTalk, WeCom, Lark, Webhook, SMS, Microsoft Teams, and SNMP trap receiver. For some resource alarms, you need to install the agent before they can work as expected.

Concepts
- Monitoring System:A monitoring system provides the following features:
- Monitor the following two types of time-series data:
- Resource utilizations such as CPU utilization of VM instances and memory utilization of hosts
- Resource capacities such as available number of IP addresses and total number of running VM instances
- Event collection: collects events predefined on the Cloud, such as host disconnection and VM HA enabling.
- Alarm: triggers alarms on time-series data or events.
- Audit: records all operations and allows queries.
- Customization: allows you to customize alarms and message templates
and use predefined alarm templates and resource groups.
- The following three types of alarms are supported:
- Resource alarm: triggers alarms on time-series data. For example, you can configure an alarm for VM instances. If the CPU utilization of a VM instance exceeds 80% by five consecutive minutes, send an alarm message to an email address.
- Event alarm: triggers alarms on events, also called event subscription. For example, you can configure an alarm for host disconnection. If a host is disconnected, an alarm message is sent to DingTalk.
- Extended alarm: receives alarm messages from message sources. For example, if a Ceph Enterprise storage pool is downgraded, an alarm message is sent to the system of the Cloud.
-
A message template specifies the text template of a
resource alarm message or event alarm message sent to an SNS system.
- A message template and message recovery template are provided by the system. If you do not create a template, the system uses the predefined templates.
- You can create multiple message templates and can set only one template as the default template. Messages are formatted by using the default template.
- You can use
${}in a template to quote variables configured in an alarm or event. - You can configure email, DingTalk, WeCom, Lark, SMS, Webhook, and Microsoft Teams as an endpoint in a message template. Messages sent by using email, DingTalk, WeCom, Lark, SMS, Webhook, or Microsoft Teams are sent in the specified format.
- A message source is used to take over extended alarm messages. If you configure alarms for message sources, extended alarm messages can be sent to various endpoints. This enables centralized management of alarm messages and improves O&M efficiencies. You can configure a message source to take over alarm messages of Ceph Enterprise.
- An alarm template is a template of alarm rules. If you associate an alarm template with a resource group, an alarm is created to monitor the resources in the group.
- A resource group consists of resources grouped based on your business needs. If you associate an alarm template with a resource group, the alarm rules specified by the template take effect on all the resources in the group.
- The following three types of alarms are supported:
- Monitor the following two types of time-series data:
- SNS:
SNS sends alarm messages to the specified endpoints. The supported types of endpoints include the system, emails, DingTalk, WeCom, Lark, Webhook, SMS, Microsoft Teams, and SNMP trap receivers.
Endpoints:- The system provides the system-type endpoint. If you associate an alarm with this endpoint, alarm messages will be displayed below the Recent Message button in the top right corner of the UI.
- You can also create an endpoint of the email, DingTalk, Lark, WeCom, Webhook, SMS, Microsoft Teams, or SNMP trap receiver type.
Characteristics
- Provides rich metric items to comprehensively monitor and alarm the core resources as well as events of the Cloud platform.
- Supports types of endpoints including system, emails, DingTalk, WeCom, Lark, Webhook, SMS, Microsoft Teams, and SNMP trap receiver for subscription topics. You can choose an appropriate endpoint to receive alarm messages according to the actual situation.
- One alarm can monitor multiple resources at the same time.
- Emails, DingTalk, WeCom, Lark, SMS, Webhook, and Microsoft Teams support customized alarm message templates. You can set alarm message templates on demand and quickly locate key information from alarm messages.
- Supports for creating a template of alarm rules. If you associate an alarm template with a resource group, an alarm is created to monitor the resources in the group.
Scenarios
The function of monitoring and alarm monitors the core resources and events of the Cloud platform and sets up an alarm receiving mechanism. When core resources are abnormal, the monitoring and alarm will make real-time responses according to the alarm level to help O&M personnel quickly locate and solve the problem.
Global Setting
- Monitoring data is retained locally for 6 months by default, and you can
customize the monitoring data retention period in the basic settings as
follows:
On the main menu of ZStack Cloud, choose . Then, the Basic tab is displayed. You can set Monitoring Data Retention Period. Enter an integer between 1 and 12. Default: 6. Unit: month.
- Monitoring data is retained locally in a size of 50GB by default, and you can
customize the monitoring data retention size in the basic settings as
follows:
On the main menu of ZStack Cloud, choose . Then, the Basic tab is displayed. You can set Monitoring Data Retention Size based on your needs. Default: 50 GB.
- ZStack Cloud supports receiving extended alarm messages. On the main menu, choose . Then, the Advanced tab is displayed. You need to turn on the Extended Alarm Notification switch to use the extended alarm function.
SNS
SNS sends alarm messages to the specified endpoints. The supported types of endpoints include the system endpoint, email, DingTalk, HTTP applications, short message service, and Microsoft Teams.
- The Cloud provides a system endpoint by default. If an alarm binds a system endpoint, you are prompted for alarm notifications displayed near the Messages button at the upper right in the UI.
- You can also create an email, DingTalk, HTTP application, short message service, or Microsoft Teams endpoint as needed.
Email Endpoint
- Messages that send to topics will be sent to a specified email address via an email server.
- You can either create an SNS text template in advance or use a system template to send emails in a unified format.
- You need to add an email server in advance under the current zone, and make sure that the email server works properly.
DingTalk Endpoint
- Messages that send to topics will be sent to a specified DingTalk robot address via DingTalk. If you appoint members, alarm notifications will be sent to corresponding DingTalk members via phone numbers.
- You can either create an SNS text template in advance or use a system template to send alarm messages in a unified format.
- If you set an SNS text template in DingTalk, follow the Markdown syntax. Currently, DingTalk only supports a subset of Markdown syntax.
HTTP Application Endpoint
- Messages that send to topics will be sent to a specified HTTP address via HTTP POST.
- If the specified HTTP application cannot be accessed without a user name and password, enter accurately the user name and the password.
Short Message Service Endpoint
- Messages that send to topics will be sent to a specified phone number via the short message service.
- You can create an SNS text template in advance and set it as the default template to send short messages according to the template that you set.
Microsoft Teams Endpoint
- Messages that send to topics will be sent to a specified Microsoft Teams group by using the webhook method.
- You can either create an SNS text template in advance or use a system template to send the Microsoft Teams messages in a unified format.
One-Click Inspection
One-Click Inspection: Comprehensively inspects the health status of key resources and services of the Cloud and scores their healthiness based on the inspection results. In addition, the one-click inspection service provides O&M suggestions and inspection reports. One-click inspection is applicable to centralized O&M scenarios.

Concepts
- Inspection Categories and Items:One-Click Inspection provides five inspection categories, including platform, compute resources, network resources, storage resources, and global settings. You can use the service to inspect key resources and services of the Cloud, such as the management node, hosts, VM instances, image storage, primary storage, physical and virtual NICs and networks, and licenses.
- Platform: Check the basic services and running status of the Cloud.
- Compute: Check the usage and running status of physical and virtual compute resources of the Cloud.
- Network: Check the configurations and status of physical and virtual networks of the Cloud.
- Storage: Check the usage and running status of physical storage resources of the Cloud.
- Global Setting: Check the configurations of key resources of the Cloud.
After you select items from certain inspection categories and launch inspection, related resources or services are inspected and their healthiness is scored. For more information about inspection items, see Inspection Items.
- Inspection Results:One-Click Inspection provides four inspection results, including Normal, Warning, Fault, Failed.
- Normal: The inspected resources or services are in normal status. This result is marked with a green icon.
- Warning: The health status of inspected resources or services is compromised, which may to some extent affect their performance and stability. This result is marked with a yellow icon.
- Fault: The inspected resources or services are in critical condition and may seriously affect business operations. This result is marked with a red icon.
- Failed: The inspection on related resources or services fails, which may seriously affect business operations. This result is marked with a grey icon.
- Healthiness Scoring:
One-Click Inspection provides an in-built healthiness scoring mechanism for Cloud resources and services. It allows you to grasp the overall running status of the Cloud in a visualized way.
Scoring on inspected resources/services: Scores resources and services based on the inspection results of related resource and service attributes.- If all attributes of a resource or service under inspection are in Normal status, the inspection result of the resource or service is Normal. The score is 100 points.
- If one attribute of a resource or service under inspection is in Warning state and the other attributes are in Normal status, the inspection result of the resource or service is Warning. The score is 50 points.
- If one attribute of a resource or service under inspection is in Fault or Failed state, the inspection result of the resource or service is Fault or Failed. The score is 0 points.
Scoring on inspection items: Scores inspection items based on the inspection results of related resources and services.-
If an inspection item does not belong to the Global Setting category, the inspection item is scored based on the following mechanism:
- Score of Inspection Item = (Score of Resource 1 + Score of Resource 2 + …… + Score of Resource N)/(N*100)*100
- For example, if an inspection item involves 3 resources, which are in Normal, Warning, and Fault/Failed status respectively, the scores of the three resources are 100, 50, and 0 points respectively. Then the score of the inspection item is (100 + 50 + 0)/(3*100)*100=50 points.
-
If an inspection item belongs to the Global Setting category, the inspection item is scored based on the following mechanism:
- The score of the inspection item is the score of the involved global setting.
- For example, if the inspection result of the involved global setting is Warning, the score of the global setting is 50 points. Then the score of the inspection item is 50 points.
Scoring on the Cloud: Scores the Cloud based on the scores of all inspection items.- Score of the Cloud = (Score of Inspection Item 1 + Score of Inspection Item 2 + …… + Score of Inspection Item N)/(N*100)*100
- For example, if you select 3 inspection items, which is scored 100 points, 50 points, and 0 points respectively, then the score of the Cloud is (100 + 50 + 0)/(3*100)*100=50 points.
- O&M Suggestions:
If resources and services are detected in Warning or Fault status, One-Click Inspection analyzes the hidden dangers and their effects on these resources and services, and provides suggestions on O&M. For more information, see Inspection Items.
- Inspection Reports:
One-Click Inspection allows you to export PDF-formatted inspection reports. An inspection report summarizes platform configurations, resource status, and inspection results. It also provides details of all abnormal inspection items and corresponding O&M suggestions.
Benefits
- Comprehensive, customized, and efficient inspection capabilities: Provides five inspection categories that cover all key resources and services of the Cloud and allows you to select inspection items based on your business scenarios. After you launch an inspection, the inspection can be completed within a few minutes.
- Multi-layered scoring mechanism: The in-built three-layer mechanism of scoring on resources/services, inspection items, and the Cloud allows you to grasp the overall picture as well as details of the Cloud running status.
- Intelligent O&M suggestions: Provides risk analysis of resources and corresponding countermeasures, facilitating efficient O&M.
Message Log
Alarm Message
An alarm message is a message sent the time when an alarm is triggered.
Operation Log
An operation log is a chronological record of operations on the specified objects and their operation results.
Current Task
A current task is an ongoing operation performed in the Cloud. You can perform centralized management over ongoing operations.
Audit
Audit monitors and records all activities on the Cloud. You can use this feature to implement operation tracking, cybersecurity classified protection compliance, security analysis, troubleshooting, and automatic O&M.
Audit
Allows you to collect with one click the log data from the Cloud and various nodes on the Cloud generated in the specified time period and download the log data.
Backup Management
Backup Service
Backup management integrates multiple disaster recovery technologies such as incremental backup and full backup that are suitable for multiple business scenarios. You can implement local backup and remote backup based on your business needs.
Backup Service is a separate feature module. To use this service, purchase both the Base License and the Plus License of Backup Service. The Plus License cannot be used independently.
Typical Backup Scenarios
Note: If
you have the Tenant Management Plus license at the same time, the project
members (project managers, project admins, and general project members) can
perform local backup for VM instances, elastic baremetal instances, and data
volumes in the project.- Local Backup
A local ImageStore image storage can act as the Local Backup Server to store scheduled backup data of the local VM instances, elastic baremetal instances, data volumes, and management node databases. Meanwhile, the seamless switchover between the primary local backup server and the secondary local backup server is supported, which effectively ensures your business continuity.
If your local data is mistakenly deleted, or data in the local primary storage is damaged, you can recover the backup data from the local backup server, as shown in Local Backup Scenario-1
Figure 3. Local Backup Scenario-1 
If you encounter a disaster in your local data center, you can rely totally on your local backup server to rebuild your data center and recover your business, as shown in Local Backup Scenario-2.
Figure 4. Local Backup Scenario-2 
- Remote Backup
A storage server in a remote data center can act as the Remote Backup Server to store the scheduled backup data of the local VM instances, elastic baremetal instances, volumes, and databases. The backup data needs to be synchronized to the remote backup server from the local backup server.
If your local data is mistakenly deleted, or data in the local primary storage is damaged, you can recover the backup data from the remote backup server, as shown in Remote Backup Scenario-1
Figure 5. Remote Backup Scenario-1 
If you encounter a disaster in your data center, you can rely totally on your remote backup server to rebuild your data center and recover your business, as shown in Remote Backup Scenario-2.
Figure 6. Remote Backup Scenario-2 
- Public Cloud BackupThe storage server in the Public Cloud can act as the Public Cloud Backup Server to store the scheduled backup data of the local VM instances, volumes, and databases. The backup data can be synchronized to the Public Cloud backup server from the local backup server.
Note: The Public Cloud backup feature does not apply to elastic
baremetal instances.If your local data is mistakenly deleted, or data in the local primary storage is damaged, you can recover the backup data from the Public Cloud backup server, as shown in Public Cloud Backup Scenario-1.
Figure 7. Public Cloud Backup Scenario-1 
If you encounter a disaster in your data center, you can rely totally on your Public Cloud backup server to rebuild your data center and recover your business, as shown in Public Cloud Backup Scenario-2.
Figure 8. Public Cloud Backup Scenario-2 
- ZStack Cloud allows you to integrate 3rd-party backup services via CBT interfaces. For details, contact the official technical support.
Backup Job
You can create backup jobs to back up VM instances, volumes, and databases on your local data center to specified local storage servers and sync the backup data to specified remote backup servers or backup servers on the public cloud.
Local Backup Data
- You can restore the backup data to your local server or synchronize the data to a remote backup server.
- When you restore a database, refreshing the browser will make the UI display improperly without affecting the restoring process.
- With the Backup Service, you can back up or restore the data stored in the management node database. Note that the operation logs and monitoring information cannot be backed up or restored.
Local Backup Server
- You can use the ImageStore deployed in the local data center as the local backup server.
- You can also deploy a new local backup server.
- You can add more than one local backup server.
- If you specify more than backup server for a backup job, these backup servers can work in the active-standby mode.
- You can clean up invalid and expired data that is completely deleted to free up storage space.
- You can view the backup data backed up to the local backup server on the details page.
Remote Backup Server
- The backup data can only be synchronized from a local backup server to the remote backup server.
- You can add only one remote backup server to the Cloud.
- You can view the backup data backed up to the remote backup server on the details page.
- Before you can restore the remote backup data of VM instances and volumes to your local server, synchronize the data from the remote backup server to a local backup server first.
- The remote backup data of databases can be restored the your local server directly.
Continuous Data Protection (CDP)
Continuous Data Protection (CDP) provides second-level and fine-grained continuous backups for important business systems in VM instances, allowing users to restore VM data to a specific time state, and retrieve files without restoring the system. ZStack Cloud provides automated solutions on resuming all applications if a hardware or operating system failure occurs.
The Continuous Data Protection (CDP) service is a separate feature module. To use this service, purchase both the Base License and the Plus License of Continuous Data Protection (CDP). The Plus License cannot be used independently.
Concepts
- CDP task: You can create a CDP task to continuously back up your VM data to a specified backup server to achieve continuous data protection and recovery.
- Recovery task: A recovery task helps you quickly restore data by specifying a CDP task and recovery point, and allows you to view the recovery progress and logs in a more friendly way.
- CDP data: The backup data generated from continuous data protection on VM instances is stored in local backup servers.
- Recovery point: A recovery point is a data point generated during continuous data protection. A recovery point corresponds to a data record within the recovery point interval specified by the user.
- Locked recovery point: You can lock or unlock a recovery point as needed. After a recovery point is locked, data of the recovery point will not be automatically cleared or deleted.
How CDP Works
ZStack Cloud provides block-level continuous data protection. You can restore CDP data according to a specified time point.
- CDP backup
After continuous data protection is performed on a VM instance, the CDP server first performs a full copy of the VM data, continuously captures I/O data changes, and timestamps, and saves the I/O data of each change, thereby achieving continuous data backup.
Figure 9. CDP Backup 
- CDP recovery
When data recovery is performed, the CDP server exposes the CDP data as a block device and then restores the I/O data at a specified time to the disk or file system of a primary storage.
Figure 10. CDP recovery 
Advantages
- Simple:
- The software-defined solution we provide is hardware-independent and scalable.
- You can preview the backup files without restoring the system. Supported file formats: pictures (.png, .jpg, .bmp, .gif, etc.), PDF files, and text files (equal to or smaller than 10 MB).
- When you create a CDP task for the first time, the Cloud intelligently recommends the desired capacity required by a CDP task based on an algorithm, helping you to plan the backup space reasonably.
- You can restore data through a wizard-style process.
- Strong:
- Agentless backup: To use the CDP service, you do not need to install agents for your VM instances or couple with other applications. This helps reduce the configuration complexity, lower the VM performance loss, and guarantee the business security.
- Second-level RPO: Provides second-level fine-grained continuous data protection for VM instances.
- Instant recovery: Supports instant recovery with the RTO in seconds, which helps to ensure the business continuity.
- Primary storage support: The CDP service applies to VM instances in different primary storage scenarios, including local storage, NFS, SharedBlock, Ceph, and CBD.
- Flexible:
- Flexible RPO settings: Provides second/minute-level RPO settings.
- Multiple recovery levels: Supports entire recovery and file-level
recovery.
- Entire recovery: You can restore data to the original VM instance or to a newly-created VM instance.
- File-level recovery: You can retrieve files without restoring the system. Both Windows and Linux file system formats are supported.
- Flexible data display and search:
- The CDP data page displays hourly data changes of a VM instance, which provides a reference for backup capacity planning.
- The CDP data page also provides a recovery point calendar, which identifies the dates with recovery points with colors. This helps you locate a recovery point quickly.
- Reliable:
- Unified O&M: You can view the critical CDP information on the CDP overview page, including the CDP task status, recovery task status, backup server usage, and unread CDP alarms.
- You can restore CDP data to the original VM instance by creating a volume for the VM instance. All of the volumes before recovery can be retained and attached to the VM instance again, which ensures the data security to the maximum extent and facilitates post-fault analysis.
- The "Create VM instance" recovery policy allows you to create a new VM instance from the selected recovery point without affecting the original VM instance. You can finish the recovery after you confirm that the data is correct. This helps to meet the recovery drill requirements.
- You can mark and lock recovery points to retain the recovery point data in long term.
- The Cloud provides a list of recovery tasks, allowing you to view the recovery records and progress in a more friendly way.
- The RPO latency policy and alarm help to effectively relieve data transmission pressure of backup servers in heavy I/O scenarios.
Scenarios
- Continuous protection of critical business data
You can use the CDP service to continuously back up and protect critical business data, such as banking system data and financial transaction data, minimizing data loss and ensuring your business continuity.
- Anti-virus and data recovery
In case of virus invasion, such as a ransomware attack, you can use CDP to restore data to any point in time to reduce the loss caused by the virus.
Limits
Currently, if a VM instance has volumes not stored on the Ceph primary storage, you could not clone the VM instance or create snapshots or images for the VM instance during CDP.
CDP Dashboard
The overview page displays the critical CDP information on different cards, including the number and status of CDP tasks and recovery tasks, the CPU and memory utilization of backup servers, top 5 backup server usage, the total disk I/O of backup servers, and unread alarm statistics in recent 7 days.
On the main menu of ZStack Cloud, choose . Then, the CDP Dashboard page is displayed.

- CDP Task:
- This card displays the number and status of CDP tasks in the Cloud.
- Task status includes running, stopped, and other (starting, running, unknown, and failed).
- You can click the number on the card to enter the CDP task page to view more information.
- Recovery Task:
- This card displays the number and status of recovery tasks in the Cloud.
- Task status includes succeeded, failed, and other (waiting, paused, recovering, canceling, and canceled).
- You can click the number on the card to enter the recovery task page to view more information.
- Total CPU Utilization of All Backup Servers: This card displays the CPU utilization of all backup servers in the Cloud.
- Total Memory Utilization of All Backup Servers: This card displays the memory utilization of all backup servers in the Cloud.
- Top 5 Backup Server Usage:
- This card displays the used capacity and total size of each backup server.
- The usage of each backup server is displayed in descending order.
- You can click the backup server name on the card to enter the details page of the backup server.
- Total Disk I/O of All Backup Servers: This page displays the disk I/O of all backup servers in the Cloud.
- Unread Alarm Statistics in Recent Seven Days:
- This card displays unread alarm statistics in recent 7 days, including the emergency level, number of alarms, and alarm name.
- You can click the More icon in the upper right corner to enter the alarm message page.
- You can view and handle the alarm messages and copy the alarm details.
- Alarm messages that you already read are not displayed here again.
CDP Task
- Before you can use the CDP service, add a local backup server to the Cloud first.
- You can create CDP tasks to continuously back up your VM data to a specified backup server to achieve continuous data protection.
- You can create CDP tasks in bulk for multiple VM instances. The Cloud support only one VM instance per CDP task.
- You can perform entire VM backup without installing an agent for your VM instances.
- The Cloud performs a full backup on the VM instances immediately after you create CDP tasks.
- The Cloud provides second-level fine-grained continuous data protection for VM instances.
- The Cloud recommends the desired capacity required by a CDP task based on an algorithm when you create a CDP task for the first time, helping you to plan the backup space reasonably.
- The CDP service applies to VM instances in different primary storage scenarios, including local, NFS, SharedBlock, and Ceph primary storages.
- You can manage the lifecycle of CDP tasks, such as creating, enabling, disabling, and deleting CDP tasks.
- You can modify the protection policy of a CDP task, including the recovery point interval, regular backup frequency, recovery point retention policy, and the backup rate when the CDP task is disabled.
- You can modify task running policy to adjust the desired size and RPO policy for the CDP task.
- You can view the creation progress of a CDP task.
- The Cloud provides CDP task resource alarms and event alarms and allows you to create these alarms.
CDP Data
- You can back up CDP data on a local backup server.
- The Cloud displays the CDP status in charts and tables and allows you to view the details by specifying a time span.
- The Cloud displays hourly data changes so that you plan the backup capacity more reasonably.
- The Cloud provides a recovery point calendar, which identifies the dates with recovery points with colors and helps you to locate recovery points quickly.
- You can lock recovery points. After a recovery point is locked, data of the recovery point will not be automatically cleared or deleted.
- The Cloud provides recovery point list and locked recovery point list and allows you to view the details by specifying a time span.
- The Cloud supports fast recovery based on selected recovery points (including locked recovery points).
- The Cloud supports instant recovery with a minimum RTO in seconds.
- The Cloud supports entire restoration and file-level restoration.
- Entire restoration allows you to restore data to the original VM
instance or to a newly-created VM instance.
- Restore data to a newly-created VM instance:
- Allows you to create a VM instance from the selected recovery point without affecting the original VM instance.
- The newly created VM instance will quickly start up for business recovery.
- Restore to the original VM instance:
- Allows you create new volumes or overwrite current
volumes.
- Create new volumes: This method allows you to retain and attach volumes before recovery to the VM instance to ensure data security.
- Overwrite current volumes: This method will overwrite the original data in the VM instance and keep the snapshots in the current volumes.
- During data restoration, the original VM instance will quickly start up for business recovery.
- Allows you create new volumes or overwrite current
volumes.
- Restore data to a newly-created VM instance:
- File-level restoration allows you to retrieve files without restoring the system. Both Windows and Linux file system formats are supported. Supported file format include picture, text, and PDF.
- Entire restoration allows you to restore data to the original VM
instance or to a newly-created VM instance.
- Allows you to clear CDP data, which will delete all the CDP data of the VM instance, including the locked recovery points. The Cloud performs full backup for the VM instance the next time the CDP task is enabled.
Recovery Task
- The Cloud provides a list of recovery tasks, allowing you to view the recovery records and progress in a more friendly way.
- The CDP service applies to VM instances in different primary storage scenarios, including local, NFS, SharedBlock, and Ceph primary storages.
- The Cloud supports instant recovery with a minimum RTO in seconds.
- The Cloud allows you to restore data to the original VM instance or to a
newly-created VM instance.
- Restore data to a newly-created VM instance:
- Allows you to create a VM instance from the selected recovery point without affecting the original VM instance.
- The newly created VM instance will quickly start up for business recovery.
- Restore to the original VM instance:
- Allows you create new volumes or overwrite current volumes.
- Create new volumes: This method allows you to retain and attach volumes before recovery to the VM instance to ensure data security.
- Overwrite current volumes: This method will overwrite the original data in the VM instance and keep the snapshots in the current volumes.
- During data restoration, the original VM instance will quickly start up for business recovery.
- Allows you create new volumes or overwrite current volumes.
- Restore data to a newly-created VM instance:
- You can manage the lifecycle of recovery tasks, such as creating, enabling, disabling, and deleting recovery tasks.
- You can rerun a failed or canceled recovery task.
- You can cancel a task only during the recovery progress. After a task is canceled, intermediate data generated during the recovery process will not be retained.
Local Backup Server
- You can use the ImageStore deployed in the local data center as the local backup server.
- You can also deploy a new local backup server.
- You can add more than one local backup server.
- You can view the CDP data backed up to the local backup server on the details page.
Scheduled O&M
Scheduled Job
ZStack Cloud provides two types of scheduled O&M resources: scheduled jobs and schedulers. These two types of resources are independent from each other. You can create schedulers and scheduled jobs based on different rules, and associate or disassociate scheduled jobs with or from schedulers.
Scheduler
ZStack Cloud provides two types of scheduled O&M resources: scheduled jobs and schedulers. These two types of resources are independent from each other. You can create schedulers and scheduled jobs based on different rules, and associate or disassociate scheduled jobs with or from schedulers.
- A scheduled job defines that a specific action be
implemented at a specified time based on a scheduler.
- You can associate any available scheduled job with a scheduler.
- You can select Disable, Enable, Attach, and Detach actions for a scheduled job based on your actual production environments.
- If you delete a scheduler, the scheduled jobs associated with the scheduler will be disassociated. You can associate the scheduled jobs with other schedulers.
- Operations triggered by scheduled jobs are all recorded by the Audit feature.
- A scheduler is used to schedule jobs. It is suitable for
business scenarios that last for a long time.
- A scheduler defines the implementation rules for a scheduled job.
- A scheduler can be used for long-term operations, for example, creating snapshots at a specified interval for a VM instance.
- If you delete a scheduler, the scheduled jobs associated with the scheduler will be disassociated. You can associate the scheduled jobs with other schedulers.
- Operations triggered by schedulers are all recorded by the Audit feature.
Tag Management
- You can create tags with different colors, simple style, and brief description. You can also attach tags to resources and search resources by using tags. This will improve your search efficiency.
- You can search for the resources without tags by clicking the option "None" when you use tags to filter resources. This is convenient for maintenance operations.
- Two types of tag are available: admin tags and tenant tags.
- Admin tags are created and owned by the administrator, and can be attached to VM instances, volumes, hosts, baremetal instances, and elastic baremetal instances.
- Tenant tags are created and owned by tenants, and can be attached to VM instances and volumes.
- Currently, you can attach tags to or detach tags from VM instances, volumes, hosts, baremetal instances, and elastic baremetal instances.
Considerations
- Admin tags are created and owned by the administrator while tenant tags are created and owned by tenants.
- Tags created by tenants can only be attached to resources of the corresponding tenants, while admin tags can be attached to all of the resources on the Cloud.
- The administrator can detach or delete tenant tags.
- Tags in a project are owned by the project. Therefore, everyone in the project, including the project admin, project manager, and project member, can perform operations on these tags.
- Currently, tag owners cannot be changed.
- When you change a resource owner, all tenant tags attached to the resource will be detached. However, the admin tags are not affected.
- After the Cloud is upgraded seamlessly, the existing tags will be updated accordingly and displayed in the latest way. If an exception occurs, refresh your browser or create a new tag.
Migration Service
- Migrate VM instances from the vCenter that you took over to the current cloud. The supported vCenter versions include 5.5, 6.0, 6.5, 6.7, and 7.0. Note that the version of the vCenter server must be consistent with that of the ESXi host.
- Migrate VM instances from a KVM cloud platform to the current cloud.
Note: If you
took over vCenter 7.0, to ensure that the VM console can open properly, we
recommend that you download the trusted root CA certificate when you log
into vCenter.


The Migration Service is a separate feature module. To use this feature, you need to purchase both the Base License and the Plus License of the Migration Service. The Plus License cannot be used independently.
- Allows you to perform one-click V2V migrations for VM instances in bulk.
- Allows you to add a conversion host and create a V2V job and lets the Cloud do the rest.
- Allows you to configure an independent migration network and a network QoS for a conversion host to control transmission bottlenecks and improve migration efficiencies.
- Allows you to customize configurations for destination VM instances when you create a V2V job.
- Monitors and manages the entire migration process in the visualized, well-designed UI.
V2V Migration
Currently, you can migrate VM instances from a VMware cloud platform or a KVM cloud platform to the current cloud.
Source Cloud Platform: VMware
- Before migrations, perform data synchronization to manually synchronize the latest status of resources in the vCenter that you took over.
- You can perform bulk V2V migrations for VM instances, and customize configurations of the migrated VM instances.
- The supported vCenter versions include 5.0, 5.1, 5.5, 6.0, 6.5, 6.7, and 7.0. Note that the version of the vCenter server must be consistent with that of the ESXi host.
- The supported VM systems of the source vCenter include RHEL/CentOS 4.x, 5.x, 6.x, 7.x, SLES 11, 12, 15, Ubuntu 12, 14, 16, 18, Windows 7, and Windows Server 2003 R2, 2008 R2, 2012 R2, 2016, 2019.
- The VM instances will be forced to shut down during the V2V migration.
Therefore, pay attention to the business impact.
Note: The system firstly
attempts to shut down the VM instances gently. If the shutdown fails,
the system will perform force shutdown. - The type of the source primary storage is not enforced. The type of the destination primary storage can be LocalStorage, NFS, Ceph, or SharedBlock.
- For Windows VM instances, the Windows VirtIO driver is automatically installed during the migration. This improves the NIC and disk efficiencies.
- You can perform V2V migration for VM instances booted by UEFI. After the migration, these VM instances are also booted by UEFI.
Source Cloud Platform: KVM
- You can perform bulk V2V migrations for VM instances, and customize configurations of the migrated VM instances.
- You can migrate the VM instances that are running or paused. Do not power off the VM instances to be migrated.
- You can perform V2V migrations for VM instances booted by UEFI. After the migration, these VM instances are also booted by UEFI.
- The type of the source primary storage is not enforced. The type of the destination primary storage can be LocalStorage, NFS, Ceph, or SharedBlock.
- For different types of source primary storages or destination primary
storages, the libvirt version and QEMU version must meet the following
requirements:
- If either the source primary storage or destination primary storage is Ceph, use libvirt 1.2.16 and QEMU 1.1 or their later versions.
- If neither the source primary storage nor destination primary storage is Ceph, use libvirt 1.2.9 and QEMU 1.1 or their later versions.
V2V Conversion Host
- A V2V conversion host must have sufficient hardware resources, such as network
bandwidth and disk space. The following table lists the minimum configuration
requirements.
Hardware Configuration Requirements CPU Minimum 8 cores Memory Minimum 16 GB Network Minimum 1 Gigabyte NIC Storage Minimum 50 GB for the rest of storage spaces
Note: You can modify the storage
configuration according to the number of VM instances to
be migrated. - The type of the V2V conversion host must be consistent with that of the source cloud platform.
- You can set an independent migration network and a network QoS for a V2V conversion host to control transmission bottlenecks and to improve migration efficiencies.
