PolarDBScheduled O&M events, such as database software upgrades and hardware maintenance and upgrades, are notified through text messages, voice calls, emails, or internal messages, and are also displayed in the console. On the Scheduled Events page, you can view the event type, task ID, cluster name, switchover time, and other details, and manually change the switchover time.
Notes
-
Events are classified into two levels of urgency:
-
[S0 Urgent] Risk Remediation: Typically involves unexpected issues requiring immediate action to prevent service failures. Notifications for these events are sent three days or less in advance, and the window for rescheduling the switchover time is shorter. Examples include urgent version replacement and upgrades, host failure recovery, and upgrades for expiring SSL certificates.
-
[S1 Scheduled] System Maintenance: Involves low-risk issue fixes or planned software and hardware upgrades. Notifications for these events are usually sent more than three days in advance, and you can cancel the event.
-
-
To receive notifications for scheduled O&M events, log on to the Message Center, and select the notification methods and configure contacts for ApsaraDB Fault or Maintenance Notifications. We recommend adding your database O&M personnel as contacts. Otherwise, you will not receive event notifications.Notification methods include Email and Internal Messages. We recommend you select Email to ensure delivery.
In the upper-right corner of the console, click the bell icon. Then, in the Internal Message Notifications panel, click Message Settings.
Figure 1. Accessing Message Settings

Figure 2. Notification settings for ApsaraDB
-
To receive real-time updates on O&M events or implement custom, event-driven automation, configure system event subscriptions in CloudMonitor. ApsaraDB pushes system events to CloudMonitor for each stage of the O&M event lifecycle, such as Scheduled, In Progress, Succeeded, and Canceled. For more information, see Manage event subscriptions (Recommended). For a list of subscribable CloudMonitor events, see Appendix 1: CloudMonitor-related system events.
Procedure
-
Log on to the console for your database product.
-
In the navigation pane on the left, click . In the top navigation bar, select the region where your instance is located.
-
The Scheduled Events page shows the details of scheduled events. By default, events with the Planned status are displayed. To view historical events, click the Completed or Canceled tab. The following table describes the event attributes.
Attribute
Example
Description
Event type
Risk remediation
Events are classified by urgency into two types: Risk remediation and system maintenance.
Status
Pending
The scheduling status of the event. Note the following key statuses:
-
Waiting Setting Time: The execution time is not set. You must set a time based on your business requirements. If you do not set a time by the Deadline, the event is automatically canceled.
-
Pending: The event is waiting to enter the scheduling phase at the scheduled Start time.
-
Executing: The event is in progress. You cannot manually intervene at this stage. To request an emergency stop, submit a ticket. Note that such non-standard operations can pose unknown risks.
-
Successful: The event has completed.
-
Canceled: The event execution failed or was canceled. Common reasons for cancellation include:
-
Canceled by user (UserCancel): The user canceled the event in the console or by calling an API operation.
-
User response timeout (UserResponseTimeout): The event was automatically canceled because the execution time was not set before the Deadline.
-
Canceled by control system (SupervisorCancel): The event initiator canceled the event.
-
Canceled as unnecessary (AvoidCancel): The event is no longer required because the risk has been resolved or the instance's current state makes the event obsolete. For example, the instance is already on the latest version and requires no update.
-
Canceled by system (AutoCancel): The system periodically inspects scheduled events. The event is canceled if the instance does not meet the execution prerequisites, for example, if the instance is in an abnormal state and cannot receive commands.
-
Execution timeout (ExecuteTimeout): The event entered the execution queue but did not finish within the expected time.
-
Execution failure (ExecuteFail): The event failed during execution due to an unknown error.
-
Event type
minor version update
The specific name of the event type. For more information, see Event types and impacts.
Cause
-
For more information, see Appendix 2: Detailed cause codes and cancellation risks.
Business impact
Transient connection interruption
The business impact varies depending on the event. For more information, see Event types and impacts.
O&M suggestions
Ensure your application has an automatic reconnection mechanism and is designed to handle the potential business impact.
O&M suggestions vary depending on the event. For more information, see Appendix 1: CloudMonitor-related system events.
Start time
-
The time when the event enters the scheduling queue. Before the start time, the event does not affect the instance. After the start time, your database remains accessible, but you cannot perform instance-level operations such as changing the configuration or migrating across availability zones. This field is empty if the event status is Waiting Setting Time.
Scheduled switchover time
-
The time when a primary/secondary switchover or link switchover occurs, if applicable. This is typically when a transient connection interruption happens. The switchover is expected to occur around this estimated time. In rare cases, such as a failback to the original availability zone, a second switchover may occur.
NoteA preparation period is required before the switchover for tasks such as event scheduling and data preparation. Therefore, a time gap exists between the start time and the switchover time. This gap varies by database product and event type.
Deadline
-
The latest time by which you can set the switchover time. The adjusted switchover time cannot be later than this Deadline.
Cancelable
Yes
You can cancel the event to prevent it from running. This option is typically available for "system maintenance" events.
ImportantScheduled events are typically issued by the cloud database management system during regular inspections. If you cancel an event, a new one may be issued in the next inspection cycle. Frequent cancellations can increase risks. To avoid these risks, schedule the event for a suitable time based on your business needs instead of canceling it. For more information about cancellation risks, see Appendix 2: Detailed cause codes and cancellation risks.
Reschedulable
Yes
You can change the execution time for most events. In rare cases involving urgent, high-risk fixes, the time window may not be long enough to allow for rescheduling.
-
-
(Optional) Modify a scheduled event.
Select the event that you want to reschedule and click Modify Scheduled Event. You can modify the event in one of two ways:
-
Immediate execution: The event's start time is set to the current time, and it enters the execution queue immediately.
-
Switchover at a specified time: Select a suitable time for the switchover from the available range. The start time is automatically calculated based on the switchover time. The new start time cannot be earlier than the current time. Otherwise, the modification fails.
-
-
(Optional) Modify the recurring time window.
In the upper-right corner of the event list, click Recurring Time Window Settings to open the configuration page.
The execution time of a scheduled event is typically calculated based on the instance's maintenance window. For more information, see the documentation about setting the maintenance window for RDS, Tair, MongoDB, and PolarDB. You can also define a custom recurring time window to meet your O&M requirements. The cloud database service then prioritizes this custom window when scheduling new events.
You can set the window on a weekly or monthly basis. For example, if you set the recurring switchover time to 02:00–03:00 on every Monday and Tuesday, and the platform's overall event window is from this Tuesday to the following Sunday, the system will identify both 02:00–03:00 on this Tuesday and 02:00–03:00 on next Monday as valid switchover times. Typically, the earlier time (this Tuesday) is prioritized.
Important-
This setting applies only to new events. To change the time of an existing event in the list, click Configure Execution Time.
-
This setting is only an aid for calculating the execution time and applies only to "system maintenance" events. The actual execution time is subject to the time displayed in the event list.
-
This is an account-level setting. Once configured, it applies to all database products that support recurring time windows.
-
-
(Optional) Cancel a scheduled event.
Select the event you want to cancel and click Cancel Scheduled Event. After reviewing the risks, click Confirm to cancel the event.
Event types and impacts
|
Event type |
Impact type |
Impact description |
|
Cluster migration Note
In a scheduled O&M operation that is initiated due to host risks, hardware warranty expiration, or operating system upgrades, the system migrates the cluster to a new server node. This applies to non-high-availability clusters and read-only clusters. |
Transient cluster disconnection |
After the scheduled switchover time begins, the following impacts occur: Note
Pending events typically trigger a cluster switchover, which is performed during the cluster maintenance window after the scheduled switchover time.
|
|
Primary/secondary switchover Note
In a scheduled O&M operation that is initiated due to host risks, hardware warranty expiration, or operating system upgrades, the system triggers a primary/secondary node switchover. This applies only to high-availability clusters. |
||
|
Cluster parameter adjustment Note
In a scheduled O&M operation that is initiated due to a known parameter risk, the system modifies the cluster parameters. If the delivered parameters include parameters that require a restart, the cluster is restarted. |
||
|
Host risk remediation Note
Remediate the fault risks on the hosts to which the cluster belongs. |
||
|
SSL certificate renewal Note
This operation is initiated when the SSL certificate of the cluster is about to expire, so that the cluster can continue to provide better security and stability. |
||
|
Backup mode upgrade Note
To provide faster backup and restoration capabilities for the cluster, the backup mode of the cluster is switched from logical backup to physical database and table backup. |
||
|
Zone migration Note
Upgrade and transform the physical infrastructure of some older regions and zones. |
||
|
Minor version upgrade Note
To improve user experience, the database service releases minor versions for clusters on an irregular basis to enrich product features or fix known defects. |
Transient cluster disconnection |
After the scheduled switchover time begins, the following impacts occur: Note
Pending events typically trigger a cluster switchover, which is performed during the cluster maintenance window after the scheduled switchover time.
|
|
Differences between minor versions |
Different minor versions (kernel versions) include different updates. Pay attention to the differences between the minor version after the upgrade and the current minor version. For more information, see the minor version release notes of the relevant product: Engine parameters |
|
|
Proxy minor version upgrade Note
To improve user experience, the database service releases minor versions for proxy nodes on an irregular basis to enrich proxy service features or fix known defects. |
Transient cluster disconnection |
After the scheduled switchover time begins, the following impacts occur: Note
Pending events typically trigger a cluster switchover, which is performed during the cluster maintenance window after the scheduled switchover time.
|
|
Differences between minor versions |
Different minor versions include different updates. Pay attention to the differences between the minor version after the upgrade and the current minor version. For more information, see the minor version release notes of the relevant product: PolarProxy release notes |
|
|
Network upgrade Note
Upgrade network hardware to improve the network performance and stability of the cluster. |
Transient cluster disconnection |
After the scheduled switchover time begins, the following impacts occur: Note
Pending events typically trigger a cluster switchover, which is performed during the cluster maintenance window after the scheduled switchover time.
|
|
Impact on VIP direct connection |
Some network upgrades may involve cross-zone migration, which changes the virtual IP address (VIP) of the cluster. If a client uses the VIP to connect to the database, the connection is interrupted. Note
To avoid impacts, use the endpoint in domain name format that is provided by the cluster, and disable DNS caching on your application and its host server. |
|
|
Storage gateway upgrade Note
Upgrade the storage gateway to improve the storage performance and stability of the cluster. |
I/O jitter |
Transient I/O jitter or increased SQL latency may occur. The impact lasts for no more than 3 seconds. |
|
Enable seamless migration Note
Enable seamless migration to improve user experience. |
Parameter modification |
No impact. Note
No restart or migration is involved. Your current business is not affected. |
|
Proxy migration Note
The host of the proxy is upgraded or maintained to improve the stability of the proxy node. |
Proxy node migration |
During the proxy node migration, the cluster endpoint and custom endpoint experience a transient disconnection that lasts for no more than 10 seconds. |
FAQ
Notifications
Start time and switchover time
Event operations
Other questions
Appendix
Appendix 1: CloudMonitor system events
|
Event code |
Event name |
Trigger condition |
Recommendations |
|
Instance:SystemMaintenance.MinorVersionUpgrade:Scheduled |
Minor version update (scheduled) |
A minor version update is scheduled. |
The event has not started. Instance availability is not affected. |
|
Instance:SystemMaintenance.MinorVersionUpgrade:ReminderNotice |
Minor version update (reminder) |
Reminders are sent 7, 3, and 1 days before execution starts. |
This is an informational notification. The event has not yet started. |
|
Instance:SystemMaintenance.MinorVersionUpgrade:Executing |
Minor version update (executing) |
The minor version update begins. |
To prevent unexpected issues, avoid manual intervention at this stage. |
|
Instance:SystemMaintenance.MinorVersionUpgrade:Executed |
Minor version update (completed) |
The minor version update is complete. |
The event completed successfully. A primary/secondary switchover may occur. Monitor for any business impact. |
|
Instance:SystemMaintenance.MinorVersionUpgrade:Canceled |
Minor version update (canceled) |
The minor version update fails or is canceled. |
The event failed or was automatically canceled, for example, because the instance is already on the latest version. Instance availability is unaffected. |
|
Instance:SystemMaintenance.Transfer:ReminderNotice |
Instance migration (reminder) |
Reminders are sent 7, 3, and 1 days before execution starts. |
This is an informational notification. The event has not yet started. |
|
Instance:SystemMaintenance.Transfer:Scheduled |
Instance migration (scheduled) |
An instance migration is scheduled. |
The event has not started. Instance availability is not affected. |
|
Instance:SystemMaintenance.Transfer:Executing |
Instance migration (executing) |
The instance migration begins. |
To prevent unexpected issues, avoid manual intervention at this stage. |
|
Instance:SystemMaintenance.Transfer:Executed |
Instance migration (completed) |
The instance migration is complete. |
The event completed successfully. A primary/secondary switchover may occur. Monitor for any business impact. |
|
Instance:SystemMaintenance.Transfer:Canceled |
Instance migration (canceled) |
The instance migration fails or is canceled. |
The event failed or was automatically canceled, for example, because the instance had already been manually migrated. Instance availability is unaffected. |
|
Instance:SystemMaintenance.ScheduledOperation:ReminderNotice |
Scheduled event (reminder) |
Reminders are sent 7, 3, and 1 days before execution starts. |
This is an informational notification. The event has not yet started. |
|
Instance:SystemMaintenance.ScheduledOperation:Scheduled |
Scheduled event (scheduled) |
A scheduled event is scheduled. |
The event has not started. Instance availability is not affected. |
|
Instance:SystemMaintenance.ScheduledOperation:Executing |
Scheduled event (executing) |
The scheduled event begins. |
To prevent unexpected issues, avoid manual intervention at this stage. |
|
Instance:SystemMaintenance.ScheduledOperation:Executed |
Scheduled event (completed) |
The scheduled event is complete. |
The event completed successfully. A primary/secondary switchover may occur. Monitor for any business impact. |
|
Instance:SystemMaintenance.ScheduledOperation:Canceled |
Scheduled event (canceled) |
The scheduled event fails or is canceled. |
Instance availability is unaffected. |
Appendix 2: Reason codes and cancellation risks
|
Cause code |
Cause description |
Risk code |
Risk description |
Notes |
Frequency |
|
InfraArchUpgrade |
Replacement or upgrade of the underlying infrastructure architecture. |
OutOfGoodPerfByHardwareUpgrade |
You cannot benefit from the improved performance and stability of the upgraded software. |
We upgrade or migrate instances to improve service quality and stability as our underlying compute, storage, and network architecture evolves. |
Monthly/Quarterly |
|
EnhanceStabilityAndResUtil |
Improves instance stability and resource utilization. |
ImpactStabAndResContention |
Affects instance stability. Potential impacts include resource contention, engine vulnerabilities, and lower-than-expected performance. |
- |
Irregularly |
|
EnableHotReplica |
Enabling seamless migration significantly improves the high availability of your database instance. This feature also speeds up maintenance tasks like instance scaling, minor version upgrades, and migrations, without affecting your services. |
- |
- |
- |
- |
|
KernelExceptionRepair |
Fixes instance exceptions caused by the engine. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
This is common for urgent risk remediation in an engine version. |
Irregularly |
|
OldKernelVersionWithHardwareUpgrade |
The outdated engine version is upgraded along with hardware resources. |
KernelVersionEndOfLife |
The engine version's lifecycle has ended. The instance cannot use new features or benefit from performance optimizations. |
This is common for routine version updates and upgrades. |
Monthly/Quarterly |
|
KernelBugFix |
Fixes engine vulnerabilities. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
This is common for urgent bug fixes in an engine version. |
Irregularly |
|
HostLoadHigh |
Excessive host load. |
HostLoadHighAffectStability |
Excessive host load affects the performance and stability of the instance. |
This is common for mitigating host hardware risks. |
Irregularly |
|
SoftwareUpgrade |
Host software upgrade. |
OutOfGoodPerfByHardwareUpgrade |
You cannot benefit from the improved performance and stability of the upgraded software. |
A cold upgrade of the host's operating system or dependent plugins. |
Monthly/Quarterly |
|
HardwareUpgrade |
Replacement or upgrade of the underlying hardware. |
OutOfGoodPerfBySoftwareUpgrade |
You cannot benefit from the improved performance and stability of the upgraded software. |
Host hardware upgrade. |
Monthly/Quarterly |
|
HostSoftHardwareUpgrade |
Host software or hardware upgrade. |
OutOfGoodPerfBySoftHardwareUpgrade |
You cannot benefit from the improved performance and stability of the upgraded software. |
Host software and hardware upgrade. |
Monthly/Quarterly |
|
ProxyNodeHostMaintain |
Maintenance or upgrade of the proxy node's host. |
Proxy node host maintenance/upgrade |
Host risks affect the stability of the proxy node. |
- |
- |
|
HostCPUException |
Host CPU exception. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
- |
Irregularly |
|
HostMemException |
Host memory exception. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
- |
Irregularly |
|
HostDiskException |
Host disk exception. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
- |
Irregularly |
|
KernelVersionWithServerlessUpgrade |
The engine is upgraded, and the instance moves from public preview to general availability (GA). |
BetaVersionEndOfLife |
The lifecycle of the public preview version has ended. The instance cannot use new features or benefit from performance optimizations. |
- |
Monthly/Quarterly |
|
ParamRiskRepairOrOptimize |
Fixes or optimizes parameters to address associated risks. |
UnknownRisks |
May lead to unknown risks. |
This is common for automatic tuning events triggered by suboptimal parameter settings in a cloud database. |
Monthly/Quarterly |
|
PGOldKernelVersionWithHardwareUpgrade |
The outdated engine version is upgraded along with the hardware resources. This may change the database port and cross-database connection string. The TimescaleDB, PostGIS, and Ganos plugins will also be upgraded to their latest versions, as older versions are incompatible. |
KernelVersionEndOfLife |
The engine version's lifecycle has ended. The instance cannot use new features or benefit from performance optimizations. |
- |
Monthly/Quarterly |
|
MaxScaleExceptionRepair |
Remediates risks in the proxy component. |
RiskEscatateToFailure |
This risk can escalate to a failure, impacting instance availability. |
This is common for urgent risk remediation in a proxy service version. |
Irregularly |
|
OriginalNetWorkHasFlawWithSqlTimeoutAndDIsconnection |
The current network type has flaws that cause slow SQL query timeouts and intermittent disconnections. This upgrade improves stability by resolving these issues. |
FlawNotResolvedAndAbnormalConnectionMayOccur |
If not resolved, these network flaws may cause connection issues. |
- |
Irregularly |
Appendix 3: Event types
|
Event type |
Description |
|
instance migration |
instance migration |
|
minor version update |
minor version update |
|
network upgrade |
network upgrade |
|
primary/secondary switchover |
primary/secondary switchover |
|
SSL certificate update |
SSL certificate update |
|
proxy minor version update |
proxy minor version update |
|
instance parameter modification |
instance parameter modification |
|
major version update |
major version update |
|
live migration |
live migration |
|
proxy migration |
proxy migration |
Related APIs
|
API |
Description |
|
Queries the number of pending events by task type. |
|
|
Modifies the task switchover time of a pending event. |
|
|
Queries the details of a pending event. |