Starting, stopping, restarting, and refreshing Cloudera Embedded Container Service Clusters
Provides information about how you can manage your Cloudera Data Services on premises Cloudera Embedded Container Service clusters.
Starting a Cloudera Embedded Container Service Cluster
Procedure to start an Cloudera Embedded Container Service cluster.
- Navigate to the ECS service, tab, click the Actions menu to the right of the Embedded Container Service cluster name and select Start.
- Click the Start button that appears in the next screen to confirm. The Command Details window shows the progress of starting services.
Stopping a Cloudera Embedded Container Service Cluster
Procedure to stop an Cloudera Embedded Container Service cluster.
- Navigate to the ECS service, tab, click the menu to the right of the Cloudera Embedded Container Service cluster name and select Stop.
- Click the Stop button in the confirmation screen. The Command Details window shows the progress of stopping services.
Stopping the Cloudera Embedded Container Service Agent causes yunikorn deployment to scale down to 0
To overcome this scenario, you must stop the ECS Agent at the role level.
- Log into your Cloudera Manager instance.
- Navigate to the ECS service.
- Click the Instances tab.
- Locate the instance that needs to be stopped and click ECS Agent in the Role Type column.
- On the ECS Role page, open the Actions menu.
- Select Stop this ECS Agent.
Before restarting the High Availability enabled Cloudera Embedded Container Service Cluster
You must review the following information prior to restarting your High Availability (HA) enabled Cloudera Embedded Container Service cluster.
About Full Cluster Restart
When a full cluster restart is required (rolling restart is not recommended), do not restart all roles at once. Employ one of the procedures.
Option A: Start Primary Master First, then All Others
- Start the primary master (the bootstrap node that existed before other nodes were added).
- Start all other roles (For example, additional masters, workers, and others).
Option B: Restart Non-Master First, then Rolling Restart Masters
- Restart all roles except master nodes (For example, workers, agents).
- Perform a rolling restart after selecting all the master nodes.
| Scenario | Procedure |
|---|---|
| Full cluster start | Start primary master first, then all other roles |
| Full cluster restart | Restart non-masters first, then rolling restart after selecting all master nodes |
Rolling Restart of an Cloudera Embedded Container Service Cluster
Procedure to perform the rolling restart.
Restarting a Cloudera Embedded Container Service Cluster
Procedure to restart your Cloudera Embedded Container Service cluster.
- Navigate to the ECS service, tab, click the menu to the right of the cluster name and select Restart.
- Click the Restart button that appears in the next screen to confirm.
- Click
Configuring Restart for an Cloudera Embedded Container Service cluster
Follow these instructions to configure a cluster restart.
Refreshing a Cloudera Embedded Container Service Cluster
Refresh your Cloudera Embedded Container Service cluster with a single step process.
Terminate All RKE2 Processes on Cloudera Embedded Container Service Hosts
The Terminate all RKE2 processes commands stop RKE2 completely on one Cloudera Embedded Container Service host or on all Cloudera Embedded Container Service hosts
in a service. Cloudera Manager runs the same
rke2-killall.sh script shipped with the Cloudera Embedded Container Service parcel, with confirmation and logging.
Use these commands when a normal Stop or Bring Down does not fully terminate RKE2—for example, when containerd shims, Pods, or systemd units leave processes running.
- Scope: One Cloudera Embedded Container Service Server or ECS Agent role
- UI path: Clusters → cluster → ECS → role → Actions → Terminate all RKE2 processes.
- What it does:
- Brings down the selected role (bring-down failures are ignored).
- On the role’s host, stops RKE2
systemdunits if present (Cloudera Embedded Container Service 1.5.5 SP4+). - Runs
rke2-killall.shto terminate all remaining RKE2-related processes.
- Scope: Entire Cloudera Embedded Container Service.
- UI path: Clusters → cluster → ECS → Actions → Terminate all RKE2 processes on all ECS hosts
- What it does:
- Brings down the Cloudera Embedded Container Service (bring-down failures are ignored).
- Runs the per-role killall command on every Cloudera Embedded Container Service server and agent host in parallel.
- Affected roles and the Cloudera Embedded Container Service are stopped in Cloudera Manager.
- RKE2 is not running on the affected host(s).
- To restore the cluster, run Bring Up on the Cloudera Embedded Container Service (or individual roles) through Cloudera Manager.
