OpenSearch/docs/reference/upgrade/rolling_upgrade.asciidoc

[[rolling-upgrades]]
== Rolling upgrades

A rolling upgrade allows an {es} cluster to be upgraded one node at
a time so upgrading does not interrupt service. Running multiple versions of
{es} in the same cluster beyond the duration of an upgrade is
not supported, as shards cannot be replicated from upgraded nodes to nodes
running the older version.

It is best to upgrade the master-eligible nodes in your cluster after all of
the other nodes. Once you have started to upgrade the master-eligible nodes
they may form a cluster that nodes of older versions cannot join. If you
upgrade the master-eligible nodes last then all the other nodes will not be
running an older version and so they will be able to join the cluster.

Rolling upgrades are supported:

* Between minor versions
* {stack-ref-68}/upgrading-elastic-stack.html[From 5.6 to 6.8]
* From 6.8 to {version}

Upgrading directly to {version} from 6.7 or earlier requires a
<<restart-upgrade, full cluster restart>>.

include::preparing_to_upgrade.asciidoc[]

[float]
=== Upgrading your cluster

To perform a rolling upgrade to {version}:

. *Disable shard allocation*.
+
--
include::disable-shard-alloc.asciidoc[]
--

. *Stop non-essential indexing and perform a synced flush.* (Optional)
+
--
While you can continue indexing during the upgrade, shard recovery
is much faster if you temporarily stop non-essential indexing and perform a
<<indices-synced-flush-api, synced-flush>>.

include::synced-flush.asciidoc[]

--

. *Temporarily stop the tasks associated with active {ml} jobs and {dfeeds}.* (Optional)
+
--
include::close-ml.asciidoc[]
--

. [[upgrade-node]] *Shut down a single node*.
+
--
include::shut-down-node.asciidoc[]
--

. *Upgrade the node you shut down.*
+
--
include::upgrade-node.asciidoc[]
include::set-paths-tip.asciidoc[]

[[rolling-upgrades-bootstrapping]]
NOTE: You should leave `cluster.initial_master_nodes` unset while performing a
rolling upgrade. Each upgraded node is joining an existing cluster so there is
no need for <<modules-discovery-bootstrap-cluster,cluster bootstrapping>>. You
must configure <<built-in-hosts-providers,either `discovery.seed_hosts` or
`discovery.seed_providers`>> on every node.
--

. *Upgrade any plugins.*
+
Use the `elasticsearch-plugin` script to install the upgraded version of each
installed {es} plugin. All plugins must be upgraded when you upgrade
a node.

. If you use {es} {security-features} to define realms, verify that your realm
settings are up-to-date. The format of realm settings changed in version 7.0, in
particular, the placement of the realm type changed. See
<<realm-settings,Realm settings>>. 

. *Start the upgraded node.*
+
--

Start the newly-upgraded node and confirm that it joins the cluster by checking
the log file or by submitting a `_cat/nodes` request:

[source,console]
--------------------------------------------------
GET _cat/nodes
--------------------------------------------------
--

. *Reenable shard allocation.*
+
--

Once the node has joined the cluster, remove the `cluster.routing.allocation.enable`
setting to enable shard allocation and start using the node:

[source,console]
--------------------------------------------------
PUT _cluster/settings
{
  "persistent": {
    "cluster.routing.allocation.enable": null
  }
}
--------------------------------------------------
--

. *Wait for the node to recover.*
+
--

Before upgrading the next node, wait for the cluster to finish shard allocation.
You can check progress by submitting a <<cat-health,`_cat/health`>> request:

[source,console]
--------------------------------------------------
GET _cat/health?v
--------------------------------------------------

Wait for the `status` column to switch from `yellow` to `green`. Once the
node is `green`, all primary and replica shards have been allocated.

[IMPORTANT]
====================================================
During a rolling upgrade, primary shards assigned to a node running the new
version cannot have their replicas assigned to a node with the old
version. The new version might have a different data format that is
not understood by the old version.

If it is not possible to assign the replica shards to another node
(there is only one upgraded node in the cluster), the replica
shards remain unassigned and status stays `yellow`.

In this case, you can proceed once there are no initializing or relocating shards
(check the `init` and `relo` columns).

As soon as another node is upgraded, the replicas can be assigned and the
status will change to `green`.
====================================================

Shards that were not <<indices-synced-flush-api,sync-flushed>> might take longer to
recover.  You can monitor the recovery status of individual shards by
submitting a <<cat-recovery,`_cat/recovery`>> request:

[source,console]
--------------------------------------------------
GET _cat/recovery
--------------------------------------------------

If you stopped indexing, it is safe to resume indexing as soon as
recovery completes.
--

. *Repeat*
+
--

When  the node has recovered and the cluster is stable, repeat these steps
for each node that needs to be updated. You can monitor the health of the cluster
with a <<cat-health,`_cat/health`>> request:

[source,console]
--------------------------------------------------
GET /_cat/health?v
--------------------------------------------------

And check which nodes have been upgraded with a <<cat-nodes,`_cat/nodes`>> request:

[source,console]
--------------------------------------------------
GET /_cat/nodes?h=ip,name,version&v
--------------------------------------------------

--

. *Restart machine learning jobs.*
+
--
include::open-ml.asciidoc[]
--


[IMPORTANT]
====================================================

During a rolling upgrade, the cluster continues to operate normally. However,
any new functionality is disabled or operates in a backward compatible mode
until all nodes in the cluster are upgraded. New functionality becomes
operational once the upgrade is complete and all nodes are running the new
version. Once that has happened, there's no way to return to operating in a
backward compatible mode. Nodes running the previous major version will not be
allowed to join the fully-updated cluster.

In the unlikely case of a network malfunction during the upgrade process that
isolates all remaining old nodes from the cluster, you must take the old nodes
offline and upgrade them to enable them to join the cluster.

If you stop half or more of the master-eligible nodes all at once during the
upgrade then the cluster will become unavailable, meaning that the upgrade is
no longer a _rolling_ upgrade. If this happens, you should upgrade and restart
all of the stopped master-eligible nodes to allow the cluster to form again, as
if performing a <<restart-upgrade,full-cluster restart upgrade>>. It may also
be necessary to upgrade all of the remaining old nodes before they can join the
cluster after it re-forms.

====================================================
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`[[rolling-upgrades]]`
[DOCS] Added link to upgrade guide and bumped the upgrade topic up to the top level (#27621) * [DOCS] Added link to the upgrade guide & tweaked the intro. * [DOCS] Bumped upgrade topic up to the top level of the TOC 2017-12-05 13:58:52 -05:00			`== Rolling upgrades`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`A rolling upgrade allows an {es} cluster to be upgraded one node at`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`a time so upgrading does not interrupt service. Running multiple versions of`
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`{es} in the same cluster beyond the duration of an upgrade is`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`not supported, as shards cannot be replicated from upgraded nodes to nodes`
			`running the older version.`

Clearer language around upgrade sequence (#47422) 2019-10-02 04:22:02 -04:00			`It is best to upgrade the master-eligible nodes in your cluster after all of`
			`the other nodes. Once you have started to upgrade the master-eligible nodes`
			`they may form a cluster that nodes of older versions cannot join. If you`
			`upgrade the master-eligible nodes last then all the other nodes will not be`
			`running an older version and so they will be able to join the cluster.`
Clarify rolling-upgrade docs (#47279) Note to upgrade the master-eligible nodes last, and note that `cluster.initial_master_nodes` should not be set. 2019-09-30 11:58:55 -04:00
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`Rolling upgrades are supported:`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`* Between minor versions`
Update docs to refer to 6.8 instead of 6.7 (#43685) A few places in the documentation had mentioned 6.7 as the version to upgrade from, when doing an upgrade to 7.0. While this is technically possible, this commit will replace all those mentions to 6.8, as this is the latest version with the latest bugfixes, deprecation checks and ugprade assistant features - which should be the one used for upgrades. Co-Authored-By: James Rodewig <james.rodewig@elastic.co> 2019-07-02 03:06:14 -04:00			`* {stack-ref-68}/upgrading-elastic-stack.html[From 5.6 to 6.8]`
fix assumption that 6.7 is last 6.x release (#42255) 2019-05-20 15:32:44 -04:00			`* From 6.8 to {version}`
[DOCS] Adds TLS warning to rolling upgrades (#35841) 2018-11-28 12:38:58 -05:00
fix assumption that 6.7 is last 6.x release (#42255) 2019-05-20 15:32:44 -04:00			`Upgrading directly to {version} from 6.7 or earlier requires a`
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`<<restart-upgrade, full cluster restart>>.`
[DOCS] Rolling upgrade with old internal indices (#37184) Upgrading the Elastic Stack perfectly documents the process to upgrade ES from 5 to 6 when internal indices are present. However, the rolling upgrade docs do not mention anything about internal indices. This adds a warning in the rolling upgrade procedure, highlighting that internal indices should be upgraded before the rolling upgrade procedure can be started. 2019-01-09 11:14:22 -05:00
Clarify that you cannot abort an upgrade (#47342) We do mention that rolling back an upgrade requires a restore from a snapshot, but it's hidden at the bottom of the "preparing to upgrade" instructions on a different page from the actual upgrade instructions. This commit duplicates the preparatory instructions onto the pages containing the actual upgrade instructions and rewords the point about rollbacks a bit. 2019-10-02 04:27:17 -04:00			`include::preparing_to_upgrade.asciidoc[]`

			`[float]`
			`=== Upgrading your cluster`

			`To perform a rolling upgrade to {version}:`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
			`. Disable shard allocation.`
			`+`
			`--`
			`include::disable-shard-alloc.asciidoc[]`
			`--`

			`. Stop non-essential indexing and perform a synced flush. (Optional)`
			`+`
			`--`
			`While you can continue indexing during the upgrade, shard recovery`
			`is much faster if you temporarily stop non-essential indexing and perform a`
[DOCS] Reformat flush API docs (#46875) (#47230) 2019-09-30 08:42:52 -04:00			`<<indices-synced-flush-api, synced-flush>>.`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
			`include::synced-flush.asciidoc[]`

			`--`

[DOCS] Simplify ML upgrade step (#40006) 2019-03-25 13:23:10 -04:00			`. Temporarily stop the tasks associated with active {ml} jobs and {dfeeds}. (Optional)`
[DOCS] Updates methods for upgrading machine learning (#38876) (#38967) 2019-02-15 12:29:45 -05:00			`+`
			`--`
			`include::close-ml.asciidoc[]`
			`--`
[DOCS] Add X-Pack upgrade details (#29038) 2018-03-15 14:40:20 -04:00
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`. [[upgrade-node]] Shut down a single node.`
			`+`
			`--`
			`include::shut-down-node.asciidoc[]`
			`--`

			`. Upgrade the node you shut down.`
			`+`
			`--`
			`include::upgrade-node.asciidoc[]`
			`include::set-paths-tip.asciidoc[]`
Clarify rolling-upgrade docs (#47279) Note to upgrade the master-eligible nodes last, and note that `cluster.initial_master_nodes` should not be set. 2019-09-30 11:58:55 -04:00
			`[[rolling-upgrades-bootstrapping]]`
			NOTE: You should leave `cluster.initial_master_nodes` unset while performing a
			`rolling upgrade. Each upgraded node is joining an existing cluster so there is`
Drop snapshot instructions for autobootstrap fix (#49755) The "Restore any snapshots as required" step is a trap: it's somewhere between tricky and impossible to restore multiple clusters into a single one. Also add a note about configuring discovery during a rolling upgrade to proscribe any rare cases where you might accidentally autobootstrap during the upgrade. 2019-12-02 07:43:18 -05:00			`no need for <<modules-discovery-bootstrap-cluster,cluster bootstrapping>>. You`
			must configure <<built-in-hosts-providers,either `discovery.seed_hosts` or
			`discovery.seed_providers`>> on every node.
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--`

			`. Upgrade any plugins.`
			`+`
			Use the `elasticsearch-plugin` script to install the upgraded version of each
[DOCS] First pass at upgrade updates for 7.0. (#39944) * [DOCS] First pass at upgrade updates for 7.0. * [DOCS] Updates X-Pack terminology * [DOCS] Incorporated feedback from lcawl. 2019-03-13 17:38:13 -04:00			`installed {es} plugin. All plugins must be upgraded when you upgrade`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`a node.`

[DOCS] Added upgrade step for realm settings (#39672) 2019-03-12 16:02:00 -04:00			`. If you use {es} {security-features} to define realms, verify that your realm`
			`settings are up-to-date. The format of realm settings changed in version 7.0, in`
			`particular, the placement of the realm type changed. See`
			`<<realm-settings,Realm settings>>.`

WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`. Start the upgraded node.`
			`+`
			`--`

			`Start the newly-upgraded node and confirm that it joins the cluster by checking`
			the log file or by submitting a `_cat/nodes` request:

[DOCS] Replace "// CONSOLE" comments with [source,console] (#46159) (#46332) 2019-09-05 10:11:25 -04:00			`[source,console]`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--------------------------------------------------`
			`GET _cat/nodes`
			`--------------------------------------------------`
			`--`

			`. Reenable shard allocation.`
			`+`
			`--`

Remove usage of transient settings to enable allocations in rolling upgrade docs (#29671) Since we disable allocation using persistent settings, we should be consistent and remove the setting from the persistent storage. Otherwise an accidental restart will lead for shards not being allocated. Relates to #28757 2018-05-01 09:51:21 -04:00			Once the node has joined the cluster, remove the `cluster.routing.allocation.enable`
			`setting to enable shard allocation and start using the node:`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
[DOCS] Replace "// CONSOLE" comments with [source,console] (#46159) (#46332) 2019-09-05 10:11:25 -04:00			`[source,console]`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--------------------------------------------------`
			`PUT _cluster/settings`
			`{`
Remove usage of transient settings to enable allocations in rolling upgrade docs (#29671) Since we disable allocation using persistent settings, we should be consistent and remove the setting from the persistent storage. Otherwise an accidental restart will lead for shards not being allocated. Relates to #28757 2018-05-01 09:51:21 -04:00			`"persistent": {`
			`"cluster.routing.allocation.enable": null`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`}`
			`}`
			`--------------------------------------------------`
			`--`

			`. Wait for the node to recover.`
			`+`
			`--`

			`Before upgrading the next node, wait for the cluster to finish shard allocation.`
			You can check progress by submitting a <<cat-health,`_cat/health`>> request:

[DOCS] Replace "// CONSOLE" comments with [source,console] (#46159) (#46332) 2019-09-05 10:11:25 -04:00			`[source,console]`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--------------------------------------------------`
[Docs] Change example to show col headers (#40822) Command needs `?v` so user can see the column headers. Otherwise the instructions in the note about checking the init and relo columns don't make sense 2019-04-05 14:55:21 -04:00			`GET _cat/health?v`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--------------------------------------------------`

			Wait for the `status` column to switch from `yellow` to `green`. Once the
			node is `green`, all primary and replica shards have been allocated.

			`[IMPORTANT]`
			`====================================================`
			`During a rolling upgrade, primary shards assigned to a node running the new`
			`version cannot have their replicas assigned to a node with the old`
			`version. The new version might have a different data format that is`
			`not understood by the old version.`

			`If it is not possible to assign the replica shards to another node`
			`(there is only one upgraded node in the cluster), the replica`
			shards remain unassigned and status stays `yellow`.

			`In this case, you can proceed once there are no initializing or relocating shards`
			(check the `init` and `relo` columns).

			`As soon as another node is upgraded, the replicas can be assigned and the`
			status will change to `green`.
			`====================================================`

[DOCS] Reformat flush API docs (#46875) (#47230) 2019-09-30 08:42:52 -04:00			`Shards that were not <<indices-synced-flush-api,sync-flushed>> might take longer to`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`recover. You can monitor the recovery status of individual shards by`
			submitting a <<cat-recovery,`_cat/recovery`>> request:

[DOCS] Replace "// CONSOLE" comments with [source,console] (#46159) (#46332) 2019-09-05 10:11:25 -04:00			`[source,console]`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`--------------------------------------------------`
			`GET _cat/recovery`
			`--------------------------------------------------`

			`If you stopped indexing, it is safe to resume indexing as soon as`
			`recovery completes.`
			`--`

			`. Repeat`
			`+`
			`--`

			`When the node has recovered and the cluster is stable, repeat these steps`
[Docs] Enhance rolling upgrade guide (#49686) Add a couple of pointers for the user to check the overall cluster health and the version of ES running on every node. Fixes: #49670 (cherry picked from commit 8ca11f54cd839f41632c556601e94da67e91a3d1) 2019-11-28 11:00:41 -05:00			`for each node that needs to be updated. You can monitor the health of the cluster`
			with a <<cat-health,`_cat/health`>> request:

			`[source,console]`
			`--------------------------------------------------`
			`GET /_cat/health?v`
			`--------------------------------------------------`

			And check which nodes have been upgraded with a <<cat-nodes,`_cat/nodes`>> request:

			`[source,console]`
			`--------------------------------------------------`
			`GET /_cat/nodes?h=ip,name,version&v`
			`--------------------------------------------------`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
			`--`

[DOCS] Clarified that you must remove X-Pack plugin when upgrading from pre-6.3. (#32016) 2018-07-20 17:17:48 -04:00			`. Restart machine learning jobs.`
[DOCS] Updates methods for upgrading machine learning (#38876) (#38967) 2019-02-15 12:29:45 -05:00			`+`
			`--`
			`include::open-ml.asciidoc[]`
			`--`

[DOCS] Add X-Pack upgrade details (#29038) 2018-03-15 14:40:20 -04:00
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00			`[IMPORTANT]`
			`====================================================`

			`During a rolling upgrade, the cluster continues to operate normally. However,`
			`any new functionality is disabled or operates in a backward compatible mode`
Clarify rolling upgrade fallback to restart upgrade (#42161) Adds a note that restarting half-or-more of the master-eligible nodes means you're no longer doing a rolling upgrade, and may need to upgrade all the things before the cluster returns to health. 2019-05-16 13:36:09 -04:00			`until all nodes in the cluster are upgraded. New functionality becomes`
			`operational once the upgrade is complete and all nodes are running the new`
			`version. Once that has happened, there's no way to return to operating in a`
			`backward compatible mode. Nodes running the previous major version will not be`
			`allowed to join the fully-updated cluster.`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
			`In the unlikely case of a network malfunction during the upgrade process that`
Clarify rolling upgrade fallback to restart upgrade (#42161) Adds a note that restarting half-or-more of the master-eligible nodes means you're no longer doing a rolling upgrade, and may need to upgrade all the things before the cluster returns to health. 2019-05-16 13:36:09 -04:00			`isolates all remaining old nodes from the cluster, you must take the old nodes`
			`offline and upgrade them to enable them to join the cluster.`

			`If you stop half or more of the master-eligible nodes all at once during the`
			`upgrade then the cluster will become unavailable, meaning that the upgrade is`
			`no longer a _rolling_ upgrade. If this happens, you should upgrade and restart`
			`all of the stopped master-eligible nodes to allow the cluster to form again, as`
			`if performing a <<restart-upgrade,full-cluster restart upgrade>>. It may also`
			`be necessary to upgrade all of the remaining old nodes before they can join the`
			`cluster after it re-forms.`
WIP: Edits to upgrade docs (#26155) * [DOCS] Updated and edited upgrade information. * Incorporated Nik's feedback. 2017-08-23 17:03:14 -04:00
[DOCS] Changed to use transient setting to reenabled allocation. Closes #27677 2018-02-20 21:18:34 -05:00			`====================================================`