OpenSearch

Commit Graph

Author	SHA1	Message	Date
Armin Braun	51f0e941d3	Reduce Number of List Calls During Snapshot Create and Delete (#44088 ) (#44209 ) * Reduce Number of List Calls During Snapshot Create and Delete Some obvious cleanups I found when investigation the API call count metering: * No need to get the latest generation id after loading latest repository data * Loading RepositoryData already requires fetching the latest generation so we can reuse it * Also, reuse list of all root blobs when fetching latest repo generation during snapshot delete like we do for shard folders * Lastly, don't try and load `index--1` (N = -1) repository data, it doesn't exist -> just return the empty repo data initially	2019-07-11 13:52:36 +02:00
Armin Braun	8ce8c627dd	Some Cleanup in o.e.i.shard (#44097 ) (#44208 ) * Some Cleanup in o.e.i.shard * Extract one duplicated method * Cleanup obviously unused code	2019-07-11 13:52:06 +02:00
Yannick Welsch	2ee07f1ff4	Simplify port usage in transport tests (#44157 ) Simplifies AbstractSimpleTransportTestCase to use JVM-local ports and also adds an assertion so that cases like #44134 can be more easily debugged. The likely reason for that one is that a test, which was repeated again and again while always spawning a fresh Gradle worker (due to Gradle daemon) kept increasing Gradle worker IDs, causing an overflow at some point.	2019-07-11 13:35:37 +02:00
David Roberts	5886aefeed	[ML] Wait for .ml-config primary before assigning persistent tasks (#44170 ) Now that ML job configs are stored in an index rather than cluster state, availability of the .ml-config index is very important to the operation of ML. When a cluster starts up the ML persistent tasks will be considered for node assignment very early on. It is best in this case if assignment is deferred until after the .ml-config index is available. The introduction of data frame analytics jobs has made this problem worse, because anomaly detection jobs already waited for the primary shards of the .ml-state, .ml-anomalies-shared and .ml-meta indices to be available before doing node assignment, and by coincidence this would probably lead to the primary shards of .ml-config also being searchable. But data frame analytics jobs had no other index checks prior to this change. This fixes problem 2 of #44156	2019-07-11 11:43:39 +01:00
Armin Braun	c0ed64bb92	Improve Repository Consistency Check in Tests (#44204 ) * Improve Repository Consistency Check in Tests (#44099) * Check that index metadata as well as snapshot metadata always exists when referenced by other metadata * Fix SnapshotResiliencyTests on ExtraFS (#44113) * As a result of #44099 we're now checking more directories and have to ignore the `extraN` folders for those like we do for indices already * Closes #44112	2019-07-11 11:14:37 +02:00
Armin Braun	8a554f9737	Remove IncompatibleSnapshots Logic from Codebase (#44096 ) (#44183 ) * The incompatible snapshots logic was created to track 1.x snapshots that became incompatible with 2.x * It serves no purpose at this point * It adds an additional GET request to every loading of RepositoryData (from loading the incompatible snapshots blob)	2019-07-11 07:15:51 +02:00
lcawl	4e6cbc2890	[DOCS] Fixes formatting in data frame analytics API	2019-07-10 18:01:47 -07:00
Mark Vieira	7c2e4b2857	[Backport] Enable caching of rest tests which use integ-test distribution (#44181 )	2019-07-10 15:42:28 -07:00
Lisa Cawley	00b16e332d	[DOCS] Reformat rollup APIs to use new API format (#44131 )	2019-07-10 15:15:02 -07:00
Lisa Cawley	fa36f82277	[DOCS] Minor edits to data frame APIs (#44138 )	2019-07-10 14:46:03 -07:00
Lisa Cawley	aaf8ba9cb4	[DOCS] Adds frequency option to data frame transform resource (#44177 )	2019-07-10 14:45:33 -07:00
Igor Motov	df2e1fb43e	Geo: add validator that only checks altitude (#43893 ) By default, we don't check ranges while indexing geo_shapes. As a result, it is possible to index geoshapes that contain contain coordinates outside of -90 +90 and -180 +180 ranges. Such geoshapes will currently break SQL and ML retrieval mechanism. This commit removes these restriction from the validator is used in SQL and ML retrieval.	2019-07-10 16:55:03 -04:00
Ryan Ernst	8fda49a834	Remove unused import in TransportShardBulkAction Accidentally left from backporting #44092	2019-07-10 13:43:47 -07:00
Christoph Büscher	cbb19032df	[Test] Additional logging for RemoteClusterClientTests (#44124 )	2019-07-10 22:41:54 +02:00
Armin Braun	d6f09fdb97	Add WARN Logging if Mock Network Accepts Huge Number of Connections (#44169 ) (#44182 ) * Add WARN Logging if Mock Network Accepts Huge Number of Connections * As discussed, added warn logging to rule out endless accept loops for #43387 * Had to handle it by the relatively awkward override in the mock nio because we don't have logging in the NIO module where (`ServerChannelContext` lives)	2019-07-10 22:08:36 +02:00
Ryan Ernst	c6efb9be2a	Convert ReplicationResponse to Writeable (#43953 ) This commit convers ReplicationResponse and all its subclasses to support Writeable.Reader as a constructor. relates #34389	2019-07-10 12:45:10 -07:00
Ryan Ernst	fb77d8f461	Removed writeTo from TransportResponse and ActionResponse (#44092 ) The base classes for transport requests and responses currently implement Streamable and Writeable. The writeTo method on these base classes is implemented with an empty implementation. Not only does this complicate subclasses to think they need to call super.writeTo, but it also can lead to not implementing writeTo when it should have been implemented, or extendiong one of these classes when not necessary, since there is nothing to actually implement. This commit removes the empty writeTo from these base classes, and fixes subclasses to not call super and in some cases implement an empty writeTo themselves. relates #34389	2019-07-10 12:42:04 -07:00
Alpar Torok	bde5802ad6	Test fixtures improovements (#43956 ) * Test fixtures improovements Don't disable some of the precommit tasks on fixtures. This no longer makes sense now that a project can both produce and use a fixture. In order for this to be possible, had to add an additional configuration to make JarHell class accessible to the task even if it's not a dependency of the project and fix some of the third party audit fallout from #43671 which wasn't detected at the time due to the issue being fixed here. Closes #43918	2019-07-10 21:21:06 +03:00
Lisa Cawley	5b71340f99	[DOCS] Moves Watcher limitations (#44141 )	2019-07-10 11:17:12 -07:00
Zachary Tong	92ad588275	Remove generic on AggregatorFactory (#43664 ) (#44079 ) AggregatorFactory was generic over itself, but it doesn't appear we use this functionality anywhere (e.g. to allow the super class to declare arguments/return types generically for subclasses to override). Most places use a wildcard constraint, and even when a concrete type is specified it wasn't used. But since AggFactories are widely used, this led to the generic touching many pieces of code and making type signatures fairly complex	2019-07-10 13:20:28 -04:00
Nhat Nguyen	b158919542	Do not use mock engine in PrimaryAllocationIT (#44083 ) PrimaryAllocationIT#testForceStaleReplicaToBePromotedToPrimary relies on the flushing when a shard is no long assigned. This behavior, however, can be randomly disabled in MockInternalEngine. Closes #44049	2019-07-10 12:26:34 -04:00
David Roberts	07f53e39b3	[ML] Fix ML memory tracker lockup when inner step fails (#44158 ) When the ML memory tracker is refreshed and a refresh is already in progress the idea is that the second and subsequent refresh requests receive the same response as the currently in progress refresh. There was a bug that if a refresh failed then the ML memory tracker's view of whether a refresh was in progress was not reset, leading to every subsequent request being registered to receive a response that would never come. This change makes the ML memory tracker pass on failures as well as successes to all interested parties and reset the list of interested parties so that further refresh attempts are possible after either a success or failure. This fixes problem 1 of #44156	2019-07-10 15:46:46 +01:00
James Rodewig	4cbd028960	[DOCS] Correct `ignore_unmapped` parm typo for nested query	2019-07-10 10:10:14 -04:00
Andrei Stefan	bb3e5351b5	SQL: double quotes escaping bug fix (#43829 ) (cherry picked from commit d589dcad18c3708913e13c757b91c846aeb35bb4)	2019-07-10 16:05:22 +03:00
James Rodewig	1ae0db7053	[DOCS] Rewrite nested query to use new format (#44130 )	2019-07-10 08:52:04 -04:00
David Turner	d0f1a756d9	Comment on the extra reroute after failing shards (#44152 ) The `ShardFailedClusterStateTaskExecutor` fails some shards, which performs a reroute, but then sometimes schedules a followup reroute. It's not clear from the code why this followup is necessary, so this commit adds a short comment describing why it's necessary.	2019-07-10 13:24:21 +01:00
David Roberts	cad804df92	[TEST] Mute ShrinkIndexIT Due to https://github.com/elastic/elasticsearch/issues/44164	2019-07-10 13:22:25 +01:00
Albert Zaharovits	018d946bba	[DOC] Backup & Restore Security Configuration (#42970 ) This commit documents the backup and restore of a cluster's security configuration. It is not possible to only backup (or only restore) security configuration, independent to the rest of the cluster's conf, so this describes how a full configuration backup&restore will include security as well. Moreover, it explains how part of the security conf data resides on the special .security index and how to backup that using regular data snapshot API. Co-Authored-By: Lisa Cawley <lcawley@elastic.co> Co-Authored-By: Tim Vernum <tim@adjective.org>	2019-07-10 14:53:56 +03:00
Andrei Stefan	9567f337f5	SQL: handle SQL not being available in a more graceful way (#43665 ) * Add test for SQL not being available error message in JDBC. * Add a new qa sub-project that explicitly disables SQL XPack module in Gradle. (cherry picked from commit 8a1ac8a3a88a325ec9b99963e0fa288c18ee0ee5)	2019-07-10 14:36:24 +03:00
Andrei Stefan	9957b66d2e	SQL: concrete indices array size bug fix (#43878 ) * The created array didn't have the correct initial size while attempting to resolve multiple indices (cherry picked from commit 341006e9913e831408f5bbc7f8ad8c453a7f630e)	2019-07-10 14:36:23 +03:00
Przemysław Witek	44781e415e	[7.x] [ML] Add DatafeedTimingStats to datafeed GetDatafeedStatsAction.Response (#43045 ) (#44118 )	2019-07-10 11:51:44 +02:00
David Roberts	853ddb5a07	[ML] Fix custom timestamp override with dot-separated fractional seconds (#44127 ) Custom timestamp overrides provided to the find_file_structure endpoint produced an invalid Grok pattern if the fractional seconds separator was a dot rather than a comma or colon. This commit fixes that problem and adds tests for this sort of timestamp override. Fixes #44110	2019-07-10 10:27:08 +01:00
Martijn van Groningen	913b6a64e8	Replace Streamable w/ Writable for MultiSearchRequest (#44057 ) This commit replaces usages of Streamable with Writeable for the MultiSearchRequest class. I ran into this when developing a custom action that reuses MultiSearchRequest in the enrich branch. Relates to #34389	2019-07-10 11:13:28 +02:00
David Roberts	cb62d4acdf	[ML-DataFrame] Add a frequency option to transform config, default 1m (#44120 ) Previously a data frame transform would check whether the source index was changed every 10 seconds. Sometimes it may be desirable for the check to be done less frequently. This commit increases the default to 60 seconds but also allows the frequency to be overridden by a setting in the data frame transform config.	2019-07-10 09:59:00 +01:00
Adrien Grand	64ff895a32	Add 7.3 release notes. (#44010 )	2019-07-10 09:36:51 +02:00
Armin Braun	a23d1ed00d	Mute SearchWithRandomExceptionsIT (#44147 ) (#44149 ) * This is failing quiete often and we can reproduce it now so we don't need additional test logging on CI * Relates #40435	2019-07-10 08:12:26 +02:00
David Turner	aec44fecbc	Decouple DiskThresholdMonitor & ClusterInfoService (#44105 ) Today the `ClusterInfoService` requires the `DiskThresholdMonitor` at construction time so that it can notify it when nodes report changes in their disk usage, but this is awkward to construct: the `DiskThresholdMonitor` requires a `RerouteService` which requires an `AllocationService` which comees from the `ClusterModule` which requires the `ClusterInfoService`. Today we break the cycle with a `LazilyInitializedRerouteService` which is itself a little ugly. This commit replaces this with a more traditional subject/observer relationship between the `ClusterInfoService` and the `DiskThresholdMonitor`.	2019-07-09 18:43:32 +01:00
David Turner	e70cad4c52	Remove node conn block after connection barrier (#44114 ) Today `testOnlyBlocksOnConnectionsToNewNodes` fails (extremely rarely) if the last attempt to connect to `node0` is delayed for so long that the test runs `nodeConnectionsBlocks.clear()` before the connection attempt obtains the expected connection block. We can turn this into a reliable failure with this delay: ```diff diff --git a/server/src/main/java/org/elasticsearch/cluster/NodeConnectionsService.java b/server/src/main/java/org/elasticsearch/cluster/NodeConnectionsService.java index f48413824d3..9a1d0336bcd 100644 --- a/server/src/main/java/org/elasticsearch/cluster/NodeConnectionsService.java +++ b/server/src/main/java/org/elasticsearch/cluster/NodeConnectionsService.java @@ -300,6 +300,13 @@ public class NodeConnectionsService extends AbstractLifecycleComponent { private final Runnable connectActivity = () -> threadPool.executor(ThreadPool.Names.MANAGEMENT).execute(new AbstractRunnable() { @Override protected void doRun() { + + try { + Thread.sleep(500); + } catch (InterruptedException e) { + throw new AssertionError("unexpected", e); + } + assert Thread.holdsLock(mutex) == false : "mutex unexpectedly held"; transportService.connectToNode(discoveryNode); consecutiveFailureCount.set(0); ``` This commit reverts the extra logging introduced in #43979 and fixes this failure by waiting for the connection attempt to hit the barrier before removing it. Fixes #40170	2019-07-09 17:03:26 +01:00
Mark Vieira	a406ef1d38	Fix eclipse project file generation (#44080 )	2019-07-09 08:55:37 -07:00
marcos ramos	88ee47c9ba	Fix OIDC documentation settings (#44115 ) Current kibana setting is xpack.security.auth.oidc.realm, but the correct one is xpack.security.authc.oidc.realm	2019-07-09 18:44:35 +03:00
David Kyle	23d7e309da	Mute put job docs test Relates to #43271	2019-07-09 13:23:31 +01:00
Sachin Frayne	389c923a82	[Docs] Fix json syntax in watcher compare condition (#44032 )	2019-07-09 13:43:18 +02:00
David Turner	268971db03	Wait for blackholed connection before discovery (#44077 ) Since #42636 we no longer treat connections specially when simulating a blackholed connection. This means that at the end of the safety phase we may have just started a connection attempt which will time out, but the default timeout is 30 seconds, much longer than the 2 seconds we normally allow for post-safety-phase discovery. This commit adds time for such a connection attempt to time out. It also fixes some spurious logging of `this` that now refers to an object with an unhelpful `toString()` implementation introduced in #42636. Fixes #44073	2019-07-09 10:59:53 +01:00
Henning Andersen	748a10866d	Reindex ScrollableHitSource pump data out (#43864 ) Refactor ScrollableHitSource to pump data out and have a simplified interface (callers should no longer call startNextScroll, instead they simply mark that they are done with the previous result, triggering a new batch of data). This eases making reindex resilient, since we will sometimes need to rerun search during retries. Relates #43187 and #42612	2019-07-09 11:50:09 +02:00
Henning Andersen	859709cc94	Closed index noop recovery during upgrade (#44072 ) Test that closed indices do noop recovery during rolling upgrade.	2019-07-09 11:46:42 +02:00
David Turner	fd9eebae81	Only apply initial recovery filter to shrunk shard (#44054 ) Today the `index.routing.allocation.initial_recovery._id` setting can only be set on indices that are the result of a shrink, but the filtered allocation decider also applies this filter to shards with a recovery source of `EMPTY_STORE`. The only way to have this setting set while the recovery source is `EMPTY_STORE` is to force-allocate an empty primary, but such a forced allocation ignores this allocation decider. This commit simplifies the allocation decider so that the `initial_recovery` setting only applies to shards with a recovery source of `LOCAL_SHARDS`.	2019-07-09 08:42:18 +01:00
Armin Braun	9eac5ceb1b	Dry up inputstream to bytesreference (#43675 ) (#44094 ) * Dry up Reading InputStream to BytesReference * Dry up spots where we use the same pattern to get from an InputStream to a BytesReferences	2019-07-09 09:18:25 +02:00
Armin Braun	f1ebb82031	Update the gcs chunk_size documentation. (#38749 ) (#44098 ) Remove `1g` from the examples, as the GCS repository chunk_size can be at most 100m.	2019-07-09 09:18:03 +02:00
Armin Braun	dc8f8e40eb	Fix DedicatedClusterSnapshotRestoreIT testSnapshotWithStuckNode (#43537 ) (#44082 ) * Fix DedicatedClusterSnapshotRestoreIT testSnapshotWithStuckNode * See comment in the test: The problem is that when the snapshot delete works out partially on master failover and the retry fails on `SnapshotMissingException` no repository cleanup is run => we still failed even with repo cleanup logic in the delete path now * Fixed the test by rerunning a create snapshot and delete loop to clean up the repo before verifying file counts * Closes #39852	2019-07-09 06:32:08 +02:00
Lisa Cawley	94578a8b47	[DOCS] Defines data frame transform resources (#43996 ) Co-Authored-By: István Zoltán Szabó <istvan.szabo@elastic.co>	2019-07-08 17:53:00 -07:00

1 2 3 4 5 ...

46658 Commits All Branches Search

46658 Commits

All Branches