OpenSearch

Commit Graph

Author	SHA1	Message	Date
Armin Braun	6c02cf0241	Fix InternalTestCluster StopRandomNode Assertion (#44258 ) (#44265 ) * The assertion added in #44214 is tripped by tests running dedicated test clusters per test needlessly.This breaks existing tests like the one in #44245. * Closes #44245	2019-07-12 13:18:55 +02:00
Armin Braun	ad6dce16f4	Safer Shard Snapshot Delete (#44165 ) (#44244 ) * Safer Shard Snapshot Delete * We shouldn't delete the snapshot meta file before we update the index in the shard folder. If we fail to update the index-N after deleting the existing index-N is broken because the snap- blob it references is gone.	2019-07-12 12:45:06 +02:00
David Turner	735c897ec6	Avoid counting votes from master-ineligible nodes (#43688 ) Today if a master-eligible node is converted to a master-ineligible node it may remain in the voting configuration, meaning that the master node may count its publish responses as an indication that it has properly persisted the cluster state. However master-ineligible nodes do not properly persist the cluster state, so it is not safe to count these votes. This change adjusts `CoordinationState` to take account of this from a safety point of view, and also adjusts the `Coordinator` to prevent such nodes from joining the cluster. Instead, it triggers a reconfiguration to remove from the voting configuration a node that now appears to be master-ineligible before processing its join. Backport of #43688, see #44260.	2019-07-12 11:30:52 +01:00
Armin Braun	9e920f9612	Make Timestamps Returned by Snapshot APIs Consistent (#43148 ) (#44261 ) * We don't have to calculate the start and end times form the shards for the status API, we have the start time available from the CS or the `SnapshotInfo` in the repo and can either take the end time form the `SnapshotInfo` or take the most recent time from the shard stats for in progress snapshots * Closes #43074	2019-07-12 12:05:35 +02:00
Albert Zaharovits	e490ecb7d3	Fix X509AuthenticationToken principal (#43932 ) Fixes a bug in the PKI authentication. This manifests when there are multiple PKI realms configured in the chain, with different principal parse patterns. There are a few configuration scenarios where one PKI realm might parse the principal from the Subject DN (according to the `username_pattern` realm setting) but another one might do the truststore validation (according to the truststore.* realm settings). This is caused by the two passes through the realm chain, first to build the authentication token and secondly to authenticate it, and that the X509AuthenticationToken sets the principal during construction.	2019-07-12 11:04:50 +03:00
Alpar Torok	8d35583c43	Fix port range allocation with large worker IDs (#44213 ) * Fix port range allocation with large worker IDs Relates to #43983 The IDs gradle uses are incremented for the lifetime of the daemon which can result in port ranges that are outside the valid range. This change implements a modulo based formula to wrap the port ranges when the IDs get too large. Adresses #44134 but #44157 is also required to be able to close it.	2019-07-12 11:04:57 +03:00
Armin Braun	d2407d0ffc	Remove Redundant Setting of OP_WRITE Interest (#43653 ) (#44255 ) * Remove Redundant Setting of OP_WRITE Interest * We shouldn't have to set OP_WRITE interest before running into a partial write. Since setting OP_WRITE is handled by the `eventHandler.postHandling` logic, I think we can simply remove this operation and simplify/remove tests that were testing the setting of the write interest	2019-07-12 09:08:17 +02:00
Yogesh Gaikwad	91c342a888	fix and enable repository-hdfs secure tests (#44044 ) (#44199 ) Due to recent changes are done for converting `repository-hdfs` to test clusters (#41252), the `integTestSecure*` tasks did not depend on `secureHdfsFixture` which when running would fail as the fixture would not be available. This commit adds the dependency of the fixture to the task. The `secureHdfsFixture` is a `AntFixture` which is spawned a process. Internally it waits for 30 seconds for the resources to be made available. For my local machine, it took almost 45 seconds to be available so I have added the wait time as an input to the `AntFixture` defaults to 30 seconds and set it to 60 seconds in case of secure hdfs fixture. The integ test for secure hdfs was disabled for a long time and so the changes done in #42090 to fix the tests are also done in this commit.	2019-07-12 12:44:01 +10:00
Mark Vieira	263f76e5ea	Revert "[DOCS] Moves Watcher troubleshooting page (#44144 )" This reverts commit `11375926ec`.	2019-07-11 17:13:08 -07:00
Lisa Cawley	11375926ec	[DOCS] Moves Watcher troubleshooting page (#44144 )	2019-07-11 14:42:32 -07:00
James Rodewig	9ff8600d46	Revert "[DOCS] Relocate several APIs to REST APIs section (#44238 )" This reverts commit 6ebd59791afe2e0d55be2989fdbb594972237340.	2019-07-11 17:01:32 -04:00
Mark Vieira	3cd9606566	Mute failing test	2019-07-11 13:32:49 -07:00
James Rodewig	62b5b81fd2	[DOCS] Relocate several APIs to REST APIs section (#44238 )	2019-07-11 16:24:28 -04:00
Armin Braun	0dd06cf7a5	Remove Dead Code Around Snapshots (#44109 ) (#44236 ) * Just some random spots that have become unused with recent cleanups	2019-07-11 21:56:36 +02:00
Mark Vieira	5698f4dc27	Pass tests.jvms system property to test tasks for maxParallelForks (#44237 )	2019-07-11 12:41:32 -07:00
Mark Vieira	7a82106de6	Ignore test seed when flag is passed (#44234 )	2019-07-11 12:30:21 -07:00
John Murphy	8030d8f6dc	[DOCS] Add `lowercase` filter to phrase suggester example so searches are case insensitive (#44186 )	2019-07-11 15:27:31 -04:00
Mayya Sharipova	32cb47b91c	Add l1norm and l2norm distances for vectors (#44116 ) Add L1norm - Manhattan distance Add L2norm - Euclidean distance relates to #37947	2019-07-11 14:30:02 -04:00
Christoph Büscher	31725ef390	[Tests] Increase SimpleQueryStringIT allowed maxClauseCount (#44215 ) For this test, we randomize the CLUSTER_MAX_CLAUSE_COUNT on test setup (@BeforeClass) between 50 and 100. Some queries in the test generate 56 clauses which hasn't been an issue before LUCENE-8811, but we slightly need to increase the minimal possible clause count now. Closes #44192	2019-07-11 20:16:20 +02:00
Yannick Welsch	ae8f625d73	Report usages old child breakers when breaking on real memory (#44221 ) This will help in investigations where the real memory circuit breaker is tripped to better understand on what the actual memory is used, i.e. whether it's a temporary thing (e.g. requests) in contrast to more permanently allocated memory (e.g. accounting).	2019-07-11 19:52:12 +02:00
Benjamin Trent	40cc081ad3	[ML][Data Frame] adds index validations to _start data frame transform (#44191 ) (#44227 ) * [ML][Data Frame] adds index validations to _start data frame transform * addressing pr comments	2019-07-11 12:50:50 -05:00
Armin Braun	2768662822	Cleanup Stale Root Level Blobs in Sn. Repository (#43542 ) (#44226 ) * Cleans up all root level temp., snap-%s.dat, meta-%s.dat blobs that aren't referenced by any snapshot to deal with dangling blobs left behind by delete and snapshot finalization failures * The scenario that get's us here is a snapshot failing before it was finalized or a delete failing right after it wrote the updated index-(N+1) that doesn't reference a snapshot anymore but then fails to remove that snapshot * Not deleting other dangling blobs since that don't follow the snap-, meta- or tempfile naming schemes to not accidentally delete blobs not created by the snapshot logic * Follow up to #42189 * Same safety logic, get list of all blobs before writing index-N blobs, delete things after index-N blobs was written	2019-07-11 19:35:15 +02:00
Andrei Stefan	e9f9f00940	SQL: add pretty printing to JSON format (#43756 ) (#44220 ) (cherry picked from commit cbd9d4c259bf5a541bc49f65f7973174a36df449)	2019-07-11 20:02:24 +03:00
Christos Soulios	c091b6c004	Migrating tests from AvgIT integration test to AvgAggregatorTests (#44076 ) (#44225 ) This PR migrates most tests from AvgIT integration test to AvgAggregatorTests, as described in #42893	2019-07-11 19:20:13 +03:00
István Zoltán Szabó	2171b6b47f	[DOCS] Adds data frame analytics API and evaluate API resource documentation (#43972 ) This PR adds the resource documentation of the data frame analytics APIs and the evaluate API to the ML API doc pool.	2019-07-11 18:12:48 +02:00
Armin Braun	5f22370b6b	Fix ShrinkIndexIT (#44214 ) (#44223 ) * Fix ShrinkIndexIT * Move this test suit to cluster scope. Currently, `testShrinkThenSplitWithFailedNode` stops a random node which randomly turns out to be the only shared master node so the cluster reset fails on account of the fact that no shared master node survived. * Closes #44164	2019-07-11 17:58:00 +02:00
Benjamin Trent	c82d9c5b50	[ML] Adds support for regression.mean_squared_error to eval API (#44140 ) (#44218 ) * [ML] Adds support for regression.mean_squared_error to eval API * addressing PR comments * fixing tests	2019-07-11 09:22:52 -05:00
Igor Motov	1636701d69	CI: Disable SimpleQueryStringIT.testDocWithAllTypes Tracked by #44192	2019-07-11 09:22:18 -05:00
Nick Knize	374030a53f	Upgrade to lucene-8.2.0-snapshot-860e0be5378 (#44171 ) (#44184 ) Upgrades lucene library to lucene-8.2.0-snapshot-860e0be5378	2019-07-11 09:17:22 -05:00
Igor Motov	66a9b721f5	Add Map to XContentParser Wrapper (#44036 ) In some cases we need to parse some XContent that is already parsed into a map. This is currently happening in handling source in SQL and ingest processors as well as parsing null_value values in geo mappings. To avoid re-serializing and parsing the value again or writing another map-based parser this commit adds an iterator that iterates over a map as if it was XContent. This makes reusing existing XContent parser on maps possible. Relates to #43554	2019-07-11 09:38:31 -04:00
James Rodewig	f01a9eeb34	[DOCS] Rewrite `has_child` query to use new format (#44190 )	2019-07-11 09:11:26 -04:00
Alpar Torok	7ba18732f7	Run some REST tests against a cluster running in docker containers (#39515 ) * Run REST tests against a cluster running on docker Closes #38053	2019-07-11 15:28:33 +03:00
Yannick Welsch	ea5513f2cf	Make NodeConnectionsService non-blocking (#44211 ) With connection management now being non-blocking, we can make NodeConnectionsService avoid the use of MANAGEMENT threads that are blocked during the connection attempts. I had to fiddle a bit with the tests as testPeriodicReconnection was using both the mock Threadpool from the DeterministicTaskQueue as well as the real ThreadPool initialized at the test class level, which resulted in races.	2019-07-11 14:08:07 +02:00
Alpar Torok	47ab2bda72	Improve how log is tailed in testclusters on failure (#40600 ) * Improoce how log is tailed in testclusters on failure - only print last few lines - print all errors and warnings - compact repeating errors and warnings	2019-07-11 15:04:08 +03:00
surprisingb	eace735d24	Update discovery-ec2 docs (#43693 ) Fix `discovery.ec2.tag.TAGNAME` example with the correct parameter.	2019-07-11 12:59:38 +01:00
Armin Braun	51f0e941d3	Reduce Number of List Calls During Snapshot Create and Delete (#44088 ) (#44209 ) * Reduce Number of List Calls During Snapshot Create and Delete Some obvious cleanups I found when investigation the API call count metering: * No need to get the latest generation id after loading latest repository data * Loading RepositoryData already requires fetching the latest generation so we can reuse it * Also, reuse list of all root blobs when fetching latest repo generation during snapshot delete like we do for shard folders * Lastly, don't try and load `index--1` (N = -1) repository data, it doesn't exist -> just return the empty repo data initially	2019-07-11 13:52:36 +02:00
Armin Braun	8ce8c627dd	Some Cleanup in o.e.i.shard (#44097 ) (#44208 ) * Some Cleanup in o.e.i.shard * Extract one duplicated method * Cleanup obviously unused code	2019-07-11 13:52:06 +02:00
Yannick Welsch	2ee07f1ff4	Simplify port usage in transport tests (#44157 ) Simplifies AbstractSimpleTransportTestCase to use JVM-local ports and also adds an assertion so that cases like #44134 can be more easily debugged. The likely reason for that one is that a test, which was repeated again and again while always spawning a fresh Gradle worker (due to Gradle daemon) kept increasing Gradle worker IDs, causing an overflow at some point.	2019-07-11 13:35:37 +02:00
David Roberts	5886aefeed	[ML] Wait for .ml-config primary before assigning persistent tasks (#44170 ) Now that ML job configs are stored in an index rather than cluster state, availability of the .ml-config index is very important to the operation of ML. When a cluster starts up the ML persistent tasks will be considered for node assignment very early on. It is best in this case if assignment is deferred until after the .ml-config index is available. The introduction of data frame analytics jobs has made this problem worse, because anomaly detection jobs already waited for the primary shards of the .ml-state, .ml-anomalies-shared and .ml-meta indices to be available before doing node assignment, and by coincidence this would probably lead to the primary shards of .ml-config also being searchable. But data frame analytics jobs had no other index checks prior to this change. This fixes problem 2 of #44156	2019-07-11 11:43:39 +01:00
Armin Braun	c0ed64bb92	Improve Repository Consistency Check in Tests (#44204 ) * Improve Repository Consistency Check in Tests (#44099) * Check that index metadata as well as snapshot metadata always exists when referenced by other metadata * Fix SnapshotResiliencyTests on ExtraFS (#44113) * As a result of #44099 we're now checking more directories and have to ignore the `extraN` folders for those like we do for indices already * Closes #44112	2019-07-11 11:14:37 +02:00
Armin Braun	8a554f9737	Remove IncompatibleSnapshots Logic from Codebase (#44096 ) (#44183 ) * The incompatible snapshots logic was created to track 1.x snapshots that became incompatible with 2.x * It serves no purpose at this point * It adds an additional GET request to every loading of RepositoryData (from loading the incompatible snapshots blob)	2019-07-11 07:15:51 +02:00
lcawl	4e6cbc2890	[DOCS] Fixes formatting in data frame analytics API	2019-07-10 18:01:47 -07:00
Mark Vieira	7c2e4b2857	[Backport] Enable caching of rest tests which use integ-test distribution (#44181 )	2019-07-10 15:42:28 -07:00
Lisa Cawley	00b16e332d	[DOCS] Reformat rollup APIs to use new API format (#44131 )	2019-07-10 15:15:02 -07:00
Lisa Cawley	fa36f82277	[DOCS] Minor edits to data frame APIs (#44138 )	2019-07-10 14:46:03 -07:00
Lisa Cawley	aaf8ba9cb4	[DOCS] Adds frequency option to data frame transform resource (#44177 )	2019-07-10 14:45:33 -07:00
Igor Motov	df2e1fb43e	Geo: add validator that only checks altitude (#43893 ) By default, we don't check ranges while indexing geo_shapes. As a result, it is possible to index geoshapes that contain contain coordinates outside of -90 +90 and -180 +180 ranges. Such geoshapes will currently break SQL and ML retrieval mechanism. This commit removes these restriction from the validator is used in SQL and ML retrieval.	2019-07-10 16:55:03 -04:00
Ryan Ernst	8fda49a834	Remove unused import in TransportShardBulkAction Accidentally left from backporting #44092	2019-07-10 13:43:47 -07:00
Christoph Büscher	cbb19032df	[Test] Additional logging for RemoteClusterClientTests (#44124 )	2019-07-10 22:41:54 +02:00
Armin Braun	d6f09fdb97	Add WARN Logging if Mock Network Accepts Huge Number of Connections (#44169 ) (#44182 ) * Add WARN Logging if Mock Network Accepts Huge Number of Connections * As discussed, added warn logging to rule out endless accept loops for #43387 * Had to handle it by the relatively awkward override in the mock nio because we don't have logging in the NIO module where (`ServerChannelContext` lives)	2019-07-10 22:08:36 +02:00

... 2 3 4 5 6 ...

46843 Commits All Branches Search

46843 Commits

All Branches