OpenSearch

mirror of https://github.com/honeymoose/OpenSearch.git synced 2025-02-18 19:05:06 +00:00

Author	SHA1	Message	Date
Adam Locke	358c522f16	[DOCS] Updating doc level security limitations (#64426 ) (#64660 ) * Updating doc level security limitations. * Incorporating review feedback. * Changes from review feedback. * Remove statement about the stats API.	2020-11-05 11:54:38 -05:00
Jim Ferenczi	9e4105ec37	Validate PIT on _msearch (#63167 ) This change ensures that we validate point in times provided by individual search requests in _msearch. Relates #63132	2020-11-05 15:38:28 +01:00
Dan Hermann	38ee2da564	Add configurable op_type for index watcher action (#64590 ) (#64647 )	2020-11-05 08:21:19 -06:00
Andrei Dan	a3d9408fda	Fix DataTiersUsageTransportActionTests testCalculateMAD (#64596 ) (#64628 ) Random the compression factor starting with 1 (to elimitinate nearly 0 values) which will only use one centroid (and yield 0 for MAD as the aproximate median is the same as the single centroid mean value) (cherry picked from commit 940e0f1fde0f40f99af117dd03ab0891c9eedae6) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-11-05 10:58:40 +00:00
Andras Palinkas	3d8e17f3bd	SQL: Fix incorrect parameter resolution (#63710 ) (#64615 ) Summary of the issue and the root cause: ``` (1) SELECT 100, 100 -> success (2) SELECT ?, ? (with params: 100, 100) -> success (3) SELECT 100, 100 FROM test -> Unknown output attribute exception for the second 100 (4) SELECT ?, ? FROM test (params: 100, 100) -> Unknown output attribute exception for the second ? (5) SELECT field1 as "x", field1 as "x" FROM test -> Unknown output attribute exception for the second "x" ``` There are two separate issues at play here: 1. Construction of `AttributeMap`s keeps only one of the `Attribute`s with the same name even if the `id`s are different (see the `AttributeMapTests` in this PR). This should be fixed no matter what, we should not overwrite attributes with one another during the construction of the `AttributeMap`. 2. The `id` on the `Alias`es is not the same in case the `Alias`es have the same `name` and same `child` It was considered to simpy fix the second issue by just reassigning the same `id`s to the `Alias`es with the same name and child, but it would not solve the `unknown output attribute exception` (see notes below). This PR covers the fix for the first issue. Relates to #56013	2020-11-04 20:38:00 -05:00
Adam Locke	1b9fe120d8	Applying changes for #61089 (#64601 ) (#64605 )	2020-11-04 13:45:55 -05:00
Jay Modi	4c3300bf57	Fix job scheduling for same scheduled time (#64598 ) The SchedulerEngine used by SLM uses a custom runnable that will schedule itself for its next execution if there is one to run. For the majority of jobs, this scheduling could be many hours or days away. Due to the scheduling so far in advance, there is a chance that time drifts on the machine or even that time varies core to core so there is no guarantee that the job actually runs on or after the scheduled time. This can cause some jobs to reschedule themselves for the same scheduled time even if they ran only a millisecond prior to the scheduled time, which causes unexpected actions to be taken such as what appears as duplicated snapshots. This change resolves this by checking the triggered time against the scheduled time and using the appropriate value to ensure that we do not have unexpected job runs. Relates #63754 Backport of #64501	2020-11-04 10:15:46 -07:00
Andrei Dan	2dbc444fe5	Tests: fix testMoveToStepRereadsPolicy flakiness (#64466 ) (#64593 ) testMoveToStepRereadsPolicy relied on an updated ILM policy that had a rollover condition that enabled the index to be rolled after one second. This changes the test to use a `max_doc`:1 condition so it's under the test's control to trigger the condition. (cherry picked from commit 73ab35a411bcdf5a92eb3d2b3bae5b1132a2bb56) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-11-04 15:04:01 +00:00
Andrei Dan	835ebcfff2	Convert flaky yml tests to EsRestTestCases (#63634 ) (#64581 ) The yml "Test Invalid Move To Step With Invalid Next Step" worked based on assuming the current step is a particular one. As we can't control the timing of ILM and we can't busy assert in yml test, this converts the test to a java test and makes use of `assertBusy` This converts the explain lifecycle yml tests that depende on ILM having run at least once to a java integration test that makes use of `assertBusy`. (cherry picked from commit 6afd0422ed5ff0e3a2e5661f0e6d192bdad9af4f) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-11-04 12:39:26 +00:00
Andrei Dan	ae186a6214	Tests: Disable ilm history index for doc tests (#64495 ) (#64575 ) (cherry picked from commit d9905e6991a2f5370914989336585e711dfa5bf9) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-11-04 11:44:53 +00:00
Mayya Sharipova	8c25130a80	Disable using unsigned_long in scripts (#64552 ) (#64557 ) Backport for #64523 Relates to #64361	2020-11-03 16:38:50 -05:00
Dimitris Athanasiou	e3fee3e6df	[ML] Ignore _doc_count field in data frame analytics (#64541 ) This field is added from version 7.11 onwards. We are adding it to the list of ignored fields for data frame analytics in 7.10 to avoid failing to start an outlier detection job in a mixed cluster environment. Relates #64503	2020-11-03 19:53:17 +02:00
Ignacio Vera	4851bc7bae	Upgrade to Lucene-8.7.0 (#64532 ) (#64537 )	2020-11-03 16:57:04 +01:00
David Roberts	e0d0ac86dd	[ML] Increase forecast test timeout (#64471 ) ForecastIT.testOverflowToDisk has been observed to fail a few times in FIPS JVMs because it takes longer than the permitted 30 seconds. This PR bumps the timeout up to 60 seconds. Fixes #63793	2020-11-02 14:58:18 +00:00
Tanguy Leroux	2d0bddf428	Fix CcrRepositoryIT.testCcrRepositoryFetchesSnapshotShardSizeEtc (#64228 ) (#64479 ) This test failed sometimes for various reasons: an empty bulk request that can't be validated, a background force-merge that completes after the store stats were collected and finally an assertBusy() that waits 10 seconds while we usually wait 60s on the follower cluster in CCR tests. Closes #64167	2020-11-02 15:47:07 +01:00
Marios Trivyzas	0a9481fcaf	EQL: [Tests] enable server side debugging (#64308 ) (#64449 ) Register a new task `runEqlCorrectnessNode` which enables developers to start an ES node in debug mode, properly restore the correctness data and then run queries against it. Assert the index is restored correctly and use new snapshot. (cherry picked from commit fc8c6dd56d602b4a62ee1ff484f00caab92dc6e2)	2020-10-31 11:55:39 +01:00
Tal Levy	ff829bf197	Removes yaml circuit-breaker tests for geoshape geogrid aggs (#64420 ) These tests were added to do a proper end-to-end test of the memory usage of the geotile_grid and geohash_grid aggregations on `geo_shape` fields. Although this was asperational, the truth is — the test environment does not run these aggregations in isolation. This means that the memory overhead is variable and too flaky to rely on over time. The unit tests for circuit-breaking remain. Closes #63158.	2020-10-30 08:07:12 -07:00
Jason Tedor	1126ba4df8	Serialize can contain data with roles (#64324 ) This commit internalizes whether or not a role represents the ability to contain data. In the future, this will let us remove the compatibility role notion.	2020-10-29 20:44:39 -04:00
James Rodewig	68d72f1e01	[DOCS] Fix "the the" typos (#64344 ) (#64354 )	2020-10-29 10:47:42 -04:00
Armin Braun	fada4a1c78	Fix CachedBlobContainerIndexInputTests (#64239 ) (#64348 ) Closing the input stream happens on a separate thread now that the `CacheFile` is implemented in a lock-free fashion. Closes #64215	2020-10-29 15:11:39 +01:00
James Rodewig	a2b18e9ab9	[DOCS] Fix case for 'Boolean' (#64299 ) (#64342 )	2020-10-29 10:05:57 -04:00
Andrei Stefan	a6d8319231	* Wrap a verification_exception in case there is no valid index available (#64267 ) Wrap a verification_exception in case there is no valid index available in an index_not_found_exception providing also the original index pattern that may be lost in the chain of filters involving the Security one. (cherry picked from commit 9c9da2f2f9a4ad12704f7d3a273f067e96cd2054)	2020-10-29 10:14:50 +02:00
Armin Braun	6bd8f079a7	Enhance CacheFile#invariant Assertion (#64272 ) (#64280 ) Follow up to #64180 tightening the assertion further.	2020-10-28 13:43:39 +01:00
Armin Braun	2983584ef6	Fix #invariant Assertion in CacheFile (#64180 ) (#64264 ) Fix #invariant Assertion in CacheFile closes #64141	2020-10-28 10:22:47 +01:00
jimczi	2492f48375	Fix test compilation	2020-10-28 08:58:09 +01:00
Jim Ferenczi	dcc433c971	Fix UOE when fetching flattened field (#64241 ) The new fields option allows to fetch the value of all fields in the mapping. However, internal fields that are used by some field mappers are also shown when concrete fields retrieved through a pattern (`` or `foo`). We have a [long term plan](https://github.com/elastic/elasticsearch/issues/63446) to hide these fields in field_caps and from pattern resolution so this change is just a hot fix to ensure that they don't break the retrieval in the meantime. The `flattened._keyed field will show up as an empty field when using a pattern that match the flattened field. Relates #63446	2020-10-28 08:49:03 +01:00
Jason Tedor	b46b6d5977	Fix compilation in DataTierTests.java This commit fixes a compilation issue in DataTierTests.java that was introduced due to language-level differences between 7.10/7.x and master.	2020-10-27 13:04:55 -04:00
Jason Tedor	04a9845a49	Adjust defaults for tiered data roles (#64015 ) This commit adjusts the defaults for the tiered data roles so that they are enabled by default, or if the node has the legacy data role. This ensures that the default experience is that the tiered data roles are enabled. To fully specifiy the behavior for the tiered data roles then: - starting a new node with the defaults: enabled - starting a new node with node.roles configured: enabled if and only if the tiered data roles are explicitly configured, independently of the node having the data role - starting a new node with node.data enabled: enabled unless the tiered data roles are explicitly disabled - starting a new node with node.data disabled: disabled unless the tiered data roles are explicitly enabled	2020-10-27 12:48:31 -04:00
Costin Leau	6ca0b6ae6d	EQL: Improve request logging (#64206 ) Add logging to multi-search queries Log response count (cherry picked from commit ee9b9d58f68e2d545d5d841e2f683ec4e96f79e6) (cherry picked from commit 02a4c6b83475cebe715311eeba123ad6fc8d6ba1)	2020-10-27 17:23:43 +02:00
Costin Leau	2363c4be4b	EQL: Polish testing infra (#64205 ) Add tie-breaker inside request creation Add configurable timeout (cherry picked from commit ff281d7b6fd7b4cd2f08bac49aa0b354b6812940) (cherry picked from commit 34bd76fc2987b1ad0b6275ac4358e362a0ba7fb0)	2020-10-27 17:23:43 +02:00
Henning Andersen	0cba23e08f	XPack Usage should run on MANAGEMENT threads (#64160 ) XPack usage starts out on management threads, but depending on the implementation of the usage plugin, they could end up running on transport threads instead. Fixed to always reschedule on a management thread.	2020-10-27 16:03:26 +01:00
Nhat Nguyen	566d1fd459	Return the same point in time in search response (#64188 ) With this change, we will always return the same point in time in a search response as its input until we implement the retry mechanism for the point in times.	2020-10-27 10:17:44 -04:00
Armin Braun	6f1c8136a6	Fix CachedBlobContainerIndexInputTests Shutdown (#64181 ) (#64211 ) Same problem as in #64100 we have to safely wait for all operations to go through to not leak file handles potentially in this test.	2020-10-27 14:22:43 +01:00
Armin Braun	93b52df8c1	Fix SearchableSnapshotDirectoryTests Leaking Listeners (#64150 ) (#64156 ) We need to make sure we don't shut down the pool while ref count release operations are enqueued but not yet executing.	2020-10-26 13:46:14 +01:00
David Kyle	51f8b0b8d2	Mute MonitoringWithWatcherRestIT.testThatLocalExporterAddsWatche (#64143 )	2020-10-26 10:19:48 +00:00
Armin Braun	8161513cb6	Fix Expected Exception Check in BlobstoreCacheService (#63474 ) (#64135 ) The `NodeNotConnectedException` exception can be nested as well in the fairly unlikley case of the disconnect occuring between the connected check and actually sending the request in the transport service. Closes #63233	2020-10-26 10:43:29 +01:00
Armin Braun	17843a40ef	Fix SearchableSnapshotDirectoryTests.testClearCache (#64100 ) (#64132 ) There is a small chance that the file deletion will run on the searchable snapshot thread pool and not on the test thread now that the cache is non-blocking in which case we fail the assertion unless we wait for that thread.	2020-10-26 10:27:08 +01:00
David Roberts	adc5509eda	[ML] Support the unsigned_long type in data frame analytics (#64072 ) Adds support for the unsigned_long type to data frame analytics. This type is handled in the same way as the long type. Values sent to the ML native processes are converted to floats and hence will lose accuracy when outside the range where a float can uniquely represent long values. Backport of #64066	2020-10-26 09:05:49 +00:00
Armin Braun	bd07e44c9a	Make Searchable Snapshot's CacheFile Lock less (#63911 ) (#64125 ) Replacing the mechanism for eviction and listener references via a read-write lock by a reference counting implementation. This fixes a bug that caused test failure #63586 in which concurrently trying to acquire or release an eviction listener while doing a file operation would sometimes lead to throwing an exception since the `tryLock` call on the read lock would fail in this case. Also this removes the possibility of blocking cluster state updates as a result of them waiting on the write-lock which might take a long time if a slow read operation executes concurrently. Closes #63586	2020-10-26 09:30:22 +01:00
David Roberts	cb0c538b35	[ML] Fix rare ML daily maintenance test race condition (#64043 ) Depending on thread scheduling the ML daily maintenance tests could do one more iteration than expected, causing rare failures. Fixes #64036	2020-10-22 13:03:02 +01:00
Rory Hunter	bfd2cbed86	Remove deprecation indexing code from 7.10 (#63942 ) The deprecation indexing code was writing to a regular data stream, and it is not yet possible to hide a data stream or prefix it with a period. This functionality we be re-added once it is possible to mark a data stream as hidden, and also to not rely on the standard logs template since that can be disabled.	2020-10-21 16:28:21 +01:00
James Rodewig	7551c4dc7f	[DOCS] Avoid trailing newline in apikey base64 encoding (#63720 ) (#64002 ) Co-authored-by: James Rodewig <40268737+jrodewig@users.noreply.github.com> Co-authored-by: Jurgen Braam <g.j.m.braam@gmail.com>	2020-10-21 09:44:48 -04:00
Yang Wang	c1071b1eb4	Update invalidate api key doc for the new ids field (#63860 ) (#63994 ) Follow-up for #63224 and adds the missing doc update for the new ids field.	2020-10-22 00:19:41 +11:00
Yang Wang	428fd7218c	Use asterisk instead of empty string to clear all cached entries (#63907 ) (#63989 ) The officially supported way to clearing all entries from a cache is to use wildcard of either * or _all. Though empty string has the same effect, it was never intended. Therefore the tests should not use empty string and this PR changes them to use *.	2020-10-21 23:14:01 +11:00
Hendrik Muhs	f2517678aa	[7.10][Transform] add support for unsigned_long data type (#63957 ) add support for unsigned_long, which required a change in writing out integer results properly, because coerce is not supported for unsigned_long fixes #63871 backport #63940	2020-10-20 21:05:46 +02:00
Ignacio Vera	d0f5066310	Upgrade to lucene-8.7.0-snapshot-72d8528c3a6 (#63912 ) (#63928 ) (#63933 )	2020-10-20 15:08:06 +02:00
Benjamin Trent	eff7f06ca6	[ML] fix inference binary classification predication label and feature importance (#63688 ) (#63930 ) When calculating feature importance, the leaf values directly correlate the value of the importance. Consequently, positive leaf values -> positive feature importance negative leaf values -> negative feature importance. It follows that for binary classification, this is done such that the importance relates to the leaf values, which relate directly to the "probability of class 1". So, the feature importance calculated is always for the importance as it relates to class 1. The inverse is the importance as it relates to class 0.	2020-10-20 08:50:15 -04:00
Mayya Sharipova	1287df4074	Fix max/min aggs for unsigned_long (#63904 ) Max and min aggs were producing wrong results for unsigned_long field if field was indexed. If field is indexed for max/min aggs instead of field data, we use values from indexed Points, values of which are derived using method pointReaderIfPossible. Before UnsignedLongFieldType#pointReaderIfPossible was incorrectly producing values, as it failed to shift them back to original values. This patch fixes method pointReaderIfPossible to produce correct original values. Relates to #60050	2020-10-19 15:59:55 -04:00
Adam Locke	c28c3422bb	[DOCS] [7.10] Combining important config settings into a single page (#63849 ) (#63883 ) * [DOCS] Combining important config settings into a single page (#63849) * Combining important config settings into a single page. * Updating ids for two pages causing link errors and implementing redirects. * Updating links to use IDs instead of xrefs.	2020-10-19 12:59:44 -04:00
Julie Tibshirani	f122b88bc5	Remove dependency from version plugin.	2020-10-18 14:09:32 -07:00

1 2 3 4 5 ...

6502 Commits