OpenSearch

mirror of https://github.com/honeymoose/OpenSearch.git synced 2025-02-20 03:45:02 +00:00

Author	SHA1	Message	Date
David Turner	9ba897fbd6	Random iterations in testDataOnlyNodePersistence (#56906 ) PR #56893 was supposed to randomise the iteration count in `testDataOnlyNodePersistence` but this change was mistakenly omitted. This commit addresses this.	2020-05-18 15:16:22 +01:00
David Turner	64280b489b	Fix testDataOnlyNodePersistence (#56893 ) This test failed if all 1000 top-level `rarely()` calls in the loop returned `false`, because then we would never set the term of the persisted state. This commit fixes this by adding an earlier call to `persistedState#setCurrentTerm`. It also changes the test to clean up the threadpools it starts whether it passes or fails.	2020-05-18 13:57:36 +01:00
James Rodewig	74554f1ae8	[DOCS] Add put snapshot repo API docs (#56827 ) (#56900 )	2020-05-18 08:55:22 -04:00
Benjamin Trent	297f864884	[ML] relax throttling on expired data cleanup (#56711 ) (#56895 ) Throttling nightly cleanup as much as we do has been over cautious. Night cleanup should be more lenient in its throttling. We still keep the same batch size, but now the requests per second scale with the number of data nodes. If we have more than 5 data nodes, we don't throttle at all. Additionally, the API now has `requests_per_second` and `timeout` set. So users calling the API directly can set the throttling. This commit also adds a new setting `xpack.ml.nightly_maintenance_requests_per_second`. This will allow users to adjust throttling of the nightly maintenance.	2020-05-18 08:46:42 -04:00
David Kyle	0fac152188	Muse AsyncSearchActionIT (#56897 ) For #56765	2020-05-18 13:36:33 +01:00
Ioannis Kakavas	bb852ab2e7	Cause is tracked in #49094 (#56887 )	2020-05-18 15:03:38 +03:00
Francisco Fernández Castaño	8ab9fc10c1	Track multipart/resumable uploads GCS API calls (#56892 ) Add tracking for multipart and resumable uploads for GoogleCloudStorage. For resumable uploads only the last request is taken into account for billing, so that's the only request that's tracked. Backport of #56821	2020-05-18 13:39:26 +02:00
David Kyle	52a329fa12	Mute sql.client.VersionTests suite (#56883 ) For #56882	2020-05-18 10:15:30 +01:00
Armin Braun	e75a6f13a1	Stop Redundantly Serializing ShardId in BulkShardResponse (#56094 ) (#56866 ) When reading/writing the individual doc responses in the context of a bulk shard response there is no need to serialize the `ShardId` over and over. This can waste a lot of memory when handling large bulk requests.	2020-05-17 10:27:17 +02:00
Armin Braun	c02850f335	Fix S3ClientSettings Leak (#56703 ) (#56862 ) Fixes the fact that repository metadata with the same settings still results in multiple settings instances being cached as well as leaking settings on closing a repository. Closes #56702	2020-05-17 09:18:20 +02:00
Armin Braun	31f54c934e	Relax Assertion About SnapshotsService Listeners (#56608 ) (#56863 ) This assertion is too strict. A snapshot will be removed from the cluster state on the CS thread before it is removed from the listeners map on the snapshot thread pool. Throughout the removal from the cluster state and listener map, the snapshot is tracked in `endingSnapshots` though, so we can relax the assertion accordingly and are still able to catch leaked listeners. Closes #56607	2020-05-17 09:17:41 +02:00
Armin Braun	b9614558b9	Fix SnapshotStatusApisIT (#56859 ) (#56861 ) In the unlikely event that the data nodes started snapshotting the shards already (and hence got blocked on the data blobs) before the master has applied the cluster state to its own `SnapshotsService` on the CS applier thread, we can get a `SnapshotMissingException` here which breaks the busy assert loop so we have to deal with it explicitly. Closes #56858	2020-05-16 21:50:25 +02:00
Armin Braun	cac85a6f18	Shorter Path in Netty ByteBuf Unwrap (#56740 ) (#56857 ) In most cases we are seeing a `PooledHeapByteBuf` here now. No need to redundantly create an new `ByteBuffer` and single element array for it here when we can just directly unwrap its internal `byte[]`.	2020-05-16 11:54:36 +02:00
Bogdan Pintea	de7dd6154e	Fix range of version number generation in test (#56849 ) The version number componenent can't equal or exceed the revision multiplier. This fixes a the VersionTests unit test. (cherry picked from commit 7d2331a2818ae20024c5c3617cd4433f90e9c098)	2020-05-16 08:59:45 +02:00
Andrei Stefan	4d47d63f55	SQL: implement SUM, MIN, MAX, AVG over literals (#56786 ) (#56850 ) * Adds support for MIN, MAX, AVG, SUM aggregates acting on literals. SELECT SUM(1) FROM index and SELECT SUM(1), AVG(2) work both on indices and as local execution. (cherry picked from commit efb72907c0391612c4a2b6256e327060b4167912)	2020-05-16 02:13:55 +03:00
Jake Landis	813609b47c	Ensure that .watcher-history-11* template is in installed prior to use (#56734 ) WatcherIndexTemplateRegistry as of https://github.com/elastic/elasticsearch/pull/52962 requires all nodes to be on 7.7.0 before it allows the version 11 index template to be installed. While in a mixed cluster, nothing prevents Watcher from running on the new host before the all of the nodes are on 7.7.0. This will result in the .watcher-history-11* index without the proper mappings. Without the proper mapping a single document (for a large watch) can exceed the default 1000 field limit and cause error to show in the logs. This commit ensures the same logic for writing to the index is applied as for installing the template. In a mixed cluster, the `10` index template will continue to be written. Only once all of nodes are on 7.7.0+ will the `11` index template be installed and used. closes #56732	2020-05-15 16:29:04 -05:00
James Rodewig	e492c23944	[DOCS] Sort metric and pipeline agg docs (#56613 ) (#56846 ) Co-authored-by: Gil Raphaelli <gil@elastic.co>	2020-05-15 17:15:53 -04:00
Tim Brooks	195a5247d4	Prevent connection races in testEnsureWeReconnect (#56654 ) Currently it is possible that a sniff connection round is occurring as we enter another test loop in testEnsureWeReconnect. The problem is that once we enter another loop, closing the connection manually can cause this pre-existing connection round to fail. This round failing can fail the test. This commit fixes the issue by ensuring that there are no in-progress connections before entering another loop.	2020-05-15 14:58:46 -06:00
Gabriel Petrovay	cb4d5f5042	Fixed calendar intervals documentation (#56666 ) - the 1-letter intervals are not parseable (`m`, `h`, `d`, `w`, `M`, `q`, `y`) - fixed formatting broken by new lines	2020-05-15 16:55:57 -04:00
Nik Everett	f3e962707b	Mute TaskManagerTests#testTrackingChannelTask It fails sometimes. Tracked by #56746.	2020-05-15 16:48:33 -04:00
Nik Everett	7b626826eb	Fix sum test It was relying on the compensated sum working but the test framework was dodging it. This forces the accuracy tests to come from a single shard where we get the proper compensated sum. Closes #56757	2020-05-15 16:16:30 -04:00
Jason Tedor	da833d6cd3	Use settings infrastructure for shards and replicas (#56801 ) We get the number of shards and replicas with our bare hands in index metadata, rather than letting the settings infrastructure do the work for us. This commit switches to using the settings infrastructure.	2020-05-15 15:59:30 -04:00
David Turner	a3e845cbad	Suppress cluster UUID logs in 6.8/7.x upgrade (#56835 ) Today a 7.x node logs `cluster UUID set to [...]` on every cluster state update received from a 6.8 master, because 6.8 nodes are not able to commit the cluster UUID properly. We could try and deduplicate these logs somehow, but that would introduce a good deal of complexity. Instead, this commit suppresses these logs entirely when receiving cluster state updates from a 6.8 master.	2020-05-15 19:45:32 +01:00
Dimitris Athanasiou	54d3cc74ec	[7.x][ML] Ensure class is represented when its cardinality is low (#56783 ) (#56829 ) In DF analytics classification, it is possible to use no samples of a class if its cardinality is too low. This commit fixes this by ensuring the target sample count can never be zero. Backport of #56783	2020-05-15 20:52:06 +03:00
Bogdan Pintea	14ad733bd1	SQL: JDBC: fix access to the Manifest for non-entry JAR URLs (#56797 ) (#56839 ) * JDBC: fix access to the Manifest for non-entry JAR The JDBC driver will attempt to read its version from the Manifest file embedded into its JAR. The URL pointing to the JAR can be provided in a few ways. So far, accessing the Manfiest was attempted by getting a URLConnection out of the URL and then getting an input stream out of this connection. For file JAR URLs, this only works however if the URL points to the driver as a JAR file entry (i.e. <sub-url>!/jdbc-driver.jar!/). If that's not the case, the JarURLConnection will throw an IOException. This commit fixes that: in case the URL points to a JAR entry (jar:file:<path>/jdbc-driver.jar!/), the manifest is read directly with JarURLConnection#getManifest(). (cherry picked from commit 2175b7b01cf5fcf3ab2bb21404a9bd454a8df3f0) Co-authored-by: Elastic Machine <elasticmachine@users.noreply.github.com>	2020-05-15 19:35:54 +02:00
James Baiera	4809db3ff9	EnrichProcessorFactory should not throw NPE if missing metadata (#55977 ) (#56793 ) In some cases the Enrich processor factory may be called before it is ready to create processors. While these calls are usually made in error, the response from the Enrich processor is an NPE which is almost always an unhelpful error when debugging an issue.	2020-05-15 12:02:13 -04:00
Andrei Dan	c8278e333a	Enable decompression of response within LowLevelRestClient (#55413 ) (#56820 ) Added support for decompression at LLRC and added integration test (cherry picked from commit 2621452473e0c236aa28db749f782a24eca6c974) Signed-off-by: Andrei Dan <andrei.dan@elastic.co> Co-authored-by: Hakky54 <hakangoudberg@hotmail.com>	2020-05-15 16:50:45 +01:00
James Rodewig	c50f86fbba	[DOCS] EQL: Document `case_sensitive` param (#56697 ) (#56818 )	2020-05-15 11:47:19 -04:00
Dan Hermann	66871c5342	[7.x] Rename endpoint from plural "_data_streams" to singular "_data_stream" (#56825 )	2020-05-15 10:27:53 -05:00
Ioannis Kakavas	239ada1669	Test adjustments for FIPS 140 (#56526 ) This change aims to fix our setup in CI so that we can run 7.x in FIPS 140 mode. The major issue that we have in 7.x and did not have in master is that we can't use the diagnostic trust manager in FIPS mode in Java 8 with SunJSSE in FIPS approved mode as it explicitly disallows the wrapping of X509TrustManager. Previous attempts like #56427 and #52211 focused on disabling the setting in all of our tests when creating a Settings object or on setting fips_mode.enabled accordingly (which implicitly disables the diagnostic trust manager). The attempts weren't future proof though as nothing would forbid someone to add new tests without setting the necessary setting and forcing this would be very inconvenient for any other case ( see #56427 (comment) for the full argumentation). This change introduces a runtime check in SSLService that overrides the configuration value of xpack.security.ssl.diagnose.trust and disables the diagnostic trust manager when we are running in Java 8 and the SunJSSE provider is set in FIPS mode.	2020-05-15 18:10:45 +03:00
Dan Hermann	2a21d4d976	Docs for data stream REST APIs	2020-05-15 09:37:45 -05:00
James Rodewig	5e09762a27	[DOCS] EQL: Align comments in `between` fn examples	2020-05-15 09:20:45 -04:00
James Rodewig	24cd45345e	[DOCS] EQL: Remove references to arrays/multi-value fields (#56772 )	2020-05-15 09:09:07 -04:00
Benjamin Trent	f71c305090	[7.x] [Transform] add support for terms agg in transforms (#56696 ) (#56809 ) * [Transform] add support for terms agg in transforms (#56696) This adds support for `terms` and `rare_terms` aggs in transforms. The default behavior is that the results are collapsed in the following manner: `<AGG_NAME>.<BUCKET_NAME>.<SUBAGGS...>...` Or if no sub aggs exist `<AGG_NAME>.<BUCKET_NAME>.<_doc_count>` The mapping is also defined as `flattened` by default. This is to avoid field explosion while still providing (limited) search and aggregation capabilities.	2020-05-15 08:08:43 -04:00
David Roberts	270a23e422	[TEST] Fix log tail mocking in native process unit tests (#56804 ) This is a followup to #56632. Tests that had to be changed to mock the C++ log handler more accurately need to be more careful about when that stream ends, as ending of that stream is used to detect crashes in the production system. Fixes #56796	2020-05-15 12:46:37 +01:00
Alan Woodward	d33d13f2be	Simplify generics on Mapper.Builder (#56747 ) Mapper.Builder currently has some complex generics on it to allow fluent builder construction. However, the second parameter, a return type from the build() method, is unnecessary, as we can use covariant return types. This commit removes this second generic parameter.	2020-05-15 12:14:49 +01:00
Francisco Fernández Castaño	1530bff0cb	Move azure client logic from AzureStorageService to AzureBlobStore (#56806 ) Backport of #56782	2020-05-15 11:30:15 +02:00
David Turner	27a090232e	Suppress Kerberos tests on JDK15 (#56767 ) Somewhat convoluted AwaitsFix for #56507 that only applies on JDK15.	2020-05-15 07:41:04 +01:00
Yang Wang	c66e7ecbfe	Fix test failure of file role store auto-reload (#56398 ) (#56802 ) Ensure assertion is only performed when we can be sure that the desired changes are picked up by the file watcher.	2020-05-15 15:10:45 +10:00
Ryan Ernst	9fb80d3827	Move publishing configuration to a separate plugin (#56727 ) This is another part of the breakup of the massive BuildPlugin. This PR moves the code for configuring publications to a separate plugin. Most of the time these publications are jar files, but this also supports the zip publication we have for integ tests.	2020-05-14 20:23:07 -07:00
Tal Levy	5e90ff32f7	Add Normalize Pipeline Aggregation (#56399 ) (#56792 ) This aggregation will perform normalizations of metrics for a given series of data in the form of bucket values. The aggregations supports the following normalizations - rescale 0-1 - rescale 0-100 - percentage of sum - mean normalization - z-score normalization - softmax normalization To specify which normalization is to be used, it can be specified in the normalize agg's `normalizer` field. For example: ``` { "normalize": { "buckets_path": <>, "normalizer": "percent" } } ```	2020-05-14 17:40:15 -07:00
Lee Hinman	a73d7d9e2b	[7.x] Don't allow invalid template combinations (#56397 ) (#56795 ) Backports the following commits to 7.x: - Don't allow invalid template combinations (#56397)	2020-05-14 16:20:53 -06:00
Mark Vieira	0fd756d511	Enforce strict license distribution requirements (#56642 )	2020-05-14 13:57:56 -07:00
Jake Landis	a22aabcc15	[7.x] Reduce chance for test failure due to schedule (#56633 ) (#56695 ) If CI is running tests at exactly 0 or 5 minutes past the hour the ack-watch docs tests may fail with a 409 error if the ack test happens to run at the exact time that the schedule watch is running. This commit changes the public documentation (and the test) for the ack to a feb 29th at noon schedule. Test doc or tests do not really care about the schedule date and this is chosen since it is a valid date, but one that is extremely unlikely to cause issues.	2020-05-14 15:52:04 -05:00
James Rodewig	2a943a58a4	[DOCS] EQL: Document `number` function (#56770 ) Co-authored-by: Ross Wolf <31489089+rw-access@users.noreply.github.com>	2020-05-14 15:44:04 -04:00
Costin Leau	6f4af43405	EQL: Skip execution for filters with empty results (#56718 ) Optimize away events queries and joins/sequence that cannot match any results without having to query the backend. (cherry picked from commit 69c8ef8cfefd8fc6dcb6d1a566bfcd537068e3e4)	2020-05-14 22:38:23 +03:00
Armin Braun	14a042fbe5	Make No. of Transport Threads == Available CPUs (#56488 ) (#56780 ) We never do any file IO or other blocking work on the transport threads so no tangible benefit can be derived from using more threads than CPUs for IO. There are however significant downsides to using more threads than necessary with Netty in particular. Since we use the default setting for `io.netty.allocator.useCacheForAllThreads` which is `true` we end up using up to `16MB` of thread local buffer cache for each transport thread. Meaning we potentially waste CPUs * 16MB of heap for unnecessary IO threads in addition to obvious inefficiencies of artificially adding extra context switches.	2020-05-14 21:33:46 +02:00
Mark Tozzi	b718193a01	Clean up DocValuesIndexFieldData (#56372 ) (#56684 )	2020-05-14 12:42:37 -04:00
Nhat Nguyen	044ee380e8	Use ConcurrentSet in testTrackingChannelTask (#56775 ) We need to use a ConcurrentSet to track the canceled tasks as cancelTaskAndDescendants can be called concurrently. Closes #56746	2020-05-14 12:22:59 -04:00
Dimitris Athanasiou	ac5902624c	[7.x][ML] Improve error upon DF analytics mappings conflict (#56700 ) (#56776 ) Adds the conflicting types and an example of an index which specifies them in order to make it easier for the user to understand the conflict. Backport of #56700	2020-05-14 19:16:10 +03:00

1 2 3 4 5 ...

51709 Commits