OpenSearch

Commit Graph

Author	SHA1	Message	Date
GeChenxin	a96f526de1	Add index name to refresh mapping task (#57598 )	2020-06-17 10:49:36 -04:00
Benjamin Trent	2de242f80e	[ML] rename EnsembleSizeInfo#inputFieldNameLengths to this.featureNameLengths (#58241 ) (#58253 )	2020-06-17 10:08:55 -04:00
Armin Braun	41af7f5455	Fix Typo in Snapshot Abort Test (#58238 ) (#58247 ) Forgot the brackets here in #58214 so in the rare case where the first update seen by the listener doesn't match it will still remove itself and never be invoked again -> timeout.	2020-06-17 14:53:39 +02:00
Nik Everett	ab2c6d9696	Save memory when auto_date_histogram is not on top (backport of #57304 ) (#58190 ) This builds an `auto_date_histogram` aggregator that natively aggregates from many buckets and uses it when the `auto_date_histogram` used to use `asMultiBucketAggregator` which should save a significant amount of memory in those cases. In particular, this happens when `auto_date_histogram` is a sub-aggregator of a multi-bucketing aggregator like `terms` or `histogram` or `filters`. For the most part we preserve the original implementation when `auto_date_histogram` only collects from a single bucket. It isn't possible to "just port the aggregator" without taking a pretty significant performance hit because we used to rewrite all of the buckets every time we switched to a coarser and coarser rounding configuration. Without some major surgery to how to delay sub-aggs we'd end up rewriting the delay list zillions of time if there are many buckets. The multi-bucket version of the aggregator has a "budget" of "wasted" buckets and only rewrites all of the buckets when we exceed that budget. Now that we don't rebucket every time we increase the rounding we can no longer get an accurate count of the number of buckets! So instead the aggregator uses an estimate of the number of buckets to trigger switching to a coarser rounding. This estimate is likely to be terrible when buckets are far apart compared to the rounding. So it also uses the difference between the first and last bucket to trigger switching to a coarser rounding. Which covers for the shortcomings of the bucket estimation technique pretty well. It also causes the aggregator to emit fewer buckets in cases where they'd be reduced together on the coordinating node. This is wonderful! But probably fairly rare. All of that does buy us some speed improvements when the aggregator is a child of multi-bucket aggregator: Without metrics or time zone: 25% faster With metrics: 15% faster With time zone: 22% faster Relates to #56487	2020-06-17 08:48:41 -04:00
Benjamin Trent	69338b03d7	[ML] expand data_streams when assigning datafeed to node (#58175 ) (#58242 )	2020-06-17 08:34:34 -04:00
Ignacio Vera	2d3d7ab387	mute CentroidCalculatorTests#testPolygonAsPoint (#58249 ) (#58250 )	2020-06-17 14:32:13 +02:00
Jason Tedor	b78b3edeea	Upgrade to JNA 5.5.0 (#58183 ) This commit bumps our JNA dependency from 4.5.1 to 5.5.0, so that we are now on the latest maintained line, and pick up a large collection of bug fixes that have accumulated.	2020-06-17 07:35:08 -04:00
Ignacio Vera	b6585f2b51	Add new extensions for Lucene86 points codec to FsDirectoryFactory (#58226 ) (#58233 )	2020-06-17 12:55:33 +02:00
Dimitris Athanasiou	36dbf08d47	[7.x][ML] Improve stability of stratified splitter tests (#58180 ) (#58224 ) The main improvement here is that the total expected count of training rows in the test is calculated as the sum of the training fraction times the cardinality of each class (instead of the training fraction times the total doc count). Also relaxes slightly the error bound on the uniformity test from 0.12 to 0.13. Closes #54122 Backport of #58180	2020-06-17 12:40:21 +03:00
Armin Braun	85be78b624	Fix Snapshot Abort Not Waiting for Data Nodes (#58214 ) (#58228 ) This was a really subtle bug that we introduced a long time ago. If a shard snapshot is in aborted state but hasn't started snapshotting on a node we can only send the failed notification for it if the shard was actually supposed to execute on the local node. Without this fix, if shard snapshots were spread out across at least two data nodes (so that each data node does not have all the primaries) the abort would actually never wait on the data nodes. This isn't a big deal with uuid shard generations but could lead to potential corruption on S3 when using numeric shard generations (albeit very unlikely now that we have the 3 minute wait there). Another negative side-effect of this bug was that master would receive a lot more shard status update messages for aborted shards since each data node not assigned a primary would send one message for that primary.	2020-06-17 11:39:50 +02:00
Przemyslaw Gomulka	9894d90e0b	[doc] known issues - week based patterns not working in 7.6 (#58099 ) (#58227 ) relates #57128 # Conflicts: # docs/reference/release-notes/7.6.asciidoc	2020-06-17 10:54:22 +02:00
Andrei Dan	e17c51151b	[7.x] ILM: don't take snapshot of a data stream's write index (#58159 ) (#58222 ) We don't allow converting a data stream's writeable index into a searchable snapshot. We are currently preventing swapping a data stream's write index with the restored index. This adds another step that will not proceed with the searchable snapshot action until the managed index is not the write index of a data stream anymore. (cherry picked from commit ccd618ead7cf7f5a74b9fb34524d00024de1479a) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-06-17 09:45:16 +01:00
Armin Braun	c2b416ee31	Fix DanglingIndicesIT Failures from MasterNotDiscoveredException (#58215 ) (#58221 ) The dangling indices action is not a proper master node action so it does not retry when executed while the cluster hasn't fully formed yet. Since we use node restarts when setting up the dangling indices state we need to manually ensure a fully formed cluster before moving on with the tests to avoid failures.	2020-06-17 10:34:08 +02:00
Ignacio Vera	7080ba5b05	Check for degenerated lines when calculating the centroid (#58216 )	2020-06-17 09:34:49 +02:00
Przemysław Witek	b22e91cefc	[7.x] Delete auto-generated annotations when job is deleted. (#58169 ) (#58219 )	2020-06-17 09:17:20 +02:00
Lisa Cawley	46d797b1d9	[DOCS] Fixes license management links (#58213 )	2020-06-16 16:49:48 -07:00
debadair	cfef2b2bec	[DOCS] Removed unused pages (#58209 )	2020-06-16 15:55:56 -07:00
Przemko Robakowski	3249ee9a86	HLRC support for data streams (#58106 ) (#58202 ) This change adds high level REST client support for data streams Relates to #53100	2020-06-17 00:21:14 +02:00
Mark Vieira	5eb6692b0f	Add OpenJDK distribution to external dependency report (#58187 )	2020-06-16 15:08:04 -07:00
Stuart Tettemer	01795d1925	Revert "Scripting: Deprecate general cache settings (#55753 )" (#58201 ) This reverts commit `88e8b34fc2`.	2020-06-16 14:58:18 -06:00
Rory Hunter	03369e0980	Implement dangling indices API (#58176 ) Backport of #50920. Part of #48366. Implement an API for listing, importing and deleting dangling indices. Co-authored-by: David Turner <david.turner@elastic.co>	2020-06-16 21:50:38 +01:00
James Rodewig	ce22e951f8	[DOCS] Add resolve index API check to DS setup tutorial (#58167 ) (#58197 ) Updates the set up a data stream tutorial to include a name check using the resolve index API.	2020-06-16 16:28:42 -04:00
James Rodewig	c548a87673	[DOCS] Add 'Change DS mappings and settings' tutorial (#58148 ) (#58195 ) Adds a tutorial for updating the mappings and index settings of a data stream's backing indices.	2020-06-16 16:20:32 -04:00
Stuart Tettemer	88e8b34fc2	Scripting: Deprecate general cache settings (#55753 ) Backport: ef543b0	2020-06-16 13:06:59 -06:00
debadair	276a4898ba	[DOCS] Fixes problematic terminology (#58184 ) * [DOCS] Fixes problematic terminology (#58178) * [DOCS] Fixes problematic terminology. * Update docs/reference/snapshot-restore/register-repository.asciidoc Co-authored-by: James Rodewig <james.rodewig@elastic.co>	2020-06-16 11:43:22 -07:00
Benjamin Trent	081da09c72	Allow GET <pattern>/_rollup/data to expand data streams (#58173 ) (#58177 )	2020-06-16 14:01:54 -04:00
Benjamin Trent	3309817d18	[ML] fixing tree inference ctor to allow target_type to be optional (#58132 ) (#58165 ) The tree trained model object will set its target_type to be regression by default. This updates the inference object to behave the same way.	2020-06-16 13:29:11 -04:00
Alan Woodward	c6acc7c976	Correctly deal with aliases when retrieving lucene FieldType	2020-06-16 18:06:37 +01:00
Benjamin Trent	6c03d97419	Mute TimeSeriesDataStreamsIT.testSearchableSnapshotAction (#58127 ) (#58181 ) Co-authored-by: Andrei Dan <andrei.dan@elastic.co>	2020-06-16 12:40:38 -04:00
Alan Woodward	12a3f6dfca	MappedFieldType should not extend FieldType (#58160 ) MappedFieldType is a combination of two concerns: * an extension of lucene's FieldType, defining how a field should be indexed * a set of query factory methods, defining how a field should be searched We want to break these two concerns apart. This commit is a first step to doing this, breaking the inheritance relationship between MappedFieldType and FieldType. MappedFieldType instead has a series of boolean flags defining whether or not the field is searchable or aggregatable, and FieldMapper has a separate FieldType passed to its constructor defining how indexing should be done. Relates to #56814	2020-06-16 16:56:43 +01:00
Dan Hermann	911d46370e	Prohibit clone, shrink, and split on a data stream's write index	2020-06-16 10:53:20 -05:00
Lee Hinman	03ce0f8a4d	[7.x] Normalized prefix for rollover API (#57271 ) (69e1c066) (#58171 ) * Normalized prefix for rollover API (#57271) Co-authored-by: Elastic Machine <elasticmachine@users.noreply.github.com> Co-authored-by: Lee Hinman <lee@writequit.org> It fixes the issue #53388 by normalizing prefix at index creation request itself * Fix compilation for backport Co-authored-by: Gaurav Chandani <chngau@amazon.com>	2020-06-16 09:22:10 -06:00
Dan Hermann	7079a3b09f	[7.x] Prohibit freezing the write index of a data stream (#58168 )	2020-06-16 09:37:32 -05:00
Yannick Welsch	1e235a7f55	Fix off-by-one on CCR lease (#58158 ) The leases issued by CCR keep one extra operation around on the leader shards. This is not harmful to the leader cluster, but means that there's potentially one delete that can't be cleaned up.	2020-06-16 14:04:58 +02:00
Francisco Fernández Castaño	a5bc5ae030	Don't log on RetentionLeaseSync error handler (#58157 ) After an index has been deleted it may take some time to cancel all the maintenance tasks such as RetentionLeaseSync, it's possible that the task is already executing before the cancellation. This commit just avoids logging a warning message for those scenarios. Closes #57864 Backport of (#58098)	2020-06-16 14:04:32 +02:00
David Turner	423697f414	Default to zero replicas for searchable snapshots (#57802 ) Today a mounted searchable snapshot defaults to having the same replica configuration as the index that was snapshotted. This commit changes this behaviour so that we default to zero replicas on these indices, but allow the user to override this in the mount request. Relates #50999	2020-06-16 10:12:23 +01:00
Yannick Welsch	e046b0a8fa	Fix realtime get of numeric fields (#58121 ) Using realtime get on numeric fields when reading from the translog would yield a ClassCastException. Closes #57462	2020-06-16 09:16:26 +02:00
Tal Levy	69d5e044af	Add optional description parameter to ingest processors. (#57906 ) (#58152 ) This commit adds an optional field, `description`, to all ingest processors so that users can explain the purpose of the specific processor instance. Closes #56000.	2020-06-15 19:27:57 -07:00
debadair	2edcd064fe	[DOCS] Fix bad xref (#58150 )	2020-06-15 15:50:49 -07:00
Tal Levy	499ad6fcc4	Pre-compile inline scripts in Ingest Script processors (#57960 ) (#58130 ) This commit introduces an optimization for inline scripts. It keeps the compiled ingest script that the ScriptProcessor.Factory has been creating for validation purposes. Previously, the Script Service's cache was leveraged because it was the best way to handle caching of both stored and inline scripts. Since inline scripts are so widely used in Ingest Node, it is probably best to ensure we are using the pre-compiled version from the beginning.	2020-06-15 15:22:56 -07:00
Jake Landis	dc7ffb154a	Update hh to HH in date processor example (#58089 ) (#58144 ) Co-authored-by: Leaf-Lin <39002973+Leaf-Lin@users.noreply.github.com>	2020-06-15 17:04:14 -05:00
Lisa Cawley	554e60860f	[DOCS] Add token and HTTPS requirements for Kerberos (#57180 ) Co-authored-by: Tim Vernum <tim@adjective.org>	2020-06-15 14:30:13 -07:00
Adam Locke	ad0364dc06	[DOCS] Add documentation for near real-time search (#57560 ) (#58138 ) * Adding documentation for near real-time search. * Adding link to NRT topic and clarifying some text. * Adding diagrams and incorporating changes from David T.	2020-06-15 16:42:57 -04:00
debadair	80524098fc	[DOCS] Reformat release highlights as What's new. (#58073 )	2020-06-15 13:26:03 -07:00
Stuart Tettemer	71a42dbde9	[7.x] Rely on the computeIfAbsent logic to prevent duplicated compilation of scripts (#55467 ) (#58123 ) Instead of serializing compilation using a plain lock / mutex combined with a double check, rely on the computeIfAbsent logic to prevent duplicated compilation of scripts. Made checkCompilationLimit to be thread-safe and lock free. Backport: 865acad Co-authored-by: Michael Bischoff <michael.bischoff@elastic.co>	2020-06-15 12:01:22 -06:00
James Rodewig	e268a89ef2	[DOCS] Fix typo in data stream docs	2020-06-15 12:59:36 -04:00
Lee Hinman	d56d2dfb09	[7.x] Scope index templates put during cluster upgrade tests (#58065 ) (#58122 ) This template was added for 7.0 for what I am guessing is a BWC issue related to deprecation warnings. It unfortunately seems to cause failures because templates for these tests are not cleared after the test (because these are upgrade tests). Resolves #56363	2020-06-15 10:47:36 -06:00
Andrei Dan	3635bd741c	[DOCS] Make ILM documentation data stream aware (#58035 ) (#58110 ) Co-authored-by: James Rodewig <james.rodewig@elastic.co> Co-authored-by: Lee Hinman <dakrone@users.noreply.github.com> (cherry picked from commit 25cbbe56dd29fbee2efe8040e9c8b92d168cb670) Signed-off-by: Andrei Dan <andrei.dan@elastic.co>	2020-06-15 15:16:14 +01:00
markharwood	03dd73dc0d	Fix for wildcard fields that returned ByteRefs not Strings to scripts. (#58060 ) (#58109 ) This need some reorg of BinaryDV field data classes to allow specialisation of scripted doc values. Moved common logic to a new abstract base class and added a new subclass to return string-based representations to scripts. Closes #58044	2020-06-15 14:52:56 +01:00
James Rodewig	0bc7c4f69e	[DOCS] Fix xref in data stream docs	2020-06-15 09:49:18 -04:00

... 8 9 10 11 12 ...

52620 Commits All Branches Search

52620 Commits

All Branches