OpenSearch

mirror of https://github.com/honeymoose/OpenSearch.git synced 2025-02-15 09:25:40 +00:00

Author	SHA1	Message	Date
Jim Ferenczi	29331f1127	Fail queries with scroll that explicitely set request_cache (#27342 ) Queries that create a scroll context cannot use the cache. They modify the search context during their execution so using the cache can lead to duplicate result for the next scroll query. This change fails the entire request if the request_cache option is explictely set on a query that creates a scroll context (`scroll=1m`) and make sure internally that we never use the cache for these queries when the option is not explicitely used. For 6.x a deprecation log will be printed instead of failing the entire request and the request_cache hint will be ignored (forced to false).	2017-11-10 16:02:06 +01:00
Tanguy Leroux	9c4d6c629a	Remove S3 output stream (#27280 ) Now the blob size information is available before writing anything, the repository implementation can know upfront what will be the more suitable API to upload the blob to S3. This commit removes the DefaultS3OutputStream and S3OutputStream classes and moves the implementation of the upload logic directly in the S3BlobContainer. related #26993 closes #26969	2017-11-10 12:22:33 +01:00
Martijn van Groningen	b4048b4e7f	Use CoveringQuery to select percolate candidate matches and extract all clauses from a conjunction query. When clauses from a conjunction are extracted the number of clauses is also stored in an internal doc values field (minimum_should_match field). This field is used by the CoveringQuery and allows the percolator to reduce the number of false positives when selecting candidate matches and in certain cases be absolutely sure that a conjunction candidate match will match and then skip MemoryIndex validation. This can greatly improve performance. Before this change only a single clause was extracted from a conjunction query. The percolator tried to extract the clauses that was rarest in order (based on term length) to attempt less candidate queries to be selected in the first place. However this still method there is still a very high chance that candidate query matches are false positives. This change also removes the influencing query extraction added via #26081 as this is no longer needed because now all conjunction clauses are extracted. https://www.elastic.co/guide/en/elasticsearch/reference/6.x/percolator.html#_influencing_query_extraction Closes #26307	2017-11-10 07:44:42 +01:00
Nicholas Knize	06ff92d237	Add ignore_malformed to geo_shape fields This commit adds ignore_malformed support to geo_shape field types to skip malformed geoJson fields. closes #23747	2017-11-09 17:59:05 -06:00
Dimitris Athanasiou	66bef26495	Aggregations: bucket_sort pipeline aggregation (#27152 ) This commit adds a parent pipeline aggregation that allows sorting the buckets of a parent multi-bucket aggregation. The aggregation also offers [from] and [size] parameters in order to truncate the result as desired. Closes #14928	2017-11-09 17:59:57 +00:00
Tal Levy	d22fd4ea58	Introduce templating support to timezone/locale in DateProcessor (#27089 ) Sometimes systems like Beats would want to extract the date's timezone and/or locale from a value in a field of the document. This PR adds support for mustache templating to extract these values. Closes #24024.	2017-11-09 09:45:32 -08:00
Tanguy Leroux	184dda9eb0	Update to AWS SDK 1.11.223 (#27278 )	2017-11-09 13:25:51 +01:00
Mayya Sharipova	abbe853f1e	Add limits for ngram and shingle settings (#27211 ) (#27318 ) Relates to #25887	2017-11-08 10:12:57 -05:00
Jay Greenberg	df5c8bb3bf	Update discovery-ec2.asciidoc Changed the recommendation to use Tribe Node to Cross Cluster Search.	2017-11-07 10:18:38 -05:00
Mayya Sharipova	148376c2c5	Add limits for ngram and shingle settings (#27211 ) * Add limits for ngram and shingle settings (#27211) Create index-level settings: max_ngram_diff - maximum allowed difference between max_gram and min_gram in NGramTokenFilter/NGramTokenizer. Default is 1. max_shingle_diff - maximum allowed difference between max_shingle_size and min_shingle_size in ShingleTokenFilter. Default is 3. Throw an IllegalArgumentException when trying to create NGramTokenFilter, NGramTokenizer, ShingleTokenFilter where difference between max_size and min_size exceeds the settings value. Closes #25887	2017-11-07 08:14:55 -05:00
Zachary Tong	6e9e07d6f8	Fix profiling naming issues (#27133 ) Some code-paths use anonymous classes (such as NonCollectingAggregator in terms agg), which messes up the display name of the profiler. If we encounter an anonymous class, we need to grab the super's name. Another naming issue was that ProfileAggs were not delegating to the wrapped agg's name for toString(), leading to ugly display. This PR also fixes up the profile documentation. Some of the examples were executing against empty indices, which shows different profile results than a populated index (and made for confusing examples). Finally, I switched the agg display names from the fully qualified name to the simple name, so that it's similar to how the query profiles work. Closes #26405	2017-11-06 16:37:33 -05:00
Shubham Aggarwal	5a925cd40c	Fixed references to Multi Index Syntax (#27283 )	2017-11-06 19:15:36 +01:00
Patrice Bourgougnon	4b7b1e2706	Add an active Elasticsearch WordPress plugin link (#27279 )	2017-11-06 18:13:27 +01:00
Boris Tyukin	8e9b30417c	Update to support bulk updates by query (#27172 ) Getting started doc stated that bulk updates by query are not supported but they are now	2017-11-06 17:32:20 +01:00
Boaz Leskes	a8ff4960f3	add split index reference in indices.asciidoc Relates to #26931	2017-11-06 12:55:41 +01:00
Simon Willnauer	bd7efa908a	Add ability to split shards (#26931 ) This change adds a new `_split` API that allows to split indices into a new index with a power of two more shards that the source index. This API works alongside the `_shrink` API but doesn't require any shard relocation before indices can be split. The split operation is conceptually an inverse `_shrink` operation since we initialize the index with a _syntetic_ number of routing shards that are used for the consistent hashing at index time. Compared to indices created with earlier versions this might produce slightly different shard distributions but has no impact on the per-index backwards compatibility. For now, the user is required to prepare an index to be splittable by setting the `index.number_of_routing_shards` at index creation time. The setting allows the user to prepare the index to be splittable in factors of `index.number_of_routing_shards` ie. if the index is created with `index.number_of_routing_shards: 16` and `index.number_of_shards: 2` it can be split into `4, 8, 16` shards. This is an intermediate step until we can make this the default. This also allows us to safely backport this change to 6.x. The `_split` operation is implemented internally as a DeleteByQuery on the lucene level that is executed while the primary shards execute their initial recovery. Subsequent merges that are triggered due to this operation will not be executed immediately. All merges will be deferred unti the shards are started and will then be throttled accordingly. This change is intended for the 6.1 feature release but will not support pre-6.1 indices to be split unless these indices have been shrunk before. In that case these indices can be split backwards into their original number of shards.	2017-11-06 11:37:55 +01:00
Pablo Musa	7b03d68f9f	[Docs] Fix minor paragraph indentation error for multiple Indices params (#25535 )	2017-11-06 10:20:20 +01:00
Nhat	c7ce5a07f2	Add size-based condition to the index rollover API (#27160 ) This is to add a max_size condition to the index rollover API. We use a totalSizeInBytes from DocsStats to evaluate this condition. Closes #27004	2017-11-04 19:51:48 -04:00
Loek van Gool	67e677f443	Add an example of dynamic field names (#27255 )	2017-11-03 23:20:58 +01:00
David Turner	fbf8c3ee83	Reinstate recommendation for ≥ 3 master-eligible nodes. (#27204 ) In the docs for 1.7 ([doc][doc-1.7], [src][src-1.7]) there was a recommendation for at least 3 master-eligible nodes "in critical clusters" but this was lost when that page was updated in 2.0 ([doc][doc-2.0], [src][src-2.0]). I'd like to reinstate this. [doc-1.7]: https://www.elastic.co/guide/en/elasticsearch/reference/1.7/modules-node.html [src-1.7]: `2cbaccb2f2/docs/reference/modules/node.asciidoc` [doc-2.0]: https://www.elastic.co/guide/en/elasticsearch/reference/2.0/modules-node.html#split-brain [src-2.0]: `4799009ad7/docs/reference/modules/node.asciidoc`	2017-11-03 08:48:48 +00:00
Yannick Welsch	7791e72626	Add additional explanations around discovery.zen.ping_timeout (#27231 ) Makes it clearer that this setting should only be changed with extra care.	2017-11-02 16:52:10 +01:00
Colin Goodheart-Smithe	c1b8140c83	Upgrade to Lucene 7.1 (#27225 )	2017-11-02 13:25:33 +00:00
Martijn van Groningen	d805c41b28	Added new terms_set query This query returns documents that match with at least one ore more of the provided terms. The number of terms that must match varies per document and is either controlled by a minimum should match field or computed per document in a minimum should match script. Closes #26915	2017-11-01 10:55:18 +01:00
Toby McLaughlin	b71f7d3559	Update Docker docs for 6.0.0-rc2 (#27166 ) * Update Docker docs for 6.0.0-rc2 * Update the docs to match the new Docker "image flavours" of "basic", "platinum", and "oss". * Clarifications for Openshift and bind-mounts * Bump docker-compose 2.x format to 2.2 * Combine Docker Toolbox instructions for setting vm.max_map_count for both macOS + Windows * devicemapper is not the default storage driver any more on RHEL	2017-11-01 14:24:30 +11:00
Igor Motov	d14486bce6	Docs: restore now fails if it encounters incompatible settings (#26933 ) This change was introduced in 5.0.0, but the documentation wasn't updated to reflect it. Closes #26453	2017-10-31 20:04:00 -04:00
javanna	506a2c276d	[DOCS] Link remote info API in Cross Cluster Search docs page Closes #26327	2017-10-31 15:24:46 +01:00
Shai Erera	bd0261916c	Fix Laplace scorer to multiply by alpha (and not add) (#27125 )	2017-10-31 13:08:44 +01:00
javanna	34666844b3	[DOCS] Clarify migrate guide and search request validation Relates to #26811	2017-10-31 12:36:00 +01:00
kel	c3e2bdf20c	Raise IllegalArgumentException if query validation failed (#26811 ) Closes #26799	2017-10-31 12:17:27 +01:00
Jim Ferenczi	792641a6e3	[Docs] #26541 : add warning regarding the limit on the number of fields that can be queried at once in the multi_match query.	2017-10-30 18:03:56 +01:00
Dimitrios Athanasiou	3796471ac4	[Docs] Fix note in bucket_selector	2017-10-30 15:20:46 +00:00
Clarkie	b1ce5cf836	[Docs] Fix indentation of examples (#27168 )	2017-10-30 11:56:38 +01:00
Jim Ferenczi	a4105c6b4a	[Docs] Clarify `span_not` query behavior for non-overlapping matches (#27150 ) Closes #27134	2017-10-30 11:29:40 +01:00
Christoph Büscher	8e62314ce4	[Docs] Remove first person "I" from getting started (#27155 ) Avoid first person style and consistently switch to an unpersonal style in the getting started docs.	2017-10-30 10:45:50 +01:00
Holger Bartnick	aa03fb72b7	[Docs] Correct link target for datatype murmur3 (#27143 )	2017-10-30 09:31:55 +01:00
Jun Ohtani	77e11f6969	[Doc] Add Ingest CSV Processor Plugin to plugin as a community plugin (#27105 ) * [Doc] Add Ingest CSV Processor Plugin to plugin as a community plugin	2017-10-27 16:16:02 +09:00
Clinton Gormley	0499dc0873	Removed the beta tag from cross-cluster search	2017-10-27 08:51:36 +02:00
Martijn van Groningen	f1e944a675	docs: describe parent/child performances	2017-10-26 11:49:13 +02:00
Catalin Ursachi	8bf33241ed	Add Delete Index API support to high-level REST client (#27019 ) Relates to #25847	2017-10-26 09:52:46 +02:00
Loading Zhang	149e558dd5	Docs: Fix ingest geoip config location (#27110 )	2017-10-25 07:16:42 -07:00
markwalkom	2b864156ca	[Docs] Clarify mapping `index` option default (#27104 )	2017-10-25 12:42:29 +02:00
Luca Cavanna	8caf7d4ff8	Decouple BulkProcessor from ThreadPool (#26727 ) Introduce minimal thread scheduler as a base class for `ThreadPool`. Such a class can be used from the `BulkProcessor` to schedule retries and the flush task. This allows to remove the `ThreadPool` dependency from `BulkProcessor`, which requires to provide settings that contain `node.name` and also needed log4j for logging. Instead, it needs now a `Scheduler` that is much lighter and gets automatically created and shut down on close. Closes #26028	2017-10-25 10:30:23 +02:00
David Turner	559fc5a4de	Update numbers to reflect 4-byte UTF-8-encoded characters (#27083 ) You need 4 bytes for characters outside the BMP, which includes many emoji and a bunch of less-common writing characters too.	2017-10-24 09:50:47 +01:00
Martijn van Groningen	87c9b79b10	Return the _source of inner hit nested as is without wrapping it into its full path context Due to a change happened via #26102 to make the nested source consistent with or without source filtering, the _source of a nested inner hit was always wrapped in the parent path. This turned out to be not ideal for users relying on the nested source, as it would require additional parsing on the client side. This change fixes this, the _source of nested inner hits is now no longer wrapped by parent json objects, irregardless of whether the _source is included as is or source filtering is used. Internally source filtering and highlighting relies on the fact that the _source of nested inner hits are accessible by its full field path, so in order to now break this, the conversion of the _source into its binary form is performed in FetchSourceSubPhase, after any potential source filtering is performed to make sure the structure of _source of the nested inner hit is consistent irregardless if source filtering is performed. PR for #26944 Closes #26944	2017-10-19 12:04:56 +02:00
İsmail Arılık	71f5e2ce6b	Fix a typo. (#27043 ) `=== Instalation with Homebrew` should be `=== Installation with Homebrew`.	2017-10-18 09:46:53 -04:00
Divyum Rastogi	984731f36b	[DOCS] better formatting of ES cluster status (#26838 ) * better formatting of ES cluster status * change phrase missing data	2017-10-18 01:40:21 -06:00
Pius	400480e3b0	action.auto_create_index can be set as a dynamic cluster setting (#27026 ) Per https://github.com/elastic/elasticsearch/pull/20274, action.auto_create_index can be set as a dynamic cluster setting.	2017-10-17 20:44:18 +00:00
Anton Pozhidaev	70668dddf3	Update docs about `script` parameter (#27010 ) Added a description of short script form. Also removed references to the obsolete `script.default_lang`.	2017-10-16 05:04:43 -07:00
Simon Willnauer	8dda827ff4	Don't refresh on `_flush` `_force_merge` and `_upgrade` (#27000 ) Today all these API calls have a sideeffect of making documents visible to search requests. While this is sometimes desired it's an unnecessary sideeffect and now that we have an internal (engine-private) index reader (#26972) we artificially add a refresh call for bwc. This change removes this sideeffect in 7.0.	2017-10-16 10:16:35 +02:00
Jason Tedor	8eba1fa17c	Add docs on full_id parameter in cat nodes API This commit adds a note to the docs on the full_id parameter in the cat nodes API. This is a useful parameter but was not previously documented anywhere. Relates #27009	2017-10-13 13:49:25 -04:00

... 5 6 7 8 9 ...

5001 Commits