OpenSearch

Commit Graph

Author	SHA1	Message	Date
Yannick Welsch	cd57395c98	Use correct primary term for replicating NOOPs (#25128 ) NOOPs should be, same as for indexing operations, written on the replica using the original operation term instead of the current term of the replica.	2017-06-08 14:20:26 +02:00
Martijn van Groningen	326fa33d4e	fielddata: Binary script doc values should make a deep copy of the BytesRef before populating it in the values array. Added common base class for ScriptDocValues.Strings and ScriptDocValues.BytesRefs now that these classes are very similar. Also cleaned up the BinaryDVFieldDataTests: * Use junit assertions instead of hamcrest * Use BytesRef directly instead of byte[] Closes #24785	2017-06-08 13:20:35 +02:00
Jim Ferenczi	eeac4b9721	Fix Fast Vector Highlighter NPE on match phrase prefix (#25116 ) The FVH fails with an NPE when a match phrase prefix is rewritten in an empty phrase query. This change makes sure that the multi match query rewrites to a MatchNoDocsQuery (instead of an empty phrase query) when there is a single term and that term does not expand to any term in the index. Fixes #25088	2017-06-08 12:27:11 +02:00
Jim Ferenczi	36a5cf8f35	Automatically early terminate search query based on index sorting (#24864 ) This commit refactors the query phase in order to be able to automatically detect queries that can be early terminated. If the index sort matches the query sort, the top docs collection is early terminated on each segment and the computing of the total number of hits that match the query is delegated to a simple TotalHitCountCollector. This change also adds a new parameter to the search request called `track_total_hits`. It indicates if the total number of hits that match the query should be tracked. If false, queries sorted by the index sort will not try to compute this information and and will limit the collection to the first N documents per segment. Aggregations are not impacted and will continue to see every document even when the index sort matches the query sort and `track_total_hits` is false. Relates #6720	2017-06-08 12:10:46 +02:00
Jim Ferenczi	21a57c1494	Always use DisjunctionMaxQuery to build cross fields disjunction (#25115 ) This commit modifies query_string, simple_query_string and multi_match queries to always use a DisjunctionMaxQuery when a disjunction over multiple fields is built. The tiebreaker is set to 1 in order to behave like the boolean query in terms of scoring. The removal of the coord factor in Lucene 7 made this change mandatory to correctly handle minimum_should_match. Closes #23966	2017-06-08 11:18:17 +02:00
Simon Willnauer	d6d416cacc	Break out clear scroll logic from TransportClearScrollAction (#25125 ) This change extracts the main logic from `TransportClearScrollAction` into a new class `ClearScrollController` and adds a corresponding unit test. Relates to #25094	2017-06-08 11:13:08 +02:00
Simon Willnauer	bdc3a16fa4	Fix naminig in GroupedActionListener GroupedActionListener still had some members named from it's specialization before it was factored out in a general purpose class.	2017-06-08 10:21:15 +02:00
Adrien Grand	a8ea2f0df4	Leverage scorerSupplier when applicable. (#25109 ) The `scorerSupplier` API allows to give a hint to queries in order to let them know that they will be consumed in a random-access fashion. We should use this for aggregations, function_score and matched queries.	2017-06-08 10:19:38 +02:00
Boaz Leskes	087f182481	Translog file recovery should not rely on lucene commits (#25005 ) When we open a translog, we rely on the `translog.ckp` file to tell us what the maximum generation file should be and on the information stored in the last lucene commit to know the first file we need to recover. This requires coordination and is currently subject to a race condition: if a node dies after a lucene commit is made but before we remove the translog generations that were unneeded by it, the next time we open the translog we will ignore those files and never delete them (I have added tests for this). This PR changes the approach to have the translog store both of those numbers in the `translog.ckp`. This means it's more self contained and easier to control. This change also decouples the translog recovery logic from the specific commit we're opening. This prepares the ground to fully utilize the deletion policy introduced in #24950 and store more translog data that's needed for Lucene, keep multiple lucene commits around and be free to recover from any of them.	2017-06-08 09:21:28 +02:00
Simon Willnauer	ce24331d1f	Add helper methods to TransportActionProxy to identify proxy actions and requests (#25124 ) Downstream users of out network intercept infrastructure need this information which is hidden due to member and class visibility.	2017-06-08 09:07:22 +02:00
Jack Conradson	d187fa78fd	Generate Painless Factory for Creating Script Instances (#25120 )	2017-06-07 16:06:11 -07:00
Christoph Büscher	9e741cd13d	Tests: Add ability to generate random new fields for xContent parsing test (#23437 ) For the response parsing we want to be lenient when it comes to parsing new xContent fields. In order to ensure this in our testing, this change adds a utility method to XContentTestUtils that takes xContent bytes representation as input and recursively a random field on each object level. Sometimes we also want to exclude a whole subtree from this treatment (e.g. skipping "_source"), other times an element (e.g. "fields", "highlight" in SearchHit) can have arbitraryly named objects. Those cases can be specified as exceptions.	2017-06-07 21:01:20 +02:00
Jim Ferenczi	68f1d4df5a	bump the Lucene version for Version 5.5 and 5.6 after the upgrade to Lucene 6.6.0	2017-06-07 19:32:13 +02:00
Ryan Ernst	2057bbc6c5	Scripting: Remove unnecessary intermediate script compilation methods on QueryShardContext (#25093 ) This commit removes wrapper methods on QueryShardContext used to compile scripts. Instead, the script service is made accessible in the context, and calls to compile can be made directly. This will ease transition to each of those location becoming their own context, since they would no longer be able to expect the same script class type.	2017-06-07 08:24:18 -07:00
Yannick Welsch	26ec89173b	Remove TranslogRecoveryPerformer (#24858 ) Splits TranslogRecoveryPerformer into three parts: - the translog operation to engine operation converter - the operation perfomer (that indexes the operation into the engine) - the translog statistics (for which there is already RecoveryState.Translog) This makes it possible for peer recovery to use the same IndexShard interface as bulk shard requests (i.e. Engine operations instead of Translog operations). It also pushes the "fail on bad mapping" logic outside of IndexShard. Future pull requests could unify the BulkShard and peer recovery path even more.	2017-06-07 17:11:27 +02:00
Jim Ferenczi	c8bf7ecaed	Higlighters: Fix MultiPhrasePrefixQuery rewriting (#25103 ) The unified highlighter rewrites MultiPhrasePrefixQuery to SpanNearQuer even when there is a single term in the phrase. Though SpanNearQuery throws an exception when the number of clauses is less than 2. This change returns a simple PrefixQuery when there is a single term and builds the SpanNearQuery otherwise. Relates #25088	2017-06-07 16:14:28 +02:00
Tim Brooks	233c63fc63	Add version 5.6 to versions (#25084 ) * Add version 5.6 to versions * Fix test * Remove 5.4.2 constant	2017-06-07 09:59:27 -04:00
Boaz Leskes	8e15186293	Update `IndexShard#refreshMetric` via a `ReferenceManager.RefreshListener` (#25083 ) The PR takes a different approach to solve #24806 than currently implemented via #25052. The `refreshMetric` that IndexShard maintains is updated using the refresh listeners infrastructure in lucene. This means that we truly count all refreshes that lucene makes and not have to worry about each individual caller (like `IndexShard@refresh` and `Engine#get()`)	2017-06-07 10:54:10 +02:00
Martijn van Groningen	db8aa8e94e	Changed inner_hits to work with the new join field type and at the same time maintaining support for the `_parent` meta field type/ Relates to #20257	2017-06-07 10:52:49 +02:00
Yu	14913fdc37	keep _parent field while updating child type mapping (#24407 ) parent/child: Allow updating mapping without specifying `_parent` field on each update. Prior to this change when a mapping has a `_parent` field then any update (also updates that didn't modify the `_parent` field) to the mapping involved specifying the `_parent` field again. With this change specifying the `_parent` field on each mapping update is no longer required. Closes #23381	2017-06-07 10:51:21 +02:00
Jason Tedor	2f5f27fafa	Remove unnecessary callback interface We have a callback interface that is not needed because it is effectively the same as java.util.function.Consumer. This commit removes it. Relates #25089	2017-06-06 20:50:03 -04:00
Tim Brooks	feca0a9f33	Bumping version to v6.0.0-alpha3 (#25077 )	2017-06-06 15:47:23 -05:00
Jason Tedor	1a681a928d	Modify cluster state callback in recovery land We use a callback in recovery land during primary relocation to ensure the relocation target is on at least the same version as the relocation source. This callback is typed as a Callback<Long> which is an unnecessary custom type (we can use Consumer<T> or the appropriate primitive callbacks). Here, we can use LongConsumer. Relates #25081	2017-06-06 16:29:10 -04:00
Jason Tedor	e03c4938c5	GET aliases should 404 if aliases are missing Previously the HEAD and GET aliases endpoints were misaigned in behavior. The HEAD verb would 404 if any aliases are missing while the GET verb would not if any aliases existed. When HEAD was aligned with GET, this broke the previous usage of HEAD to serve as an existence check for aliases. It is the behavior of GET that is problematic here though, if any alias is missing the request should 404. This commit addresses this by modifying the behavior of GET to behave in this way. This fixes the behavior for HEAD to also 404 when aliases are missing. Relates #25043	2017-06-06 14:37:29 -04:00
Jim Ferenczi	7e60cf3e54	Move parent_id query to the parent-join module (#25072 ) This change moves the parent_id query to the parent-join module and handles the case when only the parent-join field can be declared on an index (index with single type on). If single type is off it uses the legacy parent join field mapper and switch to the new one otherwise (default in 6). Relates #20257	2017-06-06 19:35:14 +02:00
Ryan Ernst	7ec39acd4b	Settings: Fix setting groups to include secure settings (#25076 ) This commit fixes the group methdos of Settings to properly include grouped secure settings. Previously the secure settings were included but without the group prefix being removed. closes #25069	2017-06-06 10:13:10 -07:00
Yu	40a13345d7	Add refresh stats tracking for realtime get (#25052 ) Passes a `LongConsumer` into the `Engine` during GETs which the engine calls if it refreshed to perform the get. Closes #24806	2017-06-06 12:39:02 -04:00
olcbean	0d5f3958e7	Expand index expressions against indices only when managing aliases (#23997 ) The index parameter in the update-aliases, put-alias, and delete-alias APIs no longer accepts alias names. Instead, it accepts only index names (or wildcards which will expand to matching indices). Closes #23960	2017-06-06 11:01:38 +02:00
Ryan Ernst	ac82824d80	Settings: Fix secure settings by prefix (#25064 ) This commit fixes a bug in retrieving a sub Settings object for a given prefix with secure settings. Before this commit the returned Settings would be filtered by the prefix, but the found setting names would not have the prefix removed.	2017-06-06 00:11:33 -07:00
Lee Hinman	b6a2b8d682	Track EWMA[1] of task execution time in search threadpool executor This is the first step towards adaptive replica selection (#24915). This PR tracks the execution time, also known as the "service time" of a task in the threadpool. The `QueueResizingEsThreadPoolExecutor` then stores a moving average of these task times which can be retrieved from the executor. Currently there is no functionality using the EWMA yet (other than tests), this is only a bite-sized building block so that it's easier to review. [1]: EWMA = Exponentially Weighted Moving Average	2017-06-05 10:09:41 -06:00
Ali Beyad	f2a23e3459	Removes an invalid assert in resizing big arrays which does not always hold (resizing can result in a smaller size than the current size, while the assert attempted to verify the new size is always greater than the current).	2017-06-05 11:49:06 -04:00
Alex Benusovich	5463294ec4	Fixed NPEs caused by requests without content. (#23497 ) REST handlers that require a body will throw an an ElasticsearchParseException "request body required". REST handlers that require a body OR source param will throw an ElasticsearchParseException "request body or source param required". Replaced asserts in BulkRequest parsing code with a more descriptive IllegalArgumentException if the line contains an empty object. Updated bulk REST test to verify an empty action line is rejected properly. Updated BulkRequestTests with randomized testing for an empty action line. Used try-with-resouces for XContentParser in AbstractBulkByQueryRestHandler.	2017-06-05 09:08:14 -06:00
Nik Everett	73307a2144	Plugins can register pre-configured char filters (#25000 ) Fixes the plumbing so plugins can register char filters and moves the `html_strip` char filter into analysis-common. Relates to #23658	2017-06-05 09:25:15 -04:00
Ryan Ernst	e22a68295c	Tests: Make secure settings available from settings builder for tests (#25037 ) This commit exposes the secure settings in Settings.Builder, so that the current secure settings can be retrieved and added to when creating settings for tests. This is necessary since secure settings can only be added once to a builder, so chains of methods using settings builders must reuse the already set mock secure settings.	2017-06-03 16:55:34 -07:00
Sergey Novikov	57b4002357	Include duplicate jar when jarhell check fails When the jarhell check fails due to a duplicate jar on the classpath, the exception message includes the full classpath but not the duplicated jar. For a long classpath, this can make it difficult to find the jar that is duplicated. This commit changes the exception message to include the duplicated jar. Relates #24953	2017-06-02 18:22:01 -04:00
Lee Hinman	a32d1b91fa	Remove comma-separated feature parsing for GetIndicesAction This removes the parsing of things like `GET /idx/_aliases,_mappings`, instead, a user must choose between retriving all index metadata with `GET /idx`, or only a specific form such as `GET /idx/_settings`. Relates to (and is a prerequisite of) #24437	2017-06-02 14:43:38 -06:00
Ryan Ernst	0d8216d5af	Scripting: Convert CompiledTemplate to a ScriptContext (#25032 ) This commit creates TemplateScript and associated classes so that templates no longer need a special ScriptService.compileTemplate method. The execute() method is equivalent to the old run() method. relates #20426	2017-06-02 13:41:26 -07:00
Ali Beyad	e024c67561	Checks the circuit breaker before allocating bytes for a new big array (#25010 ) Previously, when allocating bytes for a BigArray, the array was created (or attempted to be created) and only then would the array be checked for the amount of RAM used to see if the circuit breaker should trip. This is problematic because for very large arrays, if creating or resizing the array, it is possible to attempt to create/resize and get an OOM error before the circuit breaker trips, because the allocation happens before checking with the circuit breaker. This commit ensures that the circuit breaker is checked before all big array allocations (note, this does not effect the array allocations that are less than 16kb which use the [Type]ArrayWrapper classes found in BigArrays.java). If such an allocation or resizing would cause the circuit breaker to trip, then the breaker trips before attempting to allocate and potentially running into an OOM error from the JVM. Closes #24790	2017-06-02 15:16:22 -04:00
Ali Beyad	3cb307462d	Consolidates the logic for cleaning up snapshots on master election (#24894 ) In #24605, logic was implemented to ensure that completed snapshots were properly removed from the cluster state upon a change in master nodes. This commit removes redundant logic that also attempted to clean up completed snapshots from the cluster state on master election, but only covered a limited case that was remedied in #24605. This commit also adds a test to ensure cleaning up of completed snapshots at the right moment in time when a master election happens before finalizing a snapshot, as well as adds a check to handle the case where the old master and new master could attempt to finalize the snapshot and write the same blob to the repository simultaneously.	2017-06-02 14:51:13 -04:00
Chris Earle	6ea9d83b2d	Remove @Override that doesn't exist in parent anymore from new TransportNodesUsageAction	2017-06-02 10:19:17 -04:00
Chris Earle	6464add551	Always Accumulate Transport Exceptions (#25017 ) This removes the `accumulateExceptions()` method (and its usage) from `TransportNodesAction` and `TransportTasksAction`, forcing both transport actions to always accumulate exceptions. Without this change, some transport actions, like `TransportNodesStatsAction` would respond in very unexpected ways by returning no response due to some failure, but instead of returning an error the response would simply be empty: no response and no error. This results in a very trappy response structure where users can check for an error, then attempt to blindly use the response when no error is returned.	2017-06-02 10:01:42 -04:00
Tanguy Leroux	5f3ed99c71	[Test] Reduce number of buckets in SearchResponseTests and AggregationsTests (#24964 ) This commit reduces the number of buckets that are generated for multi bucket aggregations in AggregationsTests and SearchResponseTests. The number of buckets are now limited to a maximum of 3 but before some aggregations could generate up to 10 buckets.	2017-06-02 15:59:25 +02:00
Jim Ferenczi	b8605775df	Add the ability to set eager_global_ordinals in the new parent-join field (#25019 ) Defaults to true	2017-06-02 15:34:22 +02:00
Jason Tedor	7ebba35c32	Handle already closed while filling gaps We can hit an already closed exception when filling the gaps after blocking operations when updating the primary term on a promoted replica shard. We should catch this and suppress it as it is an expected outcome instead of letting it bubble up which leads to trying to fail the shard which throws yet another already closed exception. Relates #25021	2017-06-02 08:05:33 -04:00
olcbean	6dea5f14c3	Java api: Remove unneeded getTookInMillis method (#23923 ) Some response classes in the java api expose both `getTook()` which returns a `TimeValue` and `getTookInMillis` which returns a `long` value. `getTook()` is enough as one can do `getTook().millis()` to obtain the same result as `getTookInMillis()`, which can be removed.	2017-06-02 11:11:05 +02:00
Colin Goodheart-Smithe	779fb9a1c0	Adds nodes usage API to monitor usages of actions (#24169 ) * Adds nodes usage API to monitor usages of actions The nodes usage API has 2 main endpoints /_nodes/usage and /_nodes/{nodeIds}/usage return the usage statistics for all nodes and the specified node(s) respectively. At the moment only one type of usage statistics is available, the REST actions usage. This records the number of times each REST action class is called and when the nodes usage api is called will return a map of rest action class name to long representing the number of times each of the action classes has been called. Still to do: * [x] Create usage service to store usage statistics * [x] Record usage in REST layer * [x] Add Transport Actions * [x] Add REST Actions * [x] Tests * [x] Documentation * Rafactors UsageService so counts are done by the handlers * Fixing up docs tests * Adds a name to all rest actions * Addresses review comments	2017-06-02 08:46:38 +01:00
Tanguy Leroux	528bd25fa7	Add superset size to Significant Term REST response (#24865 ) This commit adds a new bg_count field to the REST response of SignificantTerms aggregations. Similarly to the bg_count that already exists in significant terms buckets, this new bg_count field is set at the aggregation level and is populated with the superset size value.	2017-06-02 09:45:15 +02:00
Tanguy Leroux	c66be4a951	[Test] Remove unused test resources in core (#25011 ) It looks like many unnecessary files remain in the core test resources directory. This commit removes them.	2017-06-02 09:08:51 +02:00
Ryan Ernst	8d88b94372	Scripting: Add optional context parameter to put stored script requests (#25014 ) This commit adds an optional `context` url parameter to the put stored script request. When a context is specified, the script is compiled against that context before storing, as a validation the script will work when used in that context.	2017-06-01 17:53:48 -07:00
Simon Willnauer	39e59b49b1	Extract a common base class for scroll executions (#24979 ) Today there is a lot of code duplication and different handling of errors in the two different scroll modes. Yet, it's not clear if we keep both of them but this simplification will help to further refactor this code to also add cross cluster search capabilities. This refactoring also fixes bugs when shards failed due to the node dropped out of the cluster in between scroll requests and failures during the fetch phase of the scroll. Both places where simply ignoring the failure and logging to debug. This can cause issues like #16555	2017-06-01 22:23:41 +02:00

1 2 3 4 5 ...

8322 Commits