OpenSearch

Commit Graph

Author	SHA1	Message	Date
Ryan Ernst	ac82824d80	Settings: Fix secure settings by prefix (#25064 ) This commit fixes a bug in retrieving a sub Settings object for a given prefix with secure settings. Before this commit the returned Settings would be filtered by the prefix, but the found setting names would not have the prefix removed.	2017-06-06 00:11:33 -07:00
Lee Hinman	b6a2b8d682	Track EWMA[1] of task execution time in search threadpool executor This is the first step towards adaptive replica selection (#24915). This PR tracks the execution time, also known as the "service time" of a task in the threadpool. The `QueueResizingEsThreadPoolExecutor` then stores a moving average of these task times which can be retrieved from the executor. Currently there is no functionality using the EWMA yet (other than tests), this is only a bite-sized building block so that it's easier to review. [1]: EWMA = Exponentially Weighted Moving Average	2017-06-05 10:09:41 -06:00
Ali Beyad	f2a23e3459	Removes an invalid assert in resizing big arrays which does not always hold (resizing can result in a smaller size than the current size, while the assert attempted to verify the new size is always greater than the current).	2017-06-05 11:49:06 -04:00
Alex Benusovich	5463294ec4	Fixed NPEs caused by requests without content. (#23497 ) REST handlers that require a body will throw an an ElasticsearchParseException "request body required". REST handlers that require a body OR source param will throw an ElasticsearchParseException "request body or source param required". Replaced asserts in BulkRequest parsing code with a more descriptive IllegalArgumentException if the line contains an empty object. Updated bulk REST test to verify an empty action line is rejected properly. Updated BulkRequestTests with randomized testing for an empty action line. Used try-with-resouces for XContentParser in AbstractBulkByQueryRestHandler.	2017-06-05 09:08:14 -06:00
Nik Everett	73307a2144	Plugins can register pre-configured char filters (#25000 ) Fixes the plumbing so plugins can register char filters and moves the `html_strip` char filter into analysis-common. Relates to #23658	2017-06-05 09:25:15 -04:00
Ryan Ernst	e22a68295c	Tests: Make secure settings available from settings builder for tests (#25037 ) This commit exposes the secure settings in Settings.Builder, so that the current secure settings can be retrieved and added to when creating settings for tests. This is necessary since secure settings can only be added once to a builder, so chains of methods using settings builders must reuse the already set mock secure settings.	2017-06-03 16:55:34 -07:00
Sergey Novikov	57b4002357	Include duplicate jar when jarhell check fails When the jarhell check fails due to a duplicate jar on the classpath, the exception message includes the full classpath but not the duplicated jar. For a long classpath, this can make it difficult to find the jar that is duplicated. This commit changes the exception message to include the duplicated jar. Relates #24953	2017-06-02 18:22:01 -04:00
Lee Hinman	a32d1b91fa	Remove comma-separated feature parsing for GetIndicesAction This removes the parsing of things like `GET /idx/_aliases,_mappings`, instead, a user must choose between retriving all index metadata with `GET /idx`, or only a specific form such as `GET /idx/_settings`. Relates to (and is a prerequisite of) #24437	2017-06-02 14:43:38 -06:00
Ryan Ernst	0d8216d5af	Scripting: Convert CompiledTemplate to a ScriptContext (#25032 ) This commit creates TemplateScript and associated classes so that templates no longer need a special ScriptService.compileTemplate method. The execute() method is equivalent to the old run() method. relates #20426	2017-06-02 13:41:26 -07:00
Ali Beyad	e024c67561	Checks the circuit breaker before allocating bytes for a new big array (#25010 ) Previously, when allocating bytes for a BigArray, the array was created (or attempted to be created) and only then would the array be checked for the amount of RAM used to see if the circuit breaker should trip. This is problematic because for very large arrays, if creating or resizing the array, it is possible to attempt to create/resize and get an OOM error before the circuit breaker trips, because the allocation happens before checking with the circuit breaker. This commit ensures that the circuit breaker is checked before all big array allocations (note, this does not effect the array allocations that are less than 16kb which use the [Type]ArrayWrapper classes found in BigArrays.java). If such an allocation or resizing would cause the circuit breaker to trip, then the breaker trips before attempting to allocate and potentially running into an OOM error from the JVM. Closes #24790	2017-06-02 15:16:22 -04:00
Ali Beyad	3cb307462d	Consolidates the logic for cleaning up snapshots on master election (#24894 ) In #24605, logic was implemented to ensure that completed snapshots were properly removed from the cluster state upon a change in master nodes. This commit removes redundant logic that also attempted to clean up completed snapshots from the cluster state on master election, but only covered a limited case that was remedied in #24605. This commit also adds a test to ensure cleaning up of completed snapshots at the right moment in time when a master election happens before finalizing a snapshot, as well as adds a check to handle the case where the old master and new master could attempt to finalize the snapshot and write the same blob to the repository simultaneously.	2017-06-02 14:51:13 -04:00
Chris Earle	6ea9d83b2d	Remove @Override that doesn't exist in parent anymore from new TransportNodesUsageAction	2017-06-02 10:19:17 -04:00
Chris Earle	6464add551	Always Accumulate Transport Exceptions (#25017 ) This removes the `accumulateExceptions()` method (and its usage) from `TransportNodesAction` and `TransportTasksAction`, forcing both transport actions to always accumulate exceptions. Without this change, some transport actions, like `TransportNodesStatsAction` would respond in very unexpected ways by returning no response due to some failure, but instead of returning an error the response would simply be empty: no response and no error. This results in a very trappy response structure where users can check for an error, then attempt to blindly use the response when no error is returned.	2017-06-02 10:01:42 -04:00
Tanguy Leroux	5f3ed99c71	[Test] Reduce number of buckets in SearchResponseTests and AggregationsTests (#24964 ) This commit reduces the number of buckets that are generated for multi bucket aggregations in AggregationsTests and SearchResponseTests. The number of buckets are now limited to a maximum of 3 but before some aggregations could generate up to 10 buckets.	2017-06-02 15:59:25 +02:00
Jim Ferenczi	b8605775df	Add the ability to set eager_global_ordinals in the new parent-join field (#25019 ) Defaults to true	2017-06-02 15:34:22 +02:00
Jason Tedor	7ebba35c32	Handle already closed while filling gaps We can hit an already closed exception when filling the gaps after blocking operations when updating the primary term on a promoted replica shard. We should catch this and suppress it as it is an expected outcome instead of letting it bubble up which leads to trying to fail the shard which throws yet another already closed exception. Relates #25021	2017-06-02 08:05:33 -04:00
olcbean	6dea5f14c3	Java api: Remove unneeded getTookInMillis method (#23923 ) Some response classes in the java api expose both `getTook()` which returns a `TimeValue` and `getTookInMillis` which returns a `long` value. `getTook()` is enough as one can do `getTook().millis()` to obtain the same result as `getTookInMillis()`, which can be removed.	2017-06-02 11:11:05 +02:00
Colin Goodheart-Smithe	779fb9a1c0	Adds nodes usage API to monitor usages of actions (#24169 ) * Adds nodes usage API to monitor usages of actions The nodes usage API has 2 main endpoints /_nodes/usage and /_nodes/{nodeIds}/usage return the usage statistics for all nodes and the specified node(s) respectively. At the moment only one type of usage statistics is available, the REST actions usage. This records the number of times each REST action class is called and when the nodes usage api is called will return a map of rest action class name to long representing the number of times each of the action classes has been called. Still to do: * [x] Create usage service to store usage statistics * [x] Record usage in REST layer * [x] Add Transport Actions * [x] Add REST Actions * [x] Tests * [x] Documentation * Rafactors UsageService so counts are done by the handlers * Fixing up docs tests * Adds a name to all rest actions * Addresses review comments	2017-06-02 08:46:38 +01:00
Tanguy Leroux	528bd25fa7	Add superset size to Significant Term REST response (#24865 ) This commit adds a new bg_count field to the REST response of SignificantTerms aggregations. Similarly to the bg_count that already exists in significant terms buckets, this new bg_count field is set at the aggregation level and is populated with the superset size value.	2017-06-02 09:45:15 +02:00
Tanguy Leroux	c66be4a951	[Test] Remove unused test resources in core (#25011 ) It looks like many unnecessary files remain in the core test resources directory. This commit removes them.	2017-06-02 09:08:51 +02:00
Ryan Ernst	8d88b94372	Scripting: Add optional context parameter to put stored script requests (#25014 ) This commit adds an optional `context` url parameter to the put stored script request. When a context is specified, the script is compiled against that context before storing, as a validation the script will work when used in that context.	2017-06-01 17:53:48 -07:00
Simon Willnauer	39e59b49b1	Extract a common base class for scroll executions (#24979 ) Today there is a lot of code duplication and different handling of errors in the two different scroll modes. Yet, it's not clear if we keep both of them but this simplification will help to further refactor this code to also add cross cluster search capabilities. This refactoring also fixes bugs when shards failed due to the node dropped out of the cluster in between scroll requests and failures during the fetch phase of the scroll. Both places where simply ignoring the failure and logging to debug. This can cause issues like #16555	2017-06-01 22:23:41 +02:00
Nik Everett	4fcead9a65	Add backwards compatibility indices Adds backwards compatiblity indices and repos for the 5.4.1 and 5.3.3 release.	2017-06-01 12:34:03 -04:00
Jason Tedor	0435ec8ede	Add version 5.4.2 constant This commit adds the version 5.4.2 constant to master.	2017-06-01 11:25:19 -04:00
Jason Tedor	4185337df1	Add version 5.3.3 constant This commit adds the version 5.3.3 constant to master.	2017-06-01 11:18:25 -04:00
Jay Modi	7526c29a05	Provide the TransportRequest during validation of a search context (#24985 ) This commit provides the TransportRequest that caused the retrieval of a search context to the SearchOperationListener#validateSearchContext method so that implementers have access to the request.	2017-06-01 07:49:58 -06:00
Jason Tedor	9b4a189147	Add purge option to remove plugin CLI By default, the remove plugin CLI command preserves configuration files. This is so that if a user is upgrading the plugin (which is done by first removing the old version and then installing the new version) they do not lose their configuration file. Yet, there are circumstances where preserving the configuration file is not desired. This commit adds a purge option to the remove plugin CLI command. Relates #24981	2017-06-01 08:53:39 -04:00
Boaz Leskes	1775e4253e	Introducing a translog deletion policy (#24950 ) Currently, the decisions regarding which translog generation files to delete are hard coded in the interaction between the `InternalEngine` and the `Translog` classes. This PR extracts it to a dedicated class called `TranslogDeletionPolicy`, for two main reasons: 1) Simplicity - the code is easier to read and understand (no more two phase commit on the translog, the Engine can just commit and the translog will respond) 2) Preparing for future plans to extend the logic we need - i.e., retain multiple lucene commit and also introduce a size based retention logic, allowing people to always keep a certain amount of translog files around. The latter is useful to increase the chance of an ops based recovery.	2017-06-01 14:04:21 +02:00
Thomas Decaux	3eabb3acfd	Enforce validation for PathHierarchy tokenizer (#23510 ) If delimiter or replacement parameter are an empty string, the error is not clear enough to indicate how to fix it. With this change, the user knows these parameter must be a non empty string.	2017-06-01 12:54:16 +02:00
Tim Brooks	0424099674	Fix broken build from stream with zero bytes (#24993 ) This is related to #24927. There was a small possibility that a test was attempting to compress a stream with zero bytes. This was causing a failure. This test now requires at least one byte.	2017-05-31 17:33:11 -05:00
Tim Brooks	90a5574c93	Add CompressibleBytesOutputStream for compression (#24927 ) This is a follow-up to #23941. Currently there are a number of complexities related to compression. The raw DeflaterOutputStream must be closed prior to sending bytes to ensure that EOS bytes are written. But the underlying ReleasableBytesStreamOutput cannot be closed until the bytes are sent to ensure that the bytes are not reused. Right now we have three different stream references hanging around in TCPTransport to handle this complexity. This commit introduces CompressibleBytesOutputStream to be one stream implemenation that will behave properly with or without compression enabled.	2017-05-31 11:00:40 -05:00
Lee Hinman	9d6cb4cb6d	Remove unused MeterMetric and specialized EWMA (#24975 ) This metric is not used in the ES codebase at all. It's also not as likely to be used since it relies on a periodic "tick", which we don't currently use.	2017-05-31 09:05:22 -06:00
Jim Ferenczi	ec64c2c05f	Compute the took time of the query after the expand phase (#24902 ) The took time computed for search requests does not take in account the expand search phase. This change delays the computation to after the expand phase finishes. Relates #24900	2017-05-31 12:42:05 +02:00
Masaru Hasegawa	a77b38cdd1	Fix context suggester to read values from keyword type field (#24200 ) Closes #24129	2017-05-31 11:35:01 +02:00
Martijn van Groningen	258be2b135	Moved `keyword_marker`, `trim`, `snowball` and `porter_stemmer` tokenfilter factories from core to common-analysis module. Relates to #23658	2017-05-31 09:34:08 +02:00
Martijn van Groningen	a089dc9dcd	Added more unit test coverage for terms aggregation and removed terms agg integration tests that were replaced by unit tests.	2017-05-31 09:30:10 +02:00
Martijn van Groningen	9531ef25ec	Move OldIndexBackwardsCompatibilityIT#assertBasicSearchWorks over to full cluster restart qa module. Relates to #24939	2017-05-31 09:27:41 +02:00
Tanguy Leroux	8e0d6015f9	[Test] Mute SearchResponseTests.testFromXContent() And also AggregationsTests.testFromXContent() until https://github.com/elastic/elasticsearch/pull/24964 is merged.	2017-05-31 09:16:23 +02:00
Adrien Grand	36a180ec20	Eliminate array access in tight loops when profiling is enabled. (#24959 ) This makes profiling classes acquire a timer up-front that can be then reused across all calls, in order to save bound checks for methods that are called in tight loops.	2017-05-31 09:11:00 +02:00
Adrien Grand	71264c6239	PatternAnalyzer should lowercase wildcard queries when `lowercase` is true. (#24967 )	2017-05-31 09:09:53 +02:00
Ryan Ernst	7c1211d2ed	Scripting: Add StatefulFactoryType as optional intermediate factory in script contexts (#24974 ) ScriptContexts currently understand a FactoryType that can produce instances of the script InstanceType. However, for search scripts, this does not work as we have the concept of LeafSearchScript that is created per lucene segment. This commit effectively renames the existing SearchScript class into SearchScript.LeafFactory, which is a new, optional, class that can be defined within a ScriptContext. LeafSearchScript is effectively renamed back into SearchScript. This change allows the model of stateless factory -> stateful factory -> script instance to continue, but in a generic way that any script context may take advantage of. relates #20426	2017-05-30 16:32:14 -07:00
Jason Tedor	ac94253dce	Clarify acquiring index shard permit In previous work, we refactored the delay mechanism in index shard operation permits to allow for async delaying of acquisition. This refactoring made explicit when permit acquisition is disabled whereas previously we were relying on an implicit condition, namely that all permits were acquired by the thread trying to delay acquisition. When using the implicit mechanism, we tried to acquire a permit and if this failed, we returned a null releasable as an indication that our operation should be queued. Yet, now we know when we are delayed and we should not even try to acquire a permit. If we try to acquire a permit and one is not available, we know that we are not delayed, and so acquisition should be successful. If it is not successful, something is deeply wrong. This commit takes advantage of this refactoring to simplify the internal implementation. Relates #24971	2017-05-30 16:22:17 -04:00
Jason Tedor	b28141a990	Fill gaps on primary promotion When a primary is promoted, it could have gaps in its history due to concurrency and in-flight operations when it was serving as a replica. This commit fills the gaps in the history of the promoted shard after all operations from the previous term have drained, and future operations are blocked. This commit does not handle replicating the no-ops that fill the gaps to any remaining replicas, that is the responsibility of the primary/replica sync that we are laying the ground work for. Relates #24945	2017-05-30 13:19:44 -04:00
Jim Ferenczi	ce7195d81a	Terms aggregation should remap global ordinal buckets when a sub-aggregator is used to sort the terms (#24941 ) `terms` aggregations at the root level use the `global_ordinals` execution hint by default. When all sub-aggregators can be run in `breadth_first` mode the collected buckets for these sub-aggs are dense (remapped after the initial pruning). But if a sub-aggregator is not deferrable and needs to collect all buckets before pruning we don't remap global ords and the aggregator needs to deal with sparse buckets. Most (if not all) aggregators expect dense buckets and uses this information to allocate memories. This change forces the remap of the global ordinals but only when there is at least one sub-aggregator that cannot be deferred. Relates #24788	2017-05-30 19:13:07 +02:00
Jason Tedor	ddbc4687f6	Introduce clean transition on primary promotion This commit introduces a clean transition from the old primary term to the new primary term when a replica is promoted primary. To accomplish this, we delay all operations before incrementing the primary term. The delay is guaranteed to be in place before we increment the term, and then all operations that are delayed are executed after the delay is removed which asynchronously happens on another thread. This thread does not progress until in-flight operations that were executing are completed, and after these operations drain, the delayed operations re-acquire permits and are executed. Relates #24925	2017-05-30 11:39:36 -04:00
Jason Tedor	15fc71249c	Fix typo in comment in ReplicationOperation.java Within two lines of each other appears "fallthrough" and "fall through", both typed by the same person who should have been paying better attention and only one of these is correct and the inconsistency is bothersome. This commit fixes the errant one.	2017-05-30 11:16:32 -04:00
Lee Hinman	0b3be42c10	Prevent Index & Delete request primaryTerm getter/setter, setShardId setter	2017-05-30 08:48:35 -06:00
Nik Everett	6d9ce957d4	Drop name from TokenizerFactory (#24869 ) Drops `TokenizerFactory#name`, replacing it with `CustomAnalyzer#getTokenizerName` which is much better targeted at its single use case inside the analysis API. Drops a test that I would have had to refactor which is duplicated by `AnalysisModuleTests`. To keep this change from blowing up in size I've left two mostly mechanical changes to be done in followups: 1. `TokenizerFactory` can now be entirely dropped and replaced with `Supplier<Tokenizer>`. 2. `AbstractTokenizerFactory`'s ctor still takes a `String` parameter where the name once was.	2017-05-30 10:39:22 -04:00
Zachary Tong	b8d7b83f8e	Correctly set doc_count when MovAvg "predicts" values on existing buckets (#24892 ) If the bucket already exists, due to non-overlapping series or missing data, the MovAvg creates a merged bucket with the existing aggs + the new prediction. This fixes a small bug where the doc_count was not being set correctly. Relates to #24327	2017-05-30 10:36:22 -04:00
Jason Tedor	9957bdf0ad	Handle primary failure handling replica response Today if the primary throws an exception while handling the replica response (e.g., because it is already closed while updating the local checkpoint for the replica), or because of a bug that causes an exception to be thrown in the replica operation listener, this exception is caught by the underlying transport handler plumbing and is translated into a response handler failure transport exception that is passed to the onFailure method of the replica operation listener. This causes the primary to turn around and fail the replica which is a disastrous and incorrect outcome as there's nothing wrong with the replica, it is the primary that is broken and deserves a paddlin'. This commit handles this situation by failing the primary. Relates #24926	2017-05-30 10:05:11 -04:00

1 2 3 4 5 ...

8294 Commits