OpenSearch

Commit Graph

Author	SHA1	Message	Date
Simon Willnauer	f933f80902	First step towards incremental reduction of query responses (#23253 ) Today all query results are buffered up until we received responses of all shards. This can hold on to a significant amount of memory if the number of shards is large. This commit adds a first step towards incrementally reducing aggregations results if a, per search request, configurable amount of responses are received. If enough query results have been received and buffered all so-far received aggregation responses will be reduced and released to be GCed.	2017-02-21 13:02:48 +01:00
Tanguy Leroux	39ed76c58b	Add parsing method to bulk response (#23234 ) This commit adds the `fromXContent()` parsing method to BulkResponse.	2017-02-21 10:49:40 +01:00
Tanguy Leroux	c88eb00b83	Add javadoc for DocWriteResponse.Builders (#23267 )	2017-02-21 10:19:01 +01:00
Martin Scholz	24bf18b610	Upgrade HDRHistogram to 2.1.9 (#23254 )	2017-02-21 08:50:26 +01:00
Martin Scholz	3e292d5245	Migrate TermsQuery to TermInSetQuery (#23229 )	2017-02-21 08:49:43 +01:00
Jim Ferenczi	1ff5b318be	Fix for IpRangeAggregatorTests#testRanges Handle null from/to ranges. Closes #23272	2017-02-20 21:16:14 +01:00
Jason Tedor	4c2bd5feab	Introduce sequence-number-aware translog Today, the relationship between Lucene and the translog is rather simple: every document not in Lucene is guaranteed to be in the translog. We need a stronger guarantee from the translog though, namely that it can replay all operations after a certain sequence number. For this to be possible, the translog has to made sequence-number aware. As a first step, we introduce the min and max sequence numbers into the translog so that each generation knows the possible range of operations contained in the generation. This will enable future work to keep around all generations containing operations after a certain sequence number (e.g., the global checkpoint). Relates #22822	2017-02-20 15:05:24 -05:00
Jason Tedor	15f5810774	Mark IP range aggregator test as awaits fix This test reliably fails with the seed 4AC319F8A6B0329B.	2017-02-20 14:42:16 -05:00
Christoph Büscher	ea7deace5d	Adding fromXContent to Suggest and Suggestion class (#23226 ) A follow up to #23202, this adds parsing from xContent and tests to the four Suggestion implementations and the top level suggest element to be used later when parsing the entire SearchResponse.	2017-02-20 15:45:10 +01:00
Christoph Büscher	ea9d51114c	Tests: Add unit test for InternalChildren (#23261 ) Relates to #22278	2017-02-20 14:02:56 +01:00
Jim Ferenczi	76d6b872dd	Add unit tests for GeoBoundsAggregator/InternalGeoBounds (#23259 ) * Add unit tests for GeoBoundsAggregator/InternalGeoBounds Relates #22278	2017-02-20 12:04:30 +01:00
Jim Ferenczi	69b1463f7c	Add unit tests for BinaryRangeAggregator/InternalBinaryRange (#23255 ) * Add unit tests for BinaryRangeAggregator/InternalBinaryRange Relates #22278	2017-02-20 11:55:48 +01:00
Tanguy Leroux	872412f645	[Tests] Cleans up DocWriteResponse parsing tests (#23233 ) This commit cleans up some parsing tests added from the High Level Rest Client: IndexResponseTests, DeleteResponseTests, UpdateResponseTests, BulkItemResponseTests. These tests are now more uniform with the others test-from-to-XContent tests we have, they now shuffle the XContent fields before parsing, the asserting method for parsed objects does not used a Map<String, Object> anymore, and buggy equals/hasCode methods in ShardInfo and ShardInfo.Failure have been removed.	2017-02-20 09:45:33 +01:00
Nik Everett	d9c37ce195	Adds unit test for sampler aggregation Relates to #22278	2017-02-17 16:16:04 -05:00
Nik Everett	d1de9574ea	Checkstyle: Fix link lengths in sampler aggregation	2017-02-17 15:03:57 -05:00
Jay Modi	b234644035	Enforce Content-Type requirement on the rest layer and remove deprecated methods (#23146 ) This commit enforces the requirement of Content-Type for the REST layer and removes the deprecated methods in transport requests and their usages. While doing this, it turns out that there are many places where *Entity classes are used from the apache http client libraries and many of these usages did not specify the content type. The methods that do not specify a content type explicitly have been added to forbidden apis to prevent more of these from entering our code base. Relates #19388	2017-02-17 14:45:41 -05:00
Adrien Grand	3bd1d46fc7	Add unit tests for terms aggregation objects. (#23149 ) Relates #22278	2017-02-17 18:01:40 +01:00
javanna	578853f264	Remove stale comment about setting routing before parent Order does not matter anymore since we merged #15371	2017-02-17 17:10:53 +01:00
Yuhao Bi	576e698613	Minor fix of _cat output (#23211 ) (#23213 ) One line was missing a trailing "\n"	2017-02-17 10:46:20 +01:00
Jason Tedor	00a8b8799f	Fix control group pattern The file /proc/self/cgroup lists the control groups to which the process belongs. This file is a colon separated list of three fields: 1. a hierarchy ID number 2. a comma-separated list of hierarchies 3. the pathname of the control group in the hierarchy The regex pattern for this contains a bug for the second field. It allows one or two entries in the comma-separated list, but not more. This commit fixes the pattern to allow one or more entires in the comma-separated list. Relates #23219	2017-02-16 15:31:18 -05:00
Christoph Büscher	268d15ec4c	Adding fromXContent to Suggestion.Entry and subclasses (#23202 ) This adds parsing from xContent to Suggestion.Entry and its subclasses for Terms-, Phrase- and CompletionSuggestion.Entry.	2017-02-16 17:59:55 +01:00
markharwood	1cd1ff6010	Test fix - faulty assumptions about when exceptions are thrown in relation to number of failing shards. (#23205 ) Search exceptions are thrown only when all shards report failure. Fix changes assertion logic to reflect this. Closes #23203	2017-02-16 13:48:17 +00:00
Jason Tedor	0a5917d182	Fix get HEAD requests Get HEAD requests incorrectly return a content-length header of 0. This commit addresses this by removing the special handling for get HEAD requests, and just relying on the general mechanism that exists for handling HEAD requests in the REST layer. Relates #23186	2017-02-15 13:07:29 -05:00
Christoph Büscher	458ca09e70	Fix checkstyle issue with modifier order in DocWriteResponse	2017-02-15 17:53:39 +01:00
Tanguy Leroux	e8d669f50c	Add parsing methods to BulkItemResponse (#22859 ) This commit adds a parsing method to the BulkItemResponse class. In order to do that, the way DocWriteResponses are parsed has to be changed: ConstructingObjectParser/ObjectParser is removed in favor of a simpler and more readable way to parse these objects. DocWriteResponse now provides the parseInnerToXContent() method that can be used by subclasses (IndexResponse, UpdateReponse and DeleteResponse) to parse the current token/field and potentially update a DocWriteResponseBuilder. The DocWriteResponseBuilder is a simple POJO used to contain parsed values. It can be passed around from one parsing method to another parsing method. For example, this is what is done in IndexResponse: a IndexResponseBuilder is created in IndexResponse.fromXContent(), it get passed to IndexResponse.parseXContentFields() that parses fields specific to IndexResponse (like "created") and updates the context, delegating to DocWriteResponse.parseInnerToXContent() the parsing of any other field. Once all XContent is parsed, IndexResponse.fromXContent() uses the method IndexResponseBuilder.build() to create the new instance of IndexResponse. This behavior allow to reuse parsing code among the class hierarchy while keeping the current behavior. It also allows other objects like BulkItemResponse to reuse the same parsing code to parse DocWriteResponses. Finally, IndexResponseTests, UpdateResponseTests and DeleteResponseTests have been updated to introduce some random shuffling of fields before the XContent is parsed in order to ensure that the parsing code does not rely on field order.	2017-02-15 17:33:10 +01:00
Christoph Büscher	b963144254	Add xcontent parsing to completion suggestion option (#23071 ) This adds parsing from xContent to the CompletionSuggestion.Entry.Option. The completion suggestion option also inlines the xContent rendering of the containes SearchHit, so in order to reuse the SearchHit parser this also changes the way SearchHit is parsed from using a loop-based parser to using a ConstructingObjectParser that creates an intermediate map representation and then later uses this output to create either a single SearchHit or use it with additional fields defined in the parser for the completion suggestion option.	2017-02-15 16:52:17 +01:00
Jim Ferenczi	3c26754f87	Add BWC index for new released version 5.2.1	2017-02-15 11:14:37 +01:00
Jim Ferenczi	f1aaa71a7f	Create version constants for next bug fix version v5.2.2	2017-02-15 11:13:09 +01:00
Ryan Ernst	048c87d8a5	Improve setting deprecation message (#23156 ) This change modifies the deprecation log message emitted when a setting is found which is deprecated. The new message indicates docs for the deprecated settings can be found in the breaking changes docs for the next major version. closes #22849	2017-02-14 21:33:13 -08:00
Jason Tedor	6ac1cb660b	Cleanup RestGetIndicesAction.java This commit is just a code cleanup of RestGetIndicesAction.java. For example, we remove an unnecessary class, remove some unnecessary local variables, and simplify some code flow. Relates #23129	2017-02-14 16:51:27 -05:00
Jason Tedor	673754b1d5	Fix get source HEAD requests Get source HEAD requests incorrectly return a content-length header of 0. This commit addresses this by removing the special handling for get source HEAD requests, and just relying on the general mechanism that exists for handling HEAD requests in the REST layer. Relates #23151	2017-02-14 16:37:22 -05:00
Martijn van Groningen	cab43707dc	[percolator] Removed old 2.x bwc logic.	2017-02-14 22:17:17 +01:00
Areek Zillur	e178dc5493	Add request version asserting during replica operation (#23167 )	2017-02-14 15:40:55 -05:00
Simon Willnauer	a7a3729596	Add ExpandSearchPhase as a successor for the FetchSearchPhase (#23165 ) Now that we have more flexible search phases we should move the rather hacky integration of the collapse feature as a real search phase that can be tested and used by itself. This commit adds a new ExpandSearchPhase including a unittest for the phase. It's integrated into the fetch phase as an optional successor.	2017-02-14 17:14:17 +01:00
Adrien Grand	8d6a41f671	Nested queries should avoid adding unnecessary filters when possible. (#23079 ) When nested objects are present in the mappings, many queries get deoptimized due to the need to exclude documents that are not in the right space. For instance, a filter is applied to all queries that prevents them from matching non-root documents (`+: -_type:__`). Moreover, a filter is applied to all child queries of `nested` queries in order to make sure that the child query only matches child documents (`_type:__nested_path`), which is required by `ToParentBlockJoinQuery` (the Lucene query behing Elasticsearch's `nested` queries). These additional filters slow down `nested` queries. In 1.7-, the cost was somehow amortized by the fact that we cached filters very aggressively. However, this has proven to be a significant source of slow downs since 2.0 for users of `nested` mappings and queries, see #20797. This change makes the filtering a bit smarter. For instance if the query is a `match_all` query, then we need to exclude nested docs. However, if the query is `foo: bar` then it may only match root documents since `foo` is a top-level field, so no additional filtering is required. Another improvement is to use a `FILTER` clause on all types rather than a `MUST_NOT` clause on all nested paths when possible since `FILTER` clauses are more efficient. Here are some examples of queries and how they get rewritten: ``` "match_all": {} ``` This query gets rewritten to `ConstantScore(+:* -_type:__)` on master and `ConstantScore(_type:AutomatonQuery {\norg.apache.lucene.util.automaton.Automaton@4371da44})` with this change. The automaton is the complement of `_type:__` so it matches the same documents, but is faster since it is now a positive clause. Simplistic performance testing on a 10M index where each root document has 5 nested documents on average gave a latency of 420ms on master and 90ms with this change applied. ``` "term": { "foo": { "value": "0" } } ``` This query is rewritten to `+foo:0 #(ConstantScore(+: -_type:__))^0.0` on master and `foo:0` with this change: we do not need to filter nested docs out since the query cannot match nested docs. While doing performance testing in the same conditions as above, response times went from 250ms to 50ms. ``` "nested": { "path": "nested", "query": { "term": { "nested.foo": { "value": "0" } } } } ``` This query is rewritten to `+ToParentBlockJoinQuery (+nested.foo:0 #_type:__nested) #(ConstantScore(+:* -_type:__))^0.0` on master and `ToParentBlockJoinQuery (nested.foo:0)` with this change. The top-level filter (`-_type:__`) could be removed since `nested` queries only match documents of the parent space, as well as the child filter (`#_type:__nested`) since the child query may only match nested docs since the `nested` object has both `include_in_parent` and `include_in_root` set to `false`. While doing performance testing in the same conditions as above, response times went from 850ms to 270ms.	2017-02-14 16:05:19 +01:00
Adrien Grand	a969dad43e	Integrate IndexOrDocValuesQuery. (#23119 ) This gives Lucene the choice to use index/point-based queries or doc-values-based queries depending on which one is more efficient. This commit integrates this feature for: - long/integer/short/byte/double/float/half_float/scaled_float ranges, - date ranges, - geo bounding box queries, - geo distance queries.	2017-02-14 15:57:12 +01:00
Jun Ohtani	12bbe6e660	Merge pull request #23161 from johtani/support_keyword_to_analyze_api [Analyze]Support Keyword type in Analyze API	2017-02-14 23:22:32 +09:00
Christoph Büscher	abc8cd6c5f	Remove unused sourceAsBytes field in SearchHit	2017-02-14 14:08:38 +01:00
Simon Willnauer	aef0665ddb	Detach SearchPhases from AbstractSearchAsyncAction (#23118 ) Today all search phases are inner classes of AbstractSearchAsyncAction or one of it's subclasses. This makes unit testing of these classes practically impossible. This commit Extracts `DfsQueryPhase` and `FetchSearchPhase` or of the code that composes the actual query execution types and moves most of the fan-out and collect code into an `InitialSearchPhase` class that can be used to build initial search phases (phases that retry on shards). This will make modification to these classes simpler and allows to easily compose or add new search phases down the road if additional roundtrips are required.	2017-02-14 12:34:25 +01:00
Jun Ohtani	34ebb88650	[Analyze]Support Keyword type in Analyze API Add comment and clarify	2017-02-14 17:56:36 +09:00
Jun Ohtani	4d823d69f4	[Analyze]Support Keyword type in Analyze API	2017-02-14 16:41:16 +09:00
Jason Tedor	5343b87502	Handle bad HTTP requests When Netty decodes a bad HTTP request, it marks the decoder result on the HTTP request as a failure, and reroutes the request to GET /bad-request. This either leads to puzzling responses when a bad request is sent to Elasticsearch (if an index named "bad-request" does not exist then it produces an index not found exception and otherwise responds with the index settings for the index named "bad-request"). This commit addresses this by inspecting the decoder result on the HTTP request and dispatching the request to a bad request handler preserving the initial cause of the bad request and providing an error message to the client. Relates #23153	2017-02-13 17:39:25 -05:00
Jay Modi	61e383813d	Make the version of the remote node accessible on a transport channel (#23019 ) This commit adds a new method to the TransportChannel that provides access to the version of the remote node that the response is being sent on and that the request came from. This is helpful for serialization of data attached as headers.	2017-02-13 15:15:57 -05:00
Lee Hinman	b42d47770c	Fix total disk bytes returning negative value (#23093 ) * Fix total disk bytes returning negative value This adds a workaround for JDK-8162520 - https://bugs.openjdk.java.net/browse/JDK-8162520 Some filesystems can be so large that they return a negative value for their free/used/available disk bytes due to being larger than `Long.MAX_VALUE`. This adds protection for our `FsProbe` implementation and adds a test that it does the right thing.	2017-02-13 11:20:15 -07:00
jaymode	d8d03f45c2	Fix communication with 5.3.0 nodes This commit fixes communication with 5.3.0 nodes to send XContentType to these nodes since #22691 was backported to the 5.3 branch.	2017-02-13 13:15:51 -05:00
Jason Tedor	9dff5e2af7	Properly encode location header Today when trying to encode the location header to ASCII, we rely on the Java URI API. This API requires a proper URI which blows up whenever the URI contains, for example, a space (which can happen if the type, ID, or routing contain a space). This commit addresses this issue by properly encoding the URI. Additionally, we remove the need to create a URI simplifying the code flow. Relates #23133	2017-02-13 09:34:52 -05:00
Tanguy Leroux	de94c1253a	Expose WriteRequest.RefreshPolicy string representation (#23106 ) This commit changes the RefreshPolicy enum so that string representation are exposed. This will help the high level rest client to simply use refreshPolicy.getValue() to get the corresponding parameter value of a given refresh policy.	2017-02-13 10:49:46 +01:00
Boaz Leskes	29ea3059fc	Allow a cluster state applier to register an observer and wait for a better state (#23132 ) #21817 introduced the notion of a cluster state applier and banned those for sampling the cluster state directly (as it is not applied yet). Testing has exposed one exceptional use case - if the appliers want to spawn off a follow up it may require waiting for specific new cluster state (for example, the shard started action, called by the IndicesClusterStateService, may run into trouble connecting to the master and wait for a new master to be elected). This requires creating an observer which, in turn, samples the cluster state. An example failure can be seen at https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+master+periodic/1701/console This commit allows creating an observer from a cluster state applier. The observer is adapted to exclude any potential old cluster state in its logic.	2017-02-12 14:58:22 +02:00
Jason Tedor	0f21ed5b70	Fix template HEAD requests Template HEAD requests incorrectly return a content-length header of 0. This commit addresses this by removing the special handling for template HEAD requests, and just relying on the general mechanism that exists for handling HEAD requests in the REST layer. Relates #23130	2017-02-11 18:30:16 -05:00
Lee Hinman	13446937a5	Remove action.allow_id_generation setting (#23120 ) This was an undocumented and unsettable setting that allowed id generation. Resolves #23088	2017-02-10 14:04:40 -07:00
Jim Ferenczi	1ba73d9797	Fix GraphQuery expectation after Lucene upgrade to 6.5 (#23117 ) GraphQueries are now generated as simple clauses in BooleanQuery. So for instance a multi terms synonym will generate a GraphQuery but only for the side paths, the other part of the query will not be impacted. This means that we cannot apply `minimum_should_match` or `cutoff_frequency` on GraphQuery anymore (only ES 5.3 does that because we generate all possible paths if a query has at least one multi terms synonym). Starting in 5.4 multi terms synonym will now be treated as a single term when `minimum_should_match` is computed and will be ignored when `cutoff_frequency` is set. Fixes #23102	2017-02-10 18:20:00 +01:00
sabi0	09c7c5c82f	Limit IndexRequest toString() length (#22832 ) Limits the length of `IndexRequest#toString` which also limits the size of the task description generated for `IndexRequest`s. If the document being written is larger than 2kb we skip logging the _source entirely. This is because truncating the source is tricky and it isn't worth it.	2017-02-10 10:42:08 -05:00
Sebastian	976da87e8f	Fix some Javadoc typos (#23111 )	2017-02-10 15:53:30 +01:00
Jason Tedor	a6158398dd	Fix index HEAD requests Index HEAD requests incorrectly return a content-length header of 0. This commit addresses this by removing the special handling for index HEAD requests, and just relying on the general mechanism that exists for handling HEAD requests in the REST layer. Relates #23112	2017-02-10 09:44:01 -05:00
Jason Tedor	7ac44656df	Fix alias HEAD requests Alias HEAD requests incorrectly return a content-length header of 0. This commit addresses this by removing the special handling for alias HEAD requests, and just relying on the general mechanism that exists for handling HEAD requests in the REST layer. Relates #23094	2017-02-10 09:19:35 -05:00
Adrien Grand	709cc9ba65	Upgrade to lucene-6.5.0-snapshot-f919485. (#23087 )	2017-02-10 15:08:47 +01:00
Jay Modi	7018b6ac6f	Add BulkProcessor methods with XContentType parameter (#23078 ) This commit adds methods to the BulkProcessor that accept bytes and a XContentType to avoid content type detection. The methods that do not accept XContentType with bytes have been deprecated by this commit. Relates #22691	2017-02-10 08:59:37 -05:00
Jason Tedor	4f2b4724be	Cleanup RestGetAliasesAction.java This commit is just a code cleanup of RestGetAliasesAction.java. For example, we remove an unnecessary class, simplify a convenience method, and simplify some code flow. Relates #23095	2017-02-10 08:37:05 -05:00
Tanguy Leroux	e2e5937455	Use `typed_keys` parameter to prefix suggester names by type in search responses (#23080 ) This pull request reuses the typed_keys parameter added in #22965, but this time it applies it to suggesters. When set to true, the suggester names in the search response will be prefixed with a prefix that reflects their type.	2017-02-10 10:53:38 +01:00
Boaz Leskes	e0c8a6a3eb	Relax WaitActiveShardCountIT check of exception messages So ti wouldn't depend on BulkShardRequest.toString()	2017-02-09 23:14:09 +02:00
Areek Zillur	990918a655	fix failing tests for BulkShardRequest.tostring	2017-02-09 15:34:22 -05:00
Boaz Leskes	033defee9a	fix BulkShardRequestTests after changes to BulkShardRequest.toString	2017-02-09 21:05:21 +02:00
Boaz Leskes	cd1cb41603	Move EvilPeerRecoveryIT to a unit test in RecoveryDuringReplicationTests (#22900 ) EvillPeerRecoveryIT checks scenario where recovery is happening while there are on going indexing operation that already have been assigned a seq# . This is fairly hard to achieve and the test goes through a couple of hoops via the plugin infra to achieve that. This PR extends the unit tests infra to allow for those hoops to happen in unit tests. This allows the test to be moved to RecoveryDuringReplicationTests Relates to #22484	2017-02-09 20:14:03 +02:00
Jim Ferenczi	94087b3274	Removes ExpandCollapseSearchResponseListener, search response listeners and blocking calls This changes removes the SearchResponseListener that was used by the ExpandCollapseSearchResponseListener to expand collapsed hits. The removal of SearchResponseListener is not a breaking change because it was never released. This change also replace the blocking call in ExpandCollapseSearchResponseListener by a single asynchronous multi search request. The parallelism of the expand request can be set via CollapseBuilder#max_concurrent_group_searches Closes #23048	2017-02-09 18:06:10 +01:00
Boaz Leskes	33915aefd8	Improve BulkShardRequest.toString when it has only 1 internal request Now that we use bulk for single item indexing, this is often the case. Having an indicator of the id of the indexed document helps debugging. It now looks like this `BulkShardRequest to [[test][0]] containing [index {[test][type][AVojzy9ZxfWASZ-ysmN7], source[{"auto":true}]}]`	2017-02-09 18:59:49 +02:00
Luca Cavanna	90ea778c17	Cluster allocation explain to never return empty response body (#23054 ) Empty response bodies should only be sent for HEAD requests, otherwise we should always send back info about the exception that was thrown. Removed some manual exception handling in the REST action that should be rather bubbled up and handled by our rest action infra like every other rest action does.	2017-02-09 17:46:39 +01:00
Luca Cavanna	9f60924ed5	Remove redundant reads of human flag (#23074 ) The human flag is centrally handled in RestChannel, no need to have Rest actions manually read it and set it to the builder	2017-02-09 14:58:01 +01:00
Christoph Büscher	b85fa54ee7	Tests: Renaming InternalSearchHitsTests to SearchHitsTests The class under test changed its name from InternalSearchHit(s) to just SearchHit(s), renaming the tests accordingly.	2017-02-09 14:17:21 +01:00
Tanguy Leroux	3553522328	Add parameter to prefix aggs name with type in search responses (#22965 ) This pull request adds a new parameter to the REST Search API named `typed_keys`. When set to true, the aggregation names in the search response will be prefixed with a prefix that reflects the internal type of the aggregation. Here is a simple example: ``` GET /_search?typed_keys { "aggs": { "tweets_per_user": { "terms": { "field": "user" } } }, "size": 0 } ``` And the response: ``` { "aggs": { "sterms:tweets_per_user": { ... } } } ``` This parameter is intended to make life easier for REST clients that could parse back the prefix and could detect the type of the aggregation to parse. It could also be implemented for suggesters.	2017-02-09 11:19:04 +01:00
Simon Willnauer	e02d5563f4	Harden ops counting in AbstractSearchAsyncAction (#23045 ) Today we account for too many response with an `IllegalStateException` in `AbstractSearchAsyncAction` while this is something that should never happen we should rather assert that we are always have less or equal the number of expected ops when waiting for responses.	2017-02-09 09:30:13 +01:00
Luca Cavanna	b5f5356c4a	Remove getDefaultScriptingLanguage from QueryParseContext (#23043 ) The method is not needed anymore, was needed only when we supported setting a legacy default lang, which was removed with #21607 Relates to #21607	2017-02-09 09:03:26 +01:00
Nik Everett	f7071325c4	Fix generics on LeadDocLookup (#23060 ) All the warnings were upsetting me. This doesn't change behavior.	2017-02-08 18:59:24 -05:00
Christoph Büscher	e09f3ecbb3	Add xcontent parsing to suggestion options (#23018 ) This adds parsing from xContent to Suggestion.Entry.Option and Termsuggestion.Entry.Option.	2017-02-08 19:03:12 +01:00
Jay Modi	7f3769c745	Remove ldjson support and document ndjson for bulk/msearch (#23049 ) This commit removes support for the `application/x-ldjson` Content-Type header as this was only used in the first draft of the spec and had very little uptake. Additionally, the docs for bulk and msearch have been updated to specifically call out ndjson and mention that the newline character may be preceded by a carriage return. Finally, the bulk request handling of the carriage return has been improved to remove this character from the source. Closes #23025	2017-02-08 11:55:50 -05:00
Simon Willnauer	df932ef68f	Fix line len	2017-02-08 16:41:41 +01:00
Simon Willnauer	d45761e488	Fork off a search thread before sending back fetched responses This is just a temporary fix until #23048 is fixed. FieldCollapsing is executing blocking calls on a network thread which causes potential deadlocks and trips assertions. Relates to #23048	2017-02-08 15:27:08 +01:00
Simon Willnauer	ecb01c15b9	Fold InternalSearchHits and friends into their interfaces (#23042 ) We have a bunch of interfaces that have only a single implementation for 6 years now. These interfaces are pretty useless from a SW development perspective and only add unnecessary abstractions. They also require lots of casting in many places where we expect that there is only one concrete implementation. This change removes the interfaces, makes all of the classes final and removes the duplicate `foo` `getFoo` accessors in favor of `getFoo` from these classes.	2017-02-08 14:40:08 +01:00
Simon Willnauer	2d6d871f5c	Raise a phase failure if fetch phase gets rejected	2017-02-08 12:52:18 +01:00
Boaz Leskes	0161edae10	MasterFaultDetection can start after the initial cluster state has been processed and the NodeConnectionService connect to the new master (#23037 ) After the first cluster state from a new master is processed, NodeConnectionService guarantees we connect to the new master. This removes the need to explicitly connect to the master in the MasterFaultDetection code making it simpler and bypasses the assertion triggered due to the blocking operation on the cluster state thread. Relates to #22828	2017-02-08 13:49:06 +02:00
Simon Willnauer	a8b376670c	Separate reduce (aggs, suggest and profile) from merging fetched hits (#23017 ) Today we carry on all search results including aggs, suggest and profile results until we have successfully fetched all hits for the search request. This can potentially hold on to a large amount of memory if there are heavy aggregations involved. With this change aggs and profiles are entirely consumed an released for GC before the fetch phase is executing. This is a first step towards reducing results on-the-fly if the number of non-empty response are large.	2017-02-08 10:11:51 +01:00
Yannick Welsch	9154686623	Remove legacy primary shard allocation mode based on versions (#23016 ) Elasticsearch v5.0.0 uses allocation IDs to safely allocate primary shards whereas prior versions of ES used a version-based mode instead. Elasticsearch v5 still has support for version-based primary shard allocation as it needs to be able to load 2.x shards. ES v6 can drop the legacy support.	2017-02-08 10:00:55 +01:00
Boaz Leskes	a512ab32fb	Increase time out tolerance in NoMasterNodeIT. see https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+master+multijob-intake/746/console	2017-02-08 08:50:26 +02:00
Lee Hinman	b3c27a7fdd	Disallow include_in_all for 6.0+ indices Since `_all` is now deprecated and cannot be set for new indices, we should also disallow any field that has the `include_in_all` parameter set. Resolves #22923	2017-02-07 19:31:51 -07:00
Tim Brooks	fcc568fd8d	Add methods requiring connect to forbidden apis (#22964 ) This is related to #22116. This commit adds calls that require SocketPermission connect to forbidden APIs. The following calls are now forbidden: - java.net.URL#openStream() - java.net.URLConnection#connect() - java.net.URLConnection#getInputStream() - java.net.Socket#connect(java.net.SocketAddress) - java.net.Socket#connect(java.net.SocketAddress, int) - java.nio.channels.SocketChannel#open(java.net.SocketAddress) - java.nio.channels.SocketChannel#connect(java.net.SocketAddress)	2017-02-07 14:41:50 -06:00
Boaz Leskes	ba06c14a97	TransportService.connectToNode should validate remote node ID (#22828 ) #22194 gave us the ability to open low level temporary connections to remote node based on their address. With this use case out of the way, actual full blown connections should validate the node on the other side, making sure we speak to who we think we speak to. This helps in case where multiple nodes are started on the same host and a quick node restart causes them to swap addresses, which in turn can cause confusion down the road.	2017-02-07 22:11:32 +02:00
Tim Brooks	adc1184dd0	Fix broken test in FileSystemUtilsTests Commit `ee84ce09d7` changed an exception message without changing the corresponding test. This commit fixes the related test.	2017-02-07 12:50:07 -06:00
Tim Brooks	ee84ce09d7	Allow openFileURLStream(URL) to open jars This is related to #23020. There are some cases for where this method might be called with a URL to a file inside a jar. This commit allows this method to read URLs with a protocol of 'jar:/'.	2017-02-07 11:42:27 -06:00
Ryan Ernst	470ad1ae4a	Settings: Add secure settings validation on startup (#22894 ) Secure settings from the elasticsearch keystore were not yet validated. This changed improves support in Settings so that secure settings more seamlessly blend in with normal settings, allowing the existing settings validation to work. Note that the setting names are still not validated (yet) when using the elasticsearc-keystore tool.	2017-02-07 09:34:41 -08:00
Tim Brooks	27b7d9bd8d	Add FileSystemUtil method to read 'file:/' URLs (#23020 ) As part of #22116 we are going to forbid usage of api java.net.URL#openStream(). However in a number of places across the we use this method to read files from the local filesystem. This commit introduces a helper method openFileURLStream(URL url) to read files from URLs. It does specific validation to only ensure that file:/ urls are read. Additionlly, this commit removes unneeded method FileSystemUtil.newBufferedReader(URL, Charset). This method used the openStream () method which will soon be forbidden. Instead we use the Files.newBufferedReader(Path, Charset).	2017-02-07 10:24:22 -06:00
Jay Modi	c898e8ab83	Add support for newline delimited JSON Content-Type (#22947 ) This commit adds support for the newline delimited JSON Content-Type, which is how the bulk, multi-search, and multi-search template APIs expect data to be formatted. The `elasticsearch-js` client has also been using this content type for these types of requests. Closes #22943	2017-02-07 09:20:06 -05:00
Simon Willnauer	dc659feeb4	Add a setting to disable remote cluster connections on a node (#23005 ) Today either all nodes in the cluster connect to remote clusters of only nodes that have remote clusters configured in their node config. To allow global remote cluster configuration but restrict connections to a set of nodes in the cluster this change adds a new setting `search.remote.connect` (defaults to `true`) to allow to disable remote cluster connections on a per node basis.	2017-02-07 09:59:24 +01:00
Nik Everett	0d6e622242	Make dates be ReadableDateTimes in scripts (#22948 ) Instead of longs. If you want millis since epoch you can call doc.date_field.value.millis. Relates to #22875	2017-02-06 16:44:56 -05:00
Nicholas Knize	1c9fdfd1b3	Remove GeoPointFieldMapper abstraction In order to support the evolving GeoPoint encodings in Lucene 5 and 6, ES 2.x and 5.x implements an abstraction layer to the GeoPointFieldMapper classes. As of 5.x the geo_point field mapper settled on using Lucene's more performant LatLonPoint field type and deprecated all other encodings. In 6.0 all encodings except LatLonPoint have been removed rendering this abstraction layer useless. This commit removes the abstraction layer and renames the LatLonPointFieldMapper back to GeoPointFieldMapper to mantain consistency with ES field naming.	2017-02-06 14:17:21 -06:00
Christoph Büscher	033f03109f	[Tests] Adding tests for AvgAggregator and InternalAvg (#23000 )	2017-02-06 20:05:40 +01:00
Ali Beyad	42a9f95fde	This commit changes the exception type thrown when trying to (#22921 ) create a snapshot with a name that already exists in the repository. Instead of throwing a SnapshotCreateException, which results in a generic 500 status code, a duplicate snapshot name will throw a InvalidSnapshotNameException, which will result in a 400 status code (bad request).	2017-02-06 11:39:59 -06:00
Adrien Grand	eb26e1a292	Add unit tests to histogram aggregations. (#22961 )	2017-02-06 18:18:21 +01:00
Simon Willnauer	f09c4e1cdb	Expose `search.highlight.term_vector_multi_value` as a node level setting (#22999 ) This setting was missed in the great settings refactoring and should be exposed via node level settings.	2017-02-06 18:17:34 +01:00
Simon Willnauer	7513c6e4eb	Remove QUERY_AND_FETCH search type (#22996 ) `QUERY_AND_FETCH` has been treated as an internal optimization for 2 major versions. This commit removes the search type and it's implementation details and folds the optimization in the case of a single shard into the search controller such that every search with a single shard (non DFS) will receive this optimization.	2017-02-06 17:10:03 +01:00
Boaz Leskes	5e7d22357f	Connect to new nodes concurrently (#22984 ) When a node receives a new cluster state from the master, it opens up connections to any new node in the cluster state. That has always been done serially on the cluster state thread but it has been a long standing TODO to do this concurrently, which is done by this PR. This is spin off of #22828, where an extra handshake is done whenever connecting to a node, which may slow down connecting. Also, the handshake is done in a blocking fashion which triggers assertions w.r.t blocking requests on the cluster state thread. Instead of adding an exception, I opted to implement concurrent connections which both side steps the assertion and compensates for the extra handshake.	2017-02-06 16:32:41 +01:00
Martijn van Groningen	e4663d6263	added comment	2017-02-06 15:16:16 +01:00
Martijn van Groningen	c8d470f190	Change `org.elasticsearch.bootstrap.JNAKernel32Library$SizeT` constructor's modifier to public. Otherwise `NativeMappedConverter` can't construct this class. Closes #22991	2017-02-06 15:16:16 +01:00
Christoph Büscher	d02170b277	Add parsing from xContent to MainResponse (#22934 ) Add parsing from xContent to MainResponse	2017-02-06 12:30:42 +01:00
Yannick Welsch	6f6596cfb5	Revert "Reduce log-level of IndexPrimaryRelocationIT to hunt Heisenbug" This reverts commit `d0fa6a9bd8`.	2017-02-06 11:40:39 +01:00
Adrien Grand	76f779486b	5.2.1 is now on Lucene 6.4.1 too.	2017-02-06 10:02:31 +01:00
Adrien Grand	c8496fc4f4	Upgrade to Lucene 6.4.1. (#22978 )	2017-02-06 09:28:43 +01:00
Martijn van Groningen	9201ee82f6	[TEST] Added unit tests for sum aggs. Relates to #22278	2017-02-06 08:32:10 +01:00
Lee Hinman	39e7c30912	Change certain replica failures not to fail the replica shard This changes the way that replica failures are handled such that not all failures will cause the replica shard to be failed or marked as stale. In some cases such as refresh operations, or global checkpoint syncs, it is "okay" for the operation to fail without the shard being failed (because no data is out of sync). In these cases, instead of failing the shard we should simply fail the operation, and, in the event it is a user-facing operation, return a 5xx response code including the shard-specific failures. This was accomplished by having two forms of the `Replicas` proxy, one that is for non-write operations that does not fail the shard, and one that is for write operations that will fail the shard when an operation fails. Relates to #10708	2017-02-03 14:39:46 -07:00
Nik Everett	70e3cce904	Fix name of `enable_position_increments` (#22895 ) It was accidentally renamed `enabled_position_increment` in the cleanups for 5.0. This adds `enable_position_increment` as a deprecated alias so it will continue to work.	2017-02-03 16:28:27 -05:00
Nicholas Knize	b1a6b227e1	Remove deprecated geo query parameters, and GeoPointDistanceRangeQuery This commit removes the following queries and parameters (which were deprecated in 5.0): * GeoPointDistanceRangeQuery * coerce, and ignore_malformed for GeoBoundingBoxQuery, GeoDistanceQuery, GeoPolygonQuery, and GeoDistanceSort	2017-02-03 10:08:00 -06:00
Tim Brooks	f70188ac58	Remove connect SocketPermissions from core (#22797 ) This is related to #22116. Core no longer needs `SocketPermission` `connect`. This permission is relegated to these modules/plugins: - transport-netty4 module - reindex module - repository-url module - discovery-azure-classic plugin - discovery-ec2 plugin - discovery-gce plugin - repository-azure plugin - repository-gcs plugin - repository-hdfs plugin - repository-s3 plugin And for tests: - mocksocket jar - rest client - httpcore-nio jar - httpasyncclient jar	2017-02-03 09:39:56 -06:00
Jason Tedor	9a0b216c36	Upgrade checkstyle to version 7.5 This commit upgrades the checkstyle configuration from version 5.9 to version 7.5, the latest version as of today. The main enhancement obtained via this upgrade is better detection of redundant modifiers. Relates #22960	2017-02-03 09:46:44 -05:00
Jason Tedor	01871e4def	Fix compilation in RecoverySourceHandlerTests This error arose after the signature of a method was changed.	2017-02-03 08:48:29 -05:00
Jason Tedor	6e9940283b	Avoid losing ops in file-based recovery When a primary is relocated from an old node to a new node, it can have ops in its translog that do not have a sequence number assigned. When a file-based recovery is started, this can lead to skipping these ops when replaying the translog due to a bug in the recovery logic. This commit addresses this bug and adds a test in the BWC tests. Relates #22945	2017-02-03 08:11:57 -05:00
Jim Ferenczi	4876448e39	Consilify get-field-mapping docs (#22936 ) This change also removes the reference to the difference bewteen full name and index name. They are always the same since 2.x and `name` does not refer anymore to `author.name` automatically. A simple pattern must be used instead. Remove redundant code that checks the field name twice.	2017-02-03 10:04:31 +01:00
Yannick Welsch	d0fa6a9bd8	Reduce log-level of IndexPrimaryRelocationIT to hunt Heisenbug	2017-02-03 09:46:07 +01:00
Chris Buonocore	365d33efe3	Handle missing plugin name in remove command Today if a user invokes the remove plugin command without specifying the name of a plugin to remove, we arrive at a null pointer exception. This commit adds logic to cleanly handle this situation and provide clear feedback to the user. Relates #22930	2017-02-02 19:39:56 -05:00
Nicholas Knize	58c34f0da9	Fix NPE in RangeFieldMapper.doXContentBody RangeFieldMapper.doXContentBody should only serialize format and locale when type is set to 'date_range'. closes #22925	2017-02-02 16:52:33 -06:00
Nik Everett	71b2655bb3	Fix outdated example in javadoc The code changed and the example in the javadoc didn't. Closes #22932	2017-02-02 14:28:48 -05:00
Areek Zillur	ba8ad397a1	Use bulk action interally for update action (#22915 ) Currently, update action internally uses deprecated index and delete transport actions. As of #21964, these tranport actions were deprecated in favour of using single item bulk request. In this commit, update action uses single item bulk action.	2017-02-02 14:21:53 -05:00
Jay Modi	7520a107be	Optionally require a valid content type for all rest requests with content (#22691 ) This change adds a strict mode for xcontent parsing on the rest layer. The strict mode will be off by default for 5.x and in a separate commit will be enabled by default for 6.0. The strict mode, which can be enabled by setting `http.content_type.required: true` in 5.x, will require that all incoming rest requests have a valid and supported content type header before the request is dispatched. In the non-strict mode, the Content-Type header will be inspected and if it is not present or not valid, we will continue with auto detection of content like we have done previously. The content type header is parsed to the matching XContentType value with the only exception being for plain text requests. This value is then passed on with the content bytes so that we can reduce the number of places where we need to auto-detect the content type. As part of this, many transport requests and builders were updated to provide methods that accepted the XContentType along with the bytes and the methods that would rely on auto-detection have been deprecated. In the non-strict mode, deprecation warnings are issued whenever a request with body doesn't provide the Content-Type header. See #19388	2017-02-02 14:07:13 -05:00
Nicholas Knize	b41d5747f0	Reduce GeoDistance insanity GeoDistance query, sort, and scripts make use of a crazy GeoDistance enum for handling 4 different ways of computing geo distance: SLOPPY_ARC, ARC, FACTOR, and PLANE. Only two of these are necessary: ARC, PLANE. This commit removes SLOPPY_ARC, and FACTOR and cleans up the way Geo distance is computed.	2017-02-02 12:39:42 -06:00
Igor Motov	c34b63dadd	Expand AbstractSerializingTestCase and AbstractWireSerializingTestCase to test diff serialization This commit adds two additional test cases that can be used to verify correct diff serialization in additional to binary and xcontent serialization.	2017-02-02 12:19:53 -05:00
Tanguy Leroux	f86fd62821	Parse elasticsearch exception's root causes (#22924 ) This commit change ElasticsearchException.failureFromXContent() method so that it now parses root causes which were ignored before, and adds them as suppressed exceptions of the returned exception.	2017-02-02 17:00:16 +01:00
Nik Everett	dacc150934	Expose multi-valued dates to scripts and document painless's date functions (#22875 ) Implemented by wrapping an array of reused `ModuleDateTime`s that we grow when needed. The `ModuleDateTime`s are reused when we move to the next document. Also improves the error message returned when attempting to modify the `ScriptdocValues`, removes a couple of allocations, and documents that the date functions are available in Painless. Relates to #22162	2017-02-01 21:57:07 -05:00
Ali Beyad	9f97eec12e	Fixes the Version constants for 5.2.0, 5.2.1, and 5.3.0 to have the correct Lucene version (6.4.0)	2017-02-01 12:47:36 -05:00
Lee Hinman	8d83edc4a5	Disallow introducing illegal object mappings (double '..') This disallows object mappings that would accidentally create something like `foo..bar`, which is then unparsable for the `bar` field as it does not know what its parent is. Resolves #22794	2017-02-01 09:03:20 -07:00
Ali Beyad	5c1410d031	Removes premature addition of v5.4.0 constant	2017-02-01 10:43:53 -05:00
Tanguy Leroux	3cfcd1acb7	[TEST] Fix BytesRestResponseTests.testNoErrorFromXContent	2017-02-01 10:34:42 +01:00
Boaz Leskes	5bf9cb9f70	remove await fix from testCannotAllocateStaleReplicaExplanation	2017-02-01 10:31:26 +01:00
Boaz Leskes	06b8a1ada7	fix testCannotAllocateStaleReplicaExplanation node management The test tried to create a situation where a stale replica is the only shard available. It did so by stopping the node with the replica, indexing some, stopping the primary node, starting a new node. This is flawed because the newly started node may reuse the data path of the primary node and things go back to green. Instead we should make sure that the replica is on the path that will be selected when the new node is started (i.e., the path with the smaller ordinal)	2017-02-01 10:30:42 +01:00
Tanguy Leroux	c74679b6b9	Add parsing method to BytesRestResponse's error (#22873 ) This commit adds a BytesRestResponse.errorFromXContent() method to parse the error returned by BytesRestResponse. It returns a ElasticsearchStatusException instance.	2017-02-01 10:11:17 +01:00
Ali Beyad	f436a06971	[TEST] adds AwaitsFix on two failing tests	2017-01-31 22:42:23 -05:00
Ali Beyad	6f2222f8fb	[TEST] fix node allocation result check in explain API test	2017-01-31 20:11:01 -05:00
Ali Beyad	547eb5c22f	Include stale replica shard info when explaining an unassigned primary (#22826 ) Currently, if a previously allocated shard has no in-sync copy in the cluster, but there is a stale replica copy, the explain API does not include information about the stale replica copies in its output. This commit includes any shard copy information available (even for stale copies) when explaining an unassigned primary shard that was previously allocated in the cluster. This situation can arise as follows: imagine an index with 1 primary and 1 replica and a cluster with 2 nodes. If the node holding the replica is shut down, and data continues to be indexed, only the primary will have the latest data and the replica that has gone offline will be marked as stale. Now, suppose the node holding the primary is shut down. There are no copies of the shard data in the cluster. Now, start the first stopped node (holding the stale replica) back up. The cluster is red because there is no in-sync copy available. Running the explain API before would inform the user that there is no valid shard copy in the cluster for that shard, but it would not provide any information about the existence of the stale replica that exists on the restarted node. With this commit, the explain API provides information about all the stale replica copies when explaining the unassigned primary.	2017-01-31 16:31:55 -06:00
Ali Beyad	c223457ba1	Adds v5.2.1 and v5.4.0 constants and bwc index for 5.2.0	2017-01-31 17:12:02 -05:00
Jack Conradson	3d2626c4c6	Change Namespace for Stored Script to Only Use Id (#22206 ) Currently, stored scripts use a namespace of (lang, id) to be put, get, deleted, and executed. This is not necessary since the lang is stored with the stored script. A user should only have to specify an id to use a stored script. This change makes that possible while keeping backwards compatibility with the previous namespace of (lang, id). Anywhere the previous namespace is used will log deprecation warnings. The new behavior is the following: When a user specifies a stored script, that script will be stored under both the new namespace and old namespace. Take for example script 'A' with lang 'L0' and data 'D0'. If we add script 'A' to the empty set, the scripts map will be ["A" -- D0, "A#L0" -- D0]. If a script 'A' with lang 'L1' and data 'D1' is then added, the scripts map will be ["A" -- D1, "A#L1" -- D1, "A#L0" -- D0]. When a user deletes a stored script, that script will be deleted from both the new namespace (if it exists) and the old namespace. Take for example a scripts map with {"A" -- D1, "A#L1" -- D1, "A#L0" -- D0}. If a script is removed specified by an id 'A' and lang null then the scripts map will be {"A#L0" -- D0}. To remove the final script, the deprecated namespace must be used, so an id 'A' and lang 'L0' would need to be specified. When a user gets/executes a stored script, if the new namespace is used then the script will be retrieved/executed using only 'id', and if the old namespace is used then the script will be retrieved/executed using 'id' and 'lang'	2017-01-31 13:27:02 -08:00
Boaz Leskes	eb36b82de4	Seq Number based recovery should validate last lucene commit max seq# (#22851 ) The seq# base recovery logic relies on rolling back lucene to remove any operations above the global checkpoint. This part of the plan is not implemented yet but have to have these guarantees. Instead we should make the seq# logic validate that the last commit point (and the only one we have) maintains the invariant and if not, fall back to file based recovery. This commit adds a test that creates situation where rollback is needed (primary failover with ops in flight) and fixes another issue that was surfaced by it - if a primary can't serve a seq# based recovery request and does a file copy, it still used the incoming `startSeqNo` as a filter. Relates to #22484 & #10708	2017-01-31 20:27:31 +01:00
Ryan Ernst	29f63c78cc	Internal: Convert empty and size checks of settings to not use getAsMap() (#22890 ) With the new secure settings, methods like getAsMap() no longer work correctly as a means of checking for empty settings, or the total size. This change converts the existing uses of that method to use methods directly on Settings. Note this does not update the implementations to account for SecureSettings, as that will require a followup which changes how secure settings work.	2017-01-31 10:44:09 -08:00
Jim Ferenczi	f6d38d480a	Integrate UnifiedHighlighter (#21621 ) * Integrate UnifiedHighlighter This change integrates the Lucene highlighter called "unified" in the list of supported highlighters for ES. This highlighter can extract offsets from either postings, term vectors, or via re-analyzing text. The best strategy is picked automatically at query time and depends on the field and the query to highlight.	2017-01-31 19:06:03 +01:00
Ryan Ernst	a4f6edec52	Settings: Fix settings reading to account for defaults (#22871 ) In #22762, settings preparation during bootstrap was changed slightly to account for SecureSettings, by starting with a fresh settings builder after reading the initial configuration. However, this the defaults from system properties were never re-read. This change fixes that bug (which was never released). closes #22861	2017-01-30 14:42:40 -08:00
Christoph Büscher	4e613139dc	Ensure fixed serialization order of InnerHitBuilder (#22820 ) Usually the order in which we serialize sets and maps of things doesn't matter, but since InnerHitBuilder is part of SearchSourceBuilder, which is in turn used as a cache key in its bytes serialization, we need to ensure the order of all these fields when writing them to an output stream. This adds tests and makes sure we iterate over the scriptField set and the childInnerHits map in a fixed order. Closes #22808	2017-01-30 19:28:55 +01:00
Alex Bumbu	41abf6e81d	Add used memory amount to CircuitBreakingException message (#22521 )	2017-01-28 16:54:13 +00:00
Nik Everett	e042c77301	Add tests for reducing top hits (#22837 ) Also adds many `equals` and `hashCode` implementations and moves the failure printing in `MatchAssertion` into a common spot and exposes it over `assertEqualsWithErrorMessageFromXContent` which does an object equality test but then uses `toXContent` to print the differences. Relates to #22278	2017-01-27 20:54:11 -05:00
Boaz Leskes	b1f0d8f4cf	await fix testRecoveryWaitsForOps	2017-01-27 23:26:23 +01:00
Boaz Leskes	204df2a199	fix @TestLogging annotation	2017-01-27 23:15:51 +01:00
Simon Willnauer	e946ec0c33	Don't convert source to UTF-8 it might not be valid UTF-8	2017-01-27 22:31:44 +01:00
Nik Everett	2e48fb8294	Move delete by query helpers into core (#22810 ) This moves the building blocks for delete by query into core. This should enabled two thigns: 1. Plugins other than reindex to implement "bulk by scroll" style operations. 2. Plugins to directly call delete by query. Those plugins should be careful to make sure that task cancellation still works, but this should be possible. Notes: 1. I've mostly just moved classes and moved around tests methods. 2. I haven't been super careful about cohesion between these core classes and reindex. They are quite interconnected because I wanted to make the change as mechanical as possible. Closes #22616	2017-01-27 16:09:18 -05:00
Ryan Ernst	aad51d44ab	S3 repository: Add named configurations (#22762 ) * S3 repository: Add named configurations This change implements named configurations for s3 repository as proposed in #22520. The access/secret key secure settings which were added in #22479 are reverted, and the only secure settings are those with the new named configs. All other previously used settings for the connection are deprecated. closes #22520	2017-01-27 10:42:45 -08:00
Boaz Leskes	0f58f3f34b	fix TestLogging instructions for the remove of single doc indexing actions + add some seq# related info	2017-01-27 19:21:11 +01:00
Nik Everett	e36e5fc994	Remove annotation Not allowed.	2017-01-27 12:37:35 -05:00
Nik Everett	8abd4101eb	Add tests for reducing top hits Also adds many `equals` and `hashCode` implementations and moves the failure printing in `MatchAssertion` into a common spot and exposes it over `assertEqualsWithErrorMessageFromXContent` which does an object equality test but then uses `toXContent` to print the differences. Relates to #22278	2017-01-27 12:32:17 -05:00
Jason Tedor	930282e161	Introduce sequence-number-based recovery This commit introduces sequence-number-based recovery. When a replica has fallen out of sync, rather than performing a file-based recovery we first attempt to replay operations since the last local checkpoint on the replica. To do this, at the start of recovery the replica tells the primary what its local checkpoint is. The primary will then wait for all operations between that local checkpoint and the current maximum sequence number to complete; this is to ensure that there are no gaps in the operations that will be replayed from the primary to the replica. This is a best-effort attempt as we currently have no guarantees on the primary that these operations will be available; if we are not able to replay all operations in the desired range, we just fallback to file-based recovery. Later work will strengthen the guarantees. Relates #22484	2017-01-27 08:16:38 -08:00
Simon Willnauer	417c93c570	First step towards separating individual search phases (#22802 ) At this point AbstractSearchAsyncAction is just a base-class for the first phase of a search where we have multiple replicas for each shardID. If one of them is not available we move to the next one. Yet, once we passed that first stage we have to work with the shards we succeeded on the initial phase. Unfortunately, subsequent phases are not fully detached from the initial phase since they are all non-static inner classes. In future changes this will be changed to detach the inner classes to test them in isolation and to simplify their creation. The AbstractSearchAsyncAction should be final and it should just get a factory for the next phase instead of requiring subclasses etc.	2017-01-27 15:53:41 +01:00
Igor Motov	b068814d10	Fix hanging cancelling task with no children Cancelling tasks with no cancellable children can cause the cancellation operation to hang. This commit fixes this issue.	2017-01-27 08:03:02 -05:00
Tanguy Leroux	ea7077fb1b	Add parsing method for ElasticsearchException.generateFailureXContent() (#22815 ) This commit adds a ElasticsearchException.failureFromXContent() that can be used to parse the result of ElasticsearchException.generateFailureXContent().	2017-01-27 10:12:58 +01:00
Nik Everett	1baa884ab7	Fix TophitsAggregatorTests It needs a DirectoryReader so it has to be careful. Closes #22818	2017-01-26 14:08:30 -05:00
Nik Everett	f8c28711be	Merge some equivalent interfaces (#22816 ) Remove `FromXContent` and use `CheckedFunction` instead. Remove `FromXContentWithContext` and use `ContentParser` instead.	2017-01-26 13:15:29 -05:00
Simon Willnauer	a475323aa1	Invalidate cached query results if query timed out (#22807 ) Today we cache query results even if the query timed out. This is obviously problematic since results are not complete. Yet, the decision if a query timed out or not happens too late to simply not cache the result since if we'd just throw an exception all currently waiting requests with the same request / cache key would fail with the same exception without the option to access the result or to re-execute. Instead, this change will allow the request to enter the cache but invalidates it immediately. Concurrent request might not get executed and return the timed out result which is not absolutely correct but very likely since identical requests will likely timeout as well. As a side-effect we won't hammer the node with concurrent slow searches but rather only execute one of them and return shortly cached result. Closes #22789	2017-01-26 16:45:29 +01:00
Tanguy Leroux	1fa2734566	[TEST] Fix ElasticsearchExceptionTests Some test failures can happen in ElasticsearchExceptionTests, this commit fixes them.	2017-01-26 16:33:56 +01:00
Tanguy Leroux	be96278c95	Add parsing method for ElasticsearchException.generateThrowableXContent() (#22783 ) The output of the ElasticsearchException.generateThrowableXContent() method can be parsed back by the ElasticsearchException.fromXContent() method. This commit adds unit tests in the style of the to-and-from-xcontent tests we already have for other parsing methods. It also relax the strict parsing of the ElasticsearchException.fromXContent() so that it does not throw an exception when custom metadata and headers are parsed, as long as they are either strings or arrays of strings. Every other type is ignored at parsing time.	2017-01-26 15:17:07 +01:00
Simon Willnauer	f128b7a7fe	Improve connection closing in `RemoteClusterConnection` (#22804 ) Some tests verify that all connection have been closed but due to the async / concurrent nature of `RemoteClusterConnection` there are situations where we notify listeners that trigger tests to finish before we actually closed all connections. The race is very very small and has no impact on the code correctness. This commit documents and improves the way we close connections to ensure test won't fail with false positives. Closes #22803	2017-01-26 13:58:26 +01:00
Simon Willnauer	281250dec9	Remove DFS_QUERY_AND_FETCH as a search type (#22787 ) This commit removes the search type `dfs_query_and_fetch` without a replacement. We don't allow to use this type via REST since 2.x but still keep it around for no particular reason. There we no users complaining about the availability. This should now be removed from the codebase. `query_and_fetch` is still used internally to safe a roundtrip if there is only one shard but it can't be used via the rest interface.	2017-01-26 09:14:44 +01:00
Tim Brooks	719e75bb3f	Add repository-url module and move URLRepository (#22752 ) This is related to #22116. URLRepository requires SocketPermission connect. This commit introduces a new module called "repository-url" where URLRepository will reside. With the new module, permissions can be removed from core.	2017-01-25 17:09:25 -06:00
Nik Everett	d704a880e7	Add tests for top_hits aggregation (#22754 ) Add unit tests for `TopHitsAggregator` and convert some snippets in docs for `top_hits` aggregation to `// CONSOLE`. Relates to #22278 Relates to #18160	2017-01-25 16:15:50 -05:00
Martijn van Groningen	f6ed39aa08	Merge branch 'pr/22772'	2017-01-25 17:15:24 +01:00
Martijn van Groningen	81e40e3139	[TEST] Added this for `93a28b0acf` submitted via #22772	2017-01-25 17:08:17 +01:00
Jason Tedor	cb822b4670	Fix typo in comment in OsProbe.java This commit fixes a silly typo in a comment relating to cgroups in OsProbe.java.	2017-01-25 06:30:51 -05:00
Colin Goodheart-Smithe	a9135cd636	RangeQuery WITHIN case now normalises query (#22431 ) Previous to his change when the range query was rewritten to an unbounded range (`[* TO *]`) it maintained the timezone and format for the query. This means that queries with different timezones and format which are rewritten to unbounded range queries actually end up as different entries in the search request cache. This is inefficient and unnecessary so this change nulls the timezone and format in the rewritten query so that regardless of the timezone or format the rewritten query will be the same. Although this does not fix #22412 (since it deals with the WITHIN case rather than the INTERSECTS case) it is born from the same arguments	2017-01-25 10:37:15 +00:00
Boaz Leskes	ed94f75a15	Remove EngineClosedException All usage has been removed in https://github.com/elastic/elasticsearch/pull/22631, which is back ported to 5.x. This means 6.x will never get it on the wire and we can remove it	2017-01-25 11:00:50 +01:00
javanna	5103b76610	update version checks in ElasticsearchException serialization code 5.3.0 is the first version that contains the split from headers to metadata, updated the check to reflect that. It was previously after to be able to commit to master first, and only after that backport the change. Otherwise master tests would have failed until the change was backported.	2017-01-24 20:40:17 +01:00
Lee Hinman	304296ea6a	Fix BulkItemResponse serialization for 6.x <-> 5.3.x Previously the behavior where the `OpType` byte was serialized was only in master, but it was recently backported to 5.x, so the serialization version checks need to be updated as well.	2017-01-24 12:04:11 -07:00
srgclr	93a28b0acf	skip parentid if child document is an orphan #22770	2017-01-24 17:49:53 +00:00
Luca Cavanna	47c0e13a3b	Stop returning "es." internal exception headers as http response headers (#22703 ) move "es." internal headers to separate metadata set in ElasticsearchException and stop returning them as response headers Closes #17593 * [TEST] remove ESExceptionTests, move its methods to ElasticsearchExceptionTests or ExceptionSerializationTests	2017-01-24 16:12:45 +01:00
Jason Tedor	bcffc6fa49	Add hack for Docker cgroups Docker cgroups are mounted in the wrong place (i.e., inconsistently with /proc/self/cgroup). This commit adds an undocumented hack for working around, for now. Relates #22757	2017-01-24 06:36:03 -05:00
Christoph Büscher	59aefe5a38	Include human readable responses in response parsing tests (#22717 ) As a follow up to #22649, this changes the resent tests for parsing parts of search responses to randomly set the humanReadable() flag of the XContentBuilder that is used to render the responses. This should help to test that we can parse back thoses classes if the user specifies `?human=true` in the request url.	2017-01-24 11:17:58 +01:00
Jim Ferenczi	b0c2a5da30	Remove unused field in CollapseBuilder	2017-01-24 09:26:23 +01:00
Jim Ferenczi	868b12b548	Add BWC tests for field collapsing Field collapsing is supported from version 5.3	2017-01-24 08:34:16 +01:00
Nik Everett	2e399e5505	Rename constant It deserves a new name after the cleanup in #22749	2017-01-23 16:30:42 -05:00
javanna	8065531236	[TEST] add test to verify that the SMILE format works within the _bulk api	2017-01-23 19:40:24 +01:00
Nik Everett	ee264c6957	Fix parsing for `max_determinized_states` (#22749 ) There was a typo in the `ParseField` declaration. I know we want to port these parsers to `ObjectParser` eventually but I don't have the energy for that today and want to get this fixed. Closes #22722	2017-01-23 11:57:43 -05:00
Jim Ferenczi	e48bc2eed7	Add field collapsing for search request (#22337 ) * Add top hits collapsing to search request The field collapsing is done with a custom top docs collector that "collapse" search hits with same field value. The distributed aspect is resolve using the two passes that the regular search uses. The first pass "collapse" the top hits, then the coordinating node merge/collapse the top hits from each shard. ``` GET _search { "collapse": { "field": "category", } } ``` This change also adds an ExpandCollapseSearchResponseListener that intercepts the search response and expands collapsed hits using the CollapseBuilder#innerHit} options. The retrieval of each inner_hits is done by sending a query to all shards filtered by the collapse key. ``` GET _search { "collapse": { "field": "category", "inner_hits": { "size": 2 } } } ```	2017-01-23 16:33:51 +01:00
Tanguy Leroux	11164b394b	Add unit tests for ValueCountAggregator and InternalValueCount (#22741 ) Adds unit tests for the value count aggregator. Relates #22278	2017-01-23 16:24:55 +01:00
Simon Willnauer	27b5c2ad54	Pass `forceExecution` flag to transport interceptor (#22739 ) To effectively allow a plugin to intercept a transport handler it needs to know if the handler must be executed even if there is a rejection on the thread pool in the case the wrapper forks a thread to execute the actual handler.	2017-01-23 11:04:27 +01:00
Alexander Reelsen	6159ca28ae	Version: Add missing releases from 2.x in Version.java (#22594 )	2017-01-23 09:53:21 +01:00
Tim Brooks	a4ac29c005	Add single static instance of SpecialPermission (#22726 ) This commit adds a SpecialPermission constant and uses that constant opposed to introducing new instances everywhere. Additionally, this commit introduces a single static method to check that the current code has permission. This avoids all the duplicated access blocks that exist currently.	2017-01-21 12:03:52 -06:00
Simon Willnauer	3ad6d6ebcc	Simplify InternalEngine#innerIndex (#22721 ) Today `InternalEngine#innerIndex` is a pretty big method (> 150 SLoC). This commit merged `#index` and `#innerIndex` and splits it up into smaller contained methods.	2017-01-21 08:51:35 +01:00
Jim Ferenczi	8028578305	Upgrade to Lucene 6.4.0 (#22724 ) * Upgrade to Lucene 6.4.0 `ValueSource`s are now converted to `DoubleValueSource`s using the Lucene adapter made for the migration to the new API in 6.4.0.	2017-01-21 04:48:01 +01:00
Igor Motov	cfb415de7b	Fix broken TaskInfo.toString() Related to #22387	2017-01-20 20:57:42 -05:00
Tim Brooks	d86f97c428	Add CheckedSupplier and CheckedRunnable to core (#22725 ) Introduce CheckedSupplier and CheckedRunnable functional interfaces into core. These offer a checked version of the Supplier and Runnable interfaces for use with lambda apis.	2017-01-20 19:17:44 -06:00
Ali Beyad	3bf06d1440	Fixes retrieval of the latest snapshot index blob (#22700 ) This commit ensures that the index.latest blob is first examined to determine the latest index-N blob id, before attempting to list all index-N blobs and picking the blob with the highest N. It also fixes the MockRepository#move so that tests are able to handle non-atomic moves. This is done by adding a special setting to the MockRepository that requires the test to specify if it can handle non-atomic moves. If so, then the MockRepository#move operation will be non-atomic to allow testing for against such repositories.	2017-01-20 17:00:46 -06:00
Jim Ferenczi	4ec4bad908	Fix script score function that combines _score and weight (#22713 ) The weight factor function does not check if the delegate score function needs to access the score of the query. This results in a _score equals to 0 for all score function that set a weight. This change modifies the WeightFactorFunction#needsScore to delegate the call to its underlying score function. Fix #21483	2017-01-20 19:50:57 +01:00
Nik Everett	6265ef1c1b	Deguice rest handlers (#22575 ) There are presently 7 ctor args used in any rest handlers: * `Settings`: Every handler uses it to initialize a logger and some other strange things. * `RestController`: Every handler registers itself with it. * `ClusterSettings`: Used by `RestClusterGetSettingsAction` to render the default values for cluster settings. * `IndexScopedSettings`: Used by `RestGetSettingsAction` to get the default values for index settings. * `SettingsFilter`: Used by a few handlers to filter returned settings so we don't expose stuff like passwords. * `IndexNameExpressionResolver`: Used by `_cat/indices` to filter the list of indices. * `Supplier<DiscoveryNodes>`: Used to fill enrich the response by handlers that list tasks. We probably want to reduce these arguments over time but switching construction away from guice gives us tighter control over the list of available arguments. These parameters are passed to plugins using `ActionPlugin#initRestHandlers` which is expected to build and return that handlers immediately. This felt simpler than returning an reference to the ctors given all the different possible args. Breaks java plugins by moving rest handlers off of guice.	2017-01-20 11:48:51 -05:00
Simon Willnauer	824beea89d	Fix handling of document failure expcetion in InternalEngine (#22718 ) Today we try to be smart and make a generic decision if an exception should be treated as a document failure but in some cases concurrency in the index writer make this decision very difficult since we don't have a consistent state in the case another thread is currently failing the IndexWriter/InternalEngine due to a tragic event. This change simplifies the exception handling and makes specific decisions about document failures rather than using a generic heuristic. This prevent exceptions to be treated as document failures that should have failed the engine but backed out of failing since since some other thread has already taken over the failure procedure but didn't finish yet.	2017-01-20 16:55:00 +01:00
markharwood	f01784205f	New AdjacencyMatrix aggregation Similar to the Filters aggregation but only supports "keyed" filter buckets and automatically "ANDs" pairs of filters to produce a form of adjacency matrix. The intersection of buckets "A" and "B" is named "A&B" (the choice of separator is configurable). Empty intersection buckets are removed from the final results. Closes #22169	2017-01-20 15:49:31 +00:00
Tim Brooks	bc16162d21	Remove accept SocketPermissions from core (#22622 ) This is related to #22116. Core no longer needs SocketPermission accept. This permission is relegated to the transport-netty4 module and (for tests) to the mocksocket jar.	2017-01-20 09:27:45 -06:00
Tanguy Leroux	239ed0c912	Add unit tests for DateHistogramAggregator (#22714 ) Adds unit tests for the date histogram aggregator. Relates #22278	2017-01-20 14:18:30 +01:00
Christoph Büscher	54105f3ddd	Add parsing from xContent to ShardSearchFailure (#22699 ) In preparation for being able to parse SearchResponse from its rest representation, this adds fromXContent to ShardSearchFailure.	2017-01-20 12:49:54 +01:00
Yannick Welsch	1f0e0a2170	Close InputStream when receiving cluster state in PublishClusterStateAction (#22711 ) Not closing the InputStream will leak native memory as the DeflateCompressor/Inflater won't be closed.	2017-01-20 12:26:07 +01:00
Boaz Leskes	5d806bf93e	Index creation and setting update may not return deprecation logging (#22702 ) Those services validate their setting before submitting an AckedClusterStateUpdateTask to the cluster state service. An acked cluster state may be completed by a networking thread when the last acks as received. As such it needs special care to make sure that thread context headers are handled correctly.	2017-01-20 10:14:13 +01:00
David Pilato	fc4dc5ef21	Fix comment	2017-01-20 10:13:13 +01:00

... 2 3 4 5 6 ...

7745 Commits