OpenSearch

Commit Graph

Author	SHA1	Message	Date
Artur Nowosielski	726f5dccc0	Rewrite filter queries in FiltersAggregationBuilder (#22076 ) Queries must be rewritten before the query phase executes otherwise non-executable queries like `wrapper` query or `terms` will fail or queries that require resources like script service can't access these service unless rewritten. Relates to #21303	2016-12-11 14:37:12 +01:00
Masaru Hasegawa	3df2a086d4	Resolve index names in indices_boost This change allows specifying alias/wildcard expression in indices_boost. And added another format for specifying indices_boost. It accepts array of index name and boost pair. If an index is included in multiple aliases/wildcard expressions, the first match will be used. With new format, old format is marked as deprecated. Closes #4756	2016-12-11 21:41:49 +09:00
Simon Willnauer	20ff703e07	Fix IncludeExclude parsing `include` / `exclude` in terms / sig-terms aggs seems completely broken and massively untested. This commit makes the TermsTests pass again that randomly use `include` / `exclude`. This class must be tested individually and we need real integ tests that use xcontent that use this feature.	2016-12-11 09:55:53 +01:00
Yannick Welsch	68f30cae0c	[TEST] Unmute testRestoreUnsupportedSnapshots	2016-12-10 13:49:47 +01:00
Yannick Welsch	4831632a6f	[TEST] Handle legacy snapshots as if they don't exist anymore An earlier commit removed BWC for pre-5.0 snapshots, which also meant removing the capability to load pre-5.0 snapshots. In 6.0, such snapshots are now invisible and must be treated by the BWC tests in that way.	2016-12-10 13:49:47 +01:00
Yannick Welsch	4ad85c38c3	Throw NoSuchFileException to correctly adhere to readBlob contract URLBlobContainer can in certain situations throw a FileNotFoundException. To fulfill the contract of the readBlob method it should throw a NoSuchFileException instead when the given blob cannot be found.	2016-12-10 13:49:47 +01:00
Simon Willnauer	01d67e09b9	Detach handshake from connect to node (#22037 ) Today we connect and publish the nodes connection before we execute a handshake with the node we connect to. In the case of connecting to a node that won't pass the handshake this connection is already `published` and other code paths can use it. This commit detaches the connection and the publish of the connection such that `TransportService` can do a handshake before actually connect and publish the connection.	2016-12-10 10:03:26 +01:00
Nik Everett	3adefb7b4a	Begin centralizing XContentParser creation into RestRequest (#22041 ) To get #22003 in cleanly we need to centralize as much `XContentParser` creation as possible into `RestRequest`. That'll mean we have to plumb the `NamedXContentRegistry` into fewer places. This removes `RestAction.hasBody`, `RestAction.guessBodyContentType`, and `RestActions.getRestContent`, moving callers over to `RestRequest.hasContentOrSourceParam`, `RestRequest.contentOrSourceParam`, and `RestRequest.contentOrSourceParamParser` and `RestRequest.withContentOrSourceParamParserOrNull`. The idea is to use `withContentOrSourceParamParserOrNull` if you need to handle requests without any sort of body content and to use `contentOrSourceParamParser` otherwise. I believe the vast majority of this PR to be purely mechanical but I know I've made the following behavioral change (I'll add more if I think of more): * If you make a request to an endpoint that requires a request body and has cut over to the new APIs instead of getting `Failed to derive xcontent` you'll get `Body required`. * Template parsing is now non-strict by default. This is important because we need to be able to deprecate things without requests failing.	2016-12-09 20:23:02 -05:00
Nik Everett	ddade1b5ac	Improve the error message if task and node isn't found (#22062 ) Improves the error message returned when looking up a task that belongs to a node that is no longer part of the cluster. The new error message tells the user that the node isn't part of the cluster. This is useful because if you start a task and the node goes down there isn't a record of the task at all. This hints to the user that the task might have died with the node. Relates to #22027	2016-12-09 15:50:46 -05:00
Igor Motov	93b5e55660	Restores the original default format of search slow log In 5.0, the search slow log switched to the multi-line format with no option to get back to the origin single-line format that was used prior to 5.0 by default. This commit removes the reformat option from the search slow log and returns the search slow log back to the single-line format. Closes #21711	2016-12-09 12:38:28 -05:00
Yannick Welsch	b20b160a5e	Allow flush/force_merge/upgrade on shard marked as relocated (#22078 ) A shard that is locally marked as relocated, but where the relocation target shard has not been activated yet by the master, can still receive index operations, which in return can lead to flushes being triggered. Flushing is currently (wrongly) prohibited on shards marked as relocated, which makes the flushing process go into an endless retry loop and log warnings until the shard is closed. This commit fixes this situation by allowing flush, force_merge and upgrade operations to run on shards that are marked as relocated.	2016-12-09 17:56:40 +01:00
Nik Everett	bcef1e7452	Better error message when _parent isn't an object (#21987 ) If you make a mistake and specify a mapping like: ``` { "parent": { "properties": {} }, "child": { "_parent": "parent", "properties": {} } } ``` then the error message you get back amounts to `Failed to parse mapping for [child]: can't cast a String to a Map`. Since it doens't tell you which string can't be cast to a map you have to dig through the stack trace to figure out what to fix. This replaces the error message with: ``` Failed to parse mapping [child]: [_parent] must be an object containing [type] ``` so you can tell that the problem is with the `parent` field.	2016-12-09 11:33:31 -05:00
Yannick Welsch	a724f4eb61	Don't update nodes list when stepping down as master (#22049 ) This commit simplifies the node update logic so that nodes are never removed from the cluster state when the cluster state is not published.	2016-12-09 14:55:48 +01:00
Christoph Büscher	2592ff86ce	Add fromXContent to InternalNestedIdentity This adds a fromXContent method and unit test to InternalNestedIdentity so we can parse it as part of a search response. This is part of the preparation for parsing search responses on the client side.	2016-12-09 14:52:06 +01:00
Yannick Welsch	db0660a7ea	Reject external versioning and explicit version numbers on create (#21998 ) Fixes an issue where indexing requests with operation type "create" auto-convert external versioning to internal versioning and silently ignore the version number instead of failing with an error message.	2016-12-09 14:21:22 +01:00
Michael McCandless	613a1a6a18	Add stored binary fields to static backwards compatibility indices tests (#22054 ) Add stored binary fields to static backwards compatibility indices tests	2016-12-09 05:32:40 -05:00
Adrien Grand	6714e02bef	Mute RestoreBackwardsCompatIT.testRestoreUnsupportedSnapshots.	2016-12-09 10:41:07 +01:00
Adrien Grand	787519ee4c	Fix `other_bucket` on the `filters` agg to be enabled if a key is set. (#21994 ) Closes #21951	2016-12-09 09:48:48 +01:00
Adrien Grand	1bdf4a2c5b	Partition-based include-exclude does not implement equals/hashcode/serialization correctly. (#22051 )	2016-12-09 09:48:16 +01:00
Adrien Grand	9524c81af9	Document the `locale` option of the `date` field. (#22050 ) This also adds another level of protection against using the default locale. Relates to https://discuss.elastic.co/t/mapping-for-12h-date-format/68433/3.	2016-12-09 09:45:53 +01:00
Adrien Grand	36f598138a	Start using `ObjectParser` for aggs. (#22048 ) This is an attempt to start moving aggs parsing to `ObjectParser`. There is still A LOT to do, but ObjectParser is way better than the way aggregations parsing works today. For instance in most cases, we reject numbers that are provided as strings, which we are supposed to accept since some client languages (looking at you Perl) cannot make sure to use the appropriate types. Relates to #22009	2016-12-09 09:45:16 +01:00
Ryan Ernst	b1cef5fdf8	Remove 2.0 prerelease version constants (#22004 ) * Remove 2.0 prerelease version constants This is a start to addressing #21887. This removes: * pre 2.0 snapshot format support * automatic units addition to cluster settings * bwc check for delete by query in pre 2.0 indexes	2016-12-08 21:48:35 -08:00
Igor Motov	7f79c99e9a	Add descriptions to bulk tasks Related to #21768	2016-12-08 21:59:52 -05:00
Lee Hinman	ef64d230e7	Merge remote-tracking branch 'dakrone/index-seq-id-and-primary-term'	2016-12-08 19:47:21 -07:00
Lee Hinman	ee22a477df	Add internal _primary_term doc values field, fix _seq_no indexing This adds the `_primary_term` field internally to the mappings. This field is populated with the current shard's primary term. It is intended to be used for collision resolution when two document copies have the same sequence id, therefore, doc_values for the field are stored but the filed itself is not indexed. This also fixes the `_seq_no` field so that doc_values are retrievable (they were previously stored but irretrievable) and changes the `stats` implementation to more efficiently use the points API to retrieve the min/max instead of iterating on each doc_value value. Additionally, even though we intend to be able to search on the field, it was previously not searchable. This commit makes it searchable. There is no user-visible `_primary_term` field. Instead, the fields are updated by calling: ```java index.parsedDoc().updateSeqID(seqNum, primaryTerm); ``` This includes example methods in `Versions` and `Engine` for retrieving the sequence id values from the index (see `Engine.getSequenceID`) that are only used in unit tests. These will be extended/replaced by actual implementations once we make use of sequence numbers as a conflict resolution measure. Relates to #10708 Supercedes #21480 P.S. As a side effect of this commit, `SlowCompositeReaderWrapper` cannot be used for documents that contain `_seq_no` because it is a Point value and SCRW cannot wrap documents with points, so the tests have been updated to loop through the `LeafReaderContext`s now instead.	2016-12-08 19:47:03 -07:00
Jason Tedor	c9882dd1a0	Avoid NPE in NodeService#stats if HTTP is disabled This commit adds safety against an NPE if HTTP stats are requested but HTTP is disabled on a node. Relates #22060	2016-12-08 19:59:02 -05:00
Jason Tedor	f713106827	Bump version to 5.1.2 This commit bumps the version to 5.1.2. Relates #22057	2016-12-08 16:40:39 -05:00
Ali Beyad	3da04293f3	Cannot force allocate primary to a node where the shard already exists (#22031 ) Before, it was possible that the SameShardAllocationDecider would allow force allocation of an unassigned primary to the same node on which an active replica is assigned. This could only happen with shadow replica indices, because when a shadow replica primary fails, the replica gets promoted to primary but in the INITIALIZED state, not in the STARTED state (because the engine has specific reinitialization that must take place in the case of shadow replicas). Therefore, if the now promoted primary that is initializing fails also, the primary will be in the unassigned state, because replica to primary promotion only happens when the failed shard was in the started state. The now unassigned primary shard will go through the allocation deciders, where the SameShardsAllocationDecider would return a NO decision, but would still permit force allocation on the primary if all deciders returned NO. This commit implements canForceAllocatePrimary on the SameShardAllocationDecider, which ensures that a primary cannot be force allocated to the same node on which an active replica already exists.	2016-12-08 12:21:19 -05:00
Adrien Grand	182e119699	IP range masks exclude the maximum address of the range. (#22018 ) Closes #22005	2016-12-08 15:58:32 +01:00
Ali Beyad	30bcb06606	When shard data is still being fetched from nodes in the cluster, the ReplicaShardAllocator, when in explain mode, would get the node decisions for all nodes in the cluster. The PrimaryShardAllocator neglected to do this and tried to use the shard fetch data in explain mode, which had not yet been fully fetched. This commit fixes this by ensuring the PrimaryShardAllocator gets node decisions in the same way the ReplicaShardAllocator does in explain mode, if shard data is still being fetched.	2016-12-07 22:21:09 -05:00
Ali Beyad	e6e7bab58c	Prepares allocator decision objects for use with the allocation explain API (#21691 ) This commit enhances the allocator decision result objects (namely, AllocateUnassignedDecision, MoveDecision, and RebalanceDecision) to enable them to be used directly by the cluster allocation explain API. In particular, this commit does the following: - Adds serialization and toXContent methods to the response objects, which will form the explain API responses. - Moves the calculation of the final explanation to the response object itself, removing it from the responsibility of the allocators. - Adds shard store information to the NodeAllocationResult, so that store information is available for each node, when explaining a shard allocation by the PrimaryShardAllocator or the ReplicaShardAllocator. - Removes RebalanceDecision in favor of using MoveDecision for both moving and rebalancing shards. - Removes NodeRebalanceResult in favor of using NodeAllocationResult. - Changes the notion of weight ranking to be relative to the current node, instead of an absolute weight that doesn't convey any added value to the API user and can be confusing. - Introduces a new enum AllocationDecision to convey the decision type, which enables conveying unassigned, moving, and rebalancing scenarios with more detail as opposed to just Decision.Type and AllocationStatus.	2016-12-07 17:37:51 -05:00
Ali Beyad	05f64c550a	[TEST] fixes line length issue in BulkRequestModifierTests	2016-12-07 13:11:55 -05:00
Ryan Ernst	f02a2b6546	Ingest: Moved ingest invocation into index/bulk actions (#22015 ) * Ingest: Moved ingest invocation into index/bulk actions Ingest was originally setup as a plugin, and in order to hook into the index and bulk actions, action filters were used. However, ingest was later moved into core, but the action filters were never removed. This change moves the execution of ingest into the index and bulk actions. * Address PR comments * Remove forwarder direct dependency on ClusterService	2016-12-07 08:43:26 -08:00
Christoph Büscher	7454a9647b	Add fromXContent to HighlightField This adds a fromXContent method and unit test to the HighlightField class so we can parse it as part of a serch response. This is part of the preparation for parsing search responses on the client side.	2016-12-07 16:32:44 +01:00
Yannick Welsch	c87cc15d49	Add toString() for TransportReplicationAction.ConcreteShardRequest	2016-12-07 15:58:22 +01:00
Christoph Büscher	31a1c2e240	Remove redundant source setters from IndexRequestBuilder	2016-12-07 15:20:36 +01:00
Yannick Welsch	9630b1a6e7	Promote shadow replica to primary when initializing primary fails (#22021 ) Failing an initializing primary when shadow replicas are enabled for the index can leave the primary unassigned with replicas being active. Instead, a replica should be promoted to primary, which is fixed by this commit.	2016-12-07 13:59:43 +01:00
Yannick Welsch	13e1a6fd40	Trim in-sync allocations set only when it grows (#21976 ) This commit makes two changes to how the in-sync allocations set is updated: - the set is only trimmed when it grows. This prevents trimming too eagerly when the number of replicas was decreased while shards were unassigned. - the allocation id of an active primary that failed is only removed from the in-sync set if another replica gets promoted to primary. This prevents the situation where the only available shard copy in the cluster gets removed the in-sync set. Closes #21719	2016-12-07 10:59:11 +01:00
Adrien Grand	c746854e03	Pre-built analysis factories do not implement MultiTermAware correctly. (#21981 ) We had tests for the regular factories, but not for the pre-built ones, that ship by default without requiring users to define them in the analysis settings.	2016-12-07 10:32:25 +01:00
Adrien Grand	33b8d7a19d	Expose `ip` fields as strings in scripts. (#21997 ) Currently we expose the internal representation that we use for ip addresses, which are the ipv6 bytes. However, this is not really usable, exposes internal implementation details and also does not work fine with other APIs that expect that the values can be `toString`'d. Closes #21977	2016-12-07 10:32:11 +01:00
Boaz Leskes	4519bdfeb0	InternalTestCluster shouldn't auto heal an active disruption when a new one is set Instead people should explicitly clear the existing one so it's clear what's going on.	2016-12-06 19:58:11 +01:00
shaie	6da44c8164	Fix _termvectors with preference to not hit NPE (#21959 ) When you submit a _termvectors request for an artificial document and specify the 'preference' parameter to send the request to a particular shard, the request sometimes hits NPE. Fix this case by ignoring the auto-generated artificial document ID and pick a shard per the preference parameter, or a random shard. This closes #21928	2016-12-06 17:29:09 +01:00
Jim Ferenczi	b42ca6bcc9	Include unindexed field in FieldStats response (#21821 ) * Include unindexed field in FieldStats response This change adds non-searchable fields to the FieldStats response. These fields do not have min/max informations but they can be aggregatable. Fields that are only stored in _source (store:no, index:no, doc_values:no) will still be missing since they do not have any useful information to show. Indices and clients must be at least on V_5_2_0 to see this change.	2016-12-06 13:32:57 +01:00
Boaz Leskes	a7050b2d56	Remove `InternalTestCluster.startNode(s)Async` (#21846 ) Since the removal of local discovery of #https://github.com/elastic/elasticsearch/pull/20960 we rely on minimum master nodes to be set in our test cluster. The settings is automatically managed by the cluster (by default) but current management doesn't work with concurrent single node async starting. On the other hand, with `MockZenPing` and the `discovery.initial_state_timeout` set to `0s` node starting and joining is very fast making async starting an unneeded complexity. Test that still need async starting could, in theory, still do so themselves via background threads. Note that this change also removes the usage of `INITIAL_STATE_TIMEOUT_SETTINGS` as the starting of nodes is done concurrently (but building them is sequential)	2016-12-06 12:06:15 +01:00
Daniel Mitterdorfer	a02bc8ed1c	Document thread-safety for ingest processors With this commit we document that ingest processors need to be thread-safe. Previously this could be inferred from reading the source code but we got several user questions about this so it is stated explicitly in the Javadocs of Processor now.	2016-12-06 10:07:51 +01:00
Adrien Grand	26cbda41ea	AsciiFoldingFilter's multi-term component should never preserve the original token. (#21982 ) This ports the fix of https://issues.apache.org/jira/browse/LUCENE-7536 to Elasticsearch's ASCIIFoldingTokenFilterFactory.	2016-12-06 10:01:04 +01:00
Ryan Ernst	c8f241f284	Plugins: Remove response action filters (#21950 ) Action filters currently have the ability to filter both the request and response. But the response side was not actually used. This change removes support for filtering responses with action filters.	2016-12-05 16:14:04 -08:00
Jim Ferenczi	03a0a0aebb	Undeprecate GetResponse#getFields and GetResponse#getField These functions should not have been deprecated as they can be used to retrieve stored and doc-value field.	2016-12-05 15:31:53 +01:00
Ali Beyad	ff9959c865	Don't output null source node in RecoveryFailedException (#21963 ) The RecoveryFailedException's output prints the source and target nodes for the recovery. However, sometimes there is no source node for the recovery, only a target node (such as when recovering a primary shard from disk). In this case, we don't want to display the source node. This commit fixes this by displaying "Recovery failed on target node.." instead of "Recovery failed from null to target node" which is what the output currently displays.	2016-12-04 15:23:35 -05:00
Jason Tedor	60aa14f48e	Increase test logging on test simple pings test This commit increases the test logging on the unicast zeng ping test of simple pings to gather more info for chasing a race condition that is happening in this test.	2016-12-04 08:06:01 -05:00
Jason Tedor	040c05df36	Increase timeouts in UnicastZenPingTests Sadly, the timeouts here need to be increased to reduce the likelihood of spurious test failures (test hosts under load are especially prone to this). This does slow down this test suite a bit, but it's still not as slow as it was before this endeavor of lowering these timeouts started.	2016-12-03 22:19:55 -05:00
Jason Tedor	2c8229fcaf	Cleanup unicast zen ping unknown hosts cached test This commit cleans up the unicast zen ping unknown hosts cached test: - send pings from the same node to more clearly indicate DNS lookups are not cached (within the same UnicastZenPing instance) - increase ping and wait timeout to 500ms to address race conditions (on a test host under load, the timeout was too short for the connect/handshake/ping cycle to complete)	2016-12-03 22:00:30 -05:00
Jason Tedor	460e787049	Increase resolve timeout in unknown hosts test The port limit test is a simple test that fakes that resolving an address with a port range results the correct address collection. This test is subject to a race condition where the timeout on the resolve request can fire before the resolve code finishes executing (this race is exceptionally rare, because there are not actually any DNS lookups being done here since we are just resolving addresses). This commit increases the timeout here to significantly reduce the chance of a losing race causing a spurious test failure. This increased timeout should not increase the runtime of the test, just make failures less likely.	2016-12-03 09:02:44 -05:00
Jason Tedor	f5cbc36896	Increase resolve timeout in unknown hosts test The unknown hosts test is a simple test that fakes that resolving an address results in an unknown host exception. The main purpose of this test is to ensure that we log (and do not silently drop) when a host fails to resolve. This test is subject to a race condition where the timeout on the resolve request can fire before the resolve code finishes executing (this race is exceptionally rare, because there are not actually any DNS lookups being done here, just a mock resolve implementation that throws an exception and that's where losing the race can arise). This commit increases the timeout here to significantly reduce the chance of a losing race causing a spurious test failure. This increased timeout should not increase the runtime of the test, just make failures less likely.	2016-12-03 08:46:24 -05:00
Igor Motov	bb9317253a	Add descriptions to create snapshot and restore snapshot tasks. Related to #21768	2016-12-02 21:13:54 -05:00
Jason Tedor	c6efd4eb42	Rename method in InternalEngine This commit renames InternalEngine#loadSeqNoStatsLucene to InternalEngine#loadSeqNoStatsFromLucene to make this name consistent with the method InternalEngine#loadSeqNoStatsFromLuceneAndTranslog.	2016-12-02 20:46:26 -05:00
Ryan Ernst	34eb23e98e	Plugins: Replace Rest filters with RestHandler wrapper (#21905 ) * Plugins: Replace Rest filters with RestHandler wrapper RestFilters are a complex way of allowing plugins to add extra code before rest actions are executed. This change removes rest filters, and replaces with a wrapper which a single plugin may provide.	2016-12-02 14:54:51 -08:00
Jason Tedor	b0e8696143	Clarify global checkpoint recovery Today when starting a new engine, we read the global checkpoint from the translog only if we are opening an existing translog. This commit clarifies this situation by distinguishing the three cases of engine creation in the constructor leading to clearer code. Relates #21934	2016-12-02 15:00:16 -05:00
Jason Tedor	0afef53a17	Add system call filter bootstrap check Today if system call filters fail to install on startup, we log a message but otherwise march on. This might leave users without system call filters installed not knowing that they have implicitly accepted the additional risk. We should not be lenient like this, instead clearly informing the user that they have to either fix their configuration or accept the risk of not having system call filters installed. This commit adds a bootstrap check that if system call filters are enabled, they must successfully install. Relates #21940	2016-12-02 14:27:54 -05:00
Nik Everett	0c724b1878	Keep context during reindex's retries (#21941 ) * Keep context during reindex's retries This fixes reindex and friend's retries to keep the context. * Docs	2016-12-02 13:48:51 -05:00
Jay Modi	429e517476	Do not lose host information when pinging In #21828, serialization of the host string was added to preserve this information when a TransportAddress gets serialized. However, there is still a case where this did not always work. In UnicastZenPings, DiscoveryNode instances are created for the ping hosts with the minimum compatibility version, which is currently less than the version required to preserve the host information. This means that when a node is received from a PingResponse that the host information is no longer set correctly on the InetSocketAddress contained in the DiscoveryNode. This commit adds a workaround for this situation by allowing the host string to be passed into the TransportAddress constructor that takes a StreamInput and using that as the host for the InetAddress that is created during deserialization.	2016-12-02 12:21:53 -05:00
Ke Li	7cc9833606	Avoid some redundant unboxing and object creation (#21909 )	2016-12-02 16:11:41 +01:00
shaie	8fd3637891	Return correct term statistics when a field is not found in a shard (#21922 ) If you ask for the term vectors of an artificial document with term_statistics=true, but a shard does not have any terms of the doc's field(s), it returns the doc's term vectors values as the shard-level term statistics. This commit fixes that to return 0 for `ttf` and also field-level aggregated statistics. Closes #21906	2016-12-02 08:14:45 +01:00
Simon Willnauer	adf9bd90a4	Remove legacy BWC test infrastructure and tests (#21915 ) We don't use the test infra nor do we run the tests. They might all be entirely out of date. We also have a different BWC test infra in-place. This change removes all of the legacy infra.	2016-12-02 08:06:20 +01:00
makeyang	3f1d7be07a	Refactor shard limit allocation decider This commit simplifies the shard limit allocation decider, removing some duplicated code into a common method. Relates #21845	2016-12-01 21:27:02 -05:00
Ryan Ernst	a6ad89bee0	Mappings: Fix get mapping when no indexes exist to not fail in response generation (#21924 ) When there are no indexes, get mapping has a series of special cases. Two of those expect the response object already started, and the other two respond with an exception. Those two cases (types passed in but no indexes and vice versa) would fail in their error response generation because it did not expect an object to already be started in the json generator. This change moves the object start to where it is needed for the empty responses. closes #21916	2016-12-01 16:57:12 -08:00
Simon Willnauer	6522538033	Add validation for supported index version on node join, restore, upgrade & open index (#21830 ) Today we can easily join a cluster that holds an index we don't support since we currently allow rolling upgrades from 5.x to 6.x. Along the same lines we don't check if we can support an index based on the nodes in the cluster when we open, restore or metadata-upgrade and index. This commit adds additional safety that fails cluster state validation, open, restore and /or upgrade if there is an open index with an incompatible index version created in the cluster. Realtes to #21670	2016-12-01 15:40:35 +01:00
Simon Willnauer	155de53fe3	Add a connect timeout to the ConnectionProfile to allow per node connect timeouts (#21847 ) Timeouts are global today across all connections this commit allows to specify a connection timeout per node such that depending on the context connections can be established with different timeouts. Relates to #19719	2016-12-01 15:39:49 +01:00
Boaz Leskes	92fa9149f3	rename more before() methods that now conflict with ESTestCase	2016-12-01 13:40:27 +01:00
Simon Willnauer	dd5256c324	Reduce number of connections per node depending on the nodes role (#21849 ) We currently treat every node equally when we establish connections to a node. Yet, if we are not master eligible or can't hold any data there is no point in creating a dedicated connection for sending the cluster state or running remote recoveries respectively. The usage of STATE and RECOVERY connections on non-master and/or non-data nodes will result in an IllegalStateException.	2016-12-01 08:00:48 +01:00
Jim Ferenczi	fc9b63877e	Handle specialized term queries in MappedFieldType.extractTerm(TermQuery) (#21889 ) For some fields we have a specialized implementation of a TermQuery that is specific for the field. When these kind of fields are used in a wildcard query or a span term query it fails with an exception because they don't recognize the specialized form. The impacted fields are [_all] and [_type] and the impacted queries are [span_term] and [wilcard]. This change handles these forms and correctly extracts the term inside them for further use. Fixes #21882	2016-11-30 23:11:38 +01:00
Jason Tedor	92f05e796e	Remove traces during connect with handshake This commit removes two trace logging statements during connection with handshake as they are just clutter.	2016-11-30 15:29:33 -05:00
Jason Tedor	761325bf94	Throw exception on ping from another cluster When we receive a ping from another cluster, we should throw an exception so as to not leak the channel.	2016-11-30 15:28:56 -05:00
Jason Tedor	c90ba67abb	Do not reply to pings from another cluster Today when sending responses to discovery pings, we unconditionally reply. Instead, this commit modifies the response handler to not reply when the cluster names do not match. This addresses a race condition identified after reducing the timeout in UnicastZenPingTests#testSimplePings. In particular, we send pings in the following way: - if not connected to the node, connect to the node and after successful handshake, send a ping - if connected to the node, send a ping When the ping timeout is set low, a subsequent batch of pings can race against a connect/disconnect cycle from a prior batch of pings. In particular, consider the following scenario: - node A from cluster X - node B from cluster Y - pings are initiated from node A with node B in the hosts list - node A will try to connect and handshake with B - the connection will succeed, and the handshake will eventually fail due to mismatched cluster names - on a short timeout, a second batch of pings will fire, and on this batch node A will see that it is still connected to node B; thus, it will immediately fire a ping to node B and node B will dutifully respond Relates #21894	2016-11-30 15:09:42 -05:00
Luca Cavanna	103984a4a1	Remove indices query (#21837 ) The indices query is deprecated since 5.0.0 (#17710). It can now be removed in master (future 6.0 version).	2016-11-30 19:37:01 +01:00
Adrien Grand	117944093e	Remove testing of 2.x indices in DecayFunctionScoreIT. Such old indices will not be supported in 6.0.	2016-11-30 17:16:13 +01:00
Jason Tedor	6c45695d52	Add version 5.1.1 This commit removes the version constant for 5.1.0 (due to an inadvertent release) and adds the version constant for 5.1.1. Relates #21890	2016-11-30 11:14:17 -05:00
Adrien Grand	f5ac27a20d	Fix TermsQueryBuilderTests expectations.	2016-11-30 17:07:53 +01:00
Adrien Grand	c5b9c98b99	Remove the `default` store type. (#21616 ) It used to be a hybrid store between `niofs` and `mmapfs`, which we removed when we switched to `fs` by default (which is `mmapfs` on 64-bits systems).	2016-11-30 15:33:26 +01:00
Adrien Grand	90ab477f19	The `terms` query should always map to a Lucene `TermsQuery`. (#21786 ) Currently, the `terms` query is just syctactic sugar for a `bool` query when used in a query context. This change proposes to always generate the same query in query and filter contexts, which is less confusing.	2016-11-30 15:29:09 +01:00
Luca Cavanna	5b8bdba12e	Remove subrequests method from CompositeIndicesRequest (#21873 )	2016-11-30 15:03:58 +01:00
Matt Weber	1e722c060b	Remove forked XRollingBuffer and XQueryBuilder. (#21866 ) Remove the forked versions now that we are on lucene-6.4.0-snapshot.	2016-11-30 13:45:54 +01:00
Adrien Grand	a3ef674992	Reduce memory pressure when sending large terms queries. (#21776 ) When users send large `terms` query to Elasticsearch, every value is stored in an object. This change does not reduce the amount of created objects, but makes sure these objects die young by optimizing the list storage in case all values are either non-null instances of Long objects or BytesRef objects, which seems to help the JVM significantly.	2016-11-30 13:35:56 +01:00
Adrien Grand	6231009a8f	Remove 2.x backward compatibility of mappings. (#21670 ) For the record, I also had to remove the geo-hash cell and geo-distance range queries to make the code compile. These queries already throw an exception in all cases with 5.x indices, so that does not hurt any more. I also had to rename all 2.x bwc indices from `index-${version}` to `unsupported-${version}` to make `OldIndexBackwardCompatibilityIT` happy.	2016-11-30 13:34:46 +01:00
Jason Tedor	072007c759	Speed up UnicastZenPingTests These tests using ping timeouts on the order of seconds, but this is unnecessary since all the sockets are within the same JVM it really should not take that long. Relates #21874	2016-11-29 23:27:25 -05:00
Jason Tedor	b6ba4ae34b	Add version 5.0.3 This commit adds version 5.0.3 and the BWC indices for version 5.0.2. Relates #21867	2016-11-29 18:34:55 -05:00
Jay Modi	404b42ee95	DiscoveryNode and TransportAddress should preserve host information In some cases, such as the creation of DiscoveryNode instances for unicast ping requests, the host information was not being populated properly and instead the address string was being used. Additionally, when serializing a DiscoveryNode and in turn a transport address, the host was not being set on the InetAddress when deserializing the object, so even if the address was created from a hostname, the address in the deserialized instance had no knowledge of the hostname that was originally used.	2016-11-29 16:18:08 -05:00
Luca Cavanna	6eaff9432d	SearchTemplateRequest to implement CompositeIndicesRequest (#21865 ) SearchTemplateRequest to implement CompositeIndicesRequest Given that SearchTemplateRequest effectively delegates to search when a search is being executed, it should implement the CompositeIndicesRequest interface. The subrequests method should return a single search request. When a search is not going to be executed, because we are in simulate mode, there are no inner requests, and there are no corresponding indices to that request either. Closes #21747	2016-11-29 20:52:43 +01:00
Boaz Leskes	be4074e13d	improve debug logging when node waits for initial cluster state And enabled debug logging in InternalTestClusterTests so we can see it.	2016-11-29 20:38:19 +01:00
Luca Cavanna	f253621feb	Remove deprecated query names: in, geo_bbox, mlt, fuzzy_match and match_fuzzy (#21852 ) These query names were all deprecated in 5.0.0: - in is removed in favour of terms - geo_bbox is removed in favour of geo_bounding_box - mlt is removed in favour of more_like_this - fuzzy_match and match_fuzzy are removed in favour of match	2016-11-29 19:07:01 +01:00
Jim Ferenczi	d791ddf704	Upgrade to lucene-6.4.0-snapshot-ec38570 (#21853 ) Set lucene version to 6.4.0-snapshot-ec38570 and update all the sha1s/license Fix invalid combo after upgrade in query_string query. split_on_whitespace=false is disallowed if auto_generate_phrase_queries=true Adapt the expectations of some tests to the new format of the Lucene explain output	2016-11-29 18:40:31 +01:00
Nicholas Knize	af1ab68b64	Add RangeFieldMapper for numeric and date range types Lucene 6.2 added index and query support for numeric ranges. This commit adds a new RangeFieldMapper for indexing numeric (int, long, float, double) and date ranges and creating appropriate range and term queries. The design is similar to NumericFieldMapper in that it uses a RangeType enumerator for implementing the logic specific to each type. The following range types are supported by this field mapper: int_range, float_range, long_range, double_range, date_range. Lucene does not provide a DocValue field specific to RangeField types so the RangeFieldMapper implements a CustomRangeDocValuesField for handling doc value support. When executing a Range query over a Range field, the RangeQueryBuilder has been enhanced to accept a new relation parameter for defining the type of query as one of: WITHIN, CONTAINS, INTERSECTS. This provides support for finding all ranges that are related to a specific range in a desired way. As with other spatial queries, DISJOINT can be achieved as a MUST_NOT of an INTERSECTS query.	2016-11-29 10:10:14 -06:00
Simon Willnauer	f5ff69fabe	Remove connectToNodeLight and replace it with a connection profile (#21799 ) The Transport#connectToNodeLight concepts is confusing and not very flexible. neither really testable on a unittest level. This commit cleans up the code used to connect to nodes and simplifies transport implementations to share more code. This also allows to connect to nodes with custom profiles if needed, for instance future improvements can be added to connect to/from nodes that are non-data nodes without dedicated bulks and recovery connections.	2016-11-29 09:35:07 +01:00
Ali Beyad	a884573898	[TEST] fixes FilterAllocationDecider test for decision explanation when the initial recovery is LOCAL_SHARDS	2016-11-28 20:37:19 -05:00
Ali Beyad	07bd0a30f0	Improves allocation decider decision explanation messages (#21771 ) This commit improves the decision explanation messages, particularly for NO decisions, in the various AllocationDecider implementations by including the setting(s) in the explanation message that led to the decision. This commit also returns a THROTTLE decision instead of a NO decision when the concurrent rebalances limit has been reached in ConcurrentRebalanceAllocationDecider, because it more accurately reflects a temporary throttling that will turn into a YES decision once the number of concurrent rebalances lessens, as opposed to a more permanent NO decision (e.g. due to filtering).	2016-11-28 20:23:16 -05:00
Matt Weber	04e07bcdb6	Synonym Graph Support (LUCENE-6664) (#21517 ) Integrate the patch from LUCENE-6664 into elasticsearch and add support for handling a graph token stream in match/multi-match queries. This fixes longstanding bugs with multi-token synonyms returning incorrect results with proximity queries.	2016-11-28 09:25:49 -08:00
Jim Ferenczi	8affb7c845	Fix FiltersFunctionScoreQuery highlighting (#21827 ) This is a cleanup of the fix pushed in https://github.com/elastic/elasticsearch/pull/20400. FiltersFunctionScoreQuery sub query should be extracted in CustomQueryScorer.extract (and not in CustomQueryScorer.extractUnknownQuery). This does not fix any bug in this branch (it's just a cleanup) but the intent is first to clean up and then to backport in 2.x where there is a real bug. The bug is in 2.x only because the backport of https://github.com/elastic/elasticsearch/pull/20400 in 2.x mistakenly renamed the FiltersFunctionScoreQuery to FunctionScoreQuery. This leads to incorrect highlighting on FiltersFunctionScoreQuery in 2.x.	2016-11-28 17:56:24 +01:00
Nik Everett	145d0813b5	Log ScriptException's xcontent if file script compilation fails (#21767 ) When a file script fails to compile, rather than logging the exception that caused the failure this logs the xcontent of that exception. This is both shorter and has the script stack which is useful for figuring out why the compilation failed. Still logs the entire stacktrace at debug level just in case you need it. Relates to #21733	2016-11-28 11:36:06 -05:00
Ali Beyad	db7362da67	Fixes shard level snapshot metadata loading when index-N file is missing (#21813 ) In making changes for the 5.0 version of snapshots, a bug was introduced where if an index-N file could not be found for an individual shard, the backup was to iterate over all snap-.dat files in the shard folder to know which snapshots contain that shard's data, but in 5.0, reading the snap-.dat files as backup was incorrectly passing in the blob name for the snap-.dat file, thereby failing to load all index files for a given snapshot when the index-N file is missing. This condition should be rare as there is no reason an index-N file should be absent (unless it was deleted or there was corruption reading the file), but nevertheless, this situation can be encountered and this commit fixes the bug by reading the correct snap-.dat blob name in the shard data folder.	2016-11-28 10:46:33 -05:00
Simon Willnauer	b7292a6005	Remove TcpTransport#addressSupported since TransportAddress is now final TransportAddress used to be customizable per transport but this has been removed a while ago. Therefore we can remove all usage of this method as well. Relates to #20695	2016-11-28 16:06:59 +01:00

1 2 3 4 5 ...

7052 Commits