OpenSearch

Commit Graph

Author	SHA1	Message	Date
Simon Willnauer	dc659feeb4	Add a setting to disable remote cluster connections on a node (#23005 ) Today either all nodes in the cluster connect to remote clusters of only nodes that have remote clusters configured in their node config. To allow global remote cluster configuration but restrict connections to a set of nodes in the cluster this change adds a new setting `search.remote.connect` (defaults to `true`) to allow to disable remote cluster connections on a per node basis.	2017-02-07 09:59:24 +01:00
Nik Everett	0d6e622242	Make dates be ReadableDateTimes in scripts (#22948 ) Instead of longs. If you want millis since epoch you can call doc.date_field.value.millis. Relates to #22875	2017-02-06 16:44:56 -05:00
Nicholas Knize	1c9fdfd1b3	Remove GeoPointFieldMapper abstraction In order to support the evolving GeoPoint encodings in Lucene 5 and 6, ES 2.x and 5.x implements an abstraction layer to the GeoPointFieldMapper classes. As of 5.x the geo_point field mapper settled on using Lucene's more performant LatLonPoint field type and deprecated all other encodings. In 6.0 all encodings except LatLonPoint have been removed rendering this abstraction layer useless. This commit removes the abstraction layer and renames the LatLonPointFieldMapper back to GeoPointFieldMapper to mantain consistency with ES field naming.	2017-02-06 14:17:21 -06:00
Christoph Büscher	033f03109f	[Tests] Adding tests for AvgAggregator and InternalAvg (#23000 )	2017-02-06 20:05:40 +01:00
Ali Beyad	42a9f95fde	This commit changes the exception type thrown when trying to (#22921 ) create a snapshot with a name that already exists in the repository. Instead of throwing a SnapshotCreateException, which results in a generic 500 status code, a duplicate snapshot name will throw a InvalidSnapshotNameException, which will result in a 400 status code (bad request).	2017-02-06 11:39:59 -06:00
Adrien Grand	eb26e1a292	Add unit tests to histogram aggregations. (#22961 )	2017-02-06 18:18:21 +01:00
Simon Willnauer	f09c4e1cdb	Expose `search.highlight.term_vector_multi_value` as a node level setting (#22999 ) This setting was missed in the great settings refactoring and should be exposed via node level settings.	2017-02-06 18:17:34 +01:00
Simon Willnauer	7513c6e4eb	Remove QUERY_AND_FETCH search type (#22996 ) `QUERY_AND_FETCH` has been treated as an internal optimization for 2 major versions. This commit removes the search type and it's implementation details and folds the optimization in the case of a single shard into the search controller such that every search with a single shard (non DFS) will receive this optimization.	2017-02-06 17:10:03 +01:00
Boaz Leskes	5e7d22357f	Connect to new nodes concurrently (#22984 ) When a node receives a new cluster state from the master, it opens up connections to any new node in the cluster state. That has always been done serially on the cluster state thread but it has been a long standing TODO to do this concurrently, which is done by this PR. This is spin off of #22828, where an extra handshake is done whenever connecting to a node, which may slow down connecting. Also, the handshake is done in a blocking fashion which triggers assertions w.r.t blocking requests on the cluster state thread. Instead of adding an exception, I opted to implement concurrent connections which both side steps the assertion and compensates for the extra handshake.	2017-02-06 16:32:41 +01:00
Martijn van Groningen	e4663d6263	added comment	2017-02-06 15:16:16 +01:00
Martijn van Groningen	c8d470f190	Change `org.elasticsearch.bootstrap.JNAKernel32Library$SizeT` constructor's modifier to public. Otherwise `NativeMappedConverter` can't construct this class. Closes #22991	2017-02-06 15:16:16 +01:00
Christoph Büscher	d02170b277	Add parsing from xContent to MainResponse (#22934 ) Add parsing from xContent to MainResponse	2017-02-06 12:30:42 +01:00
Yannick Welsch	6f6596cfb5	Revert "Reduce log-level of IndexPrimaryRelocationIT to hunt Heisenbug" This reverts commit `d0fa6a9bd8`.	2017-02-06 11:40:39 +01:00
Adrien Grand	76f779486b	5.2.1 is now on Lucene 6.4.1 too.	2017-02-06 10:02:31 +01:00
Adrien Grand	c8496fc4f4	Upgrade to Lucene 6.4.1. (#22978 )	2017-02-06 09:28:43 +01:00
Martijn van Groningen	9201ee82f6	[TEST] Added unit tests for sum aggs. Relates to #22278	2017-02-06 08:32:10 +01:00
Lee Hinman	39e7c30912	Change certain replica failures not to fail the replica shard This changes the way that replica failures are handled such that not all failures will cause the replica shard to be failed or marked as stale. In some cases such as refresh operations, or global checkpoint syncs, it is "okay" for the operation to fail without the shard being failed (because no data is out of sync). In these cases, instead of failing the shard we should simply fail the operation, and, in the event it is a user-facing operation, return a 5xx response code including the shard-specific failures. This was accomplished by having two forms of the `Replicas` proxy, one that is for non-write operations that does not fail the shard, and one that is for write operations that will fail the shard when an operation fails. Relates to #10708	2017-02-03 14:39:46 -07:00
Nik Everett	70e3cce904	Fix name of `enable_position_increments` (#22895 ) It was accidentally renamed `enabled_position_increment` in the cleanups for 5.0. This adds `enable_position_increment` as a deprecated alias so it will continue to work.	2017-02-03 16:28:27 -05:00
Nicholas Knize	b1a6b227e1	Remove deprecated geo query parameters, and GeoPointDistanceRangeQuery This commit removes the following queries and parameters (which were deprecated in 5.0): * GeoPointDistanceRangeQuery * coerce, and ignore_malformed for GeoBoundingBoxQuery, GeoDistanceQuery, GeoPolygonQuery, and GeoDistanceSort	2017-02-03 10:08:00 -06:00
Tim Brooks	f70188ac58	Remove connect SocketPermissions from core (#22797 ) This is related to #22116. Core no longer needs `SocketPermission` `connect`. This permission is relegated to these modules/plugins: - transport-netty4 module - reindex module - repository-url module - discovery-azure-classic plugin - discovery-ec2 plugin - discovery-gce plugin - repository-azure plugin - repository-gcs plugin - repository-hdfs plugin - repository-s3 plugin And for tests: - mocksocket jar - rest client - httpcore-nio jar - httpasyncclient jar	2017-02-03 09:39:56 -06:00
Jason Tedor	9a0b216c36	Upgrade checkstyle to version 7.5 This commit upgrades the checkstyle configuration from version 5.9 to version 7.5, the latest version as of today. The main enhancement obtained via this upgrade is better detection of redundant modifiers. Relates #22960	2017-02-03 09:46:44 -05:00
Jason Tedor	01871e4def	Fix compilation in RecoverySourceHandlerTests This error arose after the signature of a method was changed.	2017-02-03 08:48:29 -05:00
Jason Tedor	6e9940283b	Avoid losing ops in file-based recovery When a primary is relocated from an old node to a new node, it can have ops in its translog that do not have a sequence number assigned. When a file-based recovery is started, this can lead to skipping these ops when replaying the translog due to a bug in the recovery logic. This commit addresses this bug and adds a test in the BWC tests. Relates #22945	2017-02-03 08:11:57 -05:00
Jim Ferenczi	4876448e39	Consilify get-field-mapping docs (#22936 ) This change also removes the reference to the difference bewteen full name and index name. They are always the same since 2.x and `name` does not refer anymore to `author.name` automatically. A simple pattern must be used instead. Remove redundant code that checks the field name twice.	2017-02-03 10:04:31 +01:00
Yannick Welsch	d0fa6a9bd8	Reduce log-level of IndexPrimaryRelocationIT to hunt Heisenbug	2017-02-03 09:46:07 +01:00
Chris Buonocore	365d33efe3	Handle missing plugin name in remove command Today if a user invokes the remove plugin command without specifying the name of a plugin to remove, we arrive at a null pointer exception. This commit adds logic to cleanly handle this situation and provide clear feedback to the user. Relates #22930	2017-02-02 19:39:56 -05:00
Nicholas Knize	58c34f0da9	Fix NPE in RangeFieldMapper.doXContentBody RangeFieldMapper.doXContentBody should only serialize format and locale when type is set to 'date_range'. closes #22925	2017-02-02 16:52:33 -06:00
Nik Everett	71b2655bb3	Fix outdated example in javadoc The code changed and the example in the javadoc didn't. Closes #22932	2017-02-02 14:28:48 -05:00
Areek Zillur	ba8ad397a1	Use bulk action interally for update action (#22915 ) Currently, update action internally uses deprecated index and delete transport actions. As of #21964, these tranport actions were deprecated in favour of using single item bulk request. In this commit, update action uses single item bulk action.	2017-02-02 14:21:53 -05:00
Jay Modi	7520a107be	Optionally require a valid content type for all rest requests with content (#22691 ) This change adds a strict mode for xcontent parsing on the rest layer. The strict mode will be off by default for 5.x and in a separate commit will be enabled by default for 6.0. The strict mode, which can be enabled by setting `http.content_type.required: true` in 5.x, will require that all incoming rest requests have a valid and supported content type header before the request is dispatched. In the non-strict mode, the Content-Type header will be inspected and if it is not present or not valid, we will continue with auto detection of content like we have done previously. The content type header is parsed to the matching XContentType value with the only exception being for plain text requests. This value is then passed on with the content bytes so that we can reduce the number of places where we need to auto-detect the content type. As part of this, many transport requests and builders were updated to provide methods that accepted the XContentType along with the bytes and the methods that would rely on auto-detection have been deprecated. In the non-strict mode, deprecation warnings are issued whenever a request with body doesn't provide the Content-Type header. See #19388	2017-02-02 14:07:13 -05:00
Nicholas Knize	b41d5747f0	Reduce GeoDistance insanity GeoDistance query, sort, and scripts make use of a crazy GeoDistance enum for handling 4 different ways of computing geo distance: SLOPPY_ARC, ARC, FACTOR, and PLANE. Only two of these are necessary: ARC, PLANE. This commit removes SLOPPY_ARC, and FACTOR and cleans up the way Geo distance is computed.	2017-02-02 12:39:42 -06:00
Igor Motov	c34b63dadd	Expand AbstractSerializingTestCase and AbstractWireSerializingTestCase to test diff serialization This commit adds two additional test cases that can be used to verify correct diff serialization in additional to binary and xcontent serialization.	2017-02-02 12:19:53 -05:00
Tanguy Leroux	f86fd62821	Parse elasticsearch exception's root causes (#22924 ) This commit change ElasticsearchException.failureFromXContent() method so that it now parses root causes which were ignored before, and adds them as suppressed exceptions of the returned exception.	2017-02-02 17:00:16 +01:00
Nik Everett	dacc150934	Expose multi-valued dates to scripts and document painless's date functions (#22875 ) Implemented by wrapping an array of reused `ModuleDateTime`s that we grow when needed. The `ModuleDateTime`s are reused when we move to the next document. Also improves the error message returned when attempting to modify the `ScriptdocValues`, removes a couple of allocations, and documents that the date functions are available in Painless. Relates to #22162	2017-02-01 21:57:07 -05:00
Ali Beyad	9f97eec12e	Fixes the Version constants for 5.2.0, 5.2.1, and 5.3.0 to have the correct Lucene version (6.4.0)	2017-02-01 12:47:36 -05:00
Lee Hinman	8d83edc4a5	Disallow introducing illegal object mappings (double '..') This disallows object mappings that would accidentally create something like `foo..bar`, which is then unparsable for the `bar` field as it does not know what its parent is. Resolves #22794	2017-02-01 09:03:20 -07:00
Ali Beyad	5c1410d031	Removes premature addition of v5.4.0 constant	2017-02-01 10:43:53 -05:00
Tanguy Leroux	3cfcd1acb7	[TEST] Fix BytesRestResponseTests.testNoErrorFromXContent	2017-02-01 10:34:42 +01:00
Boaz Leskes	5bf9cb9f70	remove await fix from testCannotAllocateStaleReplicaExplanation	2017-02-01 10:31:26 +01:00
Boaz Leskes	06b8a1ada7	fix testCannotAllocateStaleReplicaExplanation node management The test tried to create a situation where a stale replica is the only shard available. It did so by stopping the node with the replica, indexing some, stopping the primary node, starting a new node. This is flawed because the newly started node may reuse the data path of the primary node and things go back to green. Instead we should make sure that the replica is on the path that will be selected when the new node is started (i.e., the path with the smaller ordinal)	2017-02-01 10:30:42 +01:00
Tanguy Leroux	c74679b6b9	Add parsing method to BytesRestResponse's error (#22873 ) This commit adds a BytesRestResponse.errorFromXContent() method to parse the error returned by BytesRestResponse. It returns a ElasticsearchStatusException instance.	2017-02-01 10:11:17 +01:00
Ali Beyad	f436a06971	[TEST] adds AwaitsFix on two failing tests	2017-01-31 22:42:23 -05:00
Ali Beyad	6f2222f8fb	[TEST] fix node allocation result check in explain API test	2017-01-31 20:11:01 -05:00
Ali Beyad	547eb5c22f	Include stale replica shard info when explaining an unassigned primary (#22826 ) Currently, if a previously allocated shard has no in-sync copy in the cluster, but there is a stale replica copy, the explain API does not include information about the stale replica copies in its output. This commit includes any shard copy information available (even for stale copies) when explaining an unassigned primary shard that was previously allocated in the cluster. This situation can arise as follows: imagine an index with 1 primary and 1 replica and a cluster with 2 nodes. If the node holding the replica is shut down, and data continues to be indexed, only the primary will have the latest data and the replica that has gone offline will be marked as stale. Now, suppose the node holding the primary is shut down. There are no copies of the shard data in the cluster. Now, start the first stopped node (holding the stale replica) back up. The cluster is red because there is no in-sync copy available. Running the explain API before would inform the user that there is no valid shard copy in the cluster for that shard, but it would not provide any information about the existence of the stale replica that exists on the restarted node. With this commit, the explain API provides information about all the stale replica copies when explaining the unassigned primary.	2017-01-31 16:31:55 -06:00
Ali Beyad	c223457ba1	Adds v5.2.1 and v5.4.0 constants and bwc index for 5.2.0	2017-01-31 17:12:02 -05:00
Jack Conradson	3d2626c4c6	Change Namespace for Stored Script to Only Use Id (#22206 ) Currently, stored scripts use a namespace of (lang, id) to be put, get, deleted, and executed. This is not necessary since the lang is stored with the stored script. A user should only have to specify an id to use a stored script. This change makes that possible while keeping backwards compatibility with the previous namespace of (lang, id). Anywhere the previous namespace is used will log deprecation warnings. The new behavior is the following: When a user specifies a stored script, that script will be stored under both the new namespace and old namespace. Take for example script 'A' with lang 'L0' and data 'D0'. If we add script 'A' to the empty set, the scripts map will be ["A" -- D0, "A#L0" -- D0]. If a script 'A' with lang 'L1' and data 'D1' is then added, the scripts map will be ["A" -- D1, "A#L1" -- D1, "A#L0" -- D0]. When a user deletes a stored script, that script will be deleted from both the new namespace (if it exists) and the old namespace. Take for example a scripts map with {"A" -- D1, "A#L1" -- D1, "A#L0" -- D0}. If a script is removed specified by an id 'A' and lang null then the scripts map will be {"A#L0" -- D0}. To remove the final script, the deprecated namespace must be used, so an id 'A' and lang 'L0' would need to be specified. When a user gets/executes a stored script, if the new namespace is used then the script will be retrieved/executed using only 'id', and if the old namespace is used then the script will be retrieved/executed using 'id' and 'lang'	2017-01-31 13:27:02 -08:00
Boaz Leskes	eb36b82de4	Seq Number based recovery should validate last lucene commit max seq# (#22851 ) The seq# base recovery logic relies on rolling back lucene to remove any operations above the global checkpoint. This part of the plan is not implemented yet but have to have these guarantees. Instead we should make the seq# logic validate that the last commit point (and the only one we have) maintains the invariant and if not, fall back to file based recovery. This commit adds a test that creates situation where rollback is needed (primary failover with ops in flight) and fixes another issue that was surfaced by it - if a primary can't serve a seq# based recovery request and does a file copy, it still used the incoming `startSeqNo` as a filter. Relates to #22484 & #10708	2017-01-31 20:27:31 +01:00
Ryan Ernst	29f63c78cc	Internal: Convert empty and size checks of settings to not use getAsMap() (#22890 ) With the new secure settings, methods like getAsMap() no longer work correctly as a means of checking for empty settings, or the total size. This change converts the existing uses of that method to use methods directly on Settings. Note this does not update the implementations to account for SecureSettings, as that will require a followup which changes how secure settings work.	2017-01-31 10:44:09 -08:00
Jim Ferenczi	f6d38d480a	Integrate UnifiedHighlighter (#21621 ) * Integrate UnifiedHighlighter This change integrates the Lucene highlighter called "unified" in the list of supported highlighters for ES. This highlighter can extract offsets from either postings, term vectors, or via re-analyzing text. The best strategy is picked automatically at query time and depends on the field and the query to highlight.	2017-01-31 19:06:03 +01:00
Ryan Ernst	a4f6edec52	Settings: Fix settings reading to account for defaults (#22871 ) In #22762, settings preparation during bootstrap was changed slightly to account for SecureSettings, by starting with a fresh settings builder after reading the initial configuration. However, this the defaults from system properties were never re-read. This change fixes that bug (which was never released). closes #22861	2017-01-30 14:42:40 -08:00
Christoph Büscher	4e613139dc	Ensure fixed serialization order of InnerHitBuilder (#22820 ) Usually the order in which we serialize sets and maps of things doesn't matter, but since InnerHitBuilder is part of SearchSourceBuilder, which is in turn used as a cache key in its bytes serialization, we need to ensure the order of all these fields when writing them to an output stream. This adds tests and makes sure we iterate over the scriptField set and the childInnerHits map in a fixed order. Closes #22808	2017-01-30 19:28:55 +01:00
Alex Bumbu	41abf6e81d	Add used memory amount to CircuitBreakingException message (#22521 )	2017-01-28 16:54:13 +00:00
Nik Everett	e042c77301	Add tests for reducing top hits (#22837 ) Also adds many `equals` and `hashCode` implementations and moves the failure printing in `MatchAssertion` into a common spot and exposes it over `assertEqualsWithErrorMessageFromXContent` which does an object equality test but then uses `toXContent` to print the differences. Relates to #22278	2017-01-27 20:54:11 -05:00
Boaz Leskes	b1f0d8f4cf	await fix testRecoveryWaitsForOps	2017-01-27 23:26:23 +01:00
Boaz Leskes	204df2a199	fix @TestLogging annotation	2017-01-27 23:15:51 +01:00
Simon Willnauer	e946ec0c33	Don't convert source to UTF-8 it might not be valid UTF-8	2017-01-27 22:31:44 +01:00
Nik Everett	2e48fb8294	Move delete by query helpers into core (#22810 ) This moves the building blocks for delete by query into core. This should enabled two thigns: 1. Plugins other than reindex to implement "bulk by scroll" style operations. 2. Plugins to directly call delete by query. Those plugins should be careful to make sure that task cancellation still works, but this should be possible. Notes: 1. I've mostly just moved classes and moved around tests methods. 2. I haven't been super careful about cohesion between these core classes and reindex. They are quite interconnected because I wanted to make the change as mechanical as possible. Closes #22616	2017-01-27 16:09:18 -05:00
Ryan Ernst	aad51d44ab	S3 repository: Add named configurations (#22762 ) * S3 repository: Add named configurations This change implements named configurations for s3 repository as proposed in #22520. The access/secret key secure settings which were added in #22479 are reverted, and the only secure settings are those with the new named configs. All other previously used settings for the connection are deprecated. closes #22520	2017-01-27 10:42:45 -08:00
Boaz Leskes	0f58f3f34b	fix TestLogging instructions for the remove of single doc indexing actions + add some seq# related info	2017-01-27 19:21:11 +01:00
Nik Everett	e36e5fc994	Remove annotation Not allowed.	2017-01-27 12:37:35 -05:00
Nik Everett	8abd4101eb	Add tests for reducing top hits Also adds many `equals` and `hashCode` implementations and moves the failure printing in `MatchAssertion` into a common spot and exposes it over `assertEqualsWithErrorMessageFromXContent` which does an object equality test but then uses `toXContent` to print the differences. Relates to #22278	2017-01-27 12:32:17 -05:00
Jason Tedor	930282e161	Introduce sequence-number-based recovery This commit introduces sequence-number-based recovery. When a replica has fallen out of sync, rather than performing a file-based recovery we first attempt to replay operations since the last local checkpoint on the replica. To do this, at the start of recovery the replica tells the primary what its local checkpoint is. The primary will then wait for all operations between that local checkpoint and the current maximum sequence number to complete; this is to ensure that there are no gaps in the operations that will be replayed from the primary to the replica. This is a best-effort attempt as we currently have no guarantees on the primary that these operations will be available; if we are not able to replay all operations in the desired range, we just fallback to file-based recovery. Later work will strengthen the guarantees. Relates #22484	2017-01-27 08:16:38 -08:00
Simon Willnauer	417c93c570	First step towards separating individual search phases (#22802 ) At this point AbstractSearchAsyncAction is just a base-class for the first phase of a search where we have multiple replicas for each shardID. If one of them is not available we move to the next one. Yet, once we passed that first stage we have to work with the shards we succeeded on the initial phase. Unfortunately, subsequent phases are not fully detached from the initial phase since they are all non-static inner classes. In future changes this will be changed to detach the inner classes to test them in isolation and to simplify their creation. The AbstractSearchAsyncAction should be final and it should just get a factory for the next phase instead of requiring subclasses etc.	2017-01-27 15:53:41 +01:00
Igor Motov	b068814d10	Fix hanging cancelling task with no children Cancelling tasks with no cancellable children can cause the cancellation operation to hang. This commit fixes this issue.	2017-01-27 08:03:02 -05:00
Tanguy Leroux	ea7077fb1b	Add parsing method for ElasticsearchException.generateFailureXContent() (#22815 ) This commit adds a ElasticsearchException.failureFromXContent() that can be used to parse the result of ElasticsearchException.generateFailureXContent().	2017-01-27 10:12:58 +01:00
Nik Everett	1baa884ab7	Fix TophitsAggregatorTests It needs a DirectoryReader so it has to be careful. Closes #22818	2017-01-26 14:08:30 -05:00
Nik Everett	f8c28711be	Merge some equivalent interfaces (#22816 ) Remove `FromXContent` and use `CheckedFunction` instead. Remove `FromXContentWithContext` and use `ContentParser` instead.	2017-01-26 13:15:29 -05:00
Simon Willnauer	a475323aa1	Invalidate cached query results if query timed out (#22807 ) Today we cache query results even if the query timed out. This is obviously problematic since results are not complete. Yet, the decision if a query timed out or not happens too late to simply not cache the result since if we'd just throw an exception all currently waiting requests with the same request / cache key would fail with the same exception without the option to access the result or to re-execute. Instead, this change will allow the request to enter the cache but invalidates it immediately. Concurrent request might not get executed and return the timed out result which is not absolutely correct but very likely since identical requests will likely timeout as well. As a side-effect we won't hammer the node with concurrent slow searches but rather only execute one of them and return shortly cached result. Closes #22789	2017-01-26 16:45:29 +01:00
Tanguy Leroux	1fa2734566	[TEST] Fix ElasticsearchExceptionTests Some test failures can happen in ElasticsearchExceptionTests, this commit fixes them.	2017-01-26 16:33:56 +01:00
Tanguy Leroux	be96278c95	Add parsing method for ElasticsearchException.generateThrowableXContent() (#22783 ) The output of the ElasticsearchException.generateThrowableXContent() method can be parsed back by the ElasticsearchException.fromXContent() method. This commit adds unit tests in the style of the to-and-from-xcontent tests we already have for other parsing methods. It also relax the strict parsing of the ElasticsearchException.fromXContent() so that it does not throw an exception when custom metadata and headers are parsed, as long as they are either strings or arrays of strings. Every other type is ignored at parsing time.	2017-01-26 15:17:07 +01:00
Simon Willnauer	f128b7a7fe	Improve connection closing in `RemoteClusterConnection` (#22804 ) Some tests verify that all connection have been closed but due to the async / concurrent nature of `RemoteClusterConnection` there are situations where we notify listeners that trigger tests to finish before we actually closed all connections. The race is very very small and has no impact on the code correctness. This commit documents and improves the way we close connections to ensure test won't fail with false positives. Closes #22803	2017-01-26 13:58:26 +01:00
Simon Willnauer	281250dec9	Remove DFS_QUERY_AND_FETCH as a search type (#22787 ) This commit removes the search type `dfs_query_and_fetch` without a replacement. We don't allow to use this type via REST since 2.x but still keep it around for no particular reason. There we no users complaining about the availability. This should now be removed from the codebase. `query_and_fetch` is still used internally to safe a roundtrip if there is only one shard but it can't be used via the rest interface.	2017-01-26 09:14:44 +01:00
Tim Brooks	719e75bb3f	Add repository-url module and move URLRepository (#22752 ) This is related to #22116. URLRepository requires SocketPermission connect. This commit introduces a new module called "repository-url" where URLRepository will reside. With the new module, permissions can be removed from core.	2017-01-25 17:09:25 -06:00
Nik Everett	d704a880e7	Add tests for top_hits aggregation (#22754 ) Add unit tests for `TopHitsAggregator` and convert some snippets in docs for `top_hits` aggregation to `// CONSOLE`. Relates to #22278 Relates to #18160	2017-01-25 16:15:50 -05:00
Martijn van Groningen	f6ed39aa08	Merge branch 'pr/22772'	2017-01-25 17:15:24 +01:00
Martijn van Groningen	81e40e3139	[TEST] Added this for `93a28b0acf` submitted via #22772	2017-01-25 17:08:17 +01:00
Jason Tedor	cb822b4670	Fix typo in comment in OsProbe.java This commit fixes a silly typo in a comment relating to cgroups in OsProbe.java.	2017-01-25 06:30:51 -05:00
Colin Goodheart-Smithe	a9135cd636	RangeQuery WITHIN case now normalises query (#22431 ) Previous to his change when the range query was rewritten to an unbounded range (`[* TO *]`) it maintained the timezone and format for the query. This means that queries with different timezones and format which are rewritten to unbounded range queries actually end up as different entries in the search request cache. This is inefficient and unnecessary so this change nulls the timezone and format in the rewritten query so that regardless of the timezone or format the rewritten query will be the same. Although this does not fix #22412 (since it deals with the WITHIN case rather than the INTERSECTS case) it is born from the same arguments	2017-01-25 10:37:15 +00:00
Boaz Leskes	ed94f75a15	Remove EngineClosedException All usage has been removed in https://github.com/elastic/elasticsearch/pull/22631, which is back ported to 5.x. This means 6.x will never get it on the wire and we can remove it	2017-01-25 11:00:50 +01:00
javanna	5103b76610	update version checks in ElasticsearchException serialization code 5.3.0 is the first version that contains the split from headers to metadata, updated the check to reflect that. It was previously after to be able to commit to master first, and only after that backport the change. Otherwise master tests would have failed until the change was backported.	2017-01-24 20:40:17 +01:00
Lee Hinman	304296ea6a	Fix BulkItemResponse serialization for 6.x <-> 5.3.x Previously the behavior where the `OpType` byte was serialized was only in master, but it was recently backported to 5.x, so the serialization version checks need to be updated as well.	2017-01-24 12:04:11 -07:00
srgclr	93a28b0acf	skip parentid if child document is an orphan #22770	2017-01-24 17:49:53 +00:00
Luca Cavanna	47c0e13a3b	Stop returning "es." internal exception headers as http response headers (#22703 ) move "es." internal headers to separate metadata set in ElasticsearchException and stop returning them as response headers Closes #17593 * [TEST] remove ESExceptionTests, move its methods to ElasticsearchExceptionTests or ExceptionSerializationTests	2017-01-24 16:12:45 +01:00
Jason Tedor	bcffc6fa49	Add hack for Docker cgroups Docker cgroups are mounted in the wrong place (i.e., inconsistently with /proc/self/cgroup). This commit adds an undocumented hack for working around, for now. Relates #22757	2017-01-24 06:36:03 -05:00
Christoph Büscher	59aefe5a38	Include human readable responses in response parsing tests (#22717 ) As a follow up to #22649, this changes the resent tests for parsing parts of search responses to randomly set the humanReadable() flag of the XContentBuilder that is used to render the responses. This should help to test that we can parse back thoses classes if the user specifies `?human=true` in the request url.	2017-01-24 11:17:58 +01:00
Jim Ferenczi	b0c2a5da30	Remove unused field in CollapseBuilder	2017-01-24 09:26:23 +01:00
Jim Ferenczi	868b12b548	Add BWC tests for field collapsing Field collapsing is supported from version 5.3	2017-01-24 08:34:16 +01:00
Nik Everett	2e399e5505	Rename constant It deserves a new name after the cleanup in #22749	2017-01-23 16:30:42 -05:00
javanna	8065531236	[TEST] add test to verify that the SMILE format works within the _bulk api	2017-01-23 19:40:24 +01:00
Nik Everett	ee264c6957	Fix parsing for `max_determinized_states` (#22749 ) There was a typo in the `ParseField` declaration. I know we want to port these parsers to `ObjectParser` eventually but I don't have the energy for that today and want to get this fixed. Closes #22722	2017-01-23 11:57:43 -05:00
Jim Ferenczi	e48bc2eed7	Add field collapsing for search request (#22337 ) * Add top hits collapsing to search request The field collapsing is done with a custom top docs collector that "collapse" search hits with same field value. The distributed aspect is resolve using the two passes that the regular search uses. The first pass "collapse" the top hits, then the coordinating node merge/collapse the top hits from each shard. ``` GET _search { "collapse": { "field": "category", } } ``` This change also adds an ExpandCollapseSearchResponseListener that intercepts the search response and expands collapsed hits using the CollapseBuilder#innerHit} options. The retrieval of each inner_hits is done by sending a query to all shards filtered by the collapse key. ``` GET _search { "collapse": { "field": "category", "inner_hits": { "size": 2 } } } ```	2017-01-23 16:33:51 +01:00
Tanguy Leroux	11164b394b	Add unit tests for ValueCountAggregator and InternalValueCount (#22741 ) Adds unit tests for the value count aggregator. Relates #22278	2017-01-23 16:24:55 +01:00
Simon Willnauer	27b5c2ad54	Pass `forceExecution` flag to transport interceptor (#22739 ) To effectively allow a plugin to intercept a transport handler it needs to know if the handler must be executed even if there is a rejection on the thread pool in the case the wrapper forks a thread to execute the actual handler.	2017-01-23 11:04:27 +01:00
Alexander Reelsen	6159ca28ae	Version: Add missing releases from 2.x in Version.java (#22594 )	2017-01-23 09:53:21 +01:00
Tim Brooks	a4ac29c005	Add single static instance of SpecialPermission (#22726 ) This commit adds a SpecialPermission constant and uses that constant opposed to introducing new instances everywhere. Additionally, this commit introduces a single static method to check that the current code has permission. This avoids all the duplicated access blocks that exist currently.	2017-01-21 12:03:52 -06:00
Simon Willnauer	3ad6d6ebcc	Simplify InternalEngine#innerIndex (#22721 ) Today `InternalEngine#innerIndex` is a pretty big method (> 150 SLoC). This commit merged `#index` and `#innerIndex` and splits it up into smaller contained methods.	2017-01-21 08:51:35 +01:00
Jim Ferenczi	8028578305	Upgrade to Lucene 6.4.0 (#22724 ) * Upgrade to Lucene 6.4.0 `ValueSource`s are now converted to `DoubleValueSource`s using the Lucene adapter made for the migration to the new API in 6.4.0.	2017-01-21 04:48:01 +01:00
Igor Motov	cfb415de7b	Fix broken TaskInfo.toString() Related to #22387	2017-01-20 20:57:42 -05:00
Tim Brooks	d86f97c428	Add CheckedSupplier and CheckedRunnable to core (#22725 ) Introduce CheckedSupplier and CheckedRunnable functional interfaces into core. These offer a checked version of the Supplier and Runnable interfaces for use with lambda apis.	2017-01-20 19:17:44 -06:00
Ali Beyad	3bf06d1440	Fixes retrieval of the latest snapshot index blob (#22700 ) This commit ensures that the index.latest blob is first examined to determine the latest index-N blob id, before attempting to list all index-N blobs and picking the blob with the highest N. It also fixes the MockRepository#move so that tests are able to handle non-atomic moves. This is done by adding a special setting to the MockRepository that requires the test to specify if it can handle non-atomic moves. If so, then the MockRepository#move operation will be non-atomic to allow testing for against such repositories.	2017-01-20 17:00:46 -06:00
Jim Ferenczi	4ec4bad908	Fix script score function that combines _score and weight (#22713 ) The weight factor function does not check if the delegate score function needs to access the score of the query. This results in a _score equals to 0 for all score function that set a weight. This change modifies the WeightFactorFunction#needsScore to delegate the call to its underlying score function. Fix #21483	2017-01-20 19:50:57 +01:00
Nik Everett	6265ef1c1b	Deguice rest handlers (#22575 ) There are presently 7 ctor args used in any rest handlers: * `Settings`: Every handler uses it to initialize a logger and some other strange things. * `RestController`: Every handler registers itself with it. * `ClusterSettings`: Used by `RestClusterGetSettingsAction` to render the default values for cluster settings. * `IndexScopedSettings`: Used by `RestGetSettingsAction` to get the default values for index settings. * `SettingsFilter`: Used by a few handlers to filter returned settings so we don't expose stuff like passwords. * `IndexNameExpressionResolver`: Used by `_cat/indices` to filter the list of indices. * `Supplier<DiscoveryNodes>`: Used to fill enrich the response by handlers that list tasks. We probably want to reduce these arguments over time but switching construction away from guice gives us tighter control over the list of available arguments. These parameters are passed to plugins using `ActionPlugin#initRestHandlers` which is expected to build and return that handlers immediately. This felt simpler than returning an reference to the ctors given all the different possible args. Breaks java plugins by moving rest handlers off of guice.	2017-01-20 11:48:51 -05:00
Simon Willnauer	824beea89d	Fix handling of document failure expcetion in InternalEngine (#22718 ) Today we try to be smart and make a generic decision if an exception should be treated as a document failure but in some cases concurrency in the index writer make this decision very difficult since we don't have a consistent state in the case another thread is currently failing the IndexWriter/InternalEngine due to a tragic event. This change simplifies the exception handling and makes specific decisions about document failures rather than using a generic heuristic. This prevent exceptions to be treated as document failures that should have failed the engine but backed out of failing since since some other thread has already taken over the failure procedure but didn't finish yet.	2017-01-20 16:55:00 +01:00
markharwood	f01784205f	New AdjacencyMatrix aggregation Similar to the Filters aggregation but only supports "keyed" filter buckets and automatically "ANDs" pairs of filters to produce a form of adjacency matrix. The intersection of buckets "A" and "B" is named "A&B" (the choice of separator is configurable). Empty intersection buckets are removed from the final results. Closes #22169	2017-01-20 15:49:31 +00:00
Tim Brooks	bc16162d21	Remove accept SocketPermissions from core (#22622 ) This is related to #22116. Core no longer needs SocketPermission accept. This permission is relegated to the transport-netty4 module and (for tests) to the mocksocket jar.	2017-01-20 09:27:45 -06:00
Tanguy Leroux	239ed0c912	Add unit tests for DateHistogramAggregator (#22714 ) Adds unit tests for the date histogram aggregator. Relates #22278	2017-01-20 14:18:30 +01:00
Christoph Büscher	54105f3ddd	Add parsing from xContent to ShardSearchFailure (#22699 ) In preparation for being able to parse SearchResponse from its rest representation, this adds fromXContent to ShardSearchFailure.	2017-01-20 12:49:54 +01:00
Yannick Welsch	1f0e0a2170	Close InputStream when receiving cluster state in PublishClusterStateAction (#22711 ) Not closing the InputStream will leak native memory as the DeflateCompressor/Inflater won't be closed.	2017-01-20 12:26:07 +01:00
Boaz Leskes	5d806bf93e	Index creation and setting update may not return deprecation logging (#22702 ) Those services validate their setting before submitting an AckedClusterStateUpdateTask to the cluster state service. An acked cluster state may be completed by a networking thread when the last acks as received. As such it needs special care to make sure that thread context headers are handled correctly.	2017-01-20 10:14:13 +01:00
David Pilato	fc4dc5ef21	Fix comment	2017-01-20 10:13:13 +01:00
David Pilato	ad5b8def26	Merge branch 'pr/delete-from-xcontent'	2017-01-20 09:16:34 +01:00
Lee Hinman	eb8a41ef94	Add missing serialization BWC for disk usage estimates Relates to #22081	2017-01-19 15:37:06 -07:00
Lee Hinman	4eb32e9d86	Expose disk usage estimates in nodes stats This exposes the least and most used disk usage estimates within the "fs" nodes stats output: ```json GET /_nodes/stats/fs?pretty&human { "nodes" : { "34fPVU0uQ_-wWitDzDXX_g" : { "fs" : { "timestamp" : 1481238723550, "total" : { "total" : "396.1gb", "total_in_bytes" : 425343254528, "free" : "140.6gb", "free_in_bytes" : 151068725248, "available" : "120.5gb", "available_in_bytes" : 129438912512 }, "least_usage_estimate" : { "path" : "/home/hinmanm/es/elasticsearch/distribution/build/cluster/run node0/elasticsearch-6.0.0-alpha1-SNAPSHOT/data/nodes/0", "total" : "396.1gb", "total_in_bytes" : 425343254528, "available" : "120.5gb", "available_in_bytes" : 129438633984, "used_disk_percent" : 69.56842912023208 }, "most_usage_estimate" : { "path" : "/home/hinmanm/es/elasticsearch/distribution/build/cluster/run node0/elasticsearch-6.0.0-alpha1-SNAPSHOT/data/nodes/0", "total" : "396.1gb", "total_in_bytes" : 425343254528, "available" : "120.5gb", "available_in_bytes" : 129438633984, "used_disk_percent" : 69.56842912023208 }, "data" : [{...}], "io_stats" : {...} } } } } ``` Resolves #8686 Resolves #22081	2017-01-19 13:56:52 -07:00
Jason Tedor	9781b88a38	Fix deprecation logging for lenient booleans This commit fixes an issue with deprecation logging for lenient booleans. The underlying issue is that adding deprecation logging for lenient booleans added a static deprecation logger to the Settings class. However, the Settings class is initialized very early and in CLI tools can be initialized before logging is initialized. This leads to status logger error messages. Additionally, the deprecation logging for a lot of the settings does not provide useful context (for example, in the token filter factories, the deprecation logging only produces the name of the setting, but gives no context which token filter factory it comes from). This commit addresses both of these issues by changing the call sites to push a deprecation logger through to the lenient boolean parsing. Relates #22696	2017-01-19 12:30:33 -05:00
David Pilato	5be8bd76e2	Also test found field And optimize imports	2017-01-19 17:28:31 +01:00
Tim Brooks	3deae99a34	Fix incorrect args order passed to createAggregator This commit fixes a compile issue where the arguments are passed to createAggregator in the incorrect order.	2017-01-19 10:08:38 -06:00
Christoph Büscher	e03554070c	Add parsing from xContent to SearchProfileShardResults and nested classes (#22649 ) In preparation for being able to parse SearchResponse from its rest representation for the java rest client, this adds fromXContent to SearchProfileShardResults and its nested classes.	2017-01-19 16:29:10 +01:00
Jim Ferenczi	b781a4a176	Add unit tests for FiltersAggregator (#22678 ) Adds unit tests for the `filters` aggregation. This change also adds an helper to search and reduce any aggregator in a unit test. This is done by dividing a single searcher in sub-searcher, one for each segment. Relates #22278	2017-01-19 16:22:48 +01:00
David Pilato	0315dcc306	Use now common methods with index/update Brought by #22229	2017-01-19 16:10:13 +01:00
Jim Ferenczi	3d54258de2	Don't register search response listener in transport clients Small fix for https://github.com/elastic/elasticsearch/pull/22682	2017-01-19 16:08:24 +01:00
David Pilato	718a6b9be7	Add fromxcontent methods to delete response This commit adds the parsing fromXContent() methods to the IndexResponse class. It's a pale copy of what has been done in #22229.	2017-01-19 15:59:24 +01:00
Nicholas Knize	b006636aaf	unmute FieldStatsIntegrationIT.testGeoPointNotIndexed, fix already pushed	2017-01-19 08:44:00 -06:00
Nicholas Knize	88c78833f0	Mute FieldStatsIntegrationIT.testGeoPointNotIndexed, for now	2017-01-19 08:38:17 -06:00
Jim Ferenczi	d145d459ae	Fix NPE on FieldStats with mixed cluster on version pre/post 5.2 (#22688 ) * Fix NPE on FieldStats with mixed cluster on version pre/post 5.2 In 5.2 the FieldStats API can return null min/max values. These values cannot be deserialized by a node with version pre 5.2 so if this node is pick to coordinate a FieldStats request in a mixed cluster an NPE can be thrown. This change prevents the NPE by removing the non serializable FieldStats object directly in the field stats shard request. The filtered fields will not be present in the response when a node pre 5.2 acts as a coordinating node.	2017-01-19 14:20:07 +01:00
Tanguy Leroux	833284cae2	Add parsing methods for UpdateResponse (#22586 ) This commit adds the fromXContent() method to the UpdateResponse class, so that it can be used with the high level rest client.	2017-01-19 12:49:45 +01:00
Jim Ferenczi	21dae1924f	Add the ability to define search response listeners in plugins (#22682 ) This change is a simple adaptation of https://github.com/elastic/elasticsearch/pull/19587 for the current state of master. It allows to define search response listener in the form of `BiConsumer<SearchRequest, SearchResponse>`s in a search plugin.	2017-01-19 12:48:45 +01:00
Daniel Mitterdorfer	ce765f7ad2	Use a proper boolean in FieldStatsIntegrationIT#testGeoPointNotIndexed()	2017-01-19 08:33:08 +01:00
Daniel Mitterdorfer	aece89d6a1	Make boolean conversion strict (#22200 ) This PR removes all leniency in the conversion of Strings to booleans: "true" is converted to the boolean value `true`, "false" is converted to the boolean value `false`. Everything else raises an error.	2017-01-19 07:59:18 +01:00
Nicholas Knize	51e80e7176	remove unnecessary text from exception message	2017-01-18 14:51:56 -06:00
Nicholas Knize	84e4f91253	Add geo_point to FieldStats This commit adds a new GeoPoint class to FieldStats for computing field stats over geo_point field types.	2017-01-18 14:37:03 -06:00
Nik Everett	1fe74a6b4b	Better error when can't auto create index (#22488 ) Changes the error message when `action.auto_create_index` or `index.mapper.dynamic` forbids automatic creation of an index from `no such index` to one of: * `no such index and [action.auto_create_index] is [false]` * `no such index and [index.mapper.dynamic] is [false]` * `no such index and [action.auto_create_index] contains [-<pattern>] which forbids automatic creation of the index` * `no such index and [action.auto_create_index] ([all patterns]) doesn't match` This should make it more clear why there is `no such index`. Closes #22435	2017-01-18 15:18:32 -05:00
Ali Beyad	cd52065871	[TEST] testAckedIndexing waits for all nodes to stabilize testAckedIndexing now waits for all nodes to stabilize in the cluster state through an assertBusy before final validation that all documents are found in tehir respective shards in the cluster. Before, what could happen is that the ensureGreen check passes but only after that is a ping failure from the network disruption processed by the master, thereby rendering the cluster RED again. This assertBusy waits up to 30 seconds for all nodes to have stabilized and all get document actions to succeed.	2017-01-18 13:51:25 -05:00
Michael McCandless	1d1bdd476c	Finish exposing FlattenGraphTokenFilter (#22667 )	2017-01-18 11:05:34 -05:00
Nik Everett	e71b26f480	Improve unit test coverage of aggs (#22668 ) Add tests for `GlobalAggregator`, `MaxAggregator`, and `InternalMax`. Relates to #22278	2017-01-18 10:33:45 -05:00
Simon Willnauer	24e2847af2	Streamline foreign stored context restore and allow to perserve response headers (#22677 ) Today we do not preserve response headers if they are present on a transport protocol response. While preserving these headers is not always desired, in the most cases we should pass on these headers to have consistent results for depreciation headers etc. yet, this hasn't been much of a problem since most of the deprecations are detected early ie. on the coordinating node such that this bug wasn't uncovered until #22647 This commit allow to optionally preserve headers when a context is restored and also streamlines the context restore since it leaked frequently into the callers thread context when the callers context wasn't restored again.	2017-01-18 16:17:54 +01:00
Ali Beyad	8a0a1140a9	[TEST] add logging to MockRepository to help debug index-N blob reading	2017-01-18 08:53:29 -05:00
Boaz Leskes	1227044ddd	Add a deprecation notice to shadow replicas (#22647 ) Relates to #22024 On top of documentation, the PR adds deprecation loggers and deals with the resulting warning headers. The yaml test is set exclude versions up to 6.0. This is need to make sure bwc tests pass until this is backported to 5.2.0 . Once that's done, I will change the yaml test version limits	2017-01-18 12:28:09 +01:00
Ke Li	797d105177	Remove unnecessary class cast	2017-01-18 11:09:09 +01:00
Simon Willnauer	19f9cb307a	Merge branch 'master' into feature/multi_cluster_search	2017-01-18 09:24:35 +01:00
Scott Somerville	372812da98	Allow an index to be partitioned with custom routing (#22274 ) This change makes it possible for custom routing values to go to a subset of shards rather than just a single shard. This enables the ability to utilize the spatial locality that custom routing can provide while mitigating the likelihood of ending up with an imbalanced cluster or suffering from a hot shard. This is ideal for large multi-tenant indices with custom routing that suffer from one or both of the following: - The big tenants cannot fit into a single shard or there is so many of them that they will likely end up on the same shard - Tenants often have a surge in write traffic and a single shard cannot process it fast enough Beyond that, this should also be useful for use cases where most queries are done under the context of a specific field (e.g. a category) since it gives a hint at how the data can be stored to minimize the number of shards to check per query. While a similar solution can be achieved with multiple concrete indices or aliases per value today, those approaches breakdown for high cardinality fields. A partitioned index enforces that mappings have routing required, that the partition size does not change when shrinking an index (the partitions will shrink proportionally), and rejects mappings that have parent/child relationships. Closes #21585	2017-01-18 08:51:23 +01:00
Igor Motov	500548fcda	Remove taskManager.registerChildTask Instead of forcing each task to register all nodes where its children are running, this commit runs cancellation on all nodes. The task cancellation operation doesn't run too frequently, so this optimization doesn't seem to be worth additional complexity of the interface.	2017-01-17 18:07:31 -05:00
Ali Beyad	ce811feba7	[TEST] testAckedIndexing waits for the cluster state to have propogated to all nodes in the cluster before checking the existance of documents on each node	2017-01-17 15:36:31 -05:00
Nik Everett	1169cd936e	Fix compilation in eclipse Eclipse needs a bit of extra special help with type parameters in `TransportReplicationActionTests` now.	2017-01-17 14:53:54 -05:00
Ali Beyad	554a5e3039	[TEST] add retries to MockRepository getRepositoryData to try to diagnose a NotXContentException being thrown	2017-01-17 12:17:29 -05:00
Simon Willnauer	69f1ffb1f8	fix exception message	2017-01-17 17:29:43 +01:00
Simon Willnauer	292e3a60d1	apply review comments	2017-01-17 17:20:52 +01:00
Ali Beyad	e2977889b8	Allow comma delimited array settings to have a space after each entry (#22591 ) Previously, certain settings that could take multiple comma delimited values would pick up incorrect values for all entries but the first if each comma separated value was followed by a whitespace character. For example, the multi-value "A,B,C" would be correctly parsed as ["A", "B", "C"] but the multi-value "A, B, C" would be incorrectly parsed as ["A", " B", " C"]. This commit allows a comma separated list to have whitespace characters after each entry. The specific settings that were affected by this are: cluster.routing.allocation.awareness.attributes index.routing.allocation.require.* index.routing.allocation.include.* index.routing.allocation.exclude.* cluster.routing.allocation.require.* cluster.routing.allocation.include.* cluster.routing.allocation.exclude.* http.cors.allow-methods http.cors.allow-headers For the allocation filtering related settings, this commit also provides validation of each specified entry if the filtering is done by _ip, _host_ip, or _publish_ip, to ensure that each entry is a valid IP address. Closes #22297	2017-01-17 08:51:04 -06:00
Tanguy Leroux	f5542ed47f	Simplify ElasticsearchException rendering as a XContent (#22611 ) This commit tries to simplify the way ElasticsearchException are rendered to xcontent. It adds some documentation and renames and merges some methods. Current behavior is preserved, the goal is to be more readable and centralize everything in the ElasticsearchException class.	2017-01-17 15:44:49 +01:00
Simon Willnauer	197cd7d7a9	Add test for the grouping error message if indices and cluster can't be disambiguated	2017-01-17 14:13:09 +01:00
Simon Willnauer	88f6ae55f5	Improve remote / local indices filtering by not modifying external state	2017-01-17 14:05:36 +01:00
Simon Willnauer	709cb9a39e	Merge branch 'master' into feature/multi_cluster_search	2017-01-17 12:34:36 +01:00
Simon Willnauer	1c5cc58373	apply review comments	2017-01-17 11:46:55 +01:00
Tim Brooks	16a76d9bc0	Remove blocking TCP clients and servers (#22639 ) This commit removes the option to use the blocking variants of the TCP transport server, TCP transport client, or http server.	2017-01-16 18:38:51 -06:00
Michael McCandless	ebd38e2a6a	Expose FlattenGraphTokenFilter (#22643 ) FlattenGraphTokenFilter is necessary for using graph-based token streams (e.g. the new SynonymGraphFilter) during indexing.	2017-01-16 16:53:32 -05:00
Boaz Leskes	d80e3eea6c	Replace EngineClosedException with AlreadyClosedExcpetion (#22631 ) `EngineClosedException` is a ES level exception that is used to indicate that the engine is closed when operation starts. It doesn't really add much value and we can use `AlreadyClosedException` from Lucene (which may already bubble if things go wrong during operations). Having two exception can just add confusion and lead to bugs, like wrong handling of `EngineClosedException` when dealing with document level failures. The latter was exposed by `IndexWithShadowReplicasIT`. This PR also removes the AwaitFix from the `IndexWithShadowReplicasIT` tests (which was what cause this to be discovered). While debugging the source of the issue I found some mismatches in document uid management in the tests. The term that was passed to the engine didn't correspond to the uid in the parsed doc - those are fixed as well.	2017-01-16 21:14:41 +01:00
Simon Willnauer	f30b1f82ee	Remove HttpServer and HttpServerAdapter in favor of a simple dispatch method (#22636 ) Today we have quite some abstractions that are essentially providing a simple dispatch method to the plugins defining a `HttpServerTransport`. This commit removes `HttpServer` and `HttpServerAdaptor` and introduces a simple `Dispatcher` functional interface that delegate to `RestController` by default. Relates to #18482	2017-01-16 21:06:08 +01:00
Boaz Leskes	f88ab76067	Revert "Add a deprecation notice to shadow replicas (#22025 )" This reverts commit `0da190234c`.	2017-01-16 16:15:41 +01:00
Boaz Leskes	b887681550	Revert "Don'y use `INDEX_SHARED_FS_ALLOW_RECOVERY_ON_ANY_NODE_SETTING` directly as it triggers (many) deprecation logging" This reverts commit `e976aa09bb`.	2017-01-16 16:15:32 +01:00
Boaz Leskes	e976aa09bb	Don'y use `INDEX_SHARED_FS_ALLOW_RECOVERY_ON_ANY_NODE_SETTING` directly as it triggers (many) deprecation logging #22025 deprecated this setting (pending it's removal) but it's frequent usage will spam the deprecation logs and also fails test. As temporary work around we should not use the setting object directly.	2017-01-16 16:11:59 +01:00
Boaz Leskes	0da190234c	Add a deprecation notice to shadow replicas (#22025 ) Also adds deprecation logging. See #22024	2017-01-16 15:40:05 +01:00
Christoph Büscher	59a48ffc41	ProfileResult and CollectorResult should print machine readable timing information (#22561 ) Currently both ProfileResult and CollectorResult print the time field in a human readable string format (e.g. "time": "55.20315000ms"). When trying to parse this back to a long value, for example to use in the planned high level java rest client, we can lose precision because of conversion and rounding issues. This change adds a new additional field (`time_in_nanos`) to the profile response to be able to get the original time value in nanoseconds back. The old `time` field is only printed when the `?`human=true` flag in the url is set. This follow the behaviour for all other stats-related apis. Also the format of the `time` field is slightly changed. Instead of always formatting the output as a 10-digit ms value, by using the `XContentBuilder#timeValueField()` method we now print the largest time unit present is used (e.g. "s", "ms", "micros").	2017-01-16 14:27:55 +01:00
Jason Tedor	e6dc74f2bf	Add replica ops with version conflict to translog An operation that completed successfully on a primary can result in a version conflict on a replica due to the asynchronous nature of operations. When a replica operation results in a version conflict, the operation is not added to the translog. This leads to gaps in the translog which is problematic as it can lead to situations where a replica shard can never advance its local checkpoint. As such operations are just normal course of business for a replica shard, these operations should be treated as if they completed successfully. This commit adds these operations to the translog. Relates #22626	2017-01-16 08:08:52 -05:00
javanna	8e3f1dd689	Replace custom Functional interface in ElasticsearchException with CheckedFunction	2017-01-16 13:57:58 +01:00
javanna	9a910d3c9d	Make RestChannelConsumer extend CheckedConsumer<RestChannel, Exception>	2017-01-16 13:57:58 +01:00
javanna	ab144c418e	replace ShardSearchRequest.FilterParser functional interface with CheckedFunction	2017-01-16 13:57:58 +01:00
javanna	bc22afcb2f	[TEST] replace SizeFunction with Function<Integer, Integer>	2017-01-16 13:57:58 +01:00
javanna	884302dcaa	Expose CheckedFunction	2017-01-16 13:57:58 +01:00
Jason Tedor	fc3280b3cf	Expose logs base path For certain situations, end-users need the base path for Elasticsearch logs. Exposing this as a property is better than hard-coding the path into the logging configuration file as otherwise the logging configuration file could easily diverge from the Elasticsearch configuration file. Additionally, Elasticsearch will only have permissions to write to the log directory configured in the Elasticsearch configuration file. This commit adds a property that exposes this base path. One use-case for this is configuring a rollover strategy to retain logs for a certain period of time. As such, we add an example of this to the documentation. Additionally, we expose the property es.logs.cluster_name as this is used as the name of the log files in the default configuration. Finally, we expose es.logs.node_name in cases where node.name is explicitly set in case users want to include the node name as part of the name of the log files. Relates #22625	2017-01-16 07:39:37 -05:00
Jason Tedor	9ae5410ea6	Do not configure a logger named level When logger.level is set, we end up configuring a logger named "level" because we look for all settings of the form "logger\..+" as configuring a logger. Yet, logger.level is special and is meant to only configure the default logging level. This commit causes is to avoid not configuring a logger named level. Relates #22624	2017-01-16 07:30:21 -05:00
Simon Willnauer	895124e67e	Merge branch 'master' into feature/multi_cluster_search	2017-01-16 13:20:45 +01:00
Alexander Reelsen	f6ee6e420b	Indexing: Add shard id to indexing operation listener (#22606 ) The IndexingOperationListener interface did not provide any information about the shard id when a document was indexed. This commit adds the shard id as the first parameter to all methods in the IndexingOperationListener.	2017-01-16 09:08:16 +01:00
Jason Tedor	526cf6182d	Cleanup handling of cgroup stats This commit is a simple cleanup of the code related to cgroup stats: - reduce visibility of a method - remove an unneeded logger guard - cleanup the formatting of comments	2017-01-15 12:18:16 -05:00
Simon Willnauer	5f0344a918	Pass ThreadContext to transport interceptors to allow header modification (#22618 ) TransportInterceptors are commonly used to enrich requests with headers etc. which requires access the the thread context. This is not always easily possible since threadpools are hard to access for instance if the interceptor is used on a transport client. This commit passes on the thread context to all the interceptors for further consumption. Closes #22585	2017-01-15 13:35:39 +01:00
Simon Willnauer	3f784a4424	Merge branch 'master' into feature/multi_cluster_search	2017-01-15 10:28:34 +01:00
Jason Tedor	bed719de0a	Log deleting indices at info level Deleting indices is an important event in a cluster and as such should be logged at the info level. This commit changes the logging level on index deletion to the info level. Relates #22627	2017-01-14 23:13:40 -05:00
Simon Willnauer	fde11649fb	harden tests	2017-01-13 23:59:59 +01:00
Jason Tedor	d67514606e	Fix out-of-date Javadocs on Security.java We have made the security manager non-optional, but the Javadocs for Security.java imply that it still is. This commit fixes this issue. Relates #16176	2017-01-13 17:20:45 -05:00
Simon Willnauer	63e4552c0d	Merge branch 'master' into feature/multi_cluster_search	2017-01-13 23:07:20 +01:00
Ali Beyad	0c7fc229b8	[TEST] No longer randomly block on the index-N files in the MockRepository, because the getRepositoryData() call depends on it, which is used in non-synchronized actions such as getting snapshot status.	2017-01-13 15:57:37 -05:00
Lee Hinman	cd236c4de4	Merge remote-tracking branch 'zareek/enhancement/use_shard_bulk_for_single_ops'	2017-01-13 10:09:18 -07:00
Simon Willnauer	4c1ee018f6	Remove setLocalNode from ClusterService and TransportService (#22608 ) ClusterService and TransportService expect the local discovery node to be set before they are started but this requires manual interaction and is error prone since to work absolutely correct they should share the same instance (same ephemeral ID). TransportService also has 2 modes of operation, mainly realted to transport client vs. internal to a node. This change removes the mode where we don't maintain a local node and uses a dummy local node in the transport client since we don't bind to any port in such a case. Local discovery node instances are now managed by the node itself and only suppliers and factories that allow creation only once are passed to TransportService and ClusterService.	2017-01-13 16:12:27 +01:00
Simon Willnauer	d5fa84f869	Harder close and remove reference concurrency in MockTcpTransport (#22613 ) There was still small race in MockTcpTransport where channesl that are concurrently closing are not yet removed from the reference tracking causing tests to fail. Compared to the other races before this is a rather small windown and requires very very short test durations.	2017-01-13 16:04:05 +01:00
Matt Weber	beceb4bf8a	Analyze API Position Length Support (#22574 ) Expose the position length attribute if a token has a non-standard position length greater than 1.	2017-01-13 09:12:49 -05:00
David Pilato	815c4ac4c8	Update after review and add a test	2017-01-13 12:51:35 +01:00
Simon Willnauer	e2ebabcb3c	Use a set rather than a list for connected nodes	2017-01-13 12:25:35 +01:00
Simon Willnauer	9c167cc92d	preserve original excetption	2017-01-13 12:25:18 +01:00
Simon Willnauer	a8bd57b93c	Fail request if there is a local index that matches the both a remote and local index	2017-01-13 12:22:52 +01:00
David Pilato	726c05b7c5	NPE when no setting name passed to elasticsearch-keystore ```h $ bin/elasticsearch-keystore create Created elasticsearch keystore in /Users/dpilato/Documents/Elasticsearch/apps/elasticsearch/elasticsearch-6.0.0-alpha1/config $ bin/elasticsearch-keystore add Enter value for null: xyz Exception in thread "main" java.lang.NullPointerException: invalid null input at java.security.KeyStore.setEntry(KeyStore.java:1552) at org.elasticsearch.common.settings.KeyStoreWrapper.setString(KeyStoreWrapper.java:264) at org.elasticsearch.common.settings.AddStringKeyStoreCommand.execute(AddStringKeyStoreCommand.java:83) at org.elasticsearch.cli.EnvironmentAwareCommand.execute(EnvironmentAwareCommand.java:58) at org.elasticsearch.cli.Command.mainWithoutErrorHandling(Command.java:122) at org.elasticsearch.cli.MultiCommand.execute(MultiCommand.java:69) at org.elasticsearch.cli.Command.mainWithoutErrorHandling(Command.java:122) at org.elasticsearch.cli.Command.main(Command.java:88) at org.elasticsearch.common.settings.KeyStoreCli.main(KeyStoreCli.java:39) ```	2017-01-13 12:20:20 +01:00
Simon Willnauer	6779ea9c2a	Merge branch 'master' into feature/multi_cluster_search	2017-01-13 12:10:23 +01:00
Jim Ferenczi	f18d5f22ce	Fix NPE in TermsAggregatorTests when global ordinals are needed	2017-01-13 09:16:24 +01:00
Tanguy Leroux	3a3ce61186	Update Jackson to 2.8.6 (#22596 ) closes #22266	2017-01-13 09:05:48 +01:00
Michael McCandless	568e655fdb	Source filtering: only accept array items if the previous include pattern matches (#22593 ) Source filtering was always accepting array items even if the include pattern did not match. Closes #22557	2017-01-12 19:00:39 -05:00
Nik Everett	baed02bbe2	Whitelist some ScriptDocValues in painless (#22600 ) Without this whitelist painless can't use ip or binary doc values. Closes #22584	2017-01-12 15:26:09 -05:00
Lee Hinman	58daf5fb6d	Bump version to 5.1.3 (#22597 ) * Bump version to 5.1.3 Bumps version and adds the BWC indices	2017-01-12 12:37:26 -07:00
Simon Willnauer	acf2d2f86f	Ensure new connections won't be opened if transport is closed or closing (#22589 ) Today there are several races / holes in TcpTransport and MockTcpTransport that can allow connections to be opened and remain unclosed while the actual transport implementation is closed. A recently added assertions in #22554 exposes these problems. This commit fixes several issues related to missed locks or channel creations outside of a lock not checking if the resource is still open.	2017-01-12 20:27:09 +01:00
Lee Hinman	2db01b6127	Merge remote-tracking branch 'dakrone/disable-all-by-default'	2017-01-12 10:17:51 -07:00
javanna	8e8ac5f239	Remove ParseFieldMatcher and ParseFieldMatcherSupplier Closes #19552	2017-01-12 14:43:35 +01:00
javanna	def8125e51	Remove ParseFieldMatcher usages from BaseRestHandler	2017-01-12 14:43:35 +01:00
javanna	9e680e8e51	Remove ParseFieldMatcher usages from TransportAction	2017-01-12 14:43:35 +01:00
javanna	64c3212fdb	Remove ParseFieldMatcher usages from IndexSettings	2017-01-12 14:43:35 +01:00
javanna	4449eb181b	Remove ParseFieldMatcher usages from QueryRewriteContext	2017-01-12 14:43:35 +01:00
javanna	8072f168a3	Remove ParseFieldMatcher usages from QueryParseContext	2017-01-12 14:43:35 +01:00
javanna	83a3f0e42c	fix compile error ExtendedBounds cannot yet have its ParseFieldMatcher usage removed, reverted that bit.	2017-01-12 10:21:47 +01:00
Luca Cavanna	0f7d52df68	Remove some more ParseFieldMatcher usages (#22571 )	2017-01-12 10:04:10 +01:00
Lee Hinman	7a18bb50fc	Disable _all by default This change disables the _all meta field by default. Now that we have the "all-fields" method of query execution, we can save both indexing time and disk space by disabling it. _all can no longer be configured for indices created after 6.0. Relates to #20925 and #21341 Resolves #19784	2017-01-11 16:47:13 -07:00
Simon Willnauer	bf15decf20	flush pending listeners if remote cluster connection is closed	2017-01-12 00:21:46 +01:00
Simon Willnauer	00781d24ce	Merge branch 'master' into feature/multi_cluster_search	2017-01-11 23:40:46 +01:00
Simon Willnauer	8a0393f718	Move assertion for open channels under TcpTransport lock TcpTransport has an actual mechanism to stop resources in subclasses. Instead of overriding `doStop` subclasses should override `stopInternal` that is executed under the connection lock guaranteeing that there is no concurrency etc. Relates to #22554	2017-01-11 23:37:12 +01:00
Luca Cavanna	ec73cfe937	Remove unused *XContentGenerator constructors (#22558 )	2017-01-11 20:51:23 +01:00
Ryan Ernst	8015fbbf25	Make s3 repository sensitive settings use secure settings (#22479 ) * Settings: Make s3 repository sensitive settings use secure settings This change converts repository-s3 to use the new secure settings. In order to support the multiple ways we allow aws creds to be configured, it also moves the main methods for the keystore wrapper into a SecureSettings interface, in order to allow settings prefixing to work.	2017-01-11 11:19:46 -08:00
Matt Weber	609d2aab15	QueryString and SimpleQueryString Graph Support (#22541 ) Add support for graph token streams to "query_String" and "simple_query_string" queries.	2017-01-11 18:59:43 +01:00
Lee Hinman	e93fdb8460	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2017-01-11 10:08:46 -07:00
Lee Hinman	fed2a1a822	Fix Translog.Delete serialization for sequence numbers (#22543 ) * Fix Translog.Delete serialization for sequence numbers Translog.Delete used `.writeVLong` instead of `.writeLong` for the sequence number and primary term (and their respective "read" variants). This could lead to issues where a 5.x node sent a translog operation with a negative sequence number (-2 for unassigned seq no) that tripped an assertion serializing a negative number and causing ES to exit. Adds a unit test for serialization and a mixed-cluster REST test, since that was how this was originally caught. * Use more realistic values for random seqNum and primary term * Add comment with TODO for removal in 7.0 * Change comment into an assert	2017-01-11 10:08:04 -07:00
Ali Beyad	389ffc93d8	Adds debugging information for invalid repository data x-content	2017-01-11 11:27:25 -05:00
Simon Willnauer	d3124dd62b	Merge branch 'master' into feature/multi_cluster_search	2017-01-11 17:03:30 +01:00
Simon Willnauer	6810125a8b	Prevent open channel leaks if handshake times out or is interrupted (#22554 ) The low level TCP handshake can cause channel / connection leaks if it's interrupted since the caller doesn't close the channel / connection if the handshake was not successful. This commit fixes the channel leak and adds general test infrastructure to detect channel leaks in the future.	2017-01-11 17:02:36 +01:00
Nik Everett	abb7d7841f	Remove SearchRequestParsers (#22538 ) It is empty now that we've moved all the parsing into `namedObject`.	2017-01-11 10:28:14 -05:00
Simon Willnauer	d36fc66af1	fix redundant modifier	2017-01-11 15:09:29 +01:00
Simon Willnauer	e23de3229f	add additional explaination to exception handling in AbstractSearchAsyncAction	2017-01-11 14:49:55 +01:00
Simon Willnauer	a79896674a	Simplify ActionListener helpers and add dedicated unittests	2017-01-11 14:47:45 +01:00
Simon Willnauer	2aae409508	Merge branch 'master' into feature/multi_cluster_search	2017-01-11 12:41:26 +01:00
Simon Willnauer	4c61f1d75d	Cut over to use affix setting for remote cluster configuration Instead of `search.remote.seeds.${clustername}` we now specify the seeds as: `search.remote.${clustername}.seeds` which is a real list setting compared to an unvalidated group setting before.	2017-01-11 12:38:46 +01:00
Yannick Welsch	baef86b9d3	[TEST] Disable testRamBytesUsed on JDK 9	2017-01-11 10:55:29 +01:00
Luca Cavanna	0f391336f5	Clean up SearchShardTarget (#22468 ) * unify shard target setter * Remove indexText member from SearchShardTarget * Remove duplicated indexName getter from SearchShardTarget * Remove duplicated shardId getter from SearchShardTarget * Remove duplicated nodeIde getter from SearchShardTarget * Rename SearchShardTarget#nodeIdText getter to getNodeIdText * Remove unused InternalSearchHit#internalSourceRef unused method * Remove unused InternalSearchHit#internalHighlightFields unused method * Make SearchShardTarget members final	2017-01-11 10:08:31 +01:00
Simon Willnauer	6d2d878068	Merge branch 'master' into feature/multi_cluster_search	2017-01-11 09:28:00 +01:00
Simon Willnauer	acb7f7851d	Allow affix settings to be dynamic / updatable (#22526 ) Today affix settings are not dynamic since it's required to know it's namespace in order to pull a concrete setting from it. This is not possible in practice since the namespaces are dynamic by design. This change allows to register a specialized settings consumer that consumes the namespace and the actual value if a setting gets updated.	2017-01-11 09:24:47 +01:00
Nik Everett	b71b8acf59	Remove ClusterService from ctors in reindex (#22539 ) Moves fetching the local node id into `NodeClient` which is a fairly useful place to put it so you can generate task ids from `NodeClient#executeLocally`.	2017-01-10 18:26:06 -05:00
Luca Cavanna	ddb93946aa	use ElasticsearchException#renderException in BytesRestResponse#convert (#22531 )	2017-01-10 20:58:31 +01:00
Tanguy Leroux	2dcb05fca8	Add fromxcontent methods to index response (#22229 ) This commit adds the parsing fromXContent() methods to the IndexResponse class. The method is based on a ObjectParser because it is easier to use when parsing parent abstract classes like DocWriteResponse. It also changes the ReplicationResponse.ShardInfo so that it now implements ToXContentObject. This way, the ShardInfo.fromXContent() method can be used by the IndexResponse's ObjectParser.	2017-01-10 20:25:32 +01:00
Ali Beyad	898f5a1e89	Removes remaining snapshot backwards compatibility logic (#22512 ) Previously, we removed all unneeded backward compatibility logic from the BlobStoreRepository because 6.0 does not need to support 2.x snapshot formats. During the process of removing this backward compatibility logic, some code was leftover that is no longer necessary. This commit removes all the remaining unnecessary backwards compatibility code in BlobStoreRepository.	2017-01-10 12:54:29 -06:00
Christoph Büscher	93367fb4a8	[Test] Fix test problem with potentially duplicate keys in rest response highlight section	2017-01-10 18:57:13 +01:00
Nik Everett	d50f96e122	Remove InternalAggregation.Type (#22511 ) It is no longer needed. It used to contain a lot of strings used by serialization but those have since been removed. Now it is just another thing to pass around that we don't really need.	2017-01-10 11:57:19 -05:00
Yannick Welsch	1cbb97d361	Use general cluster state batching mechanism for snapshot state updates (#22528 ) Relates to #14899	2017-01-10 17:54:49 +01:00
Yannick Welsch	8be28aa58a	Don't use Guava as compile dependency	2017-01-10 16:33:18 +01:00
Simon Willnauer	081c1ad416	Allow affix settings to delegate to actual settings (#22523 ) Affix settings are useful to namespace a certain setting. Yet, affix settings must be specialized for their concrete type which causes lot of code duplication. This commit allows to reuse an existing setting with and affix setting as soon as a concrete key is available.	2017-01-10 15:14:55 +01:00
Matt Weber	28273e0a52	Additional Graph Support in Match Query (#22503 ) Make match queries that use phrase prefix or cutoff frequency options graph aware. Closes #22490	2017-01-10 08:41:03 -05:00
Boaz Leskes	9aba49c571	ZenDiscoveryUnitTests should close objects in reverse order of creation One needs to close the higher level objects (like UnicastZenPing) before closing the transport service. The latter can throw assertions w.r.t open connections	2017-01-10 14:11:27 +01:00
Christoph Büscher	5f9dfe3186	Add parsing from xContent to InternalSearchHit and InternalSearchHits (#22429 ) This adds methods to parse InternalSearchHit and InternalSearchHits from their xContent representation. Most of the information in the original object is preserved when rendering the object to xContent and then parsing it back. However, some pieces of information are lost which we currently cannot parse back from the rest response, most notably: * the "match" property of the lucene explanation is not rendered in the "_explain" section and cannot be reconstructed on the client side * the original "shard" information (SearchShardTarget) is only rendered if the "explanation" is also set, also we loose the indexUUID of the contained ShardId because we don't write it out. As a replacement we can use ClusterState.UNKNOWN_UUID on the receiving side	2017-01-10 14:00:04 +01:00
Yannick Welsch	9fc1a735cc	Keep NodeConnectionsService in sync with current nodes in the cluster state (#22509 ) The NodeConnectionsService currently determines which nodes to connect to / disconnect from by inspecting cluster state changes and connecting to added nodes / disconnecting from removed nodes. When a master steps down (for example due to another master-eligible node shutting down which brings the number of master-eligible nodes below minimum_master_master), and the connection to other existing nodes was dropped while pinging, however, the connection to these nodes is not re-established while publishing the first cluster state that establishes the node as master. This commit changes the NodeConnectionsService connect / disconnect logic to always rely on the state that is to be / was published, looking not only at the added / removed nodes, but validating that exactly all nodes that are currently registered in NodeConnectionsService are connected (corresponds to a NOOP if the node is already connected).	2017-01-10 13:29:49 +01:00
javanna	fc815ddb75	fix generics warning	2017-01-10 13:21:03 +01:00
Tanguy Leroux	b9061c1cc9	[TESTS] Fix GetResultTests.testGetSourceAsBytes() (#22524 ) The document in the randomized GetResult can exist with no source (like if the _source was disabled in mappings), that's why the test should not always expect a non null source when the doc exists.	2017-01-10 13:12:00 +01:00
Jim Ferenczi	433c822d4f	Promote longs to doubles when a terms agg mixes decimal and non-decimal numbers (#22449 ) * Promote longs to doubles when a terms agg mixes decimal and non-decimal number This change makes the terms aggregation work when the buckets coming from different indices are a mix of decimal numbers and non-decimal numbers. In this case non-decimal number (longs) are promoted to decimal (double) which can result in a loss of precision for big numbers. Fixes #22232	2017-01-10 11:50:56 +01:00
henakamaMSFT	72ec3d2661	Fixing the error message to report the number of docs correctly for each node (#22515 ) There is a bug in the error message that is thrown if the number of docs differs between the source and target shards when recovering a shard with a syncId. The source and target doc counts are swapped around. Closes #21893	2017-01-10 09:57:57 +01:00
Lee Hinman	692ddac020	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2017-01-09 16:30:58 -07:00
Nik Everett	3fb9254b95	Replace Suggesters with namedObject (#22491 ) Removes another parser registery type thing in favor of `XContentParser#namedObject`.	2017-01-09 16:51:08 -05:00
Lee Hinman	7da6e0fb0f	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2017-01-09 14:22:00 -07:00
Simon Willnauer	7f6c89f9a8	first review round	2017-01-09 21:27:09 +01:00
Simon Willnauer	22438855d3	fix typo	2017-01-09 21:17:50 +01:00
Nik Everett	d623df9372	Remove AggregatorParsers Meant to remove it with `e3f77b4795`. It is no longer used.	2017-01-09 15:13:49 -05:00
Nik Everett	e3f77b4795	Replace AggregatorParsers with namedObject (#22397 ) Removes `AggregatorParsers`, replacing all of its functionality with `XContentParser#namedObject`. This is the third bit of payoff from #22003, one less thing to pass around the entire application.	2017-01-09 13:59:38 -05:00
Lee Hinman	bd65ad5983	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2017-01-09 11:44:14 -07:00
javanna	5ff0298576	fix typo in RemoteClusterConnection javadocs	2017-01-09 18:30:19 +01:00
Nik Everett	f75ef7adfd	Use namedObject to parse AllocationCommands (#22489 ) This removes `AllocationCommandRegistry` entirely and replaces it with `XContentParser#namedObject`, removing another class from guice.	2017-01-09 12:26:57 -05:00
javanna	77141716ba	adjust some javadocs and methods visibility	2017-01-09 18:14:26 +01:00
Boaz Leskes	be0c461b50	UnicastZenPingTests didn't properly wait for pinging rounds to be closed The test ping and waited for the ping results to be returned but since we first return the result and then close temporary connections, assertions are tripped that expects all connections to close by end of test . Closes #22497	2017-01-09 17:42:28 +01:00
javanna	4d51be1257	Adjust RemoteClusterConnection javadocs	2017-01-09 17:36:27 +01:00
javanna	4d47fafd16	Adjust RemoteClusterService javadocs	2017-01-09 17:36:27 +01:00
Simon Willnauer	b0a5212b1e	add missing license, forbidden API suppression and line length	2017-01-09 17:34:18 +01:00
Martijn van Groningen	cb8f8fc9ad	Don't ignore `ignore_unmapped` options on nested, has_child and has_parent queries if the wrapped query gets rewritten.	2017-01-09 17:17:34 +01:00
Simon Willnauer	bb5b1a022e	Fix line length	2017-01-09 17:01:57 +01:00
Jay Modi	b04f8fe159	prevent NPE when calling GetResponse#getSourceAsBytesRef This commit checks for a null BytesReference as the value for `source` in GetResult#sourceRef and simply returns null. Previously this would have resulted in a NPE. While this does seem internal at first glance, it can affect user code as a GetResponse could trigger this when the document is missing. Additionally, the CompressorFactory#uncompressIfNeeded now requires a non-null argument.	2017-01-09 09:19:05 -05:00
Nik Everett	f4884e0726	Replace SearchExtRegistry with namedObject (#22492 ) This is one of the last things in `SearchRequestParsers`.	2017-01-09 08:35:54 -05:00
Simon Willnauer	1ef98ede17	Merge branch 'master' into feature/multi_cluster_search	2017-01-09 12:09:23 +01:00
Yannick Welsch	8741691511	Fix primary relocation for shadow replicas (#22474 ) The recovery process started during primary relocation of shadow replicas accesses the engine on the source shard after it's been closed, which results in the source shard failing itself.	2017-01-08 12:18:52 +01:00
Nik Everett	12923ef896	Close and flush refresh listeners on shard close Right now closing a shard looks like it strands refresh listeners, causing tests like `delete/50_refresh/refresh=wait_for waits until changes are visible in search` to fail. Here is a build that fails: https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+multi_cluster_search+multijob-darwin-compatibility/4/console This attempts to fix the problem by implements `Closeable` on `RefreshListeners` and rejecting listeners when closed. More importantly the act of closing the instance flushes all pending listeners so we shouldn't have any stranded listeners on close. Because it was needed for testing, this also adds the number of pending listeners to the `CommonStats` object and all API to which that flows: `_cat/nodes`, `_cat/indices`, `_cat/shards`, and `_nodes/stats`.	2017-01-06 20:03:32 -05:00
Ali Beyad	b0c009ae76	Gracefully handles pre 2.x compressed snapshots In pre 2.x versions, if the repository was set to compress snapshots, then snapshots would be compressed with the LZF algorithm. In 5.x, Elasticsearch no longer supports the LZF compression algorithm. This presents an issue when retrieving snapshots in a repository or upgrading repository data to the 5.x version, because Elasticsearch throws an exception when it tries to read the snapshot metadata because it was compressed using LZF. This commit gracefully handles the situation by introducing a new incompatible-snapshots blob to the repository. For any pre-2.x snapshot that cannot be read, that snapshot is removed from the list of active snapshots, because the snapshot could not be restored anyway. Instead, the snapshot is recorded in the incompatible-snapshots blob. When listing snapshots, both active snapshots and incompatible snapshots will be listed, with incompatible snapshots showing a `INCOMPATIBLE` state. Any attempt to restore an incompatible snapshot will result in an exception.	2017-01-06 17:52:10 -05:00
javanna	ded694fc83	Make StatusToXContent extend ToXContentObject and rename it to StatusToXContentObject This also allows to make RestToXContentListener require ToXContentObject rather than ToXContent	2017-01-06 23:31:48 +01:00
javanna	8edf59c9e7	Make DocWriteResponse a ToXContentObject This involved changing also index, delete, update so that bulk can print them out in its own format.	2017-01-06 23:31:48 +01:00
javanna	d5510701a0	Make SearchResponse a ToXContentObject	2017-01-06 23:31:48 +01:00
javanna	3393d1b409	Make TermVectorsResponse a ToXContentObject	2017-01-06 23:31:48 +01:00
javanna	45d4938fcc	Migrate some more responses to ToXContentObject	2017-01-06 23:31:48 +01:00
javanna	f4aab0138d	introduce ToXContentObject interface `ToXContentObject` extends `ToXContent` without adding new methods to it, while allowing to mark classes that output complete xcontent objects to distinguish them from classes that require starting and ending an anonymous object externally. Ideally ToXContent would be renamed to ToXContentFragment, but that would be a huge change in our codebase, hence we simply document the fact that toXContent outputs fragments with no guarantees that the output is valid per se without an external ancestor. Relates to #16347	2017-01-06 23:31:48 +01:00
Ryan Ernst	cd6e3f4cea	Merge branch 'master' into keystore	2017-01-06 09:32:08 -08:00
Ryan Ernst	42ebfe7bdb	fix NPE	2017-01-06 09:11:07 -08:00
Tim B	b9c2c2f6f0	Move IfConfig.logIfNecessary call into bootstrap (#22455 ) This is related to #22116. A logIfNecessary() call makes a call to NetworkInterface.getInterfaceAddresses() requiring SocketPermission connect privileges. By moving this to bootstrap the logging call can be made before installing the SecurityManager.	2017-01-06 11:10:53 -06:00
Simon Willnauer	79093e1663	Ensure shrunk indices carry over version information from its source (#22469 ) Today when an index is shrunk the version information is not carried over from the source to the target index. This can cause major issues like mapping incompatibilities for instance if an index from a previous major version is shrunk. This commit ensures that all version information from the soruce index is preserved when a shrunk index is created. Closes #22373	2017-01-06 16:36:43 +01:00
Simon Willnauer	56082d1028	add more tests for RemoteClusterService	2017-01-06 12:35:18 +01:00
Simon Willnauer	418ec62bfb	Merge branch 'master' into feature/multi_cluster_search	2017-01-06 10:24:40 +01:00
Ryan Ernst	eb596d7270	more renames	2017-01-06 01:03:45 -08:00
Ryan Ernst	6e406aed2d	addressing more PR comments	2017-01-06 00:49:18 -08:00
javanna	5daf46286e	remove ParseFieldMatcher usages from InternalSearchHit	2017-01-05 19:33:04 +01:00
javanna	975fee402a	remove ParseFieldMatcher usages from suggesters	2017-01-05 19:33:04 +01:00
javanna	13dcb8ccbe	remove ParseFieldMatcher usages from IngestMetadata	2017-01-05 19:33:04 +01:00
javanna	d60e9bddd0	remove ParseFieldMatcher usages from IndexGraveyard	2017-01-05 19:33:04 +01:00
javanna	d87a30647b	remove ParseFieldMatcher usages from SearchAfterBuilder	2017-01-05 19:33:04 +01:00
javanna	723bdc4549	remove ParseFieldMatcher usages from FetchSourceContext	2017-01-05 19:33:04 +01:00
javanna	6102523033	remove ParseFieldMatcher usages from Script parsing code	2017-01-05 19:33:04 +01:00
javanna	6b9a8db069	fix unchecked generics warnings in ObjectParser	2017-01-05 19:33:04 +01:00
javanna	1f7960aa52	ObjectParser to no longer require ParseFieldMatcherSupplier as its Context ParseFieldMatcher as well as ParseFieldMatcherSupplier will be soon removed, hence the ObjectParser's context doesn't need to be a ParseFieldMatcherSupplier anymore. That will allow to remove ParseFieldMatcherSupplier's implementations, little by little.	2017-01-05 19:33:04 +01:00
javanna	9394792392	remove unused ParseFieldMatcher imports/arguments	2017-01-05 19:33:04 +01:00
Yannick Welsch	182e8115de	[TEST] Fix IndexRecoveryIT.testDisconnectsDuringRecovery The test currently checks that the recovering shard is not failed when it is not a primary relocation that has moved past the finalization step. Checking if it has moved past that step is done by intercepting the request between the replication source and the target and checking if it has seen then WAIT_FOR_CLUSTERSTATE action as this is the next action that is called after finalization. This action can, however, occur only after the shard was already failed, and thus trip the assertion. This commit changes the check to look out for the FINALIZE action, independently of whether it succeeded or not.	2017-01-05 19:12:21 +01:00
Yannick Welsch	cfc106d721	Don't close store under CancellableThreads (#22434 ) #22325 changed the recovery retry logic to use unique recovery ids. The change also introduced an issue, however, which made it possible for the shard store to be closed under CancellableThreads, triggering assertions in the node locking logic. This commit limits the use of CancellableThreads only to the part where we wait on the old recovery target to be closed.	2017-01-05 18:11:58 +01:00
Simon Willnauer	349ea0f9b6	cut over to use : instead of \| for cross cluster search	2017-01-05 17:03:12 +01:00
Simon Willnauer	dca54734ac	add basic docs	2017-01-05 16:10:34 +01:00
Simon Willnauer	0183b0c5a8	More cleanups	2017-01-05 15:23:55 +01:00
Simon Willnauer	1ef7115bbd	Add javadocs to TransportActionProxy	2017-01-05 14:13:08 +01:00
Simon Willnauer	7b95c2f54c	document and test concurrent remote cluster node discovery	2017-01-05 14:07:56 +01:00
Adrien Grand	97f3a9bd79	Relax LiveVersionMapTests.testRamBytesUsed. With Java9's new restrictions we cannot compute ram usage as accurately as before. See https://issues.apache.org/jira/browse/LUCENE-7595.	2017-01-05 11:33:24 +01:00
Simon Willnauer	80bf01d3c0	Merge branch 'master' into feature/multi_cluster_search	2017-01-05 08:00:03 +01:00
Simon Willnauer	a5daa5d3a2	Execute low level handshake in #openConnection (#22440 ) Today we execute the low level handshake on the TCP layer in #connectToNode. If #openConnection is used directly, which is truly expert, no handshake is executed which allows connecting to nodes that are not necessarily compatible. This change moves the handshake to #openConnection to prevent bypassing this logic.	2017-01-05 07:32:53 +01:00
Ryan Ernst	bf51522788	Add 5.3 version	2017-01-04 15:00:43 -08:00
Simon Willnauer	daf1f53c39	Make RemoteClusterConnectionIT a unit test	2017-01-04 21:36:53 +01:00
Ali Beyad	9d422c1c34	IndicesService handles all exceptions during index deletion (#22433 ) Previously, we could run into a situation where attempting to delete an index due to a cluster state update would cause an unhandled exception to bubble up to the ClusterService and cause the cluster state applier to fail. The result of this situation is that the cluster state never gets updated on the ClusterService because the exception happens before all cluster state appliers have completed and the ClusterService only updates the cluster state once all cluster state appliers have successfully completed. All other methods on IndicesService properly handle all exceptions and not just IOExceptions, but there were two instances with respect to index deletion where only IOExceptions where handled by the IndicesService. If any other exception occurred during these delete operations, the exception would be bubbled up to the ClusterService, causing the aforementioned issues. This commit ensures all methods in IndicesService properly capture all types of Exceptions, so that the ClusterService manages to update the cluster state, even in the presence of shard creation/deletion failures. Note that the lack of updating the cluster state in the presence of such exceptions can have many unintended consequences, one of them being the tripping of the assertion in IndicesClusterStateService#removeUnallocatedIndices where the assumption is that if there is an IndexService to remove with an unassigned shard, then the index must exist in the cluster state, but if the cluster state was never updated due to the aforementioned exceptions, then the cluster state will not have the index in question.	2017-01-04 13:16:49 -06:00
Adrien Grand	f8998fece5	Upgrade to lucene-6.4.0-snapshot-084f7a0. (#22413 )	2017-01-04 19:03:52 +01:00
Simon Willnauer	e642965804	Cleanup lots of code, add javadocs and tests	2017-01-04 17:26:00 +01:00
Simon Willnauer	dd0331144a	[TEST] Also register replica node in node map	2017-01-04 13:31:59 +01:00
Simon Willnauer	31499a1248	handle nodes that are not connected early in AbstractSearchAsyncAction	2017-01-04 11:23:33 +01:00
Jim Ferenczi	360ce532eb	Implement stats for geo_point and geo_shape field (#22391 ) Currently `geo_point` and `geo_shape` field are treated as `text` field by the field stats API and we try to extract the min/max values with MultiFields.getTerms. This is ok in master because a `geo_point` field is always a Point field but it can cause problem in 5.x (and 2.x) because the legacy `geo_point` are indexed as terms. As a result the min and max are extracted and then printed in the FieldStats output using BytesRef.utf8ToString which can throw an IndexOutOfBoundException since it's not valid UTF8 strings. This change ensure that we never try to extract min/max information from a `geo_point` field. It does not add a new type for geo points in the fieldstats API so we'll continue to use `text` for this kind of field. This PR is targeted to master even though we could only commit this change to 5.x. I think it's cleaner to have it in master too before we make any decision on https://github.com/elastic/elasticsearch/pull/21947. Fixes #22384	2017-01-04 10:42:22 +01:00
Ryan Ernst	4e4a40df7a	feedback	2017-01-03 15:42:38 -08:00
Jason Tedor	c6ddff757e	Cleanup some comments in IndexShard.java This commit cleans up the comments in IndexShard related to sequence numbers, making them uniform in their formatting and taking advantage of the line-length limit of 140 characters.	2017-01-03 15:10:38 -05:00
Jason Tedor	9a65d2008e	Cleanup comments in GlobalCheckpointService.java This commit cleans up the comments in GlobalCheckpointService, making them uniform in their formatting and taking advantage of the line-length limit of 140 characters.	2017-01-03 12:23:33 -05:00
Jason Tedor	64888ab1d3	Cleanup comments in SequenceNumbersService.java This commit cleans up the comments in SequenceNumbersService, making them uniform in their formatting and taking advantage of the line-length limit of 140 characters.	2017-01-03 12:06:43 -05:00
Ali Beyad	6f242920d3	[TEST] only check node decisions if not in the AWAITING_INFO state	2017-01-03 11:39:40 -05:00
Simon Willnauer	422cd1ef77	Add support for proxy nodes this commit adds full support for proxy nodes on the search layer. This allows to connection only to a small set of nodes on a remote cluster to exectue the search. The nodes will proxy the request to the correct node in the cluster while the coordinting node doesn't need to be connected to the target node.	2017-01-03 17:24:32 +01:00
javanna	ee4dde46d3	Remove ParseFieldMatcher usages from Aggregator	2017-01-03 15:52:32 +01:00
javanna	8b8ff8b9e2	Remove ParseFieldMatcher usages from SearchService	2017-01-03 15:52:32 +01:00
javanna	6329a98a97	Remove ParseFieldMatcher usages from SearchContext	2017-01-03 15:52:32 +01:00
javanna	77f4152a18	Remove ParseFieldMatcher usages from a couple of Rest Actions	2017-01-03 15:52:32 +01:00
javanna	0d67891a64	Remove ParseFieldMatcher usages from QueryParsers#parseRewriteMethod	2017-01-03 15:52:32 +01:00
javanna	40540b3f3f	Remove unused QueryParsers#setRewriteMethod	2017-01-03 15:52:32 +01:00
javanna	648ed46f01	Remove ParseFieldMatcher usages from MoreLikeThisQueryBuilder & MultiMatchQueryBuilder	2017-01-03 15:52:32 +01:00
javanna	c06d00dce1	Remove ParseFieldMatcher usage from SearchRequest	2017-01-03 15:52:32 +01:00
javanna	45c67b5ee5	Remove ParseFieldMatcher usage from AggregatorParsers	2017-01-03 15:52:32 +01:00
javanna	6f4faf5233	Remove ParseFieldMatcher usage from AllocationCommands	2017-01-03 15:52:32 +01:00
javanna	8f297ec42c	Remove ParseFieldMatcher usage from ParseFieldRegistry	2017-01-03 15:52:32 +01:00
javanna	41c7d3e092	Remove ParseFieldMatcher usage from Mappers	2017-01-03 15:52:32 +01:00
Jason Tedor	c5a8fd9719	Cleanup some whitespace in LocalCheckpointService.java This commit just fixes a couple whitespace formatting issues in o/e/i/s/LocalCheckpointService.java.	2017-01-03 09:30:31 -05:00
Jason Tedor	f086d1d3db	Cleanup comments in LocalCheckpointService.java This commit cleans up the comments in LocalCheckpointService, making them uniform in their formatting and taking advantage of the line-length limit of 140 characters.	2017-01-03 09:29:01 -05:00
Christoph Büscher	a773d46c69	Remove deprecated `minimum_number_should_match` in BoolQueryBuilder After deprecating getters and setters and the query DSL parameter in 5.x, support for `minimum_number_should_match` can be removed entirely. Also consolidated comments with the ones on 5.x branch and added an entry to the migration docs.	2017-01-03 15:14:33 +01:00
Daniel Mitterdorfer	1ed64f0551	Eliminate unneccessary declaration of IOException With this commit we remove the declaration of IOException from assertWarnings and modify all call sites. Checked with @javanna	2017-01-03 12:36:28 +01:00
Christoph Büscher	16d79842ac	Remove getters and setters for "minimumNumberShouldMatch" in BoolQueryBuilder Currently we have getters an setters for both "minimumNumberShouldMatch" and "minimumShouldMatch", which both access the same internal value (minimumShouldMatch). Since we only document the `minimum_should_match` parameter for the query DSL, I think we can deprecate the other getters and setters for 5.x and remove with 6.0, also deprecating the `minimum_number_should_match` query DSL parameter.	2017-01-03 11:29:04 +01:00
Simon Willnauer	306405fd1b	Merge branch 'master' into feature/multi_cluster_search	2017-01-03 11:17:54 +01:00
Tim Vernum	6ad5486e6b	Implement Comparable in Version (#22378 ) Supports using streams to calculate min/max of a collection of Versions, etc.	2017-01-03 12:20:17 +11:00
Ali Beyad	38427c1df0	[TEST] don't wait for all cluster info in the explain API, just assert an upper and lower bound	2017-01-02 18:31:17 -05:00
Ali Beyad	49298c16a9	[TEST] fix explain API awaiting info explanation check	2017-01-02 18:18:09 -05:00
Ali Beyad	47907b7093	[TEST] fix explain API test to allow for either awaiting info state or no valid shard copy	2017-01-02 15:24:01 -05:00
Ali Beyad	20ab4be59f	Cluster Explain API uses the allocation process to explain shard allocation decisions (#22182 ) This PR completes the refactoring of the cluster allocation explain API and improves it in the following two high-level ways: 1. The explain API now uses the same allocators that the AllocationService uses to make shard allocation decisions. Prior to this PR, the explain API would run the deciders against each node for the shard in question, but this was not executed on the same code path as the allocators, and many of the scenarios in shard allocation were not captured due to not executing through the same code paths as the allocators. 2. The APIs have changed, both on the Java and JSON level, to accurately capture the decisions made by the system. The APIs also now report on shard moving and rebalancing decisions, whereas the previous API did not report decisions for moving shards which cannot remain on their current node or rebalancing shards to form a more balanced cluster. Note: this change affects plugin developers who may have a custom implementation of the ShardsAllocator interface. The method weighShards has been removed and no longer has any utility. In order to support the new explain API, however, a custom implementation of ShardsAllocator must now implement ShardAllocationDecision decideShardAllocation(ShardRouting shard, RoutingAllocation allocation) which provides a decision and explanation for allocating a single shard. For implementations that do not support explaining a single shard allocation via the cluster allocation explain API, this method can simply return an UnsupportedOperationException.	2017-01-02 12:28:32 -06:00
javanna	a3918ad094	Remove unused ParseFieldMatcher#match method	2016-12-31 09:24:44 +01:00
javanna	cd6b569286	Remove some usages of ParseFieldMatcher in favour of using ParseField directly Relates to #19552 Relates to #22130	2016-12-31 09:24:44 +01:00
Igor Motov	f985638bba	Add a generic way of checking version before serializing custom cluster object In #22313 we added a check that prevents the SnapshotDeletionsInProgress custom cluster state objects from being sent to older elasticsearch nodes. This commits make this check generic and available to other cluster state custom objects if needed.	2016-12-30 14:27:09 -05:00
javanna	74acffaae9	fix compiler warning on access to static field using `this`	2016-12-30 18:57:47 +01:00
javanna	df2acb3d9d	Remove some more usages of ParseFieldMatcher in favour of using ParseField directly Relates to #19552 Relates to #22130	2016-12-30 18:57:47 +01:00
javanna	6c54cbade4	Remove some more usages of ParseFieldMatcher in favour of using ParseField directly Relates to #19552 Relates to #22130	2016-12-30 18:57:47 +01:00
javanna	45d010e874	Remove some usages of ParseFieldMatcher in favour of using ParseField directly Relates to #19552 Relates to #22130	2016-12-30 18:57:47 +01:00
Adrien Grand	00de5b83bd	The percentage of deleted docs needs to be strictly over 10% for deleted docs to be expunged.	2016-12-30 11:18:02 +01:00
Adrien Grand	f1d7721932	Fix TermsAggregatorTests to not use LuceneTestCase.newSearcher since it needs a DirectoryReader.	2016-12-30 10:12:24 +01:00
Adrien Grand	f89bb18a5d	Dynamic `date` fields should use the `format` that was used to detect it is a date. (#22174 ) Unless the dynamic templates define an explicit format in the mapping definition: in that case the explicit mapping should have precedence. Closes #9410	2016-12-30 09:48:24 +01:00
Adrien Grand	3f805d68cb	Add the ability to set an analyzer on keyword fields. (#21919 ) This adds a new `normalizer` property to `keyword` fields that pre-processes the field value prior to indexing, but without altering the `_source`. Note that only the normalization components that work on a per-character basis are applied, so for instance stemming filters will be ignored while lowercasing or ascii folding will be applied. Closes #18064	2016-12-30 09:36:10 +01:00
Yannick Welsch	816e1c6cc4	Free shard resources when recovery reset is cancelled Resetting a recovery consists of resetting the old recovery target and replacing it by a new recovery target object. This is done on the Cancellable threads of the new recovery target. If the new recovery target is already cancelled before or while this happens, for example due to shard closing or recovery source changing, we have to make sure that the old recovery target object frees all shard resources. Relates to #22325	2016-12-29 17:05:47 +01:00
Yannick Welsch	6e6d9eb255	Use a fresh recovery id when retrying recoveries (#22325 ) Recoveries are tracked on the target node using RecoveryTarget objects that are kept in a RecoveriesCollection. Each recovery has a unique id that is communicated from the recovery target to the source so that it can call back to the target and execute actions using the right recovery context. In case of a network disconnect, recoveries are retried. At the moment, the same recovery id is reused for the restarted recovery. This can lead to confusion though if the disconnect is unilateral and the recovery source continues with the recovery process. If the target reuses the same recovery id while doing a second attempt, there might be two concurrent recoveries running on the source for the same target. This commit changes the recovery retry process to use a fresh recovery id. It also waits for the first recovery attempt to be fully finished (all resources locally freed) to further prevent concurrent access to the shard. Finally, in case of primary relocation, it also fails a second recovery attempt if the first attempt moved past the finalization step, as the relocation source can then be moved to RELOCATED state and start indexing as primary into the target shard (see TransportReplicationAction). Resetting the target shard in this state could mean that indexing is halted until the recovery retry attempt is completed and could also destroy existing documents indexed and acknowledged before the reset. Relates to #22043	2016-12-29 10:58:15 +01:00
Igor Motov	ca90d9ea82	Remove PROTO-based custom cluster state components Switches custom cluster state components from PROTO-based de-serialization to named objects based de-serialization	2016-12-28 13:32:35 -05:00
Jim Ferenczi	e7444f7d77	Fix scaled_float numeric type in aggregations (#22351 ) `scaled_float` should be used as DOUBLE in aggregations but currently they are used as LONG. This change fixes this issue and adds a simple it test for it. Fixes #22350	2016-12-27 09:23:22 +01:00
Adrien Grand	3cb164b22e	Fix IndexShardTests.testDocStats.	2016-12-26 20:19:11 +01:00
Adrien Grand	2127db27a3	Add trace logging to CircuitBreakerServiceIT.testParentChecking.	2016-12-26 16:05:27 +01:00
Adrien Grand	2d81750a13	Make ESTestCase resilient to initialization errors.	2016-12-26 14:55:22 +01:00
Adrien Grand	f80165c374	Fix LineLength issues.	2016-12-26 11:22:09 +01:00
Adrien Grand	d89757b848	Fix mutate function to always actually modify the failure object.	2016-12-26 10:34:50 +01:00
Ali Beyad	1cb5dc42ff	Updates SnapshotDeletionsInProgress version number introduced to 5.2.0	2016-12-25 19:28:01 -05:00
Ali Beyad	8261bd358a	Synchronize snapshot deletions on the cluster state (#22313 ) Before, snapshot/restore would synchronize all operations on the cluster state except for deleting snapshots. This meant that only one snapshot/restore operation would be allowed in the cluster at any given time, except for deletions - there could be two or more snapshot deletions running at the same time, or a deletion could be running, unbeknowest to the rest of the cluster, and thus a snapshot or restore would be allowed at the same time as the snapshot deletion was still in progress. This could cause any number of synchronization issues, including the situation where a snapshot that was deleted could reappear in the index-N file, even though its data was no longer present in the repository. This commit introduces a new custom type to the cluster state to represent deletions in progress. Now, another deletion cannot start if a deletion is currently in progress. Similarily, a snapshot or restore cannot be started if a deletion is currently in progress. In each case, if attempting to run another snapshot/restore operation while a deletion is in progress, a ConcurrentSnapshotExecutionException will be thrown. This is the same exception thrown if trying to snapshot while another snapshot is in progress, or restore while a snapshot is in progress. Closes #19957	2016-12-25 19:00:20 -05:00
Jason Tedor	d5c18bf5c9	Fix doc stats test when deleting all docs This commit fixes an issue with IndexShardTests#testDocStats when the number of deleted docs is equal to the number of docs. In this case, Luence will remove the underlying segment tripping an assertion on the number of deleted docs.	2016-12-23 15:20:42 -05:00
Ryan Ernst	d4288cce79	Fix getBytes invocation to use explicit charset	2016-12-23 10:57:02 -08:00
Jason Tedor	6deb5283db	Fix delete op serialization format constant The delete op serizliation format constant for 5.x was off by one. This commit fixes this, and cleans up the handling of these formats.	2016-12-23 11:50:43 -05:00
Jason Tedor	2713549533	Use reader for doc stats Today we try to pull stats from index writer but we do not get a consistent view of stats. Under heavy indexing, this inconsistency can be very skewed indeed. In particular, it can lead to the number of deleted docs being reported as negative and this leads to serialization issues. Instead, we should provide a consistent view of the stats by using an index reader. Relates #22317	2016-12-23 09:44:56 -05:00
Boaz Leskes	c2baa5f213	TransportService should capture listener before spawning background notification task Not doing this made it difficult to establish a happens before relationship between connecting to a node and adding a listeners. Causing test code like this to fail sproadically: ``` // connection to reuse handleA.transportService.connectToNode(handleB.node); // install a listener to check that no new connections are made handleA.transportService.addConnectionListener(new TransportConnectionListener() { @Override public void onConnectionOpened(DiscoveryNode node) { fail("should not open any connections. got [" + node + "]"); } }); ``` relates to #22277	2016-12-23 13:55:12 +01:00
Yannick Welsch	baea17b53f	Separate cluster update tasks that are published from those that are not (#21912 ) This commit factors out the cluster state update tasks that are published (ClusterStateUpdateTask) from those that are not (LocalClusterUpdateTask), serving as a basis for future refactorings to separate the publishing mechanism out of ClusterService.	2016-12-23 12:23:52 +01:00
Boaz Leskes	eb7450bcdc	UnicastZenPing add trace logging on connection opening	2016-12-23 09:15:11 +01:00
Jason Tedor	faaa671fb6	Enable assertions in integration tests When starting a standalone cluster, we do not able assertions. This is problematic because it means that we miss opportunities to catch bugs. This commit enables assertions for standalone integration tests, and fixes a couple bugs that were uncovered by enabling these. Relates #22334	2016-12-22 20:08:02 -05:00
Ryan Ernst	fb690ef748	Settings: Add infrastructure for elasticsearch keystore This change is the first towards providing the ability to store sensitive settings in elasticsearch. It adds the `elasticsearch-keystore` tool, which allows managing a java keystore. The keystore is loaded upon node startup in Elasticsearch, and used by the Setting infrastructure when a setting is configured as secure. There are a lot of caveats to this PR. The most important is it only provides the tool and setting infrastructure for secure strings. It does not yet provide for keystore passwords, keypairs, certificates, or even convert any existing string settings to secure string settings. Those will all come in follow up PRs. But this PR was already too big, so this at least gets a basic version of the infrastructure in. The two main things to look at. The first is the `SecureSetting` class, which extends `Setting`, but removes the assumption for the raw value of the setting to be a string. SecureSetting provides, for now, a single helper, `stringSetting()` to create a SecureSetting which will return a SecureString (which is like String, but is closeable, so that the underlying character array can be cleared). The second is the `KeyStoreWrapper` class, which wraps the java `KeyStore` to provide a simpler api (we do not need the entire keystore api) and also extend the serialized format to add metadata needed for loading the keystore with no assumptions about keystore type (so that we can change this in the future) as well as whether the keystore has a password (so that we can know whether prompting is necessary when we add support for keystore passwords).	2016-12-22 16:28:34 -08:00
Nik Everett	55099df1cb	Support negative numbers in writeVLong (#22314 ) We don't want to use negative numbers with `writeVLong` so throw an exception when we try. On the other hand unforeseen bugs might cause us to write negative numbers (some versions of Elasticsearch don't have the exception, only an assertion) so this fixes `readVLong` so that instead of reading a wrong value and corrupting the stream it reads the negative value.	2016-12-22 13:02:39 -05:00
Boaz Leskes	13c5881f3e	UnicastZenPing's PingingRound should prevent opening connections after being closed This may cause them to leak. Provisioning for it was made in #22277 but sadly a crucial ensureOpen call was forgotten	2016-12-22 18:45:44 +01:00
Boaz Leskes	7d0dbd2082	add trace logging to UnicastZenPingTests.testResolveReuseExistingNodeConnections	2016-12-22 18:10:15 +01:00
Tal Levy	6d7261c4d3	Adds ingest processor headers to exception for unknown processor. (#22315 ) Optimistically check for `tag` of an unknown processor for better tracking of which processor declaration is to blame in an invalid configuration. Closes #21429.	2016-12-22 08:24:00 -08:00
Nik Everett	f5f2149ff2	Remove much ceremony from parsing client yaml test suites (#22311 ) * Remove a checked exception, replacing it with `ParsingException`. * Remove all Parser classes for the yaml sections, replacing them with static methods. * Remove `ClientYamlTestFragmentParser`. Isn't used any more. * Remove `ClientYamlTestSuiteParseContext`, replacing it with some static utility methods. I did not rewrite the parsers using `ObjectParser` because I don't think it is worth it right now.	2016-12-22 11:00:34 -05:00
Stéphane Campinas	e1b8528ab8	Support numeric bounds with decimal parts for long/integer/short/byte datatypes (#21972 ) Close #21600	2016-12-22 15:20:15 +01:00
Martijn van Groningen	b9a90eca20	inner hits: Don't inline inner hits if the query the inner hits is inlined into can't resolve mappings and ignore_unmapped has been set to true Closes #21620	2016-12-22 14:51:33 +01:00
Colin Goodheart-Smithe	9a73a2efb3	Fix stackoverflow error on InternalNumericMetricAggregation	2016-12-22 13:50:54 +00:00
Adrien Grand	9b3b693d15	Date detection should not rely on a hardcoded set of characters. (#22171 ) Currently we only apply date detection on strings that contain either `:`, `-` or `/`. This commit inverses the heuristic in order to only apply date detection on strings that are not parseable as a number, so that more date formats can be used as dynamic dates formats. Closes #1694	2016-12-22 14:35:59 +01:00
Adrien Grand	e39942fc02	`value_type` is useful regardless of scripting. (#22160 ) Today we only expose `value_type` in scriptable aggregations, however it is also useful with unmapped fields. I suspect we never noticed because `value_type` was not documented (fixed) and most aggregations are scriptable. Closes #20163	2016-12-22 14:35:12 +01:00
Adrien Grand	fd6e1a30de	Improve concurrency of ShardCoreKeyMap. (#22316 ) `ShardCoreKeyMap.add` is called on each segment for all search requests, which means it might become a bottleneck under a cocurrent load of cheap search requests since this method acquires a mutex. This change proposes to use a `ConcurrentHashMap` which allows to only take the mutex in the case that the `LeafReader` has never been seen before.	2016-12-22 14:34:08 +01:00
Martijn van Groningen	acd64c6ee1	fixed jdocs and removed already fixed norelease	2016-12-22 14:18:30 +01:00
Colin Goodheart-Smithe	06576ed13b	Adds abstract test classes for serialisation (#22281 ) This adds test classes that can be used to test the wire serialisation and (optionally) the XContent serialisation of objects that implement Streamable/Writeable and ToXContent. These test classes will enable classes sich as InternalAggregation (or at least its implementations) to be tested in a consistent way when is comes to testing serialisation.	2016-12-22 10:49:18 +00:00
Areek Zillur	d51f414ea3	Add ingest tests for single item bulk action	2016-12-22 01:47:59 -05:00
Areek Zillur	1c91719a88	make index and delete requests composite	2016-12-22 01:46:56 -05:00
Jason Tedor	7946396fe6	Introduce translog no-op As the translog evolves towards a full operations log as part of the sequence numbers push, there is a need for the translog to be able to represent operations for which a sequence number was assigned, but the operation did not mutate the index. Examples of how this can arise are operations that fail after the sequence number is assigned, and gaps in this history that arise when an operation is assigned a sequence number but the operation never completed (e.g., a node crash). It is important that these operations appear in the history so that they can be replicated and replayed during recovery as otherwise the history will be incomplete and local checkpoints will not be able to advance. This commit introduces a no-op to the translog to set the stage for these efforts. Relates #22291	2016-12-21 23:08:16 -05:00
Jason Tedor	91cb563247	Provide helpful error message if a plugin exists Today if an older version of a plugin exists, we fail to notify the user with a helpful error message. This happens because during plugin verification, we attempt to read the plugin descriptors for all existing plugins. When an older version of a plugin is sitting on disk, we will attempt to read this old plugin descriptor and fail due to a version mismatch. This leads to an unhelpful error message. Instead, we should check for existence of the plugin as part of the verification phase, but before attempting to read plugin descriptors for existing plugins. This enables us to provide a helpful error message to the user. Relates #22305	2016-12-21 22:37:07 -05:00
Areek Zillur	5b2393d8b4	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2016-12-21 16:33:30 -05:00
Simon Willnauer	d89b3ba2cf	add current WIP	2016-12-21 21:09:21 +01:00
Simon Willnauer	20e9e8b560	Add TransportActionProxy infrastructure	2016-12-21 20:52:42 +01:00
Boaz Leskes	0eaaee160b	better code sharding?	2016-12-21 14:24:18 -05:00
Nik Everett	8aca504c86	Clear static variable after suite This was causing test failures: https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+master+java9-periodic/1101/console https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+master+dockeralpine-periodic/513/consoleFull	2016-12-21 13:39:28 -05:00
Luca Cavanna	5c8232c03b	Restore deprecation warning for invalid match_mapping_type values (#22304 ) The deprecation warning gives now the same message as 5.x. The deprecation warning was previously removed, but given that we are still lenient with old indices we should still output the warning.	2016-12-21 16:56:55 +01:00
Adrien Grand	84edf36f11	Make `-0` compare less than `+0` consistently. (#22173 ) Our `float`/`double` fields generally assume that `-0` compares less than `+0`, except when bounds are exclusive: an exclusive lower bound on `-0` excludes `+0` and an exclusive upper bound on `+0` excludes `-0`. Closes #22167	2016-12-21 16:51:45 +01:00
Adrien Grand	3b3f9216db	Allow terms aggregations on pure boolean scripts. (#22201 ) The way aggregations on scripts work is by hiding scripts behind the same API that we use for regular fields. However, there is no native support for boolean fields, those need to be exposed as integers, with `0` standing for `false` and `1` for true. Relates #20941	2016-12-21 16:48:53 +01:00
Boaz Leskes	0e9186e137	Simplify Unicast Zen Ping (#22277 ) The `UnicastZenPing` shows it's age and is the result of many small changes. The current state of affairs is confusing and is hard to reason about. This PR cleans it up (while following the same original intentions). Highlights of the changes are: 1) Clear 3 round flow - no interleaving of scheduling. 2) The previous implementation did a best effort attempt to wait for ongoing pings to be sent and completed. The pings were guaranteed to complete because each used the total ping duration as a timeout. This did make it hard to reason about the total ping duration and the flow of the code. All of this is removed now and ping should just complete within the given duration or not be counted (note that it was very handy for testing, but I move the needed sync logic to the test). 3) Because of (2) the pinging scheduling changed a bit, to give a chance for the last round to complete. We now ping at the beginning, 1/3 and 2/3 of the duration. 4) To offset for (3) a bit, incoming ping requests are now added to on going ping collections. 5) UnicastZenPing never establishes full blown connections (but does reuse them if there). Relates to #22120 6) Discovery host providers are only used once per pinging round. Closes #21739 7) Usage of the ability to open a connection without connecting to a node ( #22194 ) and shorter connection timeouts helps with connections piling up. Closes #19370 8) Beefed up testing and sped them up. 9) removed light profile from production code	2016-12-21 15:09:58 +01:00
Nik Everett	567c65b0d5	Replace IndicesQueriesRegistry (#22289 ) * Switch query parsing to namedObject * Remove IndicesQueriesRegistry	2016-12-21 09:05:14 -05:00
Christoph Büscher	bdecbb529f	Factor out sort values from InternalSearchHit (#22080 ) This adds fromXContent method and unit test for sort values that are part of InternalSearchHit. In order to centralize serialisation and xContent parsing and rendering code, move all relevant parts to a new class which can be unit tested much better in isolation.This is part of the preparation for parsing search responses on the client side.	2016-12-21 11:19:47 +01:00
Boaz Leskes	e298180a39	IndicesStoreIntegrationIT should not use start recovery sending as an indication that the recovery started Sending a request is not a good indicator as it doesn't mean it's processed yet. Instead we should use one of the first request from source to target. This caused the cluster state block to be added to early , blocking the recovery it self	2016-12-21 10:11:56 +01:00
Simon Willnauer	dce24b5a10	make connection to nodes async and ensure that if we are not fully connected a search will fork or a reconnect	2016-12-21 10:00:04 +01:00
Martijn van Groningen	417746ca9e	Added base class for testing aggregators and some initial tests for `terms`, `top_hits` and `min` aggregations.	2016-12-21 08:44:05 +01:00
Areek Zillur	180ceef134	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2016-12-21 01:04:31 -05:00
Areek Zillur	de44584f84	incorporate feedback	2016-12-21 00:26:58 -05:00
Simon Willnauer	3625d64b7f	Add remote cluster connections manager	2016-12-20 22:23:11 +01:00
Tal Levy	5a90d9d7e6	add `ignore_missing` flag to ingest plugins (#22273 ) added `ignore_missing` flag to: - Attachment Processor - GeoIP Processor - User-Agent Processor	2016-12-20 10:53:28 -08:00
Ali Beyad	ad4405f244	Adds setting level to allocation decider explanations (#22268 ) The allocation decider explanation messages where improved in #21771 to include the specific Elasticsearch setting that contributed to the decision taken by the decider. This commit improves upon the explanation message output by including whether the setting was an index level setting or a cluster level setting. This will further help the user understand and locate the setting that is the cause of shards remaining unassigned or remaining on their current node.	2016-12-20 12:25:52 -05:00
Nik Everett	a04dcfb95b	Introduce XContentParser#namedObject (#22003 ) Introduces `XContentParser#namedObject which works a little like `StreamInput#readNamedWriteable`: on startup components register parsers under names and a superclass. At runtime we look up the parser and call it to parse the object. Right now the parsers take a context object they use to help with the parsing but I hope to be able to eliminate the need for this context as most what it is used for at this point is to move around parser registries which should be replaced by this method eventually. I make no effort to do so in this PR because it is big enough already. This is meant to the a start down a road that allows us to remove classes like `QueryParseContext`, `AggregatorParsers`, `IndicesQueriesRegistry`, and `ParseFieldRegistry`. The goal here is to reduce the amount of plumbing required to allow parsing pluggable things. With this you don't have to pass registries all over the place. Instead you must pass a super registry to fewer places and use it to wrap the reader. This is the same tradeoff that we use for NamedWriteable and it allows much, much simpler binary serialization. We think we want that same thing for xcontent serialization. The only parsing actually converted to this method is parsing `ScoreFunctions` inside of `FunctionScoreQuery`. I chose this because it is relatively self contained.	2016-12-20 11:05:24 -05:00
Yannick Welsch	710031d92f	Let ClusterStateObserver only hold onto state that's needed for change detection (#21631 ) ClusterStateObserver is a utility class that simplifies interacting with the cluster state in cases where an action takes a decision based on the current cluster state but may want to wait for a new state and retry upon failure. The ClusterStateObserver implements its functionality by keeping a reference to the last cluster state that it observed. When a new ClusterStateObserver is created, it samples a cluster state from the cluster service which is subsequently used for change detection. If actions take a long time to process, however, the cluster observer can reference very old cluster states. Due to cluster observers being created very frequently and cluster states being potentially large the referenced cluster states can waste a lot of heap space. A specific example where this can make a node go out of memory is given in point 2 of issue #21568: The action listener in TransportMasterNodeAction.AsyncSingleAction has a ClusterStateObserver to coordinate the retry mechanism if the action on the master node fails due to the node not being master anymore. The ClusterStateObserver in AsyncSingleAction keeps a reference to the full cluster state when the action was initiated. If the pending tasks queue grows quite large and has older items in it lots of cluster states can possibly be referenced. This commit changes the ClusterStateObserver to hold only onto the part of the cluster state that's needed for change detection.	2016-12-20 15:16:04 +01:00
Simon Willnauer	3515f782d1	fix compile issue	2016-12-20 15:10:58 +01:00
Christoph Büscher	bc22c86d14	SuggestionBuilder doesn't need to extend ToXContentToBytes This changes the class from extending the abstract class to implementing the ToXContent interface only. The former could lead to unexpected behaviour when trying to display the object, since the "toString()" method inherited from ToXContentToBytes would create an error message because the SuggestionBuilders toXContent() methods don't render complete json objects.	2016-12-20 14:57:28 +01:00
Simon Willnauer	7be3af1123	Merge branch 'master' into feature/multi_cluster_search	2016-12-20 14:32:09 +01:00
Tanguy Leroux	290326e73e	Add fromXContent() methods for ReplicationResponse (#22196 ) This commit adds the parsing fromXContent() methods to the ReplicationResponse.ShardInfo and ReplicationResponse.ShardInfo.Failure classes.	2016-12-20 09:29:11 +01:00
Ryan Ernst	850f51db01	Internal: Refactor SettingCommand into EnvironmentAwareCommand (#22175 ) * Internal: Refactor SettingCommand into EnvironmentAwareCommand This change renames and changes the behavior of SettingCommand to have its primary method take in a fully initialized Environment for elasticsearch instead of just a map of settings. All of the subclasses of SettingCommand already did this at some point, so this just removes duplication.	2016-12-19 15:23:44 -08:00
Nik Everett	e508f2ef6a	Fix java 9 build We removed a cast we needed to appease Java 9. I've recreated it in simpler form and left a comment about why we need it.	2016-12-19 17:34:09 -05:00
Alexander Lin	0ab3cbe3a3	Adds percent-encoding for Location headers (#21057 ) This should cause unicode elements in the location header to be percent-encoded, instead of being left alone. Closes #21016	2016-12-19 15:56:09 -05:00
Nik Everett	40b80ae104	Fix line length	2016-12-19 15:07:14 -05:00
Nik Everett	2e1d152fc0	Sub-fields should not accept `include_in_all` parameter (#21971 ) Fail to update mapping when multifield has `include_in_all`. Closes #21710	2016-12-19 15:07:00 -05:00
Grzegorz Gajos	f6b6e4e376	Added ability to remove pipelines via wildcards (#22149 ) (#22191 ) This commit is adding an ability to remove pipelines with wildcards.	2016-12-19 10:59:59 -08:00
javanna	5dae10db11	[TEST] add warnings check to ESTestCase We are currenlty checking that no deprecation warnings are emitted in our query tests. That can be moved to ESTestCase (disabled in ESIntegTestCase) as it allows us to easily catch where our tests use deprecated features and assert on the expected warnings.	2016-12-19 19:39:56 +01:00
javanna	6a27628f12	Remove support for strict parsing mode We return deprecation warnings as response headers, besides logging them. Strict parsing mode stayed around, but was only used in query tests, though we also introduced checks for deprecation warnings there that don't need strict parsing anymore (see #20993). We can then safely remove support for strict parsing mode. The final goal is to remove the ParseFieldMatcher class, but there are many many users of it. This commit prepares the field for the removal, by deprecating ParseFieldMatcher and making it effectively not needed. Strict parsing is removed from ParseFieldMatcher, and strict parsing is replaced in tests where needed with deprecation warnings checks. Note that the setting to enable strict parsing was never ported to the new settings infra hance it cannot be set in production. It is really only used in our own tests. Relates to #19552	2016-12-19 19:39:56 +01:00
javanna	38914f17ed	[TEST] improve ElasticsearchAssertions#assertEquivalent for ToXContent Rename the method to assertToXContentEquivalent to highlight that it's tailored to ToXContent comparisons. Rather than parsing into a map and replacing byte[] in both those maps, add custom equality assertions that recursively walk maps and lists and call Arrays.equals whenever a byte[] is encountered.	2016-12-19 19:32:50 +01:00
javanna	04d929ff53	add inline comments on GetField binary values parsing	2016-12-19 19:32:50 +01:00
javanna	87d8764a32	[TEST] add unit test for XContentHelper#toXContent method	2016-12-19 17:53:42 +01:00
Luca Cavanna	3421e54a42	Add fromXContent method to GetResponse (#22082 ) Moved field values `toXContent` logic to `GetField` (from `GetResult`), which outputs its own fields, and can also parse them now. Also added `fromXContent` to `GetResult` and `GetResponse`. The start object and end object for `GetResponse` output have been moved to `GetResult#toXContent`, from the corresponding rest action. This makes it possible to have `toXContent` and `fromXContent` completely symmetric, as parsing requires looping till an end object is found which is weird when the corresponding `toXContent` doesn't print that out. This also introduces the foundation for testing retrieval of _source and stored field values.	2016-12-19 17:21:26 +01:00
Areek Zillur	eb5cc1e241	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2016-12-19 10:07:02 -05:00
Yannick Welsch	63af03a104	Atomic mapping updates across types (#22220 ) This commit makes mapping updates atomic when multiple types in an index are updated. Mappings for an index are now applied in a single atomic operation, which also allows to optimize some of the cross-type updates and checks.	2016-12-19 14:39:50 +01:00
Yannick Welsch	1cabf66bd5	Use correct block levels for TRA subclasses (#22224 ) Subclasses of TransportReplicationAction can currently chose to implement block levels for which the request will be blocked. - Refresh/Flush was using the block level METADATA_WRITE although they don't operate at the cluster meta data level (but more like shard level meta data which is not represented in the block levels). Their level has been changed to null so that they can operate freely in the presence of blocks. - GlobChkptSync was using WRITE although it does not make any changes to the actual documents of a shard. The level has been changed to null so that it can operate freely in the presence of blocks. The commit also adds a check for closed indices in TRA so that the right exception is thrown if refresh/flush/checkpoint syncing is attempted on a closed index (before it was throwing an IndexNotFoundException, now it's throwing IndexClosedException).	2016-12-19 14:36:58 +01:00
Boaz Leskes	b857b316b6	Add BWC layer to seq no infra and enable BWC tests (#22185 ) Sequence BWC logic consists of two elements: 1) Wire level BWC using stream versions. 2) A changed to the global checkpoint maintenance semantics. For the sequence number infra to work with a mixed version clusters, we have to consider situation where the primary is on an old node and replicas are on new ones (i.e., the replicas will receive operations without seq#) and also the reverse (i.e., the primary sends operations to a replica but the replica can't process the seq# and respond with local checkpoint). An new primary with an old replica is a rare because we do not allow a replica to recover from a new primary. However, it can occur if the old primary failed and a new replica was promoted or during primary relocation where the source primary is treated as a replica until the master starts the target. 1) Old Primary & New Replica - this case is easy as is taken care of by the wire level BWC. All incoming requests will have their seq# set to `UNASSIGNED_SEQ_NO`, which doesn't confuse the local checkpoint logic (keeping it at `NO_OPS_PERFORMED`) 2) New Primary & Old replica - this one is trickier as the global checkpoint service currently takes all in sync replicas into consideration for the global checkpoint calculation. In order to deal with old replicas, we change the semantics to say all new node in sync replicas. That means the replicas on old nodes don't count for the global checkpointing. In this state the seq# infra is not fully operational (you can't search on it, because copies may miss it) but it is maintained on shards that can support it. The old replicas will have to go through a file based recovery at some point and will get the seq# information at that point. There is still an edge case where a new primary fails and an old replica takes over. I'lll discuss this one with @ywelsch as I prefer to avoid it completely. This PR also re-enables the BWC tests which were disabled. As such it had to fix any BWC issue that had crept in. Most notably an issue with the removal of the `timestamp` field in #21670. The commit also includes a fix for the default value of the seq number field in replicated write requests (it was 0 but should be -2), that surface some other minor bugs which are fixed as well. Last - I added some debugging tools like more sane node names and forcing replication request to implement a `toString`	2016-12-19 13:08:24 +01:00
Dimitris Athanasiou	b58bbb9e48	Allow setting aggs after parsing them elsewhere (#22238 ) This commit exposes public getters for the aggregations in AggregatorFactories.Builder. The reason is that it allows to parse the aggregation object from elsewhere (e.g. a plugin) and then be able to get the aggregation builders in order to set them in a SearchSourceBuilder. The alternative would have been to expose a setter for the AggregatorFactories.Builder object. But that would be making the API a bit trappy.	2016-12-19 09:52:07 +00:00
Simon Willnauer	ce5c094cda	Speed up filter and prefix settings operations (#22249 ) Today if a settings object has many keys ie. if somebody specifies a gazillion synonym in-line (arrays are keys ending with ordinals) operations like `Settings#getByPrefix` have a linear runtime. This can cause index creations to be very slow producing lots of garbage at the same time. Yet, `Settings#getByPrefix` is called quite frequently by group settings etc. which can cause heavy load on the system. While it's not recommended to have synonym lists with 25k entries in-line these use-cases should not have such a large impact on the cluster / node. This change introduces a view-like map that filters based on the prefixes referencing the actual source map instead of copying all values over and over again. A benchmark that adds a single key with 25k random synonyms between 2 and 5 chars takes 16 seconds to get the synonym prefix 200 times while the filtered view takes 4 ms for the 200 iterations. This relates to https://discuss.elastic.co/t/200-cpu-elasticsearch-5-index-creation-very-slow-with-a-huge-synonyms-list/69052	2016-12-19 10:48:38 +01:00
Adrien Grand	1ed2e18ded	Fix MapperService.allEnabled(). (#22227 ) It returns whether the last merged mapping has `_all` enabled rather than whether any of the types has `_all` enabled.	2016-12-19 09:55:13 +01:00
Adrien Grand	96f1739c0d	The `_all` default mapper is not completely configured. (#22236 ) In some cases, it might happen that the `_all` field gets a field type that is not totally configured, and in particular lacks analyzers. This is due to the fact that `AllFieldMapper.TypeParser.getDefault` uses `Defaults.FIELD_TYPE` as a default field type, which does not have any analyzers configured since it does not know about the default analyzers.	2016-12-19 09:54:27 +01:00
Daniel Mitterdorfer	3ce7b119d2	Enable strict duplicate checks for all XContent types (#22225 ) With this commit we enable the Jackson feature 'STRICT_DUPLICATE_DETECTION' by default for all XContent types (not only JSON). We have also changed the name of the system property to disable this feature from `es.json.strict_duplicate_detection` to the now more appropriate name `es.xcontent.strict_duplicate_detection`. Relates elastic/elasticsearch#19614 Relates elastic/elasticsearch#22073	2016-12-19 09:29:47 +01:00
Daniel Mitterdorfer	6327e35414	Change type of ingest doc meta-data field 'TIMESTAMP' to `Date` (#22234 ) With this commit we change the data type of the 'TIMESTAMP' meta-data field from a formatted date string to a plain `java.util.Date` instance. The main reason for this change is that our benchmarks have indicated that this contributes significantly to the time spent in the ingest pipeline. The overhead in terms of indexing throughput of the ingest pipeline is about 15% and breaks down roughly as follows: * 5% overhead caused by the conversion from `XContent` -> `Map` * 5% overhead caused by the timestamp formatting * 5% overhead caused by the conversion `Map` -> `XContent` Relates #22074	2016-12-19 09:10:58 +01:00
Simon Willnauer	ccfeac8dd5	Remove `doHandshake` test-only settings from TcpTransport (#22241 ) In #22094 we introduce a test-only setting to simulate transport impls that don't support handshakes. This commit implements the same logic without a setting.	2016-12-18 09:26:53 +01:00
Boaz Leskes	b78f7bc51d	InternalEngine should use global checkpoint when committing the translog relates to #22212	2016-12-18 08:05:59 +01:00
Jason Tedor	58d73bae74	Tighten sequence numbers recovery This commit touches addresses issues related to recovery and sequence numbers: - A sequence number can be assigned and a Lucene commit created with a maximum sequence number at least as large as that sequence number, yet the operation corresponding to that sequence number can be missing from both the Lucene commit and the translog. This means that upon recovery the local checkpoint will be stuck at or below this missing sequence number. To address this, we force the local checkpoint to the maximum sequence number in the Lucene commit when opening the engine. Note that there can still be gaps in the history in the translog but we do not address those here. - The global checkpoint is transferred to the target shard at the end of peer recovery. - Additionally, we reenable the relocation integration tests. Lastly, this work uncovered some bugs in the assignment of sequence numbers on replica operations: - setting the sequence number on replica write requests was missing, very likely introduced as a result of resolving merge conflicts - handling operations that arrive out of order on a replica and have a version conflict with a previous operation were never marked as processed Relates #22212	2016-12-17 09:20:46 -05:00
Simon Willnauer	1f3eb068d5	Add infrastructure to manage network connections outside of Transport/TransportService (#22194 ) Some expert users like UnicastZenPing today establishes real connections to nodes during it's ping phase that can be used by other parts of the system. Yet, this is potentially dangerous and undesirable unless the nodes have been fully verified and should be connected to in the case of a cluster state update or if we join a newly elected master. For use-cases like this, this change adds the infrastructure to manually handle connections that are not publicly available on the node ie. should not be managed by `Transport`/`TransportSerivce`	2016-12-17 11:49:57 +01:00
Simon Willnauer	0b338bf523	Cleanup random stats serialization code (#22223 ) Some of our stats serialization code duplicates complicated seriazliation logic or could use existing building blocks from StreamOutput/Input. This commit cleans up some of the serialization code.	2016-12-17 11:45:55 +01:00
Ryan Ernst	9e5cedae23	Fix line lengths in renamed seccomp file	2016-12-16 22:18:56 -08:00
Jason Tedor	f7d43132b2	Refer to system call filter instead of seccomp Today in the codebase we refer to seccomp everywhere instead of system call filter even if we are not specifically referring to Linux. This commit is a purely mechanical change to refer to system call filter where appropriate instead of the general seccomp, and only leaves seccomp in place when actually referring to the Linux implementation. Relates #22243	2016-12-16 18:30:19 -05:00
Jason Tedor	30806af6bd	Rename bootstrap.seccomp to bootstrap.system_call_filter We try to install a system call filter on various operating systems (Linux, macOS, BSD, Solaris, and Windows) but the setting (bootstrap.seccomp) to control this is named after the Linux implementation (seccomp). This commit replaces this setting with bootstrap.system_call_filter. For backwards compatibility reasons, we fallback to bootstrap.seccomp and log a deprecation message if bootstrap.seccomp is set. We intend to remove this fallback in 6.0.0. Note that now is the time to make this change it's likely that most users are not making this setting anyway as prior to version 5.2.0 (currently unreleased) it was not necessary to configure anything to enable a node to start up if the system call filter failed to install (we marched on anyway) but starting in 5.2.0 it will be necessary in this case. Relates #22226	2016-12-16 18:22:54 -05:00
Luca Cavanna	2265be69d2	Deprecate XContentType auto detection methods in XContentFactory (#22181 ) With recent changes to our parsing code we have drastically reduced the places where we auto-detect the content type from the input. The usage of these methods spread in our codebase for no reason, given that in most of the cases we know the content type upfront and we don't need any auto-detection mechanism. Deprecating these methods is a way to try and make sure that these methods are carefully used, and hopefully not introduced in newly written code. We have yet to fix the REST layer to read the Content-Type header, which is the long term solution, but for now we just want to make sure that the usage of these methods doesn't spread any further. Relates to #19388	2016-12-16 19:33:12 +01:00
Areek Zillur	d44de0cecc	Remove deprecated _suggest endpoint (#22203 ) In #20305, _suggest endpoint was deprecated in favour of using _search endpoint. This commit removes the dedicated _suggest endpoint entirely from master.	2016-12-16 12:06:02 -05:00
Areek Zillur	38060b2096	Merge branch 'master' into enhancement/use_shard_bulk_for_single_ops	2016-12-16 12:03:40 -05:00
Masaru Hasegawa	7cfa6898bf	Merge pull request #22215 from masaruh/skip_empty_boost Don't print empty indices_boost	2016-12-16 17:31:03 +09:00
Simon Willnauer	25b79cd46b	Only notify handshake handler onClose if it can be successfully removed Depending on how the connection is closed the `#onChannelClosed` callback might be invoked more than once or the handler has been processed by the response of the handshake already. This commit only notifies the handler if was removed from the pending map.	2016-12-16 09:06:57 +01:00
Masaru Hasegawa	6fe83fb524	Don't print empty indices_boost	2016-12-16 16:19:34 +09:00
Masaru Hasegawa	a0185c83a7	Merge pull request #21393 from masaruh/alias_boost Resolve index names in indices_boost	2016-12-16 15:07:51 +09:00
Nik Everett	61597f2c20	Send error_trace by default when testing (#22195 ) Sends the `error_trace` parameter with all requests sent by the yaml test framework, including the doc snippet tests. This can be overridden by settings `error_trace: false`. While this drift's core's handling of the yaml tests from the client's slightly this should only be a problem for tests that rely on the default value, both of which I've fixed by setting the value explicitly. This also escapes `\n` and `\t` in the `Stash dump on failure` so the `stack_trace` is more readable. Also fixes `RestUpdateSettingsAction` to not think of the `error_trace` parameter as a setting.	2016-12-15 13:35:14 -05:00
Boaz Leskes	b6cbcc49ba	ClusterService should expose "applied" cluster states (i.e., remove ClusterStateStatus) (#21817 ) `ClusterService` is responsible of updating the cluster state on every node (as a response to an API call on the master and when non-masters receive a new state from the master). When a new cluster state is processed, it is made visible via the `ClusterService#state` method and is sent to series of listeners. Those listeners come in two flavours - one is to change the state of the node in response to the new cluster state (call these cluster state appliers), the other is to start a secondary process. Examples for the later include an indexing operation waiting for a shard to be started or a master node action waiting for a master to be elected. The fact that we expose the state before applying it means that samplers of the cluster state had to worry about two things - working based on a stale CS and working based on a future, i.e., "being applied" CS. The `ClusterStateStatus` was used to allow distinguishing between the two. Working with a stale cluster state is not avoidable. How this PR changes things to make sure consumers don't need to worry about future CS, removing the need for the status and simplifying the waiting logic. This change does come with a price as "cluster state appliers" can't sample the cluster state from `ClusterService` whenever they want as the cluster state isn't exposed yet. However, recent clean ups made this is situation easier and this PR takes the last steps to remove such sampling. This also helps clarify the "information flow" and helps component separation (and thus potential unit testing). It also adds an assertion that will trigger if the cluster state is sampled by such listeners. Note that there are still many "appliers" that could be made a simpler, unrestricted "listener" but this can be done in smaller bits in the future. The commit also makes it clear what the `appliers` and what the `listeners` are by using dedicated interfaces. Also, since I had to change the listener types I went ahead and changed the data structure for temporary/timeout listeners (used for the observer) so addition and removal won't be an O(n) operation.	2016-12-15 17:06:25 +01:00
Tanguy Leroux	391d3a20f3	Add unit tests for toXContent methods in ReplicationResponse (#22188 ) This commit adds unit tests for the toXContent() methods of the inner classes ReplicationResponse.ShardInfo and ReplicationResponse.ShardInfo.Failure.	2016-12-15 16:12:33 +01:00

... 7 8 9 10 11 ...

7905 Commits