OpenSearch

mirror of https://github.com/honeymoose/OpenSearch.git synced 2025-03-01 16:39:11 +00:00

Author	SHA1	Message	Date
Simon Willnauer	302c7a521a	Fix analyzer alias processing (#19506 ) In the lack of tests the analyzer.alias feature was pretty much not working at all on current master. Issues like #19163 showed some serious problems for users using this feature upgrading to an alpha version. This change fixes the processing order and allows aliases to be set for existing analyzers like `default`. This change also ensures that if `default` is aliased the correct analyzer is used for `default_search` etc. Closes #19163	2016-07-21 09:32:47 +02:00
Jun Ohtani	cebad703fe	Analyze: Specify anonymous char_filters/tokenizer/token_filters in the analyze API Add parser for anonymous char_filters/tokenizer/token_filters Using Settings in AnalyzeRequest for anonymous definition Add breaking changes document Closed #8878	2016-07-21 11:06:36 +09:00
Tal Levy	f7cd86ef6d	rethrow script compilation exceptions into ingest configuration exceptions (#19318 ) * rethrow script compilation exceptions into ingest configuration exceptions * update readProcessor to rethrow any exception as an ElasticsearchException	2016-07-20 10:37:56 -07:00
Nik Everett	3a82c613e4	Migrate query registration from push to pull Remove `ParseField` constants used for names where there are no deprecated names and just use the `String` version of the registration method instead. This is step 2 in cleaning up the plugin interface for extending search time actions. Aggregations are next. This is breaking for plugins because those that register a new query should now implement `SearchPlugin` rather than `onModule(SearchModule)`.	2016-07-20 12:33:51 -04:00
Yannick Welsch	2cf94d2d8a	Fix race in testCreateIndexWaitsForAllActiveShards When index creation is not acknowledged (due to a very low request timeout) it is possible that the index is still created. If a subsequent index-exists request completes before the cluster state of the index creation has been fully applied, it might miss the newly created index.	2016-07-20 18:28:28 +02:00
Nik Everett	fc4b439635	Remove AggregationStreams and friends * Remove outdated aggregation registration method * Remove AggregationStreams * Adds StreamInput#readNamedWriteableList and StreamOutput#writeNamedWriteableList convenience methods. We strive to make the reading and writing from the streams terse so they are easier to scan visually. * Remove PipelineAggregatorStreams * Remove stream info from InternalAggreation.Type * Remove InternalAggregation#type * Remove Streamable from PipelineAggregator * Remove Streamable from MultiBucketsAggregation.Bucket	2016-07-20 09:46:04 -04:00
Daniel Mitterdorfer	a4f09d2b81	Restore parameter name auto_generate_phrase_queries (#19514 ) During query refactoring the query string query parameter 'auto_generate_phrase_queries' was accidentally renamed to 'auto_generated_phrase_queries'. With this commit we restore the old name. Closes #19512	2016-07-20 13:13:57 +02:00
Martijn van Groningen	9b1a477120	Fix ClusterInfo serialization	2016-07-20 09:16:27 +02:00
Ryan Ernst	0f2d7a84a8	Add tests for disabling positions and copy the check to text fields	2016-07-19 19:07:56 -07:00
Ryan Ernst	c85cb37cc4	Mappings: Fix not_analyzed string fields to error when position_increment_gap is set Currently if a string field is not_analyzed, but a position_increment_gap is set, it will lookup the default analyzer and set it, along with the position_increment_gap, before the code which handles setting the keyword analyzer for not_analyzed fields has a chance to run. This change adds a parsing check and test for that case.	2016-07-19 17:54:13 -07:00
Jason Tedor	128f0276d9	Fix Javadocs for ThreadPool#schedule This commit fixes an issue with an @throws tag on ThreadPool#schedule not containing a description.	2016-07-19 18:35:30 -04:00
Jason Tedor	770186f6cf	Catch the right rejected execution exception ThreadPool#schedule can throw a rejected execution exception. Yet, the rejected execution exception that it throws comes from the EsAbortPolicy which throws an EsRejectedExecutionException. This exception does not inherit from RejectedExecutionException so instead we must catch the former instead of the latter.	2016-07-19 16:45:12 -04:00
Jason Tedor	720b53b018	Handle rejected execution exception on reschedule A self-rescheduling runnable can hit a rejected execution exception but this exception goes uncaught. Instead, this exception should be caught and passed to the onRejected handler. Not catching handling this rejected execution exception can lead to test failures. Namely, a race condition can arise between the shutting down of the thread pool and cancelling of the rescheduling of the task. If another reschedule fires right as the thread pool is being terminated, the rescheduled task will be rejected leading to an uncaught exception which will cause a test failure. This commit addresses these issues. Relates #19505	2016-07-19 15:35:51 -04:00
Nik Everett	9e2221cae5	Migrate remaining aggregations to NamedWriteable After this we'll be able to remove AggregationStreams and PipelineAggregatorStreams.	2016-07-19 14:43:29 -04:00
jaymode	11389638f9	Require executor name when calling scheduleWithFixedDelay The ThreadPool#scheduleWithFixedDelay method does not make it clear that all scheduled runnable instances will be run on the scheduler thread. This becomes problematic if the actions being performed include blocking operations since there is a single thread and tasks may not get executed due to a blocking task. This change includes a few different aspects around trying to prevent this situation. The first is that the scheduleWithFixedDelay method now requires the name of the executor that should be used to execute the runnable. All existing calls were updated to use Names.SAME to preserve the existing behavior. The second aspect is the removal of using ScheduledThreadPoolExecutor#scheduleWithFixedDelay in favor of a custom runnable, ReschedulingRunnable. This runnable encapsulates the logic to deal with rescheduling a runnable with a fixed delay and mimics the behavior of executing using a ScheduledThreadPoolExecutor and provides a ScheduledFuture implementation that also mimics that of the typed returned by a ScheduledThreadPoolExecutor. Finally, an assertion was added to BaseFuture to detect blocking calls that are being made on the scheduler thread.	2016-07-19 12:47:47 -04:00
Adrien Grand	0854b03f13	Elasticsearch should reject dynamic templates with unknown `match_mapping_type`. #17285 When looking at the logstash template, I noticed that it has definitions for dynamic temilates with `match_mapping_type` equal to `byte` for instance. However elasticsearch never tries to find templates that match the byte type (only long or double as far as numbers are concerned). This commit changes template parsing in order to ignore bad values of `match_mapping_type` (given how the logstash template is popular, this would break many upgrades otherwise). Then I hope to fail the parsing on bad values in 6.0.	2016-07-19 15:38:00 +02:00
Nik Everett	a2a7ea1f17	Make ExtendedBounds immutable We used to mutate it as part of building the aggregation. That caused assertVersionSerializable to fail because it assumes that requests aren't mutated after they are sent. Closes #19481	2016-07-19 08:48:14 -04:00
Yannick Welsch	c4fe8e7bf2	Fix replica-primary inconsistencies when indexing during primary relocation with ongoing replica recoveries (#19287 ) Primary relocation violates two invariants that ensure proper interaction between document replication and peer recoveries, ultimately leading to documents not being properly replicated. Invariant 1: Document writes must be replicated based on the routing table of a cluster state that includes all shards which have ongoing or finished recoveries. This is ensured by the fact that do not start a recovery that is not reflected by the cluster state available on the primary node and we always sample a fresh cluster state before starting to replicate write operations. Invariant 2: Every operation that is not part of the snapshot taken for phase 2, must be succesfully indexed on the target replica (pending shard level errors which will cause the target shard to be failed). To ensure this, we start replicating to the target shard as soon as the recovery start and open it's engine before we take the snapshot. All operations that are indexed after the snapshot was taken are guaranteed to arrive to the shard when it's ready to index them. Note that this also means that the replication doesn't fail a shard if it's not yet ready to recieve operations - it's a normal part of a recovering shard. With primary relocations, the two invariants can be possibly violated. Let's consider a primary relocating while there is another replica shard recovering from the primary shard. Invariant 1 can be violated if the target of the primary relocation is so lagging on cluster state processing that it doesn't even know about the new initializing replica. This is very rare in practice as replica recoveries take time to copy all the index files but it is a theoretical gap that surfaces in testing scenarios. Invariant 2 can be violated even if the target primary knows about the initializing replica. This can happen if the target primary replicates an operation to the intializing shard and that operation arrives to the initializing shard before it opens it's engine but arrives to the primary source after it has taken the snapshot of the translog. Those operations will be currently missed on the new initializing replica. The fix to reestablish invariant 1 is to ensure that the primary relocation target has a cluster state with all replica recoveries that were successfully started on primary relocation source. The fix to reestablish invariant 2 is to check after opening engine on the replica if the primary has been relocated in the meanwhile and fail the recovery. Closes #19248	2016-07-19 14:07:58 +02:00
Simon Willnauer	f79fb4ada7	Create RecoveryTarget once we reset the source RecoveryTarget increments a reference on the store once it's created. If we fail to return the instance from the reset method we leak a reference causing shard locks to not be released. This change creates the reference in the return statement to ensure no references are leaked	2016-07-19 12:27:11 +02:00
Martijn van Groningen	52b1b3e31f	allocation explain: Also serialize `includeDiskInfo` field.	2016-07-19 11:54:43 +02:00
Yannick Welsch	79ab6d19af	Fix NPE when initializing replica shard has no unassignedInfo (#19491 ) An initializing replica shard might not have an UnassignedInfo object, for example when it is a relocation target. The method allocatedPostIndexCreate does not account for this situation.	2016-07-19 11:30:57 +02:00
Simon Willnauer	5b07f81fcf	Move `reset recovery` into RecoveriesCollection (#19466 ) Today when we reset a recovery because of the source not being ready or the shard is getting removed on the source (for whatever reason) we wipe all temp files and reset the recovery without respecting any reference counting or locking etc. all streams are closed and files are wiped. Yet, this is problematic since we assert that some files are on disk etc. when we finish writing a file. These assertions don't hold anymore if we concurrently wipe the tmp files. This change moves the logic out of RecoveryTarget into RecoveriesCollection which basically clones the RecoveryTarget on reset instead which allows in-flight operations to finish gracefully. This means we now have a single path for cleanups in RecoveryTarget and can safely use assertions in the class since files won't be removed unless the recovery is either canceled, failed or finished. Closes #19473	2016-07-19 10:23:02 +02:00
Adrien Grand	37e20c6f34	Automatically created indices should honor `index.mapper.dynamic`. #19478 Today they don't because the create index request that is implicitly created adds an empty mapping for the type of the document. So to Elasticsearch it looks like this type was explicitly created and `index.mapper.dynamic` is not checked. Closes #17592	2016-07-19 09:02:31 +02:00
Nik Everett	7861548786	Migrate serial_diff aggregation to NamedWriteable This is the last migration before AggregationStreams and PipelineAggregatorStreams can be removed to remove redundant code.	2016-07-18 13:00:06 -04:00
Adrien Grand	3bb6a4dea6	Try to prevent classloading deadlock. Closes #19316	2016-07-18 17:45:17 +02:00
Colin Goodheart-Smithe	e3d3f6b1f1	#19472 Enable option to use request cache for size > 0 Enable option to use request cache for size > 0	2016-07-18 16:28:07 +01:00
Yannick Welsch	4bec7ad58f	Do not throw AssertionError for expected exceptions in SearchWhileRelocatingIT (#19476 ) The test would previously catch Throwable and then decide if it was a critical exception or not. As the catch block was changed from Throwable to Exception this made the test fail for non-critical exceptions. This commit changes the test so that exceptions are only thrown when they're unexpected.	2016-07-18 16:45:07 +02:00
Martijn van Groningen	82e7f1fc43	parent/child: Make sure that no `_parent#null` gets introduces as default _parent mapping. Instead it should just be `_parent` field. Also added more tests regarding the join doc values field being added. Closes #19389	2016-07-18 16:38:13 +02:00
Nik Everett	16812cc032	Migrate moving_avg pipeline aggregation to NamedWriteable This is the first pipeline aggregation that doesn't have its own bucket type that needs serializing. It uses InternalHistogram instead. So that required reworking the new-style `registerAggregation` method to not require bucket readers. So I built `PipelineAggregationSpec` to mirror `AggregationSpec`. It allows registering any number of bucket readers or result readers.	2016-07-18 10:14:09 -04:00
Simon Willnauer	8394544548	Add a dedicated client/transport project for transport-client (#19435 ) The `client/transport` project adds a new jar build project that pulls in all dependencies and configures all required modules. Preinstalled modules are: * transport-netty * lang-mustache * reindex * percolator The `TransportClient` classes are still in core while `TransportClient.Builder` has only a protected construcutor such that users are redirected to use the new `TransportClientBuilder` from the new jar. Closes #19412	2016-07-18 15:42:24 +02:00
Colin Goodheart-Smithe	b717ad8eb6	Enable option to use request cache for size > 0 Previously if the size of the search request was greater than zero we would not cache the request in the request cache. This change retains the default behaviour of not caching requests with size > 0 but also allows the `request_cache=true` query parameter to enable the cache for requests with size > 0	2016-07-18 13:33:59 +01:00
Adrien Grand	398d70b567	Add `scaled_float`. #19264 This is a tentative to revive #15939 motivated by elastic/beats#1941. Half-floats are a pretty bad option for storing percentages. They would likely require 2 bytes all the time while they don't need more than one byte. So this PR exposes a new `scaled_float` type that requires a `scaling_factor` and internally indexes `valuescaling_factor` in a long field. Compared to the original PR it exposes a lower-level API so that the trade-offs are clearer and avoids any reference to fixed precision that might imply that this type is more accurate (actually it is less* accurate). In addition to being more space-efficient for some use-cases that beats is interested in, this is also faster that `half_float` unless we can improve the efficiency of decoding half-float bits (which is currently done using software) or until Java gets first-class support for half-floats.	2016-07-18 12:36:23 +02:00
Adrien Grand	bde99bad2e	Use a static default precision for the cardinality aggregation. #19215 Today the default precision for the cardinality aggregation depends on how many parent bucket aggregations it had. The reasoning was that the more parent bucket aggregations, the more buckets the cardinality had to be computed on. And this number could be huge depending on what the parent aggregations actually are. However now that we run terms aggregations in breadth-first mode by default when there are sub aggregations, it is less likely that we have to run the cardinality aggregation on kagilions of buckets. So we could use a static default, which will be less confusing to users.	2016-07-18 11:30:41 +02:00
Boaz Leskes	9ededa46bc	Make static Store access shard lock aware (#19416 ) We currently have concurrency issue between the static methods on the Store class and store changes that are done via a valid open store. An example of this is the async shard fetch which can reach out to a node while a local shard copy is shutting down (the fetch does check if we have an open shard and tries to use that first, but if the shard is shutting down, it will not be available from IndexService). Specifically, async shard fetching tries to read metadata from store, concurrently the shard that shuts down commits to lucene, changing the segments_N file. this causes a file not find exception on the shard fetching side. That one in turns makes the master think the shard is unusable. In tests this can cause the shard assignment to be delayed (up to 1m) which fails tests. See https://elasticsearch-ci.elastic.co/job/elastic+elasticsearch+master+java9-periodic/570 for details. This is one of the things #18938 caused to bubble up.	2016-07-18 11:22:58 +02:00
Adrien Grand	dd95dc7a0f	Fix potential AssertionError with include/exclude on terms aggregations. #19252 We call `LongBitSet.set(start, end)`, which fails when `start >= length` (0 in that case). Closes #18575	2016-07-18 11:03:24 +02:00
Martijn van Groningen	e0ebf5da1c	Template cleanup: * Removed `Template` class and unified script & template parsing logic. Templates are scripts, so they should be defined as a script. Unless there will be separate template infrastructure, templates should share as much code as possible with scripts. * Removed ScriptParseException in favour for ElasticsearchParseException * Moved TemplateQueryBuilder to lang-mustache module because this query is hard coded to work with mustache only	2016-07-18 10:16:01 +02:00
Boaz Leskes	798ee177ed	mute testAckedIndexing pending the merge of https://github.com/elastic/elasticsearch/pull/19416	2016-07-18 10:07:18 +02:00
Ali Beyad	6acb8b31fc	Removes ensureYellow() calls after index creation in the (#19452 ) integration tests, as they are no longer needed with index creation now waiting for shards to be started before returning from the index creation call (by default, it waits for the primary of each shard to be started before returning, which is what ensureYellow() was ensuring anyway). Closes #19452 Relates #19450	2016-07-15 15:37:35 -04:00
Jason Tedor	e772b6d924	Add log message about enforcing bootstrap checks This commit adds a log message when bootstrap checks are enforced informing the user that they are enforced because they are bound to an external network interface. We also log if bootstrap checks are being enforced but system checks are being ignored. Relates #19451	2016-07-15 14:29:36 -04:00
Yannick Welsch	f5b5fbcf1d	Strengthen assertions when random failures are not injected by AbstractIndicesClusterStateServiceTestCase (#19358 ) The unit tests for IndicesClusterStateService currently inject random failures upon shard creation/ routing upate / mapping update etc. This commit makes injecting failures optional so that stronger assertions can be made about the local indices / shard state in case of no failures.	2016-07-15 18:32:03 +02:00
Ali Beyad	687e2e12b3	Merge pull request #19450 from elastic/feature/friendly-index-creation Makes index creation more friendly	2016-07-15 11:48:21 -04:00
Ali Beyad	d78f40fb1e	Index creation waits for active shard copies before returning (#18985 ) Before returning, index creation now waits for the configured number of shard copies to be started. In the past, a client would create an index and then potentially have to check the cluster health to wait to execute write operations. With the cluster health semantics changing so that index creation does not cause the cluster health to go RED, this change enables waiting for the desired number of active shards to be active before returning from index creation. Relates #9126	2016-07-15 11:19:27 -04:00
Jason Tedor	917fea7c5d	Reset Priority values For historical reasons, the value associated with Priority.IMMEDIATE is -1. Yet, with a full-cluster restart required on major version upgrades, we can reset these values so they are conceptually simpler. This commit resets the values associated with Priority instances.	2016-07-15 09:34:31 -04:00
Jason Tedor	220a510d65	Make Priority an enum Today we have an abstraction Priority for representing priorities. Ideally, these values are a fixed set of constants with a well-defined ordering which sounds perfect for an enum. This commit changes Priority so that it is an enum instead of a class.	2016-07-15 08:55:49 -04:00
Jason Tedor	ac39e73183	Priority values should be unmodifiable In Priority there is a field named values that represents an ordered, by priority, list of all priorities. Yet, this collection is modifiable and this collection is exposed via the public API. This means that consumers can modify this list potentially leading to complete chaos. This commit modifies this field so that it is unmodifiable, documents that the returned collection is unmodifiable, and returns total order to the world. We also punish the bad consumer here by making them make a copy of the returned collection with which they can do as they please. This fixes a puzzling test failure which only arises if the two tests (PrioritizedExecutorsTests#testPriorityQueue and PriorityTests#testCompareTo run in the same JVM, and run in the right order). Relates #19447	2016-07-15 08:36:59 -04:00
Martijn van Groningen	d0069f0fbb	Provide access to ThreadContext in ingest plugins Also introduced a `Processor.Parameters` class that is holder for several services processors rely on, the IngestPlugin#getProcessors(...) method has been changed to accept `Processor.Parameters` instead of each service seperately.	2016-07-15 08:16:15 +02:00
Ryan Ernst	9b6e2a8e2f	Merge pull request #19440 from rjernst/rest_headers Plugins: Make rest headers registration pull based	2016-07-14 20:33:44 -07:00
Jason Tedor	a5b8cb87be	Log one plugin info per line Today we log all loaded modules and installed plugins in a single line. The number of modules has grown, and when plugins are installed a single log line containing the loaded modules and plugins is lengthy. With this commit, we log a single module or plugin per line, log these in sorted order, and also log if no modules or no plugins were loaded. Relates #19441	2016-07-14 22:46:35 -04:00
Ryan Ernst	4b9932d4a8	Merge branch 'master' into rest_headers	2016-07-14 19:03:53 -07:00
Jason Tedor	31c648eee8	Rename transport-netty to transport-netty3 This commit renames the Netty 3 transport module from transport-netty to transport-netty3. This is to make room for a Netty 4 transport module, transport-netty4. Relates #19439	2016-07-14 22:03:14 -04:00

1 2 3 4 5 ...

5754 Commits