OpenSearch

Commit Graph

Author	SHA1	Message	Date
Ryan Ernst	cca5934e9d	Tests: Pass through locale and timezone to test runner, and print in repro command line. The carrot runner currently randomizes both locale and timezone, but these are not set in the maven reproduce line. Since they aren't even printed, we have no idea what locale/timezone the tests actually ran with.	2014-11-19 22:01:26 -08:00
Simon Willnauer	5c6fe2593e	[CORE] Ban all useage of Future#cancel(true) Interrupting a thread while blocking on an NIO Read / Write Operation can cause a file to be closed due to the interrupts. This can have unpredictable effects when files are open by index readers etc. we should prevent interruptions across the board if possible. Closes #8494	2014-11-18 14:14:09 +01:00
Philip McMahon	4194a699c0	Logging: Add log4j-extras dependency Close #7927	2014-11-13 12:39:30 +00:00
Michael McCandless	87f6d6bc40	remove wrong repository	2014-11-10 14:14:11 -05:00
Michael McCandless	8aebb9656b	Core: add max_determinized_states to query_string and regexp query/filter This prevents too-difficult regular expressions from consuming excessive RAM/CPU; the default max_determinized_states is 10,000 (same as Lucene) but query_string and regepx query/filter can override per-request. The also upgrades to a new Lucene 5.0.0 snapshot. Closes #8386 Closes #8357	2014-11-10 13:43:48 -05:00
Robert Muir	610ce078fb	Upgrade master to lucene 5.0 snapshot This has a lot of improvements in lucene, particularly around memory usage, merging, safety, compressed bitsets, etc. On the elasticsearch side, summary of the larger changes: API changes: postings API became a "pull" rather than "push", collector API became per-segment, etc. packaging changes: add lucene-backwards-codecs.jar as a dependency. improvements to boolean filtering: especially ensuring it will not be slow for SparseBitSet. use generic BitSet api in plumbing so that concrete bitset type is an implementation detail. use generic BitDocIdSetFilter api for dedicated bitset cache, so there is type safety. changes to support atomic commits implement Accountable.getChildResources (detailed memory usage API) for fielddata, etc change handling of IndexFormatTooOld/New, since they no longer extends CorruptIndexException Closes #8347. Squashed commit of the following: commit d90d53f5f21b876efc1e09cbd6d63c538a16cd89 Author: Simon Willnauer <simonw@apache.org> Date: Wed Nov 5 21:35:28 2014 +0100 Make default codec/postings/docvalues format constants commit cb66c22c71cd304a36e7371b199a8c279908ae37 Merge: d4e2f6d `ad4ff43` Author: Robert Muir <rmuir@apache.org> Date: Wed Nov 5 11:41:13 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit d4e2f6dfe767a5128c9b9ae9e75036378de08f47 Merge: 4e5445c `4111d93` Author: Robert Muir <rmuir@apache.org> Date: Wed Nov 5 06:26:32 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit 4e5445c775f580730eb01360244e9330c0dc3958 Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 16:19:19 2014 -0500 FixedBitSet -> BitSet commit 9887ea73e8b857eeda7f851ef3722ef580c92acf Merge: 1bf8894 `fc84666` Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 15:26:25 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit 1bf8894430de3e566d0dc5623b0cc28b0d674ebb Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 15:22:51 2014 -0500 remove nocommit commit a9c2a2259ff79c69bae7806b64e92d5f472c18c8 Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 13:48:43 2014 -0500 turn jenkins red again commit 067baaaa4d52fce772c81654dcdb5051ea79139f Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 13:18:21 2014 -0500 unzip from stream commit 82b6fba33d362aca2313cc0ca495f28f5ebb9260 Merge: b2214bb `6523cd9` Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 13:10:59 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit b2214bb093ec2f759003c488c3c403c8931db914 Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 13:09:53 2014 -0500 go back to my URL until we can figure out what is up with jenkins commit e7d614172240175a51f580aeaefb6460d21cede9 Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 10:52:54 2014 -0500 try this jenkins commit 337a3c7704efa7c9809bf373152d711ee55f876c Author: Simon Willnauer <simonw@apache.org> Date: Tue Nov 4 16:17:49 2014 +0100 Rename temp-files under lock to prevent metadata reads while renaming commit 77d5ba80d0a76efa549dd753b9f114b2f2d2d29c Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 10:07:11 2014 -0500 continue to treat too-old/too-new as corruption for now commit 98d0fd2f4851bc50e505a94ca592a694d502c51c Author: Robert Muir <rmuir@apache.org> Date: Tue Nov 4 09:24:21 2014 -0500 fix last nocommit commit 643fceed66c8caf22b97fc489d67b4a2a90a1a1c Author: Simon Willnauer <simonw@apache.org> Date: Tue Nov 4 14:46:17 2014 +0100 remove NoSuchDirectoryException commit 2e43c4feba05cfaf451df70f946c0930cbcc4557 Merge: 93826e4 `8163107` Author: Simon Willnauer <simonw@apache.org> Date: Tue Nov 4 14:38:00 2014 +0100 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit 93826e4d56a6a97c2074669014af77ff519bde63 Merge: 7f10129 `44e24d3` Author: Simon Willnauer <simonw@apache.org> Date: Tue Nov 4 12:54:27 2014 +0100 Merge branch 'master' into enhancement/lucene_5_0_upgrade Conflicts: src/main/java/org/elasticsearch/index/store/DistributorDirectory.java src/main/java/org/elasticsearch/index/store/Store.java src/main/java/org/elasticsearch/indices/recovery/RecoveryStatus.java src/test/java/org/elasticsearch/index/store/DistributorDirectoryTest.java src/test/java/org/elasticsearch/index/store/StoreTest.java src/test/java/org/elasticsearch/indices/recovery/RecoveryStatusTests.java commit 7f10129364623620575c109df725cf54488b3abb Author: Adrien Grand <jpountz@gmail.com> Date: Tue Nov 4 11:32:24 2014 +0100 Fix TopHitsAggregator to not ignore the top-level/leaf collector split. commit 042fadc8603b997bdfdc45ca44fec70dc86774a6 Author: Adrien Grand <jpountz@gmail.com> Date: Tue Nov 4 11:31:20 2014 +0100 Remove MatchDocIdSet in favor of DocValuesDocIdSet. commit 7d877581ff5db585a674c95ac391ac78a0282826 Author: Adrien Grand <jpountz@gmail.com> Date: Tue Nov 4 11:10:08 2014 +0100 Make the and filter use the cost API. Lucene 5 ensured that cost() can safely be used, and this will have the benefit that the order in which filters are specified is not important anymore (only for slow random-access filters in practice). commit 78f1718aa2cd82184db7c3a8393e6215f43eb4a8 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 23:55:17 2014 -0500 fix previous eclipse import braindamage commit 186c40e9258ce32f22a9a714ab442a310b6376e0 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 22:32:34 2014 -0500 allow child queries to exhaust iterators again commit b0b1271305e1b6d0c4c4da51a3c54df1aa5c0605 Author: Ryan Ernst <ryan@iernst.net> Date: Mon Nov 3 14:50:44 2014 -0800 Fix nocommit for mapping output. index_options will not be printed if the field is not indexed. commit ba223eb85e399c9620a347a983e29bf703953e7a Author: Ryan Ernst <ryan@iernst.net> Date: Mon Nov 3 14:07:26 2014 -0800 Remove no commit for chinese analyzer provider. We should have a separate issue to address not using this provider on new indexes. commit ca554b03c4471797682b2fb724f25205cf040c4a Author: Ryan Ernst <ryan@iernst.net> Date: Mon Nov 3 13:41:59 2014 -0800 Fix stop tests commit de67c4653ec47dee9c671390536110749d2bb05f Author: Ryan Ernst <ryan@iernst.net> Date: Mon Nov 3 12:51:17 2014 -0800 Remove analysis nocommits, switching over to Lucene43*Filters for backcompat commit 50cae9bec72c25c33a1ab8a8931bccb3355171e2 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 15:32:25 2014 -0500 add ram accounting and TODO lazy-loading (its no worse than master, can be a followup improvement) for suggesters commit 7a7f0122f138684b312d0f0b03dc2a9c16c15f9c Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 15:11:26 2014 -0500 bump lucene version commit cd0cae5c35e7a9e049f49ae45431f658fb86676b Merge: 446bc09 `3c72073` Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 14:49:05 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit 446bc09b4e8bf4602d3c252b53ddaa0da65cce2f Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 14:46:30 2014 -0500 remove hack commit a19d85a968d82e6d00292b49630ef6ff2dbf2f32 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 12:53:11 2014 -0500 dont create exceptions with circular references on corruption (will open a PR for this) commit 0beefb9e821d97c37e90ec556d81ac7b00369b8a Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 11:47:14 2014 -0500 temporarily add craptastic detector for this horrible bug commit e9f2d298bff75f3d1591f8622441e459c3ce7ac3 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 10:56:01 2014 -0500 add nocommit commit e97f1d50a91a7129650b8effc7a9ecf74ca0569a Merge: c57a3c8 `f1f50ac` Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 10:12:12 2014 -0500 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit c57a3c8341ed61dca62eaf77fad6b8b48aeb6940 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 10:11:46 2014 -0500 fix nocommit commit dd0e77e4ec07c7011ab5f6b60b2ead33dc2333d2 Author: Robert Muir <rmuir@apache.org> Date: Mon Nov 3 09:54:09 2014 -0500 nocommit -> TODO, this is in much more places in the codebase, bigger issue commit 3cc3bf56d72d642059f8fe220d6f2fed608363e9 Author: Ryan Ernst <ryan@iernst.net> Date: Sat Nov 1 23:59:17 2014 -0700 Remove nocommit and awaitsfix for edge ngram filter test. commit 89f115245155511c0fbc0d5ee62e63141c3700c1 Author: Ryan Ernst <ryan@iernst.net> Date: Sat Nov 1 23:57:44 2014 -0700 Fix EdgeNGramTokenFilter logic for version <= 4.3, and fixed instanceof checks in corresponding tests to correctly check for reverse filter when applicable. commit 112df869cd199e36aab0e1a7a288bb1fdb2ebf1c Author: Robert Muir <rmuir@apache.org> Date: Sun Nov 2 00:08:30 2014 -0400 execute geo disjoint query/filter as intersects commit e5061273cc685f1252e9a3a9ae4877ec9bce7752 Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 22:58:59 2014 -0400 remove chinese analyzer from docs commit ea1af11b8978fcc551f198e24fe21d52806993ef Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 22:29:00 2014 -0400 fix ram accounting bug commit 53c0a42c6aa81aa6bf81d3aa77b95efd513e0f81 Merge: e3bcd3c `6011a18` Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 22:16:29 2014 -0400 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit e3bcd3cc07a4957e12c7b3affc462c31290a9186 Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 22:15:01 2014 -0400 fix url-email back compat (thanks ryan) commit 91d6b096a96c357755abee167098607223be1aad Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 22:11:26 2014 -0400 bump lucene version commit d2bb9568df72b37ec7050d25940160b8517394bc Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 20:33:07 2014 -0400 remove nocommit commit 1d049c471e19e5c457262c7399c5bad9e023b2e3 Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 20:28:58 2014 -0400 fix eclipse to group org/com imports together: without this, its madness commit 09d8c1585ee99b6e63be032732c04ef6fed84ed2 Author: Robert Muir <rmuir@apache.org> Date: Sat Nov 1 14:27:41 2014 -0400 remove nocommit, if you dont liek it, print assembly and tell me how it can be better commit 8a6a294313fdf33b50c7126ec20c07867ecd637c Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 20:01:55 2014 +0100 Remove deprecated usage of DocIdSets.newDocIDSet. commit 601bee60543610558403298124a84b1b3bbd1045 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 14:13:18 2014 -0400 maybe one of these zillions of annotations will stop thread leaks commit 9d3f69abc7267c5e455aefa26db95cb554b02d62 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 14:05:39 2014 -0400 fix some analysis nocommits commit 312e3a29c77214b8142d21c33a6b2c2b151acf9a Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 18:28:45 2014 +0100 Remove XConstantScoreQuery/XFilteredQuery/ApplyAcceptedDocsFilter. commit 5a0cb9f8e167215df7f1b1fad11eec6e6c74940f Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 17:06:45 2014 +0100 Fix misleading documentation of DocIdSets.toCacheable. commit 8b4ef2b5b476fff4c79c0c2a0e4769ead26cf82b Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 17:05:59 2014 +0100 Fix CustomRandomAccessFilterStrategy to override the right method. commit d7a9a407a615987cfffc651f724fbd8795c9c671 Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 16:21:35 2014 +0100 Better handle the special case when there is a single SHOULD clause. commit 648ad389f07e92dfc451f345549c9841ba5e4c9a Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 15:53:38 2014 +0100 Cut over XBooleanFilter to BitDocIdSet.Builder. The idea is similar to what happened to Lucene's BooleanFilter. Yet XBooleanFilter is a bit more sophisticated and I had to slightly change the way it is implemented in order to make it work. The main difference with before is that slow filters are now applied lazily, so eg. if you have 3 MUST clauses, two with a fast iterator and the third with a slow iterator, the previous implementation used to apply the fast iterators first and then only check the slow filter for bits which were set in the bit set. Now we are computing a bit set based on the fast must clauses and then basically returning a BitsFilteredDocIdSet.wrap(bitset, slowClause). Other than that, BooleanFilter still uses the bitset optimizations when or-ing and and-ind filters. Another improvement is that BooleanFilter is now aware of the cost API. commit b2dad312b4bc9f931dc3a25415dd81c0d9deee08 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 10:18:53 2014 -0400 clear nocommit commit 4851d2091e744294336dfade33906c75fbe695cd Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 15:15:16 2014 +0100 cut over to RoaringDocIdSet commit ca6aec24a901073e65ce4dd6b70964fd3612409e Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:57:30 2014 +0100 make nocommit more explicit commit d0742ee2cb7a6c48b0bbb31580b7fbcebdb6ec40 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 09:55:24 2014 -0400 fix standardtokenizer nocommit commit 7d6faccafff22a86af62af0384838391d46695ca Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:54:08 2014 +0100 fix compilation commit a038a405c1ff6458ad294e6b5bc469e622f699d0 Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:53:43 2014 +0100 fix compilation commit 30c9e307b1f5d80e2deca3392c0298682241207f Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:52:35 2014 +0100 fix compilation commit e5139bc5a0a9abd2bdc6ba0dfbcb7e3c2e7b8481 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 09:52:16 2014 -0400 clear nocommit here commit 85dd2cedf7a7994bed871ac421cfda06aaf5c0a5 Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:46:17 2014 +0100 fix CompletionPostingsFormatTest commit c0f3781f616c9b0ee3b5c4d0998810f595868649 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 09:38:00 2014 -0400 add tests for these analyzers commit 51f9999b4ad079c283ae762c862fd0e22d00445f Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 14:10:26 2014 +0100 remove nocommit - this is not an issue commit fd1388fa03e622b0738601c8aeb2dbf7949a6dd2 Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Fri Oct 31 14:07:01 2014 +0100 Remove redundant null check commit 3d6dd51b0927337ba941a235446b22e8cd500dc3 Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Fri Oct 31 14:01:37 2014 +0100 Removed the work around to prevent p/c error when invoking #iterator() twice, because the custom query filter wrapper now doesn't transform the result to a cache doc id set any more. I think the transforming to a cachable doc id set in CustomQueryWrappingFilter isn't needed at all, because we use the DocIdSet only once and because of that is just slowed things down. commit 821832a537e00cd1216064b379df3e01d2911d3a Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 13:54:33 2014 +0100 one more nocommit commit 77eb9ea4c4ea50afb2680c29682ddcb3851a9d4f Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Fri Oct 31 13:52:29 2014 +0100 Remove cast commit a400573c034ed602221f801b20a58a9186a06eae Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 13:49:24 2014 +0100 fix stop filter commit 51746087cf8ec34c4d20aa05ba8dbff7b3b43eec Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 13:21:36 2014 +0100 fix changed semantics of FBS.nextSetBit to check for NO_MORE_DOCS commit 8d0a4e2511310f1293860823fe3ba80ac771bbe3 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 08:13:44 2014 -0400 do the bogus cast differently commit 46a5cc5732dea096c0c80ae5ce42911c9c51e44e Author: Simon Willnauer <simonw@apache.org> Date: Fri Oct 31 13:00:16 2014 +0100 I hate it but P/C now passes commit 580c0c2f82bbeacf217e594f22312b11d1bdb839 Merge: a9d3c00 `1645434` Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 31 06:54:31 2014 -0400 fix nocommit/classcast commit a9d3c004d62fe04989f49a897e6ff84973c06eb9 Author: Adrien Grand <jpountz@gmail.com> Date: Fri Oct 31 08:49:31 2014 +0100 Update TODO. commit aa75af0b407792aeef32017f03a6f442ed970baa Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 19:18:25 2014 -0400 clear obselete nocommits from lucene bump commit d438534cf41fcbe2d88070e2f27c994625e082c2 Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 18:53:20 2014 -0400 throw classcastexception when ES abuses regular filtercache for nested docs commit 2c751f3a8feda43ec127c34769b069de21f3d16f Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 18:31:34 2014 -0400 bump lucene revision, fix tests commit d6ef7f6304ae262bf6228a7d661b2a452df332be Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 22:37:58 2014 +0100 fix merge problems commit de9d361f88a9ce6bb3fba85285de41f223c95767 Merge: 41f6aab `f6b37a3` Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 22:28:59 2014 +0100 Merge branch 'master' into enhancement/lucene_5_0_upgrade Conflicts: pom.xml src/main/java/org/elasticsearch/Version.java src/main/java/org/elasticsearch/gateway/local/state/meta/MetaDataStateFormat.java commit 41f6aab388aa80c40b08a2facab2617576203a0d Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 17:48:46 2014 +0100 fix potiential NPE commit c4428b12e1ae838b91e847df8b4a8be7f49e10f4 Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 17:38:46 2014 +0100 don't advance iterator in a match(doc) method commit 28ab948e99e3ea4497c9b1e468384806ba7e1790 Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 17:34:58 2014 +0100 don't advance iterator in a match(doc) method commit eb0f33f6634fadfcf4b2bf7327400e568f0427bb Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 16:55:54 2014 +0100 fix GeoUtilsTest commit 7f711fe3eaf73b6c2268cf42d5a41132a61ad831 Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 16:43:16 2014 +0100 Use a dedicated default index option if field type is not indexed by default commit 78e3f37ab779e3e1b25b45a742cc86ab5f975149 Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 10:56:14 2014 -0400 disable this test with AwaitsFix to reduce noise commit 9a590f563c8e03a99ecf0505c92d12d7ab20d11d Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 09:38:49 2014 +0100 fix lucene version commit abe3ca1d8bb6b5101b545198f59aec44bacfa741 Author: Simon Willnauer <simonw@apache.org> Date: Thu Oct 30 09:35:05 2014 +0100 fix AnalyzingCompletionLookupProvider to wrok with new codec API commit 464293b245852d60bde050c6d3feb5907dcfbf5f Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 00:26:00 2014 -0400 don't try to write stuff to tests class directory commit 031cc6c19f4fe4423a034b515f77e5a0e282a124 Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 00:12:36 2014 -0400 AwaitsFix these known issues to reduce noise commit 4600d51891e35847f2d344247d6f915a0605c0d1 Author: Robert Muir <rmuir@apache.org> Date: Thu Oct 30 00:06:53 2014 -0400 openbitset lives on commit 8492bae056249e2555d24acd55f1046b66a667c4 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 23:42:54 2014 -0400 fixes for filter tests commit 31f24ce4efeda31f97eafdb122346c7047a53bf2 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 23:12:38 2014 -0400 don't use fieldcache commit 8480789942fdff14a6d2b2cd8134502fe62f20c8 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 23:04:29 2014 -0400 ancient index no longer supported commit 02e78dc7ebdd827533009f542582e8db44309c57 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 23:37:02 2014 +0100 fix more tests commit ff746c6df23c50b3f3ec24922413b962c8983080 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 23:08:19 2014 +0100 fix all mapper commit e4fb84b517107b25cb064c66f83c9aa814a311b2 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 22:55:54 2014 +0100 fix distributor tests and cut over to FileStore API commit 20c850e2cfe3210cd1fb9e232afed8d4ac045857 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 22:42:18 2014 +0100 use DOCS_ONLY if index=true and current options == null commit 44169c108418413cfe51f5ce23ab82047463e4c2 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 22:33:36 2014 +0100 Fix index=yes\|no settings in mappers commit a3c5f77987461a18121156ed345d42ded301c566 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 21:51:41 2014 +0100 fix several field mappers conversion from setIndexed to indexOptions commit df84d736908e88a031d710f98e222be68ae96af1 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 21:33:35 2014 +0100 fix SourceFieldMapper to be not indexed commit b2bf01d12a8271a31fb2df601162d0e89924c8f5 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 21:23:08 2014 +0100 Cut over to .liv files in store and corruption tests commit 619004df436f9ef05d24bef1b6a7f084c6b0ad75 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 17:05:52 2014 +0100 fix more tests commit b7ed653a8b464de446e00456bce0a89e47627c38 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 16:19:08 2014 +0100 [STORE] Add dedicated method to write temporary files Recovery writes temporary files which might not end up in the right distributor directories today. This commit adds a dedicated API that allows specifying the target file name in order to create the tempoary file in the correct directory. commit 7d574659f6ae04adc2b857146ad0d8d56ca66f12 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 10:28:49 2014 -0400 add some leniency to temporary bogus method commit f97022ea7c2259f7a5cf97d924c59ed75ab65b32 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 10:24:17 2014 -0400 fix MultiCollector bug commit b760533128c2b4eb10ad76e9689ef714293dd819 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:56:08 2014 +0100 CheckIndex is now closeable we need to close it commit 9dae9fb6d63546a6c2427be2a2d5c8358f5b1934 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:45:11 2014 +0100 s/Lucene51/Lucene50 commit 7aea9b86856a8c1b06a08e7c312ede1168af1287 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:42:30 2014 +0100 fix BloomFilterPostingsFormat commit 16fea6fe842e88665d59cc091e8224e8dc6ce08c Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:41:16 2014 +0100 fix some codec format issues commit 3d77aa97dd2c4012b63befef3f2ba2525965e8a6 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:30:43 2014 +0100 fix CodecTests commit 6ef823b1fde25657438ace1aabd9d552d6ae215e Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:26:47 2014 +0100 make it compile commit 9991eee1fe99435118d4dd42b297ffc83fce5ec5 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 09:12:43 2014 -0400 add an ugly hack for TopHitsAggregator for now commit 03e768a01fcae6b1f4cb50bcceec7d42977ac3e6 Author: Simon Willnauer <simonw@apache.org> Date: Wed Oct 29 14:01:02 2014 +0100 cut over ES090PostingsFormat commit 463d281faadb794fdde3b469326bdaada25af048 Merge: 0f8740a `8eac79c` Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 08:30:36 2014 -0400 Merge branch 'master' into enhancement/lucene_5_0_upgrade commit 0f8740a782455a63524a5a82169f6bbbfc613518 Author: Robert Muir <rmuir@apache.org> Date: Wed Oct 29 01:00:15 2014 -0400 fix/hack remaining filter and analysis issues commit df534488569da13b31d66e581456dfd4b55156b9 Author: Robert Muir <rmuir@apache.org> Date: Tue Oct 28 23:11:47 2014 -0400 fix ngrams / openbitset usage commit 11f5dc3b9887f4da80a0fa1818e1350b30599329 Author: Robert Muir <rmuir@apache.org> Date: Tue Oct 28 22:42:44 2014 -0400 hack over sort comparators commit 4ebdc754350f512596f6a02770d223e9f5f7975a Author: Robert Muir <rmuir@apache.org> Date: Tue Oct 28 21:27:07 2014 -0400 compiler errors < 100 commit 2d60c9e29de48ccb0347dd87f7201f47b67b83a0 Author: Robert Muir <rmuir@apache.org> Date: Tue Oct 28 03:13:08 2014 -0400 clear some nocommits around ram usage commit aaf47fe6c0aabcfb2581dd456fc50edf871da758 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 12:27:34 2014 -0400 migrate fieldinfo handling commit ef6ed6d15d8def71cd880d97249678136cd29fe3 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 12:07:13 2014 -0400 more simple fixes commit f475e1048ae697dd9da5bd9da445102b0b7bc5b3 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 11:58:21 2014 -0400 more fielddata ram accounting fixes commit 16b4239eaa9b4262df258257df4f31d39f28a3a2 Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 16:47:32 2014 +0100 add missing file commit 5b542fa2a6da81e36a0c35b8e891a1d8bc58f663 Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 16:43:29 2014 +0100 cut over completion posting formats - still some nocommits commit ecdea49404c4ec4e1b78fb54575825f21b4e096e Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 11:21:09 2014 -0400 fielddata accountable fixes commit d43da265718917e20c8264abd43342069198fe9c Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 16:19:53 2014 +0100 cut over BloomFilterPostings to new API commit 29b192ba621c14820175775d01242162b88bd364 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 10:22:51 2014 -0400 fix more analyzers commit 74b4a0c5283e323a7d02490df469497c722780d2 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 09:54:25 2014 -0400 fix tests commit 554084ccb4779dd6b1c65fa7212ad1f64f3a6968 Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 14:51:48 2014 +0100 maintain supressed exceptions on CorruptIndexException commit cf882d9112c5e8ef1e9f2b0f800f7aa59001a4f2 Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 14:47:17 2014 +0100 commitOnClose=false commit ebb2a9189ab2f459b7c6c9985be610fd90dfe410 Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 14:46:06 2014 +0100 cut over indexwriter closeing in InternalEngine commit cd21b3d4706f0b562bd37792d077d60832aff65f Author: Simon Willnauer <simonw@apache.org> Date: Mon Oct 27 14:38:10 2014 +0100 fix constant commit f93f900c4a1c90af3a21a4af5735a7536423fe28 Author: Robert Muir <rmuir@apache.org> Date: Mon Oct 27 09:50:49 2014 -0400 fix test commit a9a752940b1ab4699a6a08ba8b34afca82b843fe Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Mon Oct 27 09:26:18 2014 +0100 Be explicit about the index options commit d9ee815babd030fa2ceaec9f467c105ee755bf6b Author: Simon Willnauer <simonw@apache.org> Date: Sun Oct 26 20:03:44 2014 +0100 cut over store and directory commit b3f5c8e39039dd8f5caac0c4dd1fc3b1116e64ca Author: Robert Muir <rmuir@apache.org> Date: Sun Oct 26 13:08:39 2014 -0400 more test fixes commit 8842f2684e3606aae0860c27f7a4c53e273d47fb Author: Robert Muir <rmuir@apache.org> Date: Sun Oct 26 12:14:52 2014 -0400 tests manual labor commit c43de5aec337919a3fdc3638406dff17fc80bc98 Author: Robert Muir <rmuir@apache.org> Date: Sun Oct 26 11:04:13 2014 -0400 BytesRef -> BytesRefBuilder commit 020c0d087a2f37566a1db390b0e044ebab030138 Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Sun Oct 26 15:53:37 2014 +0100 Moved over to BitSetFilter commit 48dd1b909e6c52cef733961c9ecebfe4f67109fe Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Sun Oct 26 15:53:11 2014 +0100 Left over Collector api change in ScanContext commit 6ec248ef63f262bcda400181b838fd9244752625 Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Sun Oct 26 15:47:40 2014 +0100 Moved indexed() over to indexOptions != null or indexOptions == null commit 9937aebfd8546ae4bb652cd976b3b43ac5ab7a63 Author: Martijn van Groningen <martijn.v.groningen@gmail.com> Date: Sun Oct 26 13:26:31 2014 +0100 Fixed many compile errors. Mainly around the breaking Collector api change in 5.0. commit fec32c4abc0e3309cf34260c8816305a6f820c9e Author: Robert Muir <rmuir@apache.org> Date: Sat Oct 25 11:22:17 2014 -0400 more easy fixes commit dab22531d801800d17a65dc7c9464148ce8ebffd Author: Robert Muir <rmuir@apache.org> Date: Sat Oct 25 09:33:41 2014 -0400 more progress commit 414767e9a955010076b0497cc4f6d0c1850b48d3 Author: Robert Muir <rmuir@apache.org> Date: Sat Oct 25 06:33:17 2014 -0400 more progress commit ad9d969fddf139a8830254d3eb36a908ba87cc12 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 24 14:28:01 2014 -0400 current state of fun commit 464475eecb0be15d7d084135ed16051f76a7e521 Author: Robert Muir <rmuir@apache.org> Date: Fri Oct 24 11:42:41 2014 -0400 bump to 5.0 snapshot	2014-11-05 15:48:51 -05:00
javanna	f2d545c40e	[TEST] exclude org.elasticsearch.test.test package from test-jar The package was only excluded during test-jar sources generation but ended up in the actual jar.	2014-11-04 10:46:41 +01:00
Alexander Reelsen	5eeac2fdf6	Netty: Add HTTP pipelining support This adds HTTP pipelining support to netty. Previously pipelining was not supported due to the asynchronous nature of elasticsearch. The first request that was returned by Elasticsearch, was returned as first response, regardless of the correct order. The solution to this problem is to add a handler to the netty pipeline that maintains an ordered list and thus orders the responses before returning them to the client. This means, we will always have some state on the server side and also requires some memory in order to keep the responses there. Pipelining is enabled by default, but can be configured by setting the http.pipelining property to true\|false. In addition the maximum size of the event queue can be configured. The initial netty handler is copied from this repo https://github.com/typesafehub/netty-http-pipelining Closes #2665	2014-10-31 16:30:11 +01:00
Michael McCandless	7506255263	Upgrade to Lucene 4.10.2	2014-10-29 11:21:51 -04:00
Simon Willnauer	b23e6e0593	[TEST] initialize SUITE \| GLOBAL scope cluster in a private random context Today any call to the current randomized context modifies the random sequence such that cluster initialization is context dependent. If due to an error for instance a static util method is used like `randomLong` inside the TestCluster instead of the provided Random instance all reproducibility guarantees are gone. This commit adds a safe mechanism to initialize these clusters even if a static helper is used. All none test scope clusters are now initialized in a private randomized context.	2014-10-28 23:53:25 +01:00
Chris Earle	60c16ba92c	Use groovy-x.y.z-indy jar for better scripting performance Using the Groovy jar with the indy (short for `invokedynamic`) classifier enables usage of the `invokedynamic` instruction available in Java 7+. Due to buggy JVMs, it should only be used with Java 7u60 or later.	2014-10-21 13:08:08 -04:00
Lukasz Dywicki	b0a14f0b63	Avoid shading of org.joda.convert package, fixes #3557	2014-10-17 15:44:10 +02:00
Shay Banon	361b7b16b8	Upgrade to Jackson 2.4.2 closes #7934 closes #7932	2014-10-02 15:32:04 -04:00
Adrien Grand	3b38db121b	Mappings: Make lookup structures immutable. This commit makes the lookup structures that are used for mappings immutable. When changes are required, a new instance is created while the current instance is left unmodified. This is done efficiently thanks to a hash table implementation based on a array hash trie, see org.elasticsearch.common.collect.CopyOnWriteHashMap. ManyMappingsBenchmark returns indexing times that are similar to the ones that can be observed in current master. Ultimately, I would like to see if we can make mappings completely immutable as well and updated atomically. This is not trivial however, eg. because of dynamic mappings. So here is a first baby step that should help move towards that direction. Close #7486	2014-10-02 13:42:20 +02:00
Ryan Ernst	df22e54baf	Move forbidden api signature files to dev-tools. This avoids the files showing up in the binary release, since .txt files are copied. closes #7917 closes #7921	2014-09-29 15:27:43 -07:00
Simon Willnauer	20a0c68964	[BUILD] Release version should match latest version This commit ensures that the latest version in our code is identical to the project.version specified in the pom.xml file.	2014-09-29 17:45:10 +02:00
mikemccand	6bf635039c	Core: upgrade to Lucene 4.10.1	2014-09-28 13:42:12 -04:00
Michael McCandless	5e9e2cf50c	Core: try again to upgrade to Lucene 4.10.1-snapshot	2014-09-24 13:48:49 -04:00
Michael McCandless	ab3be76644	Revert Lucene upgrade	2014-09-24 13:25:55 -04:00
Michael McCandless	15c75b1967	Core: upgrade to Lucene 4.10.1 snapshot Lucene will soon release official 4.10.1, but by upgrading sooner we can 1) sidestep the false failures due to the 1.8.0_20 JVM hotspot bug (has caused a number of false failures in recent Jenkins tests), 2) make sure none of the Lucene changes in 4.10.1 are problematic. Closes #7844	2014-09-24 13:13:07 -04:00
Simon Willnauer	38d88b2e2c	[TEST] Added basic test for InternalTestCluster reproducibility	2014-09-09 21:48:42 +02:00
Britta Weber	59ce940d43	[TESTS]: create directory for heapdumps if not exists Before the heapdump was either written in a file with the directory name if the heapdump path ended without / or not written at all if the path ended with / closes #7645	2014-09-08 21:21:05 +02:00
mikemccand	130fdef367	Core: remove built-in support for Lucene's experimental codecs Lucene's experimental codecs (from the codecs module) do not provide backwards compatibility and are free to change from release to release. When they do change, they typically cannot in general read older indices and the resulting exceptions look like index corruption. So, we are removing built-in support for them to prevent applications from choosing one and then seeing strange exceptions on upgrade. Closes #7566 Closes #7604	2014-09-08 04:55:15 -04:00
Robert Muir	223dab8921	[Lucene] Upgrade to Lucene 4.10 Closes #7584	2014-09-05 12:21:08 -04:00
Adrien Grand	b49853a619	Internal: Upgrade Guava to 18.0. 17.0 and earlier versions were affected by the following bug https://code.google.com/p/guava-libraries/issues/detail?id=1761 which caused caches that are configured with weights that are greater than 32GB to actually be unbounded. This is now fixed. Relates to #6268 Close #7593	2014-09-04 20:14:59 +02:00
Boaz Leskes	598854dd72	[Discovery] accumulated improvements to ZenDiscovery Merging the accumulated work from the feautre/improve_zen branch. Here are the highlights of the changes: __Testing infra__ - Networking: - all symmetric partitioning - dropping packets - hard disconnects - Jepsen Tests - Single node service disruptions: - Long GC / Halt - Slow cluster state updates - Discovery settings - Easy to setup unicast with partial host list __Zen Discovery__ - Pinging after master loss (no local elects) - Fixes the split brain issue: #2488 - Batching join requests - More resilient joining process (wait on a publish from master) Closes #7493	2014-09-01 16:13:57 +02:00
Boaz Leskes	183ca37dfa	Code style improvement	2014-08-29 09:01:05 +02:00
Britta Weber	44dbd9b0c9	test: write heap dump to log folder Per default the heap dump is written to target/JX/pidXYZ.hprof In order to keep them when a new test is is started, they should be written to log folder which is not cleared in a new test run. Heap dump location can be set with -Dtests.heapdump.path=/path/to/heapdump closes #7452	2014-08-28 14:51:10 +02:00
Boaz Leskes	50f852ffeb	[TEST] Added LongGCDisruption and a test simulating GC on master nodes Also rename DiscoveryWithNetworkFailuresTests to DiscoveryWithServiceDisruptions which better suites what we do.	2014-08-27 15:47:40 +02:00
Alexander Reelsen	e478f24b9e	Dependencies: Upgrade to Apache HttpComponents client 4.3.5 Closes #7342	2014-08-22 09:44:13 +02:00
Britta Weber	5bfd75e489	Test: write heap dump per default The heap dump is always interesting in case a test causes an OOM. Per default is is written to target/JX/pidXYZ.hprof	2014-08-21 15:59:32 +02:00
Shay Banon	3da5773dd9	Upgrade to Jackson Smile 2.4.1.1 This fixes issue 18 in Smile (https://github.com/FasterXML/jackson-dataformat-smile/issues/18) closes #7327	2014-08-19 11:10:54 -07:00
Shay Banon	fc405f6f8a	Upgrade to Netty 3.9.3.Final closes #7328	2014-08-19 11:09:35 -07:00
Robert Muir	f85554d6ef	Merge pull request #7158 from uschindler/forbiddenapis-1.6-update Update forbidden-apis to 1.6 Closes #7158	2014-08-15 10:28:27 -04:00
Ryan Ernst	c1b6e53cbb	Internal: Fix a very rare case of corruption in compression used for internal cluster communication. See CorruptedCompressorTests for details on how this bug can be hit. This change also removes the ability to use the unsafe variant of ChunkedEncoder, removing support for the compress.lzf.decoder setting.	2014-08-11 07:26:09 -07:00
Alexander Reelsen	9e4064a9ba	Dependencies: Version bump HPPC to 0.6.0	2014-08-08 08:20:07 +02:00
Uwe Schindler	28622e49f7	Update forbidden-apis to 1.6	2014-08-05 00:57:18 +02:00
Alexander Reelsen	3c83c0f9d7	Dependencies: Randomized testing version bump	2014-08-04 09:29:16 +02:00
uboness	5f0719fd50	Added CliToolTestCase's sub classes to the test jar	2014-08-03 22:18:33 +02:00
uboness	45714b1977	Added CliToolTestCase to the test jar	2014-08-03 21:10:25 +02:00
uboness	5ccc7beaf4	Added a cli infrastructure CliTool is a base class for command-line interface tools (such as the plugin manager and potentially others). It supports the following: - single or multi command tool - help printing infrastructure (based on help files) - consistent mechanism of parsing arguments (based on commons-cli lib) - separation of argument parsing and command execution (for easier unit testing) - terminal abstraction (will use System.console() when available)	2014-08-02 17:16:27 +02:00
Simon Willnauer	196b56d47a	[BUILD] Add profile that skips all validations This commit adds a profile that skips all validation ie. - nocommit / tabs checking - forbidden API checks - license headers It's not active by default but can easily be activated with `mvn -Pdev` or in the `~/.m2/settings.xml` for reference see: http://maven.apache.org/guides/introduction/introduction-to-profiles.html	2014-07-31 15:22:11 +02:00
Simon Willnauer	d2493ea48a	[CORE] Support parsing lucene minor version strings We parse the version that is shipped with the Lucene segments in order to find the version of lucene that wrote a particular segment. Yet, some lucene version ie: * 4.3.1 (Elasticsearch 0.90.2) * 4.5.1 (Elasticsearch 0.90.7) * 3.6.1 (pre Elasticsearch 0.90.0) wrote illegal strings containing the minor version which causes IAE exceptions being thrown from lucenes parsing method. Closes #7055	2014-07-28 13:02:00 +02:00
Adrien Grand	bf4bdcce73	Build: Remove UnsafeUtils from forbidden-apis exclusion list.	2014-07-23 09:30:51 +02:00
Ryan Ernst	64ab22816c	Scripting: Add script engine for lucene expressions. These are javascript expressions, which can only access numeric fielddata, parameters, and _score. They can only be used for searches (not document updates). closes #6818	2014-07-15 07:49:01 -07:00
Shay Banon	386a14370a	Upgrade to jackson core 2.4.1.1 Note, we had to disable the symbol overflow, since the many mapping case was tripping it closes #6789	2014-07-09 17:49:51 +02:00
Alexander Reelsen	ec92646289	Build: Add all netty classes during shading In order to have access to all codecs and handlers by netty, they need to be exposed during shading, otherwise only the classes, which are used by the built are exposed.	2014-07-07 15:09:36 +02:00
Shay Banon	5093f050ab	Upgrade Jackson to 2.4.1 closes #6757	2014-07-07 09:49:53 +02:00
Simon Willnauer	9ddfaf3aaf	[TEST] Expose `tests.filter` for elasticsearch tests. `-Dtests.filter` allows to pass filter expressions to the elasticsearch tests. This allows to filter test annotaged with TestGroup annotations like @Slow, @Nightly, @Backwards, @Integration with a boolean expresssion like: * to run only backwards tests run: `mvn -Dtests.bwc.version=X.Y.Z -Dtests.filter="@backwards"` * to run all integration tests but skip slow tests run: `mvn -Dtests.filter="@integration and not @slow" * to take defaults into account ie run all test as well as backwards: `mvn -Dtests.filter="default and @backwards" This feature is a more powerful alternative to flags like `-Dtests.nighly=true\|false` etc. Closes #6703	2014-07-03 11:40:49 +02:00
Boaz Leskes	72d2ac1328	Better support for partial buffer reads/writes in translog infrastructure Some IO api can return after writing & reading only a part of the requested data. On these rare occasions, we should call the methods again to read/write the rest of the data. This has cause rare translog corruption while writing huge documents on Windows. Noteful parts of the commit: - A new Channels class with utility methods for reading and writing to channels - Writing or reading to channels is added to the forbidden API list - Added locking to SimpleFsTranslogFile - Removed FileChannelInputStream which was not used Closes #6441 , #6576	2014-07-01 19:11:36 +02:00
Robert Muir	b55ad98d73	Upgrade to Lucene 4.9 (closes #6623 )	2014-06-26 08:18:59 -04:00
Lee Hinman	50bb274efa	Remove MVEL as a built-in scripting language	2014-06-26 10:33:28 +02:00
Lee Hinman	2708e453ac	Re-shade MVEL as a dependency	2014-06-20 11:28:50 +02:00
Lee Hinman	c70f6d0171	Add Groovy as a scripting language, add sandboxing for Groovy Sandboxes the groovy scripting language with multiple configurable whitelists: `script.groovy.sandbox.receiver_whitelist`: comma-separated list of string classes for objects that may have methods invoked. `script.groovy.sandbox.package_whitelist`: comma-separated list of packages under which new objects may be constructed. `script.groovy.sandbox.class_whitelist` comma-separated list of classes that are allowed to be constructed. As well as a method blacklist: `script.groovy.sandbox.method_blacklist`: comma-separated list of methods that are never allowed to be invoked, regardless of target object. The sandbox can be entirely disabled by setting: `script.groovy.sandbox.enabled: false`	2014-06-20 10:20:16 +02:00
Britta Weber	ef05334fdd	don't check .metadata folder for tabs and nocommits	2014-06-17 12:07:13 +02:00
Simon Willnauer	4dfa822e1b	[TEST] Add basic Backwards Compatibility Tests This commit add a basic infrastructure as well as primitive tests to ensure version backwards compatibility between the current development trunk and an arbitrary previous version. The compatibility tests are simple unit tests derived from a base class that starts and manages nodes from a provided elasticsearch release package. Use the following commandline executes all backwards compatiblity tests in isolation: ``` mvn test -Dtests.bwc=true -Dtests.bwc.version=1.2.1 -Dtests.class=org.elasticsearch.bwcompat.* ``` These tests run basic checks like rolling upgrades and routing/searching/get etc. against the specified version. The version must be present in the `./backwards` folder as `./backwards/elasticsearch-x.y.z`	2014-06-16 12:40:43 +02:00
Simon Willnauer	b8537d0e95	[BUILD] exclude target dir from tab validation	2014-06-12 12:31:18 +02:00
Nik Everett	29c10ed1bb	[BUILD] Generate source jars for tests Closes #6125	2014-06-12 12:05:54 +02:00
Simon Willnauer	5575ba1a12	[BUILD] Check for tabs and nocommits in the code on validate This commit adds checks for nocommit and tabs in the source code. The task is executed during the validate phase and can be disabled via `-Dvalidate.skip`	2014-06-12 11:11:23 +02:00
mikemccand	a71bb13563	Compilation: don't warn about using Sun proprietary APIs E.g. we use Unsafe in quite a few places and this generates lots of warnings, which we now suppress using the undocumented -XDignore.symbol.file command-line option to javac. Closes #6423	2014-06-06 05:46:06 -04:00
mikemccand	af30947b66	get -Dtests.verbose passing through Maven	2014-06-05 09:38:10 -04:00
Adrien Grand	7ab99de483	Routing: Restore shard routing. Routing has been inadvertly changed in #5562 resulting in documents going to different shards in 1.2. This is a terrible bug because an indexing request would not necessarily go to the same shard anymore, potentially leading to duplicates. Close #6391	2014-06-03 16:37:54 +02:00
Simon Willnauer	ea80b381c0	[BUILD] Java version line was missleading calling `java -version` might not print the java version that is actually used to run maven. This commit prints a more accurate version	2014-06-02 15:51:38 +02:00
Shay Banon	b2c7c8b0e7	Upgrade to netty 3.9.1 closes #6331	2014-05-30 00:20:37 +02:00
Simon Willnauer	57316bbfb3	Remove obsolet ASF repository Lucene 4.8.1 is on maven central	2014-05-19 21:17:33 +02:00
Simon Willnauer	85a0b76dbb	Upgrade to Lucene 4.8.1 This commit upgrades to the latest Lucene 4.8.1 release including the following bugfixes: * An IndexThrottle now kicks in when merges start falling behind limiting index threads to 1 until merges caught up. Closes #6066 * RateLimiter now kicks in at the configured rate where previously the limiter was limiting at ~8MB/sec almost all the time. Closes #6018	2014-05-19 20:47:55 +02:00
David Pilato	0dbc83e7b0	[TEST] Do not filter gz files	2014-05-16 15:23:09 +02:00
David Pilato	bd871f96c2	Check that a plugin is Lucene compatible with the current running node using `lucene` property in `es-plugin.properties` file. * If plugin does not provide `lucene` property, we consider that the plugin is compatible. * If plugin provides `lucene` property, we try to load related Enum org.apache.lucene.util.Version. If this fails, it means that the node is too "old" comparing to the Lucene version the plugin was built for. * We compare then two first digits of current node lucene version against two first digits of plugin Lucene version. If not equal, it means that the plugin is too "old" for the current node. Plugin developers who wants to launch plugin check only have to add a `lucene` property in `es-plugin.properties` file. If you are using maven to build your plugin, you can do it like this: In `pom.xml`: ```xml <properties> <lucene.version>4.6.0</lucene.version> </properties> <build> <resources> <resource> <directory>src/main/resources</directory> <filtering>true</filtering> </resource> </resources> </build> ``` In `es-plugin.properties`, add: ```properties lucene=${lucene.version} ``` BTW, if you don't already have it, you can add the plugin version as well: ```properties version=${project.version} ``` You can disable that check using `plugins.check_lucene: false`.	2014-05-16 13:41:20 +02:00
Simon Willnauer	13f37b3800	Shade mustache into org.elasticsearch.common package Previously we shared the jar but never rewrote the packages such that the shading had no effect. Closes #6192	2014-05-15 21:21:36 +02:00
Simon Willnauer	e47de1f809	[TEST] Randomize number of available processors We configure the threadpools according to the number of processors which is different on every machine. Yet, we had some test failures related to this and #6174 that only happened reproducibly on a node with 1 available processor. This commit does: * sometimes randomize the number of available processors * if we don't randomize we should set the actual number of available processors in the settings on the test node * always print out the num of processors when a test fails to make sure we can reproduce the thread pool settings with the reproduce info line Closes #6176	2014-05-15 12:24:53 +02:00
Adrien Grand	9425472f61	[BUILD] Shade t-digest.	2014-05-14 00:42:48 +02:00
Adrien Grand	cc530b9037	Use t-digest as a dependency. Our improvements to t-digest have been pushed upstream and t-digest also got some additional nice improvements around memory usage and speedups of quantile estimation. So it makes sense to use it as a dependency now. This also allows to remove the test dependency on Apache Mahout. Close #6142	2014-05-13 10:38:08 +02:00
David Pilato	645efa05df	Update shade-plugin to 2.3 Shade-plugin 2.2 does not work with JDK8 (see http://jira.codehaus.org/browse/MSHADE/fixforversion/19828)	2014-05-12 10:23:49 +02:00
Martijn van Groningen	e2a2f13f17	Added FilteredQuery to the list of forbidden apis	2014-05-08 09:54:10 +02:00
Holger Hoffstätte	f5c9bf6f0f	Update JNA to latest version Updating to this version allows to configure a special JNA directory, in case the /tmp directory is mounted with the noexec option, as JNA extracts some data and tries to execute parts of it. Also updated documentation to clarify mlockall and memory settings as well as pointing to the new jna.tmpdir system property. Closes #5493	2014-05-02 11:52:57 +02:00
Matt Weber	2663d04a96	Run tests through forbidden-apis.	2014-04-30 17:48:33 +02:00
Shay Banon	2076194d8f	Upgrade to Jackson 2.3.3 fixes the long value bug as well...	2014-04-29 20:13:43 -04:00
Robert Muir	8e0a479316	Upgrade to Lucene 4.8 Closes #5932	2014-04-28 06:45:50 -04:00
Shay Banon	6899e642b5	Upgrade to Guava 17 closes #5953	2014-04-28 11:02:30 +02:00
javanna	9a68e60142	[TEST] Allow to disable randomization of shards and replicas via system property Needed for REST backwards compatibility tests, since we need to run older tests with the latest runner, which randomizes shards and replicas, but the tests rely on defaults (5,1). Done in a generic way based on compatibility versions e.g. `-Dtests.compatibility=1.0.0` allows to run tests in a special manner that is compatibile with 1.0.0 version. Also moved back randomIndexTemplate to ElasticsearchIntegrationTest (from ImmutableCluster) where all the randomized aspects should be. Closes #5897	2014-04-24 22:18:31 +02:00
javanna	918da65d35	[TEST] Added blacklist to be able to skip specific REST tests The blacklist can be provided through -Dtests.rest.blacklist and supports a comma separated list of globs e.g. -Dtests.rest.blacklist=get/10_basic/,index//* Also added some missing docs and made it clearer that the suite/test descriptions effectively contains their (relative) path (api/yaml_file/test section) Closes #5881	2014-04-22 09:52:48 +02:00
Uwe Schindler	a6764e24a4	Update forbidden-apis to 1.5.1 and remove the relaxed missing classes error.	2014-04-19 08:53:45 +02:00
Shay Banon	21a3667888	Upgrade to mvel 2.2.0.Final closes #5877	2014-04-18 21:52:24 +02:00
Simon Willnauer	eaf04fc828	[BUILD] Fail release if test have AwaitsFix annotations Note this commit upgrades Forbidden APIs to version 1.5 Closes #5807	2014-04-17 16:45:46 +02:00
Shay Banon	bc5bdbc5de	Remove jsr166y now that we on Java 7, cleanup jsr166e to classes we use	2014-04-15 13:17:28 +02:00
Simon Willnauer	8bede7024f	Use TransportBulkAction for internal request from IndicesTTLService This prevents executing bulks internal autocreate indices logic and ensures that this internal request never creates an index automaticall. This fixes a bug where the TTL purger thread ran after the actual index it was purging was already closed / deleted and that re-created that index. Closes #5766	2014-04-15 12:40:25 +02:00
Simon Willnauer	754eb16835	Upgrade to Lucene 4.7.2	2014-04-14 18:30:03 +02:00
Bill Hwang	dcc6a6e138	[BUILD] Remove site dependencies generation Removed dependencies library generation. It stalls Jenkins static analysis job	2014-04-07 16:01:02 -07:00
javanna	1ec4f8f04b	[TEST] Replaced RestTestSuiteRunner with parametrized test that uses RandomizedRunner directly ElasticsearchRestTests extends now ElasticsearchIntegrationTest and makes use of our ordinary test infrastructure, in particular all randomized aspects now come for free instead of having to maintain a separate (custom) tests runner We previously parsed only the tests that needed to be run given the version of the cluster the tests are running against. This doesn't happen anymore as it didn't buy much and it would be harder to support as the tests get now parsed before the test cluster gets started. Thus all the tests are now parsed regardless of their skip sections, afterwards the ones that don't need to be run will be skipped through assume directives. Fixed REST tests that rely on a specific number of shards as this change introduces also random number of shards and replicas (through randomIndexTemplate) Closes #5654	2014-04-07 17:08:05 +02:00
Simon Willnauer	6f5b7fa086	[BUILD] Set -Dtests.jvms=auto by default to make use of multiple JVMs	2014-04-03 13:01:01 +02:00
Simon Willnauer	42b20d601f	Upgrade to Lucene 4.7.1 * Removed XTermsFilter fixed in LUCENE-5502 * Switched back to automaton queries that caused failures due to LUCENE-5532 * Fixed Highlight test that has different results due to LUCENE-5538	2014-04-01 23:50:55 +02:00
javanna	806c4e87fb	[TEST] Removed last occurences of cluster_seed, no longer used Relates to #5233	2014-04-01 17:57:57 +02:00
Simon Willnauer	71de2bc414	[BUILD] Allow to set tests memory via the commandline	2014-04-01 14:12:52 +02:00
javanna	38dd501ab5	[TEST] added ExternalTestCluster that allows to run tests against an external cluster All the ordinary test operations happen based on the ImmutableTestCluster base class and are executed via transport client. Will be used especially for the REST tests once migrated to the standard randomized runner. Added new httpAddresses method to ImmutableTestCluster to be able to retrieve the http addresses to connect to for the REST tests. Both versions will look inside the cluster to figure out which nodes are available for http calls and their addresses. The external cluster is used as global cluster if the tests.cluster system property is available. The property needs to contain a comma separated list of available elasticsearch nodes that will be used to connect to the cluster (e.g. localhost:9300,localhost:9301). Only a subset of the integration tests can currently be run successfully against the external cluster, for more precision the ones that don't modify the cluster layout (don't require cluster() functionalities but rely only on immutableCluster()). Also at least two data nodes are required otherwise the ensureGreen calls cannot succeed. Closes #5630	2014-04-01 11:45:35 +02:00
Kevin	e78bbbf3ec	add CBOR data format support	2014-03-28 20:30:39 +11:00
Simon Willnauer	f1b32c4636	[Build] use the same execution hint file across the pom file	2014-03-27 16:28:22 +01:00
Simon Willnauer	11b51b1780	[Build] use branch version in execution hit file name	2014-03-27 15:32:53 +01:00
Adrien Grand	b5b82626e7	Forbid Math.abs(int/long). We have had a couple of bugs because of the use of these methods without paying attention that it might return a negative value when provided with MIN_VALUE. There is one common and legitimate usage of this method in order to perform a modulo operation which would always return a positive number. This use-case has been extracted to MathUtils.mod. Close #5562	2014-03-27 14:50:43 +01:00
Robert Muir	5babf59813	Update to forbidden-apis 1.4.1 Closes #5492	2014-03-22 15:00:19 -04:00
Simon Willnauer	876b2592ac	[Build] Skip topN hints if tests are skipped	2014-03-20 12:45:35 +01:00
Simon Willnauer	7c494461f3	[BUILD] Add top execution hints to test phase Added a ant task that prints the Top N most expensive tests after each test run.	2014-03-20 12:29:48 +01:00
David Smiley	644fdfc4aa	Upgrade to Spatial4j 0.4.1 and JTS 1.13 Fixes #5279	2014-03-20 11:17:51 +01:00
Bill Hwang	fe487373e6	Revert "Findbug warning supression" This reverts commit `744eabad03`.	2014-03-17 13:55:39 -07:00
Bill Hwang	744eabad03	Findbug warning supression Added logic to enable findbug warnings supression via annotations	2014-03-17 13:35:37 -07:00
Dridi Boukelmoune	9500dddad3	Move systemd files from /etc to /usr/lib As documented in systemd's manual pages tmpfiles.d(5) and systemd.unit(5), a package should install its default configuration in /usr/lib, which can be overriden by system administrators in /etc. New locations in the rpm: /usr/lib/systemd/system/elasticsearch.service /usr/lib/tmpfiles.d/elasticsearch.conf	2014-03-17 14:06:34 +01:00
David Pilato	c6915ef4d6	Enforce java version 1.7 When building elasticsearch, we now require to use java 1.7. Maven will check that before compiling any class. If Java version is incorrect, you will get the following message: ``` [WARNING] Rule 0: org.apache.maven.plugins.enforcer.RequireJavaVersion failed with message: Detected JDK Version: 1.6.0-65 is not in the allowed range [1.7,). ``` Closes #5428.	2014-03-17 08:43:33 +01:00
Simon Willnauer	821173b5cf	Enforce query instance checking before it wrapper as a filter We have the default QueryWrapperFilter as well as our custom one while our wrapper is explicitly marked as no_cache such that it will never be included in a cache. This was not consistenly used and caused several problems during tests where p/c related queries were used as filters and ended up in the cache. This commit adds the QueryWrapperFilter ctor to the forbidden APIs to enforce the query instance checks.	2014-03-14 20:18:01 +01:00
javanna	d80dd00424	upgrade randomized-testing to 2.1.1 Note that the standard `atLeast` implementation has now Integer.MAX_VALUE as upper bound, thus it behaves differently from what we expect in our tests, as we never expect the upper bound to be that high. Added our own `atLeast` to `AbstractRandomizedTest` so that it has the expected behaviour with a reasonable upper bound. See https://github.com/carrotsearch/randomizedtesting/issues/131	2014-03-14 11:47:00 +01:00
Martijn van Groningen	6f8f773f8c	Disabled query size estimation in percolator, because this is too expensive cpu wise. Lucene's RamUsageEstimator.sizeOf(Object) is to expensive. Query size estimation will be enabled when a cheaper way of query size estimation can be found. Closes #5372 Relates to #5339	2014-03-14 15:29:24 +07:00
Bill Hwang	2e56253293	Added static analysis profile to pom.xml Added pmd, findbug as well as site generation logic to top pom.xml file Created customized pmd ruleset	2014-03-13 12:23:07 -07:00
Adrien Grand	e3b87926bf	[Build] Remove XReferenceManager and XSearcherManager from forbidden-apis' exclude list. These classes have been removed on the upgrade to Lucene 4.7.	2014-03-12 15:06:39 +01:00
Igor Motov	39d2377be6	Use patched version of TermsFilter to prevent using wrong cached results See LUCENE-5502 Closes #5363	2014-03-11 20:48:22 -04:00
Shay Banon	8cfff9d796	jackson: upgrade to 2.3.2	2014-03-09 23:40:43 +01:00
Shay Banon	3ea746c45e	mvel: upgrade to 2.1.9	2014-03-09 22:56:09 +01:00
Shay Banon	992747a159	Force merges to not happen when indexing a doc / flush Today, even though our merge policy doesn't return new merge specs on SEGMENT_FLUSH, merge on the scheduler is still called on flush time, and can cause merges to stall indexing during merges. Both for the concurrent merge scheduler (the default) and the serial merge scheduler. This behavior become worse when throttling kicks in (today at 20mb per sec). In order to solve it (outside of Lucene for now), we wrap the merge scheduler with an EnableMergeScheduler, where, on the thread level, using a thread local, the call to merge can be enabled/disabled. A Merges helper class is added where all explicit merges operations should go through. If the scheduler is the enabled one, it will enable merges before calling the relevant explicit method call. In order to make sure Merges is the only class that calls the explicit merge calls, the IW variant of them is added to the forbidden APIs list. closes #5319	2014-03-05 12:26:26 +00:00
Zachary Tong	7b16c5857d	Percentiles aggregation. A new metric aggregation that can compute approximate values of arbitrary percentiles. Close #5323	2014-03-03 18:06:14 +01:00
Simon Willnauer	8ceb98752d	Move master to Java 1.7 Closes #5267	2014-02-27 15:12:02 +01:00
Simon Willnauer	30d7b8de2f	Upgrade to Lucene 4.7 Closes #5104 Closes #5129 Closes #3757	2014-02-26 22:21:10 +01:00
Bill Hwang	db57f7ed0e	Add thrid party license generation profile	2014-02-20 10:05:10 -08:00
Isabel Drost-Fromm	48004ff8a5	Add mustache templating to query execution. Adds support for storing mustache based query templates that can later be filled with query parameter values at execution time. Templates may be both quoted, non-quoted and referencing templates stored in config/scripts/*.mustache by file name. See docs/reference/query-dsl/queries/template-query.asciidoc for templating examples. Implementation detail: mustache itself is being shaded as it depends directly on guava - so having it marked optional but included in the final distribution raises chances of version conflicts downstream. Fixes #4879	2014-02-20 12:21:59 +01:00
mrsolo	dffc7cd06d	code coverage hookup	2014-02-14 10:05:40 -08:00
Simon Willnauer	ab97bf0fd9	s/\t/ /	2014-02-07 13:35:58 +01:00
Simon Willnauer	0fb8d982be	Use patched version of ReferenceManager to prevent infinite loop in RefrenceManager#accquire() See LUCENE-5436	2014-02-06 21:45:38 +01:00
Simon Willnauer	9cf8251a0d	Add RamUsageEstimator#sizeOf(Object) to forbidden APIs This method can be a performance trap since it traverse the entire object tree that is referenced by the provided object. See LUCENE-5373	2014-01-31 21:43:20 +01:00
Simon Willnauer	91acca7836	Upgrade to Lucene 4.6.1 This upgrade includes a fix for RAM estimation on IndexReader that allows to expose the amount of used bytes per segment now as a setting in Elasticsearch. (LUCENE-5373) Additionally this bugfix release contained a small fix for highlighting that was already ported to Elasticsearch when reported (LUCENE-5361) Closes #4897	2014-01-28 10:35:39 +01:00
Simon Willnauer	25b49dd50b	Install SecurityManager inside ElasticsearchTestCase for easier randomization We currently run always with SecurityManager installed. To make sure we work also without we should randomly swap it out ie. run without the security manager.	2014-01-24 22:18:05 +01:00
Simon Willnauer	55e8df40cd	Move provided lucene-expression jar below the test-framework Having a dep that refrences lucene-core before the test framework confuses some runners.	2014-01-23 16:10:57 +01:00
Simon Willnauer	416e328cea	Mark 'lucene-expression' as 'provided' in pom.xml We currently pull in the lucene-expression module that is referenced by lucene-suggest. Yet, we don't make use of this dependency at all and it pulls in a bunch of unshaded libs like `antlr` and `asm` which are pretty common in other projects. We should exclude this dependency since we don't use it at all and it causes problems when Elasticsearch is used as a node client. (see #4858) If we mark the dependency as provided it won't be included in the distribution. Closes #4859 Closes #4858	2014-01-23 14:23:41 +01:00
Simon Willnauer	53192919c6	Move to [2.0] snap	2014-01-21 17:07:39 +01:00
Shay Banon	f1174eac3a	upgrade to guava 16.0 also fixes #4830	2014-01-21 15:26:03 +01:00
Simon Willnauer	7f51fbc5ab	Add SecurityManger / policy when running tests. This commit adds a security manager to the test JVMs that prevents mainly writing files outside of the JVMs current test directory.	2014-01-17 15:15:10 +01:00
Simon Willnauer	85ca6c6762	move to [1.0.0] SNAP	2014-01-16 10:28:38 +01:00
Alexander Reelsen	c6155c5142	release [1.0.0.RC1]	2014-01-15 17:02:22 +00:00
David Pilato	c386155a73	add MockPageCacheRecycler in test jar MockPageCacheRecycler is missing in test jar which makes failing tests when using test jar in plugins: ``` 1> [2014-01-11 10:51:30,531][ERROR][test ] FAILURE : testWikipediaRiver(org.elasticsearch.river.wikipedia.WikipediaRiverTest) 1> REPRODUCE WITH : mvn test -Dtests.seed=5DAFD4FBAE587363 -Dtests.class=org.elasticsearch.river.wikipedia.WikipediaRiverTest -Dtests.method=testWikipediaRiver -Dtests.prefix=tests -Dtests.network=true -Dfile.encoding=MacRoman -Duser.timezone=Europe/Paris -Des.logger.level=INFO -Des.node.local=true -Dtests.cluster_seed=134842C2D806FFC0 1> Throwable: 1> java.lang.NoClassDefFoundError: org/elasticsearch/cache/recycler/MockPageCacheRecycler 1> org.elasticsearch.test.cache.recycler.MockPageCacheRecyclerModule.configure(MockPageCacheRecyclerModule.java:30) 1> org.elasticsearch.common.inject.AbstractModule.configure(AbstractModule.java:60) ``` Closes #4694.	2014-01-13 09:06:59 +01:00
Shay Banon	ffe3537285	no need for the mvn repository anymore	2014-01-12 00:26:34 +01:00
mrsolo	2df42e4460	Added licene-maven-plugin to build The plugin adds basic checks for license headers. At this stage we only enable checks for java source files. Other files are omitted at this point.	2014-01-10 22:04:48 +01:00
Shay Banon	3262398347	give the compiler more memory, otherwise it fails sometimes	2014-01-09 11:25:59 +01:00
Shay Banon	0eaed0da26	add joda-convert so missing annotations in joda-time will not cause failures when used as dependency fixes #4660	2014-01-08 22:14:44 +01:00
Simon Willnauer	fa16969360	Cleanup comments and class names s/ElasticSearch/Elasticsearch * Clean up s/ElasticSearch/Elasticsearch on docs/* * Clean up s/ElasticSearch/Elasticsearch on src/* bin/* & pom.xml * Clean up s/ElasticSearch/Elasticsearch on NOTICE.txt and README.textile Closes #4634	2014-01-07 11:21:51 +01:00
Adrien Grand	4271d573d6	Page-based cache recycling. Refactor cache recycling so that it only caches large arrays (pages) that can later be used to build more complex data-structures such as hash tables. - QueueRecycler now takes a limit like other non-trivial recyclers. - New PageCacheRecycler (inspired of CacheRecycler) has the ability to cache byte[], int[], long[], double[] or Object[] arrays using a fixed amount of memory (either globally or per-thread depending on the Recycler impl, eg. queue is global while thread_local is per-thread). - Paged arrays in o.e.common.util can now optionally take a PageCacheRecycler to reuse existing pages. - All aggregators' data-structures now use PageCacheRecycler: - for all arrays (counts, mins, maxes, ...) - LongHash can now take a PageCacheRecycler - there is a new BytesRefHash (inspired from Lucene but quite different, still; for instance it cheats on BytesRef comparisons by using Unsafe) that also takes a PageCacheRecycler Close #4557	2014-01-06 19:02:00 +01:00
Simon Willnauer	e7a84d744a	Add ability to run certain packages with assertions disabled Test can be run with `-Dtests.assertion.disabled=org.elasticsearch` to run the tests without assertions to make sure assertions don't hide any assignements etc. that introduce bugs in production.	2013-12-30 19:36:02 +01:00
David Pilato	b29f89f7f9	We run PluginManagerTests using only node client. We also add some debug logs and fix `tests.network` (setting it to true was not working from jenkins)	2013-12-30 15:40:52 +01:00
David Pilato	7694f0b7a0	Increase MaxPermSize to 128m for tests	2013-12-30 09:56:47 +01:00
Simon Willnauer	11c4218566	Start Test nodes sometimes without mock modules We are mocking out some functionality to add assertions etc. or randomize store types. We should randomly run with our defaults to make sure we don't hide any potential problems.	2013-12-29 00:50:10 +01:00
Luca Cavanna	63cbc84393	removed rest-spec submodule and prepared project for same files added directly to the codebase (no submodule) within rest-api-spec (temporarily disabled FileUtilsTests & REST tests as there's temporarily no rest-spec dir) Relates to #4540 #4376	2013-12-27 20:36:12 +01:00
Shay Banon	e5f52ce778	update to netty 3.9.0	2013-12-23 20:06:45 +01:00
Luca Cavanna	d97a00d4a7	added REST test suites runner The REST layer can now be tested through tests that are shared between all the elasticsearch official clients. The tests are based on REST specification that can be found on the elasticsearch-rest-api-spec project and consist of YAML files that describe the operations to be executed and the obtained results that need to be tested. REST tests can be executed through the ElasticsearchRestTests class, which relies on the rest-spec git submodule that contains the rest spec and tests pulled from the elasticsearch-rest-spec-api project. The rest-spec submodule gets automatically initialized and updated through maven (generate-test-resources phase). The REST runner and the needed classes are distributed within the test artifact. The following are the options supported by the REST tests runner: - tests.rest[true\|false\|host:port]: determines whether the REST tests need to be run and if so whether to rely on an external cluster (providing host and port) or fire a test cluster (default) - tests.rest.suite: comma separated paths of the test suites to be run (by default loaded from /rest-spec/test classpath). it is possible to run only a subset of the tests providing a sub-folder or even a single yaml file (the default /rest-spec/test prefix is optional when files are loaded from classpath) e.g. -Dtests.rest.suite=index,get,create/10_with_id - tests.rest.spec: REST spec path (default /rest-spec/api from classpath) - tests.iters: runs multiple iterations - tests.seed: seed to base the random behaviours on - tests.appendseed[true\|false]: enables adding the seed to each test section's description (default false) - tests.cluster_seed: seed used to create the test cluster (if enabled) Closes #4469	2013-12-17 15:36:16 +01:00
Simon Willnauer	8f85d63b67	Enable incremental compilation using a workaround for the maven-compiler-plugin 3.1 bug	2013-12-14 21:56:01 +01:00
Shay Banon	59d85bc22a	upgrade to Jackson 2.3.0	2013-12-14 00:53:52 +01:00
Alexander Reelsen	81e13a870b	Packaging: Ensure setting of sysctl vm.max_map_count In order to be sure that memory mapped lucene directories are working one can configure the kernel about how many memory mapped areas a process may have. This setting ensure for the debian and redhat initscripts as well as the systemd startup, that this setting is set high enough. Closes #4397	2013-12-11 09:19:22 +01:00
Simon Willnauer	2b42a0f94a	Override DefaultExceptionHandler to filter out certain exceptions We have the situation that some tests fail since they don't handle EsRejectedExecutionException which gets thrown when a node shuts down. That is ok to ignore this exception and not fail. We also suffer from OOMs that can't create native threads but don't get threaddumps for those failures. This patch prints the thread stacks once we catch a OOM which can' create native threads.	2013-12-04 14:18:52 +00:00
mrsolo	3494ac252e	removing test.jvm.argline declaration <properties> declaration was overwriting environment setting.	2013-12-02 11:44:51 -08:00
Simon Willnauer	2c8ee3fbbe	Moving to 1.0.0RC1 snap	2013-12-02 17:10:07 +01:00
Kevin Kluge	296cfbe390	release [1.0.0.Beta2]	2013-12-02 15:45:30 +00:00
David Pilato	fa762f09fa	OOM when building with java6 The maven-compiler-plugin upgrade from 2.3.2 to 3.1 (see #4279) could cause out of memory issue when building the project with Maven and JDK6 and default memory settings (no `MAVEN_OPTS`). This issue does not appear with JDK7.	2013-11-28 18:22:29 +01:00
David Pilato	baaa1a6aa2	Upgrade Maven Surefire Plugin to 2.16 Closes #4275.	2013-11-28 14:56:51 +01:00
David Pilato	90c3fce4fe	Upgrade Maven Source Plugin to 2.2.1 Closes #4276.	2013-11-28 14:56:34 +01:00
David Pilato	975c43f6d5	Upgrade RPM Maven Plugin to 2.1-alpha-3 Closes #4282.	2013-11-28 14:56:20 +01:00
David Pilato	d18f0a28d9	Upgrade Maven Resources Plugin to 2.6 Closes #4280.	2013-11-28 14:56:07 +01:00
David Pilato	fa0a5e5844	Upgrade Maven Jar Plugin to 2.4 Closes #4281.	2013-11-28 14:55:53 +01:00
David Pilato	2b32121e34	Upgrade Maven Eclipse Plugin to 2.9 Closes #4277.	2013-11-28 14:55:40 +01:00
David Pilato	eac602ec70	Upgrade Maven Dependency Plugin to 2.8 Closes #4274.	2013-11-28 14:52:58 +01:00
David Pilato	b1c62b02b0	Upgrade Maven Compiler Plugin to 3.1 Closes #4279.	2013-11-28 14:52:32 +01:00
David Pilato	79d92dffa6	Upgrade Maven Assembly Plugin to 2.4 Closes #4278.	2013-11-28 14:49:51 +01:00
David Pilato	5ec0bf32d6	Remove randomizedtesting-runner explicit dependency Was added by mistake with #4266	2013-11-28 12:07:21 +01:00
David Pilato	a9655ba812	Update to shade plugin 2.2 to shade test artifact as well When we want to use test artifact in other projects, dependencies are not shaded as for core artifact. Issue opened in maven shade project: [MSHADE-158](http://jira.codehaus.org/browse/MSHADE-158) When using it in other projects, you basically need to change your `pom.xml` file: ```xml <dependency> <groupId>org.elasticsearch</groupId> <artifactId>elasticsearch</artifactId> <version>${elasticsearch.version}</version> <type>test-jar</type> <scope>test</scope> </dependency> ``` You can also define some properties: ```xml <properties> <tests.jvms>1</tests.jvms> <tests.shuffle>true</tests.shuffle> <tests.output>onerror</tests.output> <tests.client.ratio></tests.client.ratio> <es.logger.level>INFO</es.logger.level> </properties> ``` Closes #4266	2013-11-28 10:54:48 +01:00
Simon Willnauer	5f7146aab8	Run tests with reduced stack size as we run in production	2013-11-26 11:16:02 +01:00
Simon Willnauer	cf3ba7c51c	Only use -Dtests.jvm.argline instead of numbered options	2013-11-26 11:11:19 +01:00
Simon Willnauer	46ab6a1533	Include JVM Arg line in reproduce line	2013-11-25 14:35:34 +01:00
Simon Willnauer	8e17d636ef	Upgrade to Lucene 4.6 This commit upgrades to Lucene 4.6 and contains the following improvements: * Remove XIndexWriter in favor of the fixed IndexWriter * Removes patched XLuceneConstantScoreQuery * Now uses Lucene passage formatters contributed from Elasticsearch in PostingsHighlighter * Upgrades to Lucene46 Codec from Lucene45 Codec * Fixes problem in CommonTermsQueryParser where close was never called. Closes #4241	2013-11-24 21:08:38 +01:00
Simon Willnauer	eb55458e44	Upgrade RandomizedRunner Maven Plugin to 2.0.14	2013-11-22 15:00:12 +01:00
Simon Willnauer	a949a8056b	Upgrade to forbiddenapis 1.4	2013-11-22 08:01:30 +01:00
Simon Willnauer	00562d3a6f	allow passing JVM args via env variables if not defined via the cmd	2013-11-21 17:58:43 +01:00
Simon Willnauer	02522dac06	Allow passing additional parameters to the test JVM like GC settings etc.	2013-11-21 17:39:11 +01:00
colings86	76c5f53dfa	Changed pom to allow import and running from eclipse Currently when importing projects into eclipse you need to run 'mvn eclipse:eclipse' on the command line to generate the poject files. This means that when the pom changes you need to re-run the command on the command line to reflect those changes in the project in eclipse. This commit allows the developer to import the project as an existing maven project (can be shared using git after import) and then allows the application to be run inside eclipse using the .launch file in /dev-tools enabling easy debugging of the application within eclipse without requiring a maven build.	2013-11-21 10:56:52 +01:00
Simon Willnauer	c48c8fd974	Include inner classes in the test package as well	2013-11-14 13:49:37 +01:00
Simon Willnauer	16ee742682	Cleanup test framework in order to release it as a jar file This commit adds javadocs and removed unused methods from central classes like ElasticsearchIntegrationTest. It also changes visibility of many methods and classes that are only needed inside the test infrastructure.	2013-11-12 11:54:55 +01:00
Martijn van Groningen	9fdcb0ad05	Upgraded to hppc version 0.5.3	2013-11-11 14:27:52 +01:00
Shay Banon	b17732ea56	upgrade to Netty 3.8.0	2013-11-09 15:41:24 +01:00
Simon Willnauer	985916038e	Just filter test jar and not the main jar file	2013-11-08 23:14:56 +01:00
Simon Willnauer	a824f69d21	Package test-framework as individual jar This commit causes all classes under 'org.elasticsearch.test.*' to be included in a 'elasticsearch-${version}-test.jar' that can be inclued by 3rd party projects or plugins via: ``` <dependency> <groupId>org.elasticsearch</groupId> <artifactId>elasticsearch</artifactId> <version>${elasticsearch.version}</version> <type>test-jar</type> <scope>test</scope> </dependency> ```	2013-11-08 15:08:31 +01:00
Simon Willnauer	c946a69e28	Set MaxDirectMemorySize for tests to have consistent upper bounds	2013-11-07 09:44:19 +01:00
Simon Willnauer	e1b6988886	move to [1.0.0.Beta2] SNAP	2013-11-06 16:19:50 +01:00
Simon Willnauer	77bc5d5ecf	release [1.0.0.Beta1]	2013-11-06 15:32:43 +01:00
Shay Banon	75eda6e957	upgrade randomized testing to 2.0.12	2013-10-31 10:02:03 +01:00
Simon Willnauer	3a34aa735e	Upgrade to Lucene 4.5.1	2013-10-24 18:37:44 +02:00
Simon Willnauer	0eea6e8183	Enable Random TransportClients in tests	2013-10-10 13:23:02 +02:00
Adrien Grand	6b02611971	Update Lucene to version 4.5.0.	2013-10-08 15:44:03 +02:00
Martijn van Groningen	088e05b368	Migrate from Trove to Hppc.	2013-10-03 23:10:57 +02:00
Simon Willnauer	8f087e802d	Print Java Version during Validate	2013-09-27 14:39:07 +02:00
Simon Willnauer	5a365795ec	Added support for dynamic transport client ratio in tests. Currently tests only run with node clients but eventually we want to run all tests with randomly choosen node / transport clients. To enable this during development and on test servers as a transition phase this commit adds the ability to allow a fraction of the clients used in tests to be transport clients. By default this is still disabled. To enable transport clients in tests pass '-Dtests.client.ratio=[0..1]' where '1.0' will force transport clients and '0.0' completely disables them. If an empty string is passed the ratio is chosen at random for each test.	2013-09-22 22:25:30 +02:00
Simon Willnauer	cabbf7805b	Make TestCluster based integration tests more repoducible While testing an async system providing reproducible tests that use randomized components is a hard task we should at least try to reestablish the enviroment of a failing test as much as possible. This commit allows to re-establish the shared 'TestCluster' by resetting the cluster to a predefined shared state before each test. Before this commit a tests that is executed in isolation was likely using a entirely different node enviroment as the failing test since the 'TestCluster' kept intermediate nodes started by other tests around.	2013-09-17 23:07:29 +02:00
Shay Banon	e110d53b0c	change default number of test jvms to 1 the auto default can cause a lot of pressure on a machine when running the tests (each process actually runs a multi node cluster)	2013-09-17 21:30:27 +02:00
Costin Leau	dcc45070bd	Merge pull request #3702 from costin/master add elasticsearch as a service for Windows platforms	2013-09-17 05:16:29 -07:00
Costin Leau	08bf131899	rework script to handle path with spaces use service id for pid name disable filtering on *.exe (caused corruption) rename exe names and add more options to .bat start/stop operations are now supported (and expected to be called) by service.bat add more variables from the env to customize default behavior prior to installing the service add manager option fixes regarding batch flow specify service id in description minor readability improvement include .exe only in ZIP archive rename x64 service id to make it work out of the box add elasticsearch as a service for Windows platforms based on Apace Commons Daemon supports both x64 and x86	2013-09-17 15:01:09 +03:00
Shay Banon	ae0489266e	Add `node.mode` with `local` or `network` Compared to setting node.local to true, would be nicer to support node.mode with values of local or network. Note, node.local is still supported. closes #3713	2013-09-17 00:34:10 +02:00
Shay Banon	88173d762a	upgrade to jackson 2.2.3	2013-09-16 15:38:47 +02:00
Shay Banon	31097691af	upgrade to netty 3.7.0	2013-09-16 15:05:28 +02:00
Shay Banon	cd90382964	upgrade to google guava v 15	2013-09-16 09:14:00 +02:00
Simon Willnauer	0e936c99e3	Add -Dtests.integration support to pom.xml	2013-09-11 22:16:43 +02:00
Simon Willnauer	e68303d6a6	Allow test output to be configurable via CMD	2013-09-04 12:46:43 +02:00
Adrien Grand	5b6be0c456	Use a separate build directory for Eclipse. The fact that Maven and Eclipse share the same build directories can trigger race conditions when both are trying to build at the same time, eg. if you run `mvn clean test` while Eclipse is up and running: Eclipse will notice that some class files are missing and start compiling in parallel with Maven.	2013-08-29 10:29:26 +02:00
Shay Banon	f3a35ccc90	upgrade to joda 2.3	2013-08-24 00:07:57 +02:00
Boaz Leskes	9869427ef6	Allow to configure root logging level using system properties. Ex. -Des.logger.level=DEBUG . Defaults to INFO as before.	2013-08-15 14:36:00 +02:00
Simon Willnauer	ba13930b32	Fix test include pattern to include Test.class We missed Test.class which is not our convention but we could miss some tests. We should better include the *Test.class tests as well.	2013-08-13 17:46:48 +02:00
Simon Willnauer	59be83f9fc	Remove accidentially committed default values `-Dtests.maxFailures` and `-Dtests.failfast` should not be enabled by default.	2013-08-12 14:32:04 +02:00
Simon Willnauer	dbed36a13f	Added support for `tests.failfast` and `tests.maxiters` This commit adds support for failing fast when running a test case with `-Dtests.iters=N` and uses some goodness from LuceneTestCase in a new base `AbstractRandomizedTest`. This class checks among other things if a tests doesn't call `super.setup` / `super.tearDown` when it should do and checks if a large static resources are not cleaned up after the tests ie. a running node.	2013-08-12 13:18:56 +02:00
Alexander Reelsen	68b77c1ae3	Included only runtime dependencies when copying This makes sure, that no test dependencies are placed in the distribution	2013-08-06 15:13:25 +02:00
Shay Banon	976152b23e	enable sys out check + remove unused code	2013-07-26 18:57:21 +02:00
Shay Banon	70bbcb4c48	Add Git build info when we build a distribution closes #3370	2013-07-24 20:24:23 +02:00
Simon Willnauer	ed473e272d	Cut over to JUnit & Randomized Runner from TestNG For better integration with the Lucene Test Framework and the availabilty of RandomizedRunner / Randommized Testing this commit moves over from TestNG to JUnit. This commit also moves relevant places over to RandomzedRunner for reproduceability and removes copied classes from the Lucene Test Framework.	2013-07-24 16:59:36 +02:00
Simon Willnauer	2e9851138e	Upgrade to Lucene 4.4	2013-07-23 13:55:15 +02:00
Alexander Reelsen	3162f5b725	Updated maven shade plugin to version 2.1 The currently used maven shade plugin still keeps references to the original classes in their constant pools around. This is never a problem at runtime, but for dependency tools which try to use the constant pool for determining dependencies will get confused (OSGI for example). This patch simply bumps the version and will implicetely fix fix http://jira.codehaus.org/browse/MSHADE-105 Closes #3254 Closes #3255	2013-07-04 15:57:13 +02:00
Alexander Reelsen	455bc32460	Moving forbidden-api checks to compile phase instead of test phase (fail fast)	2013-06-28 13:12:52 +02:00
Shay Banon	0114fb0f58	add running the tests in headless mode in maven	2013-06-27 23:23:13 +01:00
Adrien Grand	1954f770a1	Put Eclipse settings in the root directory. This enforces that settings are taken into account whichever mean is used to import the project into Eclipse (manual import, m2e, mvn eclipse:eclipse, ...).	2013-06-26 16:51:47 +02:00
Simon Willnauer	a388588b1f	Upgrade to Lucene 4.3.1	2013-06-18 22:15:31 +02:00
Shay Banon	f155525cad	upgrade jackson to 2.2.2, netty to 3.6.6	2013-06-11 20:36:08 +02:00
Alexander Reelsen	a5f9173e14	Making deb installable by being lintian compatible According to #2515 the ubuntu software center does not allow to install debian packages which are not lintian compatible I worked on the package and made it lintian compatible by doing * Ignoring errors about arch dependent binaries as we will not split this package. The arch dependent libraries are used correctly. * Added a copyright file pointing to the apache license in debian Closes #2515 Closes #2320	2013-06-07 13:53:14 +02:00
Alexander Reelsen	183bb76371	Ensure config files are not overwritten in RPM upgrade In order to ensure that configuration files do not get overwritten when upgrading an RPM, it is not sufficient to mark them as configuration. You have to use the 'noreplace' parameter to make sure, they are never overwritten. Added this parameter for the /etc/elasticsearch directory as well as the /etc/sysconfig/elasticsearch file. In addition, the post remove script now only deletes the user in case of a package removal (and does nothing on package upgrade). Closes #3123	2013-05-31 16:26:10 +02:00
Adrien Grand	c16a46e15c	Make it easier to get started with Eclipse. This patch makes mvn eclipse:eclipse generate additional eclipse configuration files so that Eclipse: - uses Java 1.6 compliance level, - truncates lines after 140 chars, - uses 4 spaces for indentation, - automatically adds a license header when creating a new class file, - organizes imports the same way as Intellij Idea (which makes sense I guess since most of the code bas has been written with Intellij, this will prevent from having large diffs due to the fact that the order of imports has changed).	2013-05-30 16:48:58 +02:00
Simon Willnauer	31f0aca65d	Integrate forbiddenAPI checks into Maven build. This commit integrates the forbiddenAPI checks that checks Java byte code against a list of "forbidden" API signatures. The commit also contains the fixes of the current source code that didn't pass the default API checks. See https://code.google.com/p/forbidden-apis/ for details. Closes #3059	2013-05-19 23:25:44 +02:00
Shay Banon	2ab72da7d6	update to joda 2.2	2013-05-11 23:37:56 +02:00
Shay Banon	e09e3eb73b	upgrade to mvel 2.1.5	2013-05-11 23:17:00 +02:00
Shay Banon	522ca7e889	upgrade to jackson 2.2.1	2013-05-09 18:00:23 +02:00
Simon Willnauer	2219925485	Upgrade to Lucene 4.3.0 This Lucene Release introduced a new API on DocIdSetIterator that requires each implementation to return a `cost` upperbound as a function of the iterated documents. This API allows for several optimizations during query execution especially in Conjunction and Disjunction Queries with min_should_match set. Closes #2990	2013-05-06 18:03:16 +02:00
Shay Banon	6c3bb4dcdd	move to 1.0.0.Beta1 snap	2013-04-29 13:51:09 +02:00
Shay Banon	cb75ce0caa	release 0.90.0 GA	2013-04-29 13:41:43 +02:00
Shay Banon	80fc55a01d	upgrade to netty 3.6.5	2013-04-09 09:44:49 -07:00
Shay Banon	3120457bfe	move to 0.90.0.RC3 snap	2013-04-08 05:48:29 -07:00
Shay Banon	3a8cba4d50	release 0.90.0.RC2	2013-04-08 05:46:26 -07:00
Simon Willnauer	eb8b38d027	Upgrade to Lucene 4.2.1	2013-04-03 12:22:39 +02:00
Alexander Reelsen	0a466352cd	Add support for creating a fedora RPM package with maven Note: This has been disabled by default and is therefore not included in a standard build. The main reason for this is, that you need to have a RPM binary and the rpm development packages installed, which is not the case on many systems. The package contains an init.d-script as well as systemd configurations. You can build your own RPM package simply by running 'maven rpm:rpm'	2013-04-02 16:19:45 +02:00
Shay Banon	ea698add72	move to 0.90.0.RC2 snap	2013-03-20 19:06:30 +01:00
Shay Banon	a2f14b68e8	release 0.90.0.RC1	2013-03-20 19:05:08 +01:00
Shay Banon	bea7bdde4c	upgrade to guava 14.0.1	2013-03-18 23:10:14 +01:00
Simon Willnauer	11bf7a8b1a	Upgrade to Lucene 4.2	2013-03-11 08:23:01 +01:00
Shay Banon	1d45edd856	upgrade to jackson 2.1.4	2013-02-28 09:10:43 +01:00
Shay Banon	f02c3ec39a	upgrade to guava 14.0	2013-02-27 18:31:13 +01:00
Shay Banon	31c273231a	remove compress flag, as its no longer relevant	2013-02-27 18:13:48 +01:00
Shay Banon	bd75b731c6	move to 0.90.0.Beta2 snap	2013-02-26 10:33:57 +01:00
Shay Banon	ab3a59e0bf	release 0.90.0.Beta1	2013-02-26 10:32:50 +01:00
Shay Banon	358c0e35fb	upgrade to latest jackson	2013-02-25 11:15:45 +01:00
Shay Banon	7787901a2d	cleanup the pom	2013-02-23 10:52:30 +01:00
Shay Banon	a3096157f8	upgrade to netty 3.6.3	2013-02-22 17:20:41 +01:00
Simon Willnauer	06fc9c77ba	don't exclude services otherwise es.jar doesn't include the SPI information to load XBloomPostings	2013-02-18 23:06:01 +01:00
Simon Willnauer	8db436f107	Remove backported Lucene 4 spatial code in favor of the released version in Lucene 4.1	2013-02-18 18:43:55 +01:00
Igor Motov	6890c9fa62	Move action.wait_on_mapping_change setting to pom	2013-02-06 11:48:58 -05:00
Martijn van Groningen	46dd42920c	Remove scope support in query and facet dsl. Remove support for the `scope` field in facets and `_scope` field in the nested and parent/child queries. The scope support for nested queries will be replaced by the `nested` facet option and a facet filter with a nested filter. The nested filters will now support the a `join` option. Which controls whether to perform the block join. By default this enabled, but when disabled it returns the nested documents as hits instead of the joined root document. Search request with the current scope support. ``` curl -s -XPOST 'localhost:9200/products/_search' -d '{ "query" : { "nested" : { "path" : "offers", "query" : { "match" : { "offers.color" : "blue" } }, "_scope" : "my_scope" } }, "facets" : { "size" : { "terms" : { "field" : "offers.size" }, "scope" : "my_scope" } } }' ``` The following will be functional equivalent of using the scope support: ``` curl -s -XPOST 'localhost:9200/products/_search?search_type=count' -d '{ "query" : { "nested" : { "path" : "offers", "query" : { "match" : { "offers.color" : "blue" } } } }, "facets" : { "size" : { "terms" : { "field" : "offers.size" }, "facet_filter" : { "nested" : { "path" : "offers", "query" : { "match" : { "offers.color" : "blue" } }, "join" : false } }, "nested" : "offers" } } }' ``` The scope support for parent/child queries will be replaced by running the child query as filter in a global facet. Search request with the current scope support: ``` curl -s -XPOST 'localhost:9200/products/_search' -d '{ "query" : { "has_child" : { "type" : "offer", "query" : { "match" : { "color" : "blue" } }, "_scope" : "my_scope" } }, "facets" : { "size" : { "terms" : { "field" : "size" }, "scope" : "my_scope" } } }' ``` The following is the functional equivalent of using the scope support with parent/child queries: ``` curl -s -XPOST 'localhost:9200/products/_search' -d '{ "query" : { "has_child" : { "type" : "offer", "query" : { "match" : { "color" : "blue" } } } }, "facets" : { "size" : { "terms" : { "field" : "size" }, "global" : true, "facet_filter" : { "term" : { "color" : "blue" } } } } }' ``` Closes #2606	2013-01-31 15:09:57 +01:00
Martijn van Groningen	98a674fc6e	Added suggest api. # Suggest feature The suggest feature suggests similar looking terms based on a provided text by using a suggester. At the moment there the only supported suggester is `fuzzy`. The suggest feature is available since version `0.21.0`. # Fuzzy suggester The `fuzzy` suggester suggests terms based on edit distance. The provided suggest text is analyzed before terms are suggested. The suggested terms are provided per analyzed suggest text token. The `fuzzy` suggester doesn't take the query into account that is part of request. # Suggest API The suggest request part is defined along side the query part as top field in the json request. ``` curl -s -XPOST 'localhost:9200/_search' -d '{ "query" : { ... }, "suggest" : { ... } }' ``` Several suggestions can be specified per request. Each suggestion is identified with an arbitary name. In the example below two suggestions are requested. The `my-suggest-1` suggestion uses the `body` field and `my-suggest-2` uses the `title` field. The `type` field is a required field and defines what suggester to use for a suggestion. ``` "suggest" : { "suggestions" : { "my-suggest-1" : { "type" : "fuzzy", "field" : "body", "text" : "the amsterdma meetpu" }, "my-suggest-2" : { "type" : "fuzzy", "field" : "title", "text" : "the rottredam meetpu" } } } ``` The below suggest response example includes the suggestions part for `my-suggest-1` and `my-suggest-2`. Each suggestion part contains a terms array, that contains all terms outputted by the analyzed suggest text. Each term object includes the term itself, the original start and end offset in the suggest text and if found an arbitary number of suggestions. ``` { ... "suggest": { "my-suggest-1": { "terms" : [ { "term" : "amsterdma", "start_offset": 5, "end_offset": 14, "suggestions": [ ... ] } ... ] }, "my-suggest-2" : { "terms" : [ ... ] } } ``` Each suggestions array contains a suggestion object that includes the suggested term, its document frequency and score compared to the suggest text term. The meaning of the score depends on the used suggester. The fuzzy suggester's score is based on the edit distance. ``` "suggestions": [ { "term": "amsterdam", "frequency": 77, "score": 0.8888889 }, ... ] ``` # Global suggest text To avoid repitition of the suggest text, it is possible to define a global text. In the example below the suggest text is a global option and applies to the `my-suggest-1` and `my-suggest-2` suggestions. ``` "suggest" : { "suggestions" : { "text" : "the amsterdma meetpu", "my-suggest-1" : { "type" : "fuzzy", "field" : "title" }, "my-suggest-2" : { "type" : "fuzzy", "field" : "body" } } } ``` The suggest text can be specied as global option or as suggestion specific option. The suggest text specified on suggestion level override the suggest text on the global level. # Other suggest example. In the below example we request suggestions for the following suggest text: `devloping distibutd saerch engies` on the `title` field with a maximum of 3 suggestions per term inside the suggest text. Note that in this example we use the `count` search type. This isn't required, but a nice optimalization. The suggestions are gather in the `query` phase and in the case that we only care about suggestions (so no hits) we don't need to execute the `fetch` phase. ``` curl -s -XPOST 'localhost:9200/_search?search_type=count' -d '{ "suggest" : { "suggestions" : { "my-title-suggestions" : { "suggester" : "fuzzy", "field" : "title", "text" : "devloping distibutd saerch engies", "size" : 3 } } } }' ``` The above request could yield the response as stated in the code example below. As you can see if we take the first suggested term of each suggest text term we get `developing distributed search engines` as result. ``` { ... "suggest": { "my-title-suggestions": { "terms": [ { "term": "devloping", "start_offset": 0, "end_offset": 9, "suggestions": [ { "term": "developing", "frequency": 77, "score": 0.8888889 }, { "term": "deloping", "frequency": 1, "score": 0.875 }, { "term": "deploying", "frequency": 2, "score": 0.7777778 } ] }, { "term": "distibutd", "start_offset": 10, "end_offset": 19, "suggestions": [ { "term": "distributed", "frequency": 217, "score": 0.7777778 }, { "term": "disributed", "frequency": 1, "score": 0.7777778 }, { "term": "distribute", "frequency": 1, "score": 0.7777778 } ] }, { "term": "saerch", "start_offset": 20, "end_offset": 26, "suggestions": [ { "term": "search", "frequency": 1038, "score": 0.8333333 }, { "term": "smerch", "frequency": 3, "score": 0.8333333 }, { "term": "serch", "frequency": 2, "score": 0.8 } ] }, { "term": "engies", "start_offset": 27, "end_offset": 33, "suggestions": [ { "term": "engines", "frequency": 568, "score": 0.8333333 }, { "term": "engles", "frequency": 3, "score": 0.8333333 }, { "term": "eggies", "frequency": 1, "score": 0.8333333 } ] } ] } } ... } ``` # Common suggest options: * `suggester` - The suggester implementation type. The only supported value is 'fuzzy'. This is a required option. * `text` - The suggest text. The suggest text is a required option that needs to be set globally or per suggestion. # Common fuzzy suggest options * `field` - The field to fetch the candidate suggestions from. This is an required option that either needs to be set globally or per suggestion. * `analyzer` - The analyzer to analyse the suggest text with. Defaults to the search analyzer of the suggest field. * `size` - The maximum corrections to be returned per suggest text token. * `sort` - Defines how suggestions should be sorted per suggest text term. Two possible value: `score` - Sort by sore first, then document frequency and then the term itself. `frequency` - Sort by document frequency first, then simlarity score and then the term itself. * `suggest_mode` - The suggest mode controls what suggestions are included or controls for what suggest text terms, suggestions should be suggested. Three possible values can be specified: `missing` - Only suggest terms in the suggest text that aren't in the index. This is the default. `popular` - Only suggest suggestions that occur in more docs then the original suggest text term. ** `always` - Suggest any matching suggestions based on terms in the suggest text. # Other fuzzy suggest options: * `lowercase_terms` - Lower cases the suggest text terms after text analyzation. * `max_edits` - The maximum edit distance candidate suggestions can have in order to be considered as a suggestion. Can only be a value between 1 and 2. Any other value result in an bad request error being thrown. Defaults to 2. * `min_prefix` - The number of minimal prefix characters that must match in order be a candidate suggestions. Defaults to 1. Increasing this number improves spellcheck performance. Usually misspellings don't occur in the beginning of terms. * `min_query_length` - The minimum length a suggest text term must have in order to be included. Defaults to 4. * `shard_size` - Sets the maximum number of suggestions to be retrieved from each individual shard. During the reduce phase only the top N suggestions are returned based on the `size` option. Defaults to the `size` option. Setting this to a value higher than the `size` can be useful in order to get a more accurate document frequency for spelling corrections at the cost of performance. Due to the fact that terms are partitioned amongst shards, the shard level document frequencies of spelling corrections may not be precise. Increasing this will make these document frequencies more precise. * `max_inspections` - A factor that is used to multiply with the `shards_size` in order to inspect more candidate spell corrections on the shard level. Can improve accuracy at the cost of performance. Defaults to 5. * `threshold_frequency` - The minimal threshold in number of documents a suggestion should appear in. This can be specified as an absolute number or as a relative percentage of number of documents. This can improve quality by only suggesting high frequency terms. Defaults to 0f and is not enabled. If a value higher than 1 is specified then the number cannot be fractional. The shard level document frequencies are used for this option. * `max_query_frequency` - The maximum threshold in number of documents a sugges text token can exist in order to be included. Can be a relative percentage number (e.g 0.4) or an absolute number to represent document frequencies. If an value higher than 1 is specified then fractional can not be specified. Defaults to 0.01f. This can be used to exclude high frequency terms from being spellchecked. High frequency terms are usually spelled correctly on top of this this also improves the spellcheck performance. The shard level document frequencies are used for this option. Closes #2585	2013-01-24 15:41:06 +01:00

... 3 4 5 6 7 ...

554 Commits