OpenSearch

Commit Graph

Author	SHA1	Message	Date
emasab	a142e8cfd8	Build local year inside DateFormat lambda bugfix for https://github.com/elastic/elasticsearch/issues/41797 (#42120) This makes sure that the year can change between when the lambda is generated and when it is executed without causing the incorrect year to be used. Resolves #41797	2019-05-23 10:36:11 -06:00
Yannick Welsch	770d8e9e39	Remove usage of max_local_storage_nodes in test infrastructure (#41652 ) Moves the test infrastructure away from using node.max_local_storage_nodes, allowing us in a follow-up PR to deprecate this setting in 7.x and to remove it in 8.0. This also changes the behavior of InternalTestCluster so that starting up nodes will not automatically reuse data folders of previously stopped nodes. If this behavior is desired, it needs to be explicitly done by passing the data path from the stopped node to the new node that is started.	2019-05-22 11:04:55 +02:00
Alexander Reelsen	8e33a5292a	Add HTML strip processor (#41888 ) This processor uses the lucene HTMLStripCharFilter class to remove HTML entities from a field. This adds to the char filter, so that there is possibility to store the stripped version as well. Note, that the characeter filter replaces tags with a newline, so that the produced HTML will look slightly different than the incoming HTML with regards to newlines.	2019-05-09 13:01:07 +02:00
Christoph Büscher	52495843cc	[Docs] Fix common word repetitions (#39703 )	2019-04-25 20:47:47 +02:00
Alpar Torok	25944c4317	convert modules to use testclusters (#40804 ) * convert modules to use testclusters * Eliminate PluginPropertiesTask and move logic in plugin where it belongs	2019-04-04 11:45:40 +03:00
Martijn van Groningen	89837eb918	Remove -Xlint exclusions in the ingest-common module. (#40505 ) Fix the generics in processors extending AbstractStringProcessor and its factory. Relates to #40366	2019-03-29 09:43:36 +01:00
Jake Landis	797d6b8a66	Execute ingest node pipeline before creating the index (#39607 ) (#39796 ) Prior to this commit (and after 6.5.0), if an ingest node changes the _index in a pipeline, the original target index would be created. For daily indexes this could create an extra, empty index per day. This commit changes the TransportBulkAction to execute the ingest node pipeline before attempting to create the index. This ensures that the only index created is the original or one set by the ingest node pipeline. This was the execution order prior to 6.5.0 (#32786). The execution order was changed in 6.5 to better support default pipelines. Specifically the execution order was changed to be able to read the settings from the index meta data. This commit also includes a change in logic such that if the target index does not exist when ingest node pipeline runs, it will now pull the default pipeline (if one exists) from the settings of the best matched of the index template. Relates #32786 Relates #32758 Closes #36545	2019-03-07 13:31:41 -06:00
Martijn van Groningen	b8659fcb83	No need to extend from StatusToXContentObject, if RestToXContentListener is used instead of RestStatusToXContentListener	2019-03-04 13:29:10 +01:00
Martijn van Groningen	0550ead176	Cleanup GrokProcessorGetAction class (#39567 ) * Removed request builder. From 7.0, request builders are no longer used. * Use RestStatusToXContentListener instead of custom RestBuilderListener in the rest action. * Changed a few public constructor's and constants' visibility from public to package protected. (these are only used internally, so no need to for public visibility)	2019-03-04 08:51:23 +01:00
Ioannis Kakavas	ec2b64af63	Disable date parsing test in non english locale (#39052 ) This ensures we do not attempt to parse non english locale dates in FIPS mode. The error, originally assumed to affect only Joda, affects Java time in the same manner and manifests only with the version of BouncyCastle FIPS certified provider we use in tests. The upstream issue https://github.com/bcgit/bc-java/issues/405 indicates that the behavior is resolved in later versions of the BouncyCastle library and should be tested again when the new versions become FIPS 140 certified	2019-02-20 09:02:37 +02:00
Alexander Reelsen	884b5063a4	Create ISO8601 joda compatible java time formatter (#38434 ) The existing formatter being used was not on par with the joda formatter as it was missing the ability to parse a comma as a separator between seconds and milliseconds. While a real iso8601 would be much more complex, this might be sufficient for some more use-cases. The ingest date formatter now also uses the iso8601 formatter by default. Closes #38345	2019-02-11 15:11:26 +01:00
Alexander Reelsen	56edc8e37f	Fix timezone fallback in ingest processor (#38407 ) (#38664 ) If no timezone was specified in the date processor, then the conversion would lead to wrong time, as UTC was assumed by default, leading to incorrectly parsed dates. This commit does not assume a default timezone and will thus not format the dates in a wrong way.	2019-02-09 20:28:59 +01:00
Christoph Büscher	820029522b	Mute DateProcessorTests#testJodaPatternLocale (#38265 ) Only fails on FIPS 8, muting this selectively.	2019-02-03 19:52:53 +01:00
Alexander Reelsen	6c5a7387af	Replace joda time in ingest-common module (#38088 ) This commit fully replaces any remaining joda time time classes with java time implementations. Relates #27330	2019-02-01 10:15:18 +01:00
Henning Andersen	68ed72b923	Handle scheduler exceptions (#38014 ) Scheduler.schedule(...) would previously assume that caller handles exception by calling get() on the returned ScheduledFuture. schedule() now returns a ScheduledCancellable that no longer gives access to the exception. Instead, any exception thrown out of a scheduled Runnable is logged as a warning. This is a continuation of #28667, #36137 and also fixes #37708.	2019-01-31 17:51:45 +01:00
Tal Levy	e0d5de33da	fix DateIndexNameProcessorTests offset pattern (#38069 ) `XX` was being used to represent an offset pattern, it should be `ZZ` Fixes #38067.	2019-01-31 08:57:56 +01:00
Alexander Reelsen	b94acb608b	Speed up converting of temporal accessor to zoned date time (#37915 ) The existing implementation was slow due to exceptions being thrown if an accessor did not have a time zone. This implementation queries for having a timezone, local time and local date and also checks for an instant preventing to throw an exception and thus speeding up the conversion. This removes the existing method and create a new one named DateFormatters.from(TemporalAccessor accessor) to resemble the naming of the java time ones. Before this change an epoch millis parser using the toZonedDateTime method took approximately 50x longer. Relates #37826	2019-01-31 08:55:40 +01:00
Jason Tedor	89bffc25de	Mute failing date index name processor test This test is repeatedly failing, so this commit mutes it. Relates #38067	2019-01-30 20:37:52 -05:00
Colin Goodheart-Smithe	21e392e95e	Removes typed calls from YAML REST tests (#37611 ) This PR attempts to remove all typed calls from our YAML REST tests. The PR adds include_type_name: false to create index requests that use a mapping and also to put mapping requests. It also removes _type from index requests where they haven't already been removed. The PR ignores tests named *_with_types.yml since this are specifically testing typed API behaviour. The change also includes changing the test harness to add the type _doc to index, update, get and bulk requests that do not specify the document type when the test is running against a mixed 7.x/6.x cluster.	2019-01-30 16:32:58 +00:00
David Roberts	2f7776c8b7	Switch default time format for ingest from Joda to Java for v7 (#37934 ) Date formats with and without the "8" prefix are now all treated as Java time formats, so that ingest does the same as mappings in this respect.	2019-01-30 16:26:28 +00:00
Alexander Reelsen	9e350d027e	Add BWC compatible processing to ingest date processors (#37407 ) The ingest date processor is currently only able to parse joda formats. However it is not using the existing elasticsearch classes but access joda directly. This means that our existing BWC layer does not notify the user about deprecated formats. This commit switches to use the exising Elasticsearch Joda methods to acquire a date format, that includes the BWC check and the ability to parse java 8 dates. The date parsing in ingest has also another extra feature, that the fallback year, when a date format without a year is used, is the current year, and not 1970 like usual. This is currently not properly supported in the DateFormatter class. As this is the only case for this feature and java time can take care of this using the toZonedDateTime() method, a workaround just for the joda time parser has been created, that can be removed soon again from 7.0.	2019-01-25 13:50:19 +01:00
Jack Conradson	de55b4dfd1	Add types deprecation to script contexts (#37554 ) This adds deprecation to _type in the script contexts for ingest and update. This adds a DeprecationMap that wraps the ctx Map containing _type for these specific contexts.	2019-01-18 09:13:49 -08:00
Jake Landis	195873002b	ingest: compile mustache template only if field includes '{{'' (#37207 ) * ingest: compile mustache template only if field includes '{{'' Prior to this change, any field in an ingest node processor that supports script templates would be compiled as mustache template regardless if they contain a template or not. Compiling normal text as mustache templates is harmless. However, each compilation counts against the script compilation circuit breaker. A large number of processors without any templates or scripts could un-intuitively trip the too many script compilations circuit breaker. This change simple checks for '{{' in the text before it attempts to compile. fixes #37120	2019-01-09 14:47:47 -06:00
Jake Landis	384757deff	ingest: support default pipelines + bulk upserts (#36618 ) This commit adds support to enable bulk upserts to use an index's default pipeline. Bulk upsert, doc_as_upsert, and script_as_upsert are all supported. However, bulk script_as_upsert has slightly surprising behavior since the pipeline is executed _before_ the script is evaluated. This means that the pipeline only has access the data found in the upsert field of the script_as_upsert. The non-bulk script_as_upsert (existing behavior) runs the pipeline _after_ the script is executed. This commit does _not_ attempt to consolidate the bulk and non-bulk behavior for script_as_upsert. This commit also adds additional testing for the non-bulk behavior, which remains unchanged with this commit. fixes #36219	2018-12-17 16:25:11 -06:00
Jake Landis	7bf822bbbb	ingest: fix on_failure with Drop processor (#36686 ) This commit allows a document to be dropped when a Drop processor is used in the on_failure fork of the processor chain. Fixes #36151	2018-12-17 14:10:13 -06:00
Jake Landis	190ac8e9bf	ingest: support default pipeline through an alias (#36231 ) This commit allows writes that go through an alias to use the default pipeline defined on the backing index. Fixes #35817	2018-12-05 16:25:50 -06:00
Jake Landis	9150a93b1c	ingest: dot_expander_processor prevent null add/append to source document (#35106 ) * don't allow null values to be added or appended to the source document if the field does not exist.	2018-11-05 17:16:42 -06:00
Alexander Reelsen	409050e8de	Refactor: Remove settings from transport action CTOR (#35208 ) As settings are not used in the transport action constructor, this removes the passing of the settings in all the transport actions.	2018-11-05 13:08:18 +01:00
Alpar Torok	59536966c2	Add a new "contains" feature (#34738 ) The contains syntax was added in #30874 but the skips were not properly put in place. The java runner has the feature so the tests will run as part of the build, but language clients will be able to support it at their own pace.	2018-10-25 08:50:50 +03:00
Jake Landis	89dc07bdd9	ingest: better support for conditionals with simulate?verbose (#34155 ) This commit introduces two corrections to the way simulate?verbose handles conditionals on processors. 1) Prior to this change when executing simulate?verbose for processors with conditionals that evaluate to false, that processor would still be displayed in the result set. What was displayed was correct, such that no changes to the document occurred. However, if the conditional evaluates to false, the processor should not even be displayed. 2) Prior to this change when executing simulate?verbose for pipeline processors with conditionals, the individual steps would no longer be displayed. Commit `e37e5df` addressed the issue, but failed account for a conditional on the pipeline processor. Since a pipeline processor can introduce cycles and is effectively a single processor that encapsulates multiple other processors that are potentially guarded by a single conditional, special handling is needed to for pipeline and conditional pipeline processors.	2018-10-23 11:33:48 -05:00
Armin Braun	8e155b8430	INGEST: Rename Pipeline Processor Param. (#34733 ) * `name` is more readable/ergnomic than having `pipeline` twice	2018-10-23 13:43:26 +02:00
Ryan Ernst	8734540345	Ensure map keys cannot be self referencing (#34569 ) This commit improves self reference checking to map keys, as well as adds it to ingest script processing.	2018-10-17 15:16:13 -07:00
Ryan Ernst	a2c941806b	Tests: Add support for custom contexts to mock scripts (#34100 ) This commit adds the ability to plug in compilation of custom contexts in mock script engine. This is needed for testing plugins which add custom contexts like watcher.	2018-09-27 12:23:59 -07:00
Armin Braun	0ba1855740	INGEST: Tests for Drop Processor (#33430 ) * INGEST: Tests for Drop Processor * UT for behavior of dropped callback and drop processor * Moved drop processor to `server` project to enable this test * Simple IT * Relates #32278	2018-09-25 19:29:22 +02:00
Jake Landis	e37e5dfc04	ingest: support simulate with verbose for pipeline processor (#33839 ) * ingest: support simulate with verbose for pipeline processor This change better supports the use of simulate?verbose with the pipeline processor. Prior to this change any pipeline processors executed with simulate?verbose would not show all intermediate processors for the inner pipelines. This changes also moves the PipelineProcess and TrackingResultProcessor classes to enable instance checks and to avoid overly public classes. As well this updates the error message for when cycles are detected in pipelines calling other pipelines.	2018-09-20 08:33:07 -05:00
Armin Braun	ef1066d7f8	INGEST: Allow Repeated Invocation of Pipeline (#33419 ) * Allows repeated, non-recursive invocation of the same pipeline	2018-09-05 22:04:53 +02:00
Armin Braun	46774098d9	INGEST: Implement Drop Processor (#32278 ) * INGEST: Implement Drop Processor * Adjust Processor API * Implement Drop Processor * Closes #23726	2018-09-05 14:25:29 +02:00
Armin Braun	cc4d7059bf	Ingest: Add conditional per processor (#32398 ) * Ingest: Add conditional per processor * closes #21248	2018-08-30 03:46:39 +02:00
Armin Braun	f690b492e7	INGEST: Add Pipeline Processor (#32473 ) * INGEST: Add Pipeline Processor * Adds Processor capable of invoking other pipelines * Closes #31842	2018-08-29 11:03:10 +02:00
Jake Landis	e9b0807c67	ingest: minor - update test to include dissect (#33211 ) This change also includes placing the bytes processor in the correct order (helps to avoid merge conflict when back patching processors)	2018-08-28 11:55:04 -07:00
Jake Landis	79b507dbf5	ingest: Introduce the dissect processor (#32884 ) * ingest: Introduce the dissect processor The ingest node dissect processor is an alternative to Grok to split a string based on a pattern. Dissect differs from Grok such that regular expressions are not used to split the string. Dissect can be used to parse a source text field with a simpler pattern, and is often faster the Grok for basic string parsing. This processor uses the dissect library which does most of the work.	2018-08-28 07:11:20 -07:00
Armin Braun	986c55b830	INGEST: Add Configuration Except. Data to Metdata (#32322 ) * closes #27728	2018-08-15 19:02:19 +02:00
Armin Braun	be31cc642b	INGEST: Enable default pipelines (#32286 ) * INGEST: Enable default pipelines * Add `default_pipeline` index setting * `_none` is interpreted as no pipeline * closes #21101	2018-08-02 17:11:12 +02:00
Armin Braun	cf7489899a	INGEST: Clean up Java8 Stream Usage (#32059 ) * GrokProcessor: Rationalize the loop over the map to save allocations and indirection * IngestDocument: Rationalize way we append to `List`	2018-07-30 21:25:30 +02:00
Ryan Ernst	34d006f82a	Tests: Fix convert error tests to use fixed value (#32415 ) The error tests for hex values previously used a random string of digits, but this could be a valid hex value. This commit changes these tests to use a fixed invalid hex value. closes #32370	2018-07-30 10:00:55 -07:00
javanna	83d007e7be	[TEST] Mute failing testConvertLongHexError See #32370	2018-07-27 11:50:13 +02:00
Dimitris Athanasiou	de53f0123f	[TEST] Mute ConvertProcessortTests.testConvertIntHexError Relates #32370	2018-07-25 17:35:23 +01:00
Ryan Ernst	49d4b26f16	Ingest: Support integer and long hex values in convert (#32213 ) This commit adds checks for hex formatted strings in the convert processor, allowing strings like `0x1` to be parsed as integer `1`. closes #32182	2018-07-24 12:05:50 -07:00
Christoph Büscher	ff87b7aba4	Remove unnecessary warning supressions (#32250 )	2018-07-23 11:31:04 +02:00
Armin Braun	7aa8a0a927	INGEST: Extend KV Processor (#31789 ) (#32232 ) * INGEST: Extend KV Processor (#31789) Added more capabilities supported by LS to the KV processor: * Stripping of brackets and quotes from values (`include_brackets` in corresponding LS filter) * Adding key prefixes * Trimming specified chars from keys and values Refactored the way the filter is configured to avoid conditionals during execution. Refactored Tests a little to not have to add more redundant getters for new parameters. Relates #31786 * Add documentation	2018-07-20 22:32:50 +02:00

1 2 3 4

189 Commits