druid

Commit Graph

Author	SHA1	Message	Date
Ben Sykes	275c1ec64c	Fix error assuming a Complex Type that is a Number is a double (#15272 ) * Fix error assuming a Complex Type that is a Number is a double In the case where a complex type is a number, it may not be castable to double. It can safely be case as Number first to get to the doubleValue.	2023-10-30 09:52:52 +05:30
Vishesh Garg	039b05585c	Add worker status and duration metrics in live and task reports (#15180 ) Add worker status and duration metrics in live and task reports for tracking.	2023-10-30 09:43:22 +05:30
Zoltan Haindrich	f4a74710e6	Process pure ordering changes with windowing operators (#15241 ) - adds a new query build path: DruidQuery#toScanAndSortQuery which: - builds a ScanQuery without considering the current ordering - builds an operator to execute the sort - fixes a null string to "null" literal string conversion in the frame serializer code - fixes some DrillWindowQueryTest cases - fix NPE in NaiveSortOperator in case there was no input - enables back CoreRules.AGGREGATE_REMOVE - adds a processing level OffsetLimit class and uses that instead of just the limit in the rac parts - earlier window expressions on top of a subquery with an offset may have ignored the offset	2023-10-29 16:40:49 +05:30
317brian	737947754d	docs: add concurent compaction docs (#15218 ) Co-authored-by: Kashif Faraz <kashif.faraz@gmail.com>	2023-10-27 10:29:34 -07:00
kaisun2000	60c2ad597a	Enhance json parser error logging to better track Istio Proxy error message (#15176 ) Currently the inter Druid communication via rest endpoints is based on json formatted payload. Upon parsing error, there is only a generic exception stating expected json token type and current json token type. There is no detailed error log about the content of the payload causing the violation. In the micro-service world, the trend is to deploy the Druid servers in k8 with the mesh network. Often the istio proxy or other proxies is used to intercept the network connection between Druid servers. The proxy may give error messages for various reasons. These error messages are not expected by the json parser. The generic error message from Druid can be very misleading as the user may think the message is based on the response from the other Druid server. For example, this is an example of mysterious error message QueryInterruptedException{msg=Next token wasn't a START_ARRAY, was[VALUE_STRING] from url[http://xxxxx:8088/druid/v2/], code=Unknown exception, class=org.apache.druid.java.util.common.IAE, host=xxxxx:8088}" While the context of the message is the following from the proxy when it can't tunnel the network connection. pstream connect error or disconnect/reset before header So this very simple PR is just to enhance the logging and get the real underlying message printed out. This would save a lot of head scratching time if Druid is deployed with mesh network. Co-authored-by: Kai Sun <kai.sun@salesforce.com>	2023-10-27 14:20:19 +05:30
Laksh Singla	7c8e841362	Suppress CVE's in master (#15231 )	2023-10-27 09:29:18 +05:30
Simon Hofbauer	e9b7e4a0eb	fix JSON flaky tests (#15261 ) Co-authored-by: simonh5 <simonh5@illinois.edu>	2023-10-26 20:27:09 -07:00
Alexander Saydakov	f1132d20c5	use datasketches-java 4.2.0 (#15257 ) * use datasketches-java 4.2.0 * use exclusive mode * fixed issues raised by CodeQL * fixed issue raised by spotbugs * fixed issues raised by intellij * added missing import * Update QuantilesSketchKeyCollector search mode and adjust tests. * Update sizeOf functions and add unit tests * Add unit tests --------- Co-authored-by: AlexanderSaydakov <AlexanderSaydakov@users.noreply.github.com> Co-authored-by: Gian Merlino <gianmerlino@gmail.com> Co-authored-by: Adarsh Sanjeev <adarshsanjeev@gmail.com>	2023-10-26 16:28:33 -07:00
David Christle	fc0b940f78	Document the allowed range of announcer maxBytesPerNode (#15063 )	2023-10-26 14:51:01 -07:00
Pranav	e7b8e6569b	Updating plugin which has fix for corrupt nodejs pkg (#15259 )	2023-10-25 21:49:58 -07:00
Zoltan Haindrich	f48263bbb3	Report function name for unknown exceptions during execution (#14987 ) * provide function name when unknown exceptions are encountered * fix keywords/etc * fix keywrod order - regex excercise * add test * add check&fix keywords * decoupledIgnore * Revert "decoupledIgnore" This reverts commit `e922c820a7`. * unpatch Function * move to a different location * checkstyle	2023-10-25 13:37:30 -07:00
YongGang	7a25ee4fd9	Ability to send task types to k8s or worker task runner (#15196 ) * Ability to send task types to k8s or worker task runner * add more tests * use runnerStrategy to determine task runner * minor refine * refine runner strategy config * move workerType config to upper level * validate config when application start	2023-10-25 09:55:56 -07:00
Laksh Singla	207398a47d	Initialize null handling in CompressedBigDecimalAggregatorTimeseriesTestBase to fix failing test(#15252 )	2023-10-25 20:26:46 +05:30
Adarsh Sanjeev	c5fa649ea5	Rename segment load wait parameter (#15251 )	2023-10-25 18:08:37 +05:30
Zoltan Haindrich	6784e9c507	Fix summary row issues in case postaggregations are happening (#15232 ) * fix-1/2 * add message v1 * extend test to cover for IOB issue * move stuff around * change message * fix testcase string * compute postaggs (thank you Clint!) * enable feature for test * ignore tests in msq --------- Co-authored-by: Soumyava Das <soumyava@users.noreply.github.com>	2023-10-24 20:33:59 -07:00
Soumyava	06f40a0019	remove calcite AggregateRemoveRule to fix nested group by query with order by in outer query (#15237 ) * Fixing nested group by query with order by in outer query * Adding examples	2023-10-24 15:30:13 -07:00
Clint Wylie	4149c9422c	cleanup temp files for nested column serializer (#15236 ) * cleanup temp files for nested column serializer * fix style * fix tests in default value mode	2023-10-24 15:30:00 -07:00
Abhishek Radhakrishnan	63e3e9531d	Update S3 retry logic to account for the underlying cause in case of `IOException` (#15238 ) * Update S3 retry logic based on the underlying cause in case of IOException. 4xx and other errors wrapped in IOException for instance aren't retriable. * Fix CI	2023-10-24 15:04:42 -07:00
AmatyaAvadhanula	65b69cded4	Filter pending segments upgraded with transactional replace (#15169 ) * Filter pending segments upgraded with transactional replace * Push sequence name filter to metadata query	2023-10-23 21:18:47 +05:30
Zoltan Haindrich	2e31cb2901	DrillWindowQueryTest: use proper way to decide if the query is ordered (#15118 )	2023-10-23 10:54:28 -04:00
Zoltan Haindrich	b95035f183	Fix VirtualColumn related issues in window expressions (#15119 ) for some exotic queries like: SELECT '_'\|\|dim1, MIN(cast(0 as double)) OVER (), MIN(cast((cnt\|\|cnt) as bigint)) OVER () FROM foo the compilation have resulted in NPE -s mostly because VirtualColumn -s were not handled properly	2023-10-23 14:05:59 +05:30
Clint Wylie	c8e458452d	Fix native is boolean filter cache key tests to test the right thing (#15216 )	2023-10-23 11:24:46 +05:30
AmatyaAvadhanula	33fdd770f7	Consider only supervisors with append lock for concurrent transactional replace (#15220 ) A SegmentTransactionReplaceAction must only update the mapping of tasks with append locks that are running concurrently. To ensure this, we return the supervisor id only if it has the taskLockType as APPEND in its context.	2023-10-22 14:12:36 +05:30
Zoltan Haindrich	fbbb9c7730	Allow DESC ordering in window expressions (#15195 )	2023-10-20 07:55:28 -04:00
Xavier Léauté	e03f863cf6	update core Apache Kafka dependencies to 3.6.0 (#15214 ) Release notes: https://downloads.apache.org/kafka/3.6.0/RELEASE_NOTES.html https://kafka.apache.org/blog#apache_kafka_360_release_announcement	2023-10-19 20:27:09 -07:00
Xavier Léauté	352702bb25	run some integration tests with Java 21 (#15104 ) * use setup-java everywhere for consistency * add Java 21 to integration test matrix * simplify docker build containers script + add Java 21 * fix for Java versions reporting 21-ea	2023-10-20 11:18:13 +08:00
Sébastien	5752a1a383	Proper default for taskLockType in streaming ingestion (#15213 ) * Proper value for taskLockType in streaming ingestion with concurrent compaction	2023-10-19 21:01:31 +05:30
Tejaswini Bandlamudi	1f39c054a7	Fix GHA workflow bugs (#15209 )	2023-10-19 17:11:36 +05:30
Zoltan Haindrich	9fb0dbfc9f	Fix json inputs for drill windowing tests (#15148 ) This PR: adds a flag to JsonToParquet to do the fix during conversion updates the json files to more correct conents some resultset mismatches were fixed by this updates parquet to 1.13.1	2023-10-19 14:02:41 +05:30
AmatyaAvadhanula	a8febd457c	A Replacing task must read segments created before it acquired its lock (#15085 ) * Replacing tasks must read segments created before they acquired their locks	2023-10-19 11:13:07 +05:30
Laksh Singla	fa311dd0b6	Cast the values read in the EXTERN function to the type in the row signature (#15183 )	2023-10-19 10:24:23 +05:30
Atul Mohan	780207869b	Attach user identity to router request logs (#15126 ) * Attach user identity to router request logs * Add test * More tests	2023-10-18 19:40:58 -07:00
Clint Wylie	5c14b42e72	fix incorrect unnest dimension cursor value matcher implementation (#15192 )	2023-10-18 16:43:06 -07:00
Clint Wylie	061cfee224	add native filters for "(filter) is true" and "(filter) is false" (#15182 ) * add native filters for "(filter) is true" and "(filter) is false" changes: * add IsTrueDimFilter, IsFalseDimFilter, and abstract IsBooleanDimFilter for native json filter implementations of `(filter) IS TRUE` and `(filter) IS FALSE` * add IsBooleanFilter for actual filtering logic for these filters, which ignore includeUnknown to always use matches with false for true and !matches with true for false * fix test incorrectly adjusted to wrong answer in #15058 * add tests for default value mode	2023-10-18 13:07:35 -07:00
Zoltan Haindrich	c58b7f40ee	Rename windowing option (#15184 )	2023-10-18 10:54:20 +05:30
Clint Wylie	22034a1630	preserve Rows.objectToStrings behavior of translating null into "null" inside of lists and arrays (#15190 )	2023-10-17 19:49:36 -07:00
Laksh Singla	b4540ed5d4	Optimize the reading of numerical frame arrays in MSQ (#15175 )	2023-10-18 02:33:42 +05:30
Karan Kumar	953ce79439	Add undocumented taskLockType to MSQ. (#15168 ) Patch adds an undocumented parameter taskLockType to MSQ so that we can start enabling this feature for users who are interested in testing the new lock types.	2023-10-17 21:44:04 +05:30
George Shiqi Wu	dc0b163e19	Separate task lifecycle from kubernetes/location lifecycle (#15133 ) * Separate k8s and druid task lifecycles * Remove extra log lines * Fix unit tests * fix unit tests * Fix unit tests * notify listeners on task completion * Fix unit test * unused var * PR changes * Fix unit tests * Fix checkstyle * PR changes	2023-10-17 08:17:43 -07:00
Pranav	0a27a7a7ca	Update eirslett frontend (#15154 )	2023-10-16 20:16:32 +05:30
Laksh Singla	dc8d2192c3	Introduce natural comparator for types that don't have a StringComparator (#15145 ) Fixes a bug when executing queries with the ordering of arrays	2023-10-16 10:37:32 +05:30
Pranav	4b0d1b3488	Fix expression result writing of arrays in Hadoop Ingestion (#15127 )	2023-10-13 13:41:41 -07:00
Sébastien	9ca10c7bd7	Added concurrent compaction switches (#15114 ) * Added concurrent compaction switches	2023-10-13 21:03:39 +05:30
Zoltan Haindrich	6d62c75866	Fix columns with null values in windowing expressions (#15131 )	2023-10-13 10:42:45 -04:00
Karan Kumar	f0a70fe3c4	Fixing the flaky tests. (#15142 )	2023-10-13 16:20:05 +05:30
Adarsh Sanjeev	4deeb7e936	Fix issue with checking segment load status (#15147 ) This PR addresses a bug with waiting for segments to be loaded. In the case of append, segments would be created with the same version. This caused the number of segments returned to be incorrect. This PR changes this to keep track of the range of partition numbers as well for each version, which lets the task wait for the correct set of segments. The partition numbers are expected to be continuous since the task obtains the lock for the segment while running.	2023-10-13 16:06:13 +05:30
AmatyaAvadhanula	d25caaefa4	Add support for streaming ingestion with concurrent replace (#15039 ) Add support for streaming ingestion with concurrent replace --------- Co-authored-by: Kashif Faraz <kashif.faraz@gmail.com>	2023-10-13 09:09:03 +05:30
Tejaswini Bandlamudi	0a6f78c0bb	Fix GHA workflow bugs (#15138 )	2023-10-12 21:25:57 +05:30
Karan Kumar	61ea9e07c5	Limit pages size to a configurable limit (#14994 ) Adding the ability to limit the pages sizes of select queries. We piggyback on the same machinery that is used to control the numRowsPerSegment. This patch introduces a new context parameter rowsPerPage for which the default value is set to 100000 rows. This patch also optimizes adding the last selectResults stage only when the previous stages have sorted outputs. Currently for each select query with selectDestination=durableStorage, we used to add this extra selectResults stage.	2023-10-12 14:01:46 +05:30
Clint Wylie	a0fd9ec55c	fix issue with SQL boolean constants not respecting nulls when strict booleans and sql compatible null handling are enabled (#15135 )	2023-10-12 01:23:24 -07:00

1 2 3 4 5 ...

13336 Commits All Branches Search

13336 Commits

All Branches