druid

Commit Graph

Author	SHA1	Message	Date
Clint Wylie	b038a11280	fix issues with handling arrays with all null elements and arrays of booleans in strict mode (#14297 )	2023-05-17 01:33:44 -07:00
Tejaswini Bandlamudi	bbbb031057	Do not cancel old GHA workflows triggered on branch commits (#14279 ) * group and limit workflows only on PRs and not on branch commits * also apply to Static Checks CI	2023-05-16 12:13:08 +05:30
Soumyava	96a3c00754	Fixing an issue with filtering on a single dimension by converting In… (#14277 ) * Fixing an issue with filtering on a single dimension by converting In filter to a selector filter as needed with Filters.toFilter * Adding a test so that any future refactoring does not break this behavior * Made comment a bit more meaningful	2023-05-15 20:10:36 -07:00
Adarsh Sanjeev	e8ef31fe92	Fix condition for timeout in worker task launcher (#14270 ) * Fix condition for timeout in worker task launcher	2023-05-16 08:30:00 +05:30
Victoria Lim	66d4ea014c	Docs: Tutorial for streaming ingestion using Kafka + Docker file to use with Jupyter tutorials (#13984 )	2023-05-15 15:20:52 -07:00
Peter Marshall	c4aa98953b	202304-docs-removeDF (#14132 )	2023-05-15 15:08:57 -07:00
Paul Rogers	3c0983c8e9	Extend the IT framework to allow tests in extensions (#13877 ) The "new" IT framework provides a convenient way to package and run integration tests (ITs), but only for core modules. We have a use case to run an IT for a contrib extension: the proposed gRPC query extension. This PR provides the IT framework functionality to allow non-core ITs.	2023-05-15 20:29:51 +05:30
Adarsh Sanjeev	10bce22e68	Configure maxBytesPerWorker directly instead of using StageDefinition (#14257 ) * Configure maxBytesPerWorker directly instead of using StageDefinition	2023-05-15 16:51:57 +05:30
AmatyaAvadhanula	e9913abbbf	Add new lock types: APPEND and REPLACE (#14258 ) * Add new lock types: APPEND and REPLACE	2023-05-14 22:38:32 -07:00
imply-cheddar	f9861808bc	Be able to load segments on Peons (#14239 ) * Be able to load segments on Peons This change introduces a new config on WorkerConfig that indicates how many bytes of each storage location to use for storage of a task. Said config is divided up amongst the locations and slots and then used to set TaskConfig.tmpStorageBytesPerTask The Peons use their local task dir and tmpStorageBytesPerTask as their StorageLocations for the SegmentManager such that they can accept broadcast segments.	2023-05-12 16:51:00 -07:00
317brian	8bda7297e1	doc: fix unnest datasource syntax (#14272 )	2023-05-12 13:05:27 -07:00
Tejaswini Bandlamudi	9e0708f5e6	update heap size of coordinator, overlord services in docker IT environment (#14214 )	2023-05-12 23:19:48 +05:30
Kashif Faraz	ba11b3d462	Refactor: Add OverlordDuty to replace OverlordHelper and align with CoordinatorDuty (#14235 ) Changes: - Replace `OverlordHelper` with `OverlordDuty` to align with `CoordinatorDuty` - Each duty has a `run()` method and defines a `Schedule` with an initial delay and period. - Update existing duties `TaskLogAutoCleaner` and `DurableStorageCleaner` - Add utility class `Configs` - Update log, error messages and javadocs - Other minor style improvements	2023-05-12 22:39:56 +05:30
317brian	6254658f61	docs: fix links (#14111 )	2023-05-12 09:59:16 -07:00
Nicholas Lippis	58dcbf9399	queue tasks in kubernetes task runner if capacity is fully utilized (#14156 ) * queue tasks if all slots in use * Declare hamcrest-core dependency * Use AtomicBoolean for shutdown requested * Use AtomicReference for peon lifecycle state * fix uninitialized read error * fix indentations * Make tasks protected * fix KubernetesTaskRunnerConfig deserialization * ensure k8s task runner max capacity is Integer.MAX_VALUE * set job duration as task status duration * Address pr comments --------- Co-authored-by: George Shiqi Wu <george.wu@imply.io>	2023-05-12 09:41:44 -06:00
Abhishek Agarwal	9eebeead44	Tune stale bot to pick older issues first (#14267 )	2023-05-12 11:45:29 +05:30
Tejaswini Bandlamudi	8ef99f091a	Fix jdk setup in GHA (#14091 ) Instead of downloading jdk everytime we run CI, we're using inbuilt temurin jdk distributions 8, 11, 17 by settiing JAVA_HOME variable. This is not working as expected since we were not setting this as global environment variable as a result all CI builds are running on jdk11. This PR fixes the issue.	2023-05-12 10:36:59 +05:30
Clint Wylie	eae9e07ea9	suppress CVE-2021-40331 since it applies to ranger-hive-plugin which afaict we do not use (#14261 )	2023-05-11 21:58:47 -07:00
Clint Wylie	9875090bee	fix segment metadata queries for auto ingested columns that had all null values (#14262 )	2023-05-11 20:58:06 -07:00
Kashif Faraz	47a70d03e8	Docs: Minor rephrase in indexing-service.md (#14231 ) * Fix language in indexing-service * Apply suggestions from code review Co-authored-by: Katya Macedo <38017980+ektravel@users.noreply.github.com>	2023-05-12 08:22:02 +05:30
317brian	cc37987dff	docs: copyedits for MSQ join algos (#14012 )	2023-05-11 14:21:09 -07:00
Soumyava	f128b9b666	Updates to filter processing for inner query in Joins (#14237 )	2023-05-11 17:21:41 +05:30
Clint Wylie	a58cebe491	add array_to_mv function to convert arrays into mvds to assist with migration from mvds to arrays (#14236 )	2023-05-11 04:43:28 -07:00
Kashif Faraz	64e6283eca	Do not allow retention rules to be null (#14223 ) Changes: - Do not allow retention rules for any datasource or cluster to be null - Allow empty rules at the datasource level but not at the cluster level - Add validation to ensure that `druid.manager.rules.defaultRule` is always set correctly - Minor style refactors	2023-05-11 14:33:56 +05:30
AmatyaAvadhanula	47e48ee657	Remove incorrect optimization (#14246 )	2023-05-11 00:54:41 -07:00
Clint Wylie	e833a4700d	suppress hadoop3 cve that seem not applicable to us (#14252 )	2023-05-10 23:08:05 -07:00
Abhishek Agarwal	f3ff36a004	Move the stale bot to a GHA action (#14238 ) Move the stale bot to a GHA action	2023-05-11 11:31:28 +05:30
Clint Wylie	aaaff74740	fix npe regression in json_value when filtering non-existent paths (#14250 ) * fix npe regression in json_value when filtering non-existent paths * more coverage	2023-05-10 22:39:22 -07:00
Clint Wylie	6db11bfc60	suppress some cves and fix javadoc build when using java 17 (#14241 )	2023-05-10 15:47:10 -07:00
Clint Wylie	625c4745b1	add context flag "useAutoColumnSchemas" to use new auto types for MSQ segment generation (#14175 )	2023-05-10 15:37:14 -07:00
George Shiqi Wu	161d12eb44	Fix unit tests for java 17 (#14207 ) Fix a unit test that fails in java 17	2023-05-09 20:02:31 +05:30
Kashif Faraz	bd0080c4ce	Update default values in docs (#14233 )	2023-05-09 19:13:51 +05:30
Shingo Kitagawa	152e9375e2	update documentation about multiValueHandling (#14197 ) * update documentation about multiValueHandling * Update docs/ingestion/ingestion-spec.md Co-authored-by: Victoria Lim <vtlim@users.noreply.github.com> * Update docs/ingestion/ingestion-spec.md Co-authored-by: Gian Merlino <gianmerlino@gmail.com> * fix spelling --------- Co-authored-by: Victoria Lim <vtlim@users.noreply.github.com> Co-authored-by: Gian Merlino <gianmerlino@gmail.com>	2023-05-08 16:16:54 -07:00
Clint Wylie	8805d8d7db	fix issues with filtering nulls on values coerced to numeric types (#14139 ) * fix issues with filtering nulls on values coerced to numeric types * fix issues with 'auto' type numeric columns in default value mode * optimize variant typed columns without nested data * more tests for 'auto' type column ingestion	2023-05-08 13:19:02 -07:00
Vadim Ogievetsky	0a3889b192	account for auto allowing for leading and trailing spaces (#14224 )	2023-05-08 13:18:31 -07:00
minseok	3c62c00d4c	Fix Typos in DruidToGraphiteEventConverter (#14219 )	2023-05-08 17:46:32 +05:30
Clint Wylie	a7a4bfd331	modify QueryScheduler to lazily acquire lanes when executing queries to avoid leaks (#14184 ) This PR fixes an issue that could occur if druid.query.scheduler.numThreads is configured and any exception occurs after QueryScheduler.run has been called to create a Sequence. This would result in total and/or lane specific locks being acquired, but because the sequence was not actually being evaluated, the "baggage" which typically releases these locks was not being executed. An example of how this can happen is if a group-by having filter, which wraps and transforms this sequence happens to explode while wrapping the sequence. The end result is that the locks are acquired, but never released, eventually halting the ability to execute any queries.	2023-05-08 11:42:05 +05:30
Rohan Garg	4d8feeb279	Fix planning in CASE expressions with complex WHEN and ELSE expressions (#14220 )	2023-05-08 11:35:04 +05:30
George Shiqi Wu	eed5f4f291	Add labels to k8s jobs for the PodTemplateTaskAdapter (#14205 ) * Add labels * Add prefix * remove newline * fix syntax * Update prefix	2023-05-08 10:56:52 +08:00
Adarsh Sanjeev	fb38085ddb	Add wait for worker shutdown to MSQ task cancel (#14198 ) * Add wait for worker shutdown to MSQ task cancel * Fix checkstyle	2023-05-05 16:29:59 -07:00
Churro	123c4908c8	Ephemeral storage is respected from the overlod for peon tasks (#14201 )	2023-05-05 16:27:29 -07:00
Vadim Ogievetsky	4c15e978f1	Web console: misc bug fixes (#14216 ) * fixing little things * clear edit columns when switching to SQL tab * updated snapshots	2023-05-05 15:45:19 -07:00
Abhishek Radhakrishnan	6ca3fb9b08	Remove the redundant ISO-8601 text in the readme. (#14210 )	2023-05-05 11:27:29 -07:00
Abhishek Radhakrishnan	46dabab36d	Fix NPE in test parse exception report. Add more tests with different thresholds. (#14209 )	2023-05-05 10:05:41 -07:00
Clint Wylie	01e88848ce	restore .idea/misc.xml to see if it fixes intellij inspection ci (#14208 )	2023-05-05 11:47:16 +05:30
zachjsh	48cde236c4	Add columnMappings to explain plan output (#14187 ) * Add columnMappings to explain plan output * * fix checkstyle * add tests * * improve test coverage * * temporarily remove unit-test need to run ITs * * depend on build * * temporarily lower unit test threshold * * add back dependency on unit-tests * * add license headers * * fix header order * * review comments * * fix intellij inspection errors * * revert code coverage change	2023-05-04 10:36:28 -07:00
Abhishek Agarwal	edfd46ed45	Better actionable error message when druid services are not running (#14202 ) We have seen that the first-time users often don't know the next steps if druid services are unresponsive for some reason. This PR makes some of those messages a bit more clear.	2023-05-04 18:03:59 +05:30
Abhishek Radhakrishnan	68f908e511	Fix uncaught `ParseException` when reading Avro from Kafka (#14183 ) In StreamChunkParser#parseWithInputFormat, we call byteEntityReader.read() without handling a potential ParseException, which is thrown during this function call by the delegate AvroStreamReader#intermediateRowIterator. A ParseException can be thrown if an Avro stream has corrupt data or data that doesn't conform to the schema specified or for other decoding reasons. This exception if uncaught, can cause ingestion to fail.	2023-05-04 12:35:36 +05:30
Abhishek Radhakrishnan	954f3917ef	Add check for required avroBytesDecoder property that otherwise causes NPE. (#14177 )	2023-05-03 09:53:58 -07:00
AmatyaAvadhanula	ac7181bbda	Persist supervisor spec only after successful start (#14150 ) * Persist spec after successful start * Fix checkstyle. * checkstyle after mvn install	2023-05-03 18:27:39 +05:30

1 2 3 4 5 ...

12734 Commits All Branches Search

12734 Commits

All Branches