druid

Commit Graph

Author	SHA1	Message	Date
Himanshu	c078ed40fd	groupBy query: optional limit push down to segment scan (#8426 ) * groupBy query: optional limit push down to segment scan * make segment level limit push down configurable * fix teamcity errors * fix segment limit pushdown flag handling on query level config override * use equals for comparator check * fix sql and null handling * fix unused imports * handle null offset in NullableValueGroupByColumnSelectorStrategy for buffer comparator similar to RowBasedGrouperHelper.NullableRowBasedKeySerdeHelper	2019-10-08 15:35:07 -07:00
Lucas Capistrant	d801ce2f29	Update rollup table to properly reflect 0.16.0 (#8638 ) This table stated that `index_parallel` tasks were best-effort only. However, this changed with #8061 and this documentation update was simply missed.	2019-10-07 12:37:15 -07:00
Xavier Léauté	1d42551d95	Fix statsd types (#8628 ) * fix segment underReplicated/unavailable counts to be gauges instead of counters * fix jvm/gc/cpu to be a counter instead of timre jvm/gc/cpu represents the total cpu time spent for multiple gc invocations, not the time spent in each gc cycle. the number needs to be divided by jvm/gc/count to get the average gc time per cycle * update docs * fix spellcheck	2019-10-06 14:14:09 -07:00
Parag Jain	f0d74b240d	password provider for basic authentication of HttpEmitterConfig (#8618 )	2019-10-02 15:59:17 -07:00
Nishant Bangarwa	8537fbeca7	Implementing dropwizard emitter for druid (#7363 ) * Implementing dropwizard emitter for druid making metric manager and alert emitters as optional * Refactor and make things work more improvements improve docs refactrings * Fix teamcity inspections * review comments * more review comments * add limit to max number of gauges * update pom version * fix pom * review comments * review comment * review comments * fix broken doc link review comments review comments * review comments * fix checkstyle * more spell check fixes * fix travis failures	2019-10-01 14:59:30 -07:00
pdeva	db65068c42	add reference to indexer nodes (#8607 )	2019-09-30 16:45:33 -06:00
Sashidhar Thallam	51a7235ebc	Making optimal usage of multiple segment cache locations (#8038 ) * #7641 - Changing segment distribution algorithm to distribute segments to multiple segment cache locations * Fixing indentation * WIP * Adding interface for location strategy selection, least bytes used strategy impl, round-robin strategy impl, locationSelectorStrategy config with least bytes used strategy as the default strategy * fixing code style * Fixing test * Adding a method visible only for testing, fixing tests * 1. Changing the method contract to return an iterator of locations instead of a single best location. 2. Check style fixes * fixing the conditional statement * Added testSegmentDistributionUsingLeastBytesUsedStrategy, fixed testSegmentDistributionUsingRoundRobinStrategy * to trigger CI build * Add documentation for the selection strategy configuration * to re trigger CI build * updated docs as per review comments, made LeastBytesUsedStorageLocationSelectorStrategy.getLocations a synchronzied method, other minor fixes * In checkLocationConfigForNull method, using getLocations() to check for null instead of directly referring to the locations variable so that tests overriding getLocations() method do not fail * Implementing review comments. Added tests for StorageLocationSelectorStrategy * Checkstyle fixes * Adding java doc comments for StorageLocationSelectorStrategy interface * checkstyle * empty commit to retrigger build * Empty commit * Adding suppressions for words leastBytesUsed and roundRobin of ../docs/configuration/index.md file * Impl review comments including updating docs as suggested * Removing checkLocationConfigForNull(), @NotEmpty annotation serves the purpose * Round robin iterator to keep track of the no. of iterations, impl review comments, added tests for round robin strategy * Fixing the round robin iterator * Removed numLocationsToTry, updated java docs * changing property attribute value from tier to type * Fixing assert messages	2019-09-28 00:17:44 -06:00
Himanshu	9f1f5e115c	doubleMean aggregator to be used at query time (#8459 ) * doubleMean aggregator for computing mean * make docs * build fixes * address review comment: handle null args	2019-09-26 08:04:33 -07:00
Nishant Bangarwa	a75ddaad9e	Add TrustedDomain Authenticator (#8248 ) * Add TrustedDomain Authenticator update javadoc Add nullable annotations Add cautionary note fix travis failure * add IP to spell checker	2019-09-25 11:25:03 -07:00
Rye	f2a444321b	Added live reports for Kafka and Native batch task (#8557 ) * Added live reports for Kafka and Native batch task * Removed unused local variables * Added the missing unit test * Refine unit test logic, add implementation for HttpRemoteTaskRunner * checksytle fixes * Update doc descriptions for updated API * remove unnecessary files * Fix spellcheck complaints * More details for api descriptions	2019-09-23 21:08:36 -07:00
Vadim Ogievetsky	52f3f2c229	fix docs version interpolation (#8568 )	2019-09-22 17:38:55 -07:00
Vadim Ogievetsky	94298f7809	Update Kafka loading docs to use the streaming data loader (#8544 ) * fix redirects * remove useless page * fix Single server reference configurations formatting * update batch data loading * update Kafka docs * fix typos and tests * add more links * fix spelling	2019-09-22 15:00:52 -07:00
Chi Cao Minh	aeac0d4fd3	Adjust defaults for hashed partitioning (#8565 ) * Adjust defaults for hashed partitioning If neither the partition size nor the number of shards are specified, default to partitions of 5,000,000 rows (similar to the behavior of dynamic partitions). Previously, both could be null and cause incorrect behavior. Specifying both a partition size and a number of shards now results in an error instead of ignoring the partition size in favor of using the number of shards. This is a behavior change that makes it more apparent to the user that only one of the two properties will be honored (previously, a message was just logged when the specified partition size was ignored). * Fix test * Handle -1 as null * Add -1 as null tests for single dim partitioning * Simplify logic to handle -1 as null * Address review comments	2019-09-21 20:57:40 -07:00
Chi Cao Minh	99b6eedab5	Rename partition spec fields (#8507 ) * Rename partition spec fields Rename partition spec fields to be consistent across the various types (hashed, single_dim, dynamic). Specifically, use targetNumRowsPerSegment and maxRowsPerSegment in favor of targetPartitionSize and maxSegmentSize. Consistent and clearer names are easier for users to understand and use. Also fix various IntelliJ inspection warnings and doc spelling mistakes. * Fix test * Improve docs * Add targetRowsPerSegment to HashedPartitionsSpec	2019-09-20 14:59:18 -06:00
Xavier Léauté	e184d24a74	add support for dogstatsd events in statsd-emitter (#8546 ) * add support for dogstatsd events in statsd-emitter * add option to turn on alert events (off by default) * updated docs	2019-09-19 08:12:30 -07:00
Chi Cao Minh	7dcbaca658	Spellcheck docs (#8548 ) * Spellcheck docs Fix spelling mistakes in docs and add CI job for running spellcheck on docs. * Add missing license header	2019-09-17 12:47:30 -07:00
Vadim Ogievetsky	0490909ab3	Web console: Update web console docs for 0.16.0 (#8530 ) * Update webconsole docs * home view * fix annotation typo	2019-09-13 09:09:36 -07:00
Clint Wylie	75978e5b98	move google ext docs from contrib to core (#8512 ) * move google ext docs from contrib to core * fix links * revert unintended change * more links, add note to example ext doc that it was removed, unlink from sidebar	2019-09-12 09:40:39 -07:00
Jonathan Wei	0145642d8b	Move router/indexer config/API docs to main pages (#8510 ) * Move router/indexer config/API docs to main pages * Restore missing properties, fix typo * Use sentence casing * Fix broken link	2019-09-11 21:42:58 -07:00
Clint Wylie	fb078eea1e	fix web-console build in src distribution, fix kafka doc minimum version (#8502 )	2019-09-10 21:01:07 -07:00
Chi Cao Minh	14a8613d69	Exit JVM on curator unhandled errors (#8458 ) * Exit JVM on curator unhandled errors If an unhandled error occurs when curator is talking to ZooKeeper, exit the JVM in addition to stopping the lifecycle to prevent the process from being left in a zombie state. With this change, BoundedExponentialBackoffRetryWithQuit is no longer needed as when curator exceeds the configured retries, it triggers its unhandled error listeners. A new "connectionTimeoutMs" CuratorConfig setting is added mostly to facilitate testing curator unhandled errors, but it may be useful for users as well. * Address review comments	2019-09-06 16:43:59 -07:00
Clint Wylie	fd58fbc8d3	fix statds dogstatsdServiceAsTag docs example to match behavior (#8477 )	2019-09-05 19:05:25 -07:00
SeKing	6a6893b406	Fix operator mistake of expression OR (#8452 ) * Add realization for updating version of derived segments in MaterializedView * add unit test, and change code style for the sake of ease of understanding * fix document's mistake of expression	2019-09-04 21:27:18 -07:00
Lucas Capistrant	bfb02f09f8	Add druid.segmentCache.numBootstrapThreads back to the docs (#8462 )	2019-09-04 20:27:17 -07:00
legendtkl	0be4a41c06	Website Doc: fix bash command (#8442 ) * fix "gunzip -k" to "gunzip -c"	2019-08-30 22:22:09 -07:00
Clint Wylie	3baf31e9a8	add documentation for group by array based result format (#8416 )	2019-08-28 08:30:31 -07:00
Jonathan Wei	c626452b47	Add nano-quickstart single server example configuration (#8390 ) * Add nano-quickstart single server example configuration * Use two workers * Shrink processing buffers	2019-08-24 22:07:20 -07:00
Furkan KAMACI	02fe3db911	Zookeeper version is updated. (#8363 ) * Zookeeper version is updated. * Zookeeper version is updated at licenses.yaml * licenses.yaml is updated and dependencies are fixed to make the project successfully build. * Zookeeper versions are fixed at licenses.yaml	2019-08-24 22:00:43 -07:00
Jihoon Son	95fa609615	Fix wrong partitionsSpec type names in the document (#8297 ) * Fix wrong type names for partitionsSpec * add unit tests; add json properties for backward compatibility * beautify conf names * remove maxRowsPerSegment from hashed partitionsSpec * fix doc build	2019-08-23 13:44:58 -07:00
Clint Wylie	7749571a7f	order and add more ports to hadoop docker container in hadoop indexing tutorial (#8329 ) LGTM	2019-08-23 15:43:06 -05:00
Surekha	cf2a2dd917	Add group_id to the sys.tasks table (#8304 ) * Add group_id to overlord tasks API and sys.tasks table * adjust test * modify docs * Make groupId nullable * fix integration test * fix toString * Remove groupId from TaskInfo * Modify docs and tests * modify TaskMonitorTest	2019-08-22 15:28:23 -07:00
Clint Wylie	010f70b371	autogenerate NOTICE.BINARY from NOTICE and licenses.yaml (#8306 ) * migrate binary notice entries to live in licenses.yaml, use licenses.yaml and NOTICE to generate NOTICE.BINARY at distribution time * +x * move release scripts to distribution/bin, fixup notice script, trim dependencies for avro and kerberos in licenses.yaml * add missing hdfs-storage dependencies * revert to old syntax, fixes * formatting * update notices for recently updated dependencies	2019-08-21 12:46:27 -07:00
Gian Merlino	d007477742	Docusaurus build framework + ingestion doc refresh. (#8311 ) * Docusaurus build framework + ingestion doc refresh. * stick to npm instead of yarn * fix typos * restore some _bin * Adjustments. * detect and fix redirect anchors * update anchor lint * Web-console: remove specific column filters (#8343) * add clear filter * update tool kit * remove usless check * auto run * add % * Fix resource leak (#8337) * Fix resource leak * Patch comments * Enable Spotbugs NP_NONNULL_RETURN_VIOLATION (#8234) * Fixes from PR review. * Fix more anchors. * Preamble nix. * Fix more anchors, headers * clean up placeholder page * add to website lint to travis config * better broken link checking * travis fix * Fixed more broken links * better redirects * unfancy catch * fix LGTM error * link fixes * fix md issues * Addl fixes	2019-08-20 21:48:59 -07:00
Fokko Driesprong	d5a19675dd	Remove fromPigAvroStorage from the docs (#8340 ) This one has been deprecated a while ago	2019-08-20 16:34:55 -07:00
Jonathan Wei	dd2e53baf4	Clarify Avro decoder docs (#8302 )	2019-08-19 15:37:18 -05:00
Jihoon Son	31af4eb9ad	Rename maxNumSubTasks to maxNumConcurrentSubTasks for native parallel index task (#8324 )	2019-08-16 15:57:13 -07:00
Jihoon Son	5dac6375f3	Add support for parallel native indexing with shuffle for perfect rollup (#8257 ) * Add TaskResourceCleaner; fix a couple of concurrency bugs in batch tasks * kill runner when it's ready * add comment * kill run thread * fix test * Take closeable out of Appenderator * add javadoc * fix test * fix test * update javadoc * add javadoc about killed task * address comment * Add support for parallel native indexing with shuffle for perfect rollup. * Add comment about volatiles * fix test * fix test * handling missing exceptions * more clear javadoc for stopGracefully * unused import * update javadoc * Add missing statement in javadoc * address comments; fix doc * add javadoc for isGuaranteedRollup * Rename confusing variable name and fix typos * fix typos; move fetch() to a better home; fix the expiration time * add support https	2019-08-15 17:43:35 -07:00
Jihoon Son	eeae5d9365	Add a warning about experimental segment locking (#8301 ) * Add a warning about experimental segment locking * fix typo	2019-08-15 16:07:59 -07:00
Jihoon Son	a5c9c2950f	Add missing maxBytesInMemory in tuningConfig for auto compaction (#8274 ) * Add missing tuningConfigs for auto compaciton * Add doc * add test	2019-08-13 14:10:26 -05:00
Alexandre Yang	6b4d028b96	[statsd-emitter] Add config to send Druid process/service as tag (#8238 ) * [statsd-emitter] Add serviceAsTag option * [statsd-emitter] Refactor serviceAsTag option * [statsd-emitter] Update statsd.md * [statsd-emitter] add default prefix * [statsd-emitter] update statsd.md * [statsd-emitter] Remove extra spaces * [statsd-emitter] Improve docs for config `dogstatsdServiceAsTag` * [statsd-emitter] Simplify equals() for StatsDEmitterConfig.java * [statsd-emitter] Add @Nullable for StatsDEmitterConfig.java	2019-08-12 13:18:44 -07:00
Nathan	b28e252d9a	Minor Spelling Error (#8277 ) * Minor Spelling Error * Update mySQL password in docs /extensions-core/mysql update druid.metadata.storage.connector.password	2019-08-09 16:06:02 -05:00
Jonathan Wei	e88bbe71c0	Adjust default globalIngestionHeapLimitBytes for indexer, add more docs (#8255 )	2019-08-07 23:04:07 -07:00
Jonathan Wei	5e57492298	Add docs for CliIndexer as an experimental feature (#8245 ) * Experimental CliIndexer docs * PR comments	2019-08-06 15:57:17 -07:00
Lucas Capistrant	e252abedc5	Enable toggling request logging on/off for different query types (#7562 ) * Enable ability to toggle SegmentMetadata request logging on/off * Move SegmentMetadata query log filter to FilteredRequestLogger * Update documentation to reflect the segment metadata flag moving to the filtered request logger * Modify patch to allow blacklist of query types to not log to request logger * Address styling and naming requests following latest code review * Fix indentation on multiple locations per Druid style rules	2019-08-06 15:47:30 +03:00
Samarth Jain	93cf9d4ad4	SQL support for t-digest based sketch aggregators (#8100 ) * SQL support for t-digest based sketch aggregators * Fix teamcity errors * Add missing dependencies * Remove unused dependency * Address code review comments * Add checks for compression param	2019-08-05 12:01:42 -07:00
Jihoon Son	1ee828ff49	Add a cluster-wide configuration to force timeChunk lock and add a doc for segment locking (#8173 ) * Add a cluster-wide configuration to force timeChunk lock and add a doc for segment locking * add more test * javadoc for missingIntervalsInOverwriteMode * Fix test * Address comments * avoid spotbugs	2019-08-02 20:30:05 -07:00
Chi Cao Minh	4bd3bad8ba	Add IPv4 SQL functions (#8223 ) * Add IPv4 SQL functions New SQL functions for filtering IPv4 addresses: - IPV4_MATCH: Check if IP address belongs to a subnet - IPV4_PARSE: Convert string IP address to integer - IPV4_STRINGIFY: Convert integer IP address to string These are the SQL analogs of the druid expressions with the same name. Filtering is more efficient when operating on IP addresses as integers instead of strings. * Refactor operator conversions into named constants	2019-08-01 21:29:58 -07:00
Clint Wylie	01c8c82982	correct kerberos doc extension load list (#8224 )	2019-08-01 17:03:25 -07:00
Chi Cao Minh	7783b31846	Add IPv4 druid expressions (#8197 ) * Add IPv4 druid expressions New druid expressions for filtering IPv4 addresses: - ipv4address_match: Check if IP address belongs to a subnet - ipv4address_parse: Convert string IP address to long - ipv4address_stringify: Convert long IP address to string These expressions operate on IP addresses represented as either strings or longs, so that they can be applied to dimensions with mixed representation of IP addresses. The filtering is more efficient when operating on IP addresses as longs. In other words, the intended use case is: 1) Use ipv4address_parse to convert to long at ingestion time 2) Use ipv4address_match to filter (on longs) at query time 3) Use ipv4adress_stringify to convert to (readable) string at query time * Fix licenses and null handling * Simplify IPv4 expressions * Fix tests * Fix check for valid ipv4 address string	2019-08-01 11:45:04 -07:00
Surekha	f0ecdfee30	Fix `is_realtime` column behavior in sys.segments table (#8154 ) * Fix is_realtime flag * make variable final * minor changes * Modify is_realtime behavior based on review comment * Fix UT	2019-07-31 22:26:49 -06:00

1 2 3 4 5 ...

1956 Commits