Apache Druid: a high performance real-time analytics database.
Go to file
zhangyue19921010 bddacbb1c3
Dynamic auto scale Kafka-Stream ingest tasks (#10524)
* druid task auto scale based on kafka lag

* fix kafkaSupervisorIOConfig and KinesisSupervisorIOConfig

* druid task auto scale based on kafka lag

* fix kafkaSupervisorIOConfig and KinesisSupervisorIOConfig

* test dynamic auto scale done

* auto scale tasks tested on prd cluster

* auto scale tasks tested on prd cluster

* modify code style to solve 29055.10 29055.9 29055.17 29055.18 29055.19 29055.20

* rename test fiel function

* change codes and add docs based on capistrant reviewed

* midify test docs

* modify docs

* modify docs

* modify docs

* merge from master

* Extract the autoScale logic out of SeekableStreamSupervisor to minimize putting more stuff inside there &&  Make autoscaling algorithm configurable and scalable.

* fix ci failed

* revert msic.xml

* add uts to test autoscaler create && scale out/in and kafka ingest with scale enable

* add more uts

* fix inner class check

* add IT for kafka ingestion with autoscaler

* add new IT in groups=kafka-index named testKafkaIndexDataWithWithAutoscaler

* review change

* code review

* remove unused imports

* fix NLP

* fix docs and UTs

* revert misc.xml

* use jackson to build autoScaleConfig with default values

* add uts

* use jackson to init AutoScalerConfig in IOConfig instead of Map<>

* autoscalerConfig interface and provide a defaultAutoScalerConfig

* modify uts

* modify docs

* fix checkstyle

* revert misc.xml

* modify uts

* reviewed code change

* reviewed code change

* code reviewed

* code review

* log changed

* do StringUtils.encodeForFormat when create allocationExec

* code review && limit taskCountMax to partitionNumbers

* modify docs

* code review

Co-authored-by: yuezhang <yuezhang@freewheel.tv>
2021-03-06 14:36:52 +05:30
.github Add instructions for updating licenses (#10894) 2021-03-03 00:57:09 -08:00
.idea IntelliJ inspection and checkstyle rule for "Collection.EMPTY_* field accesses replaceable with Collections.empty*()" (#9690) 2020-06-18 09:47:07 -07:00
benchmarks Migrate bitmap benchmarks to JMH (#10936) 2021-03-04 12:50:55 -08:00
cloud Fix kinesis ingestion bugs (#10761) 2021-02-05 02:49:58 -08:00
codestyle Fix misspellings in druid-forbidden-apis. (#10634) 2020-12-05 15:26:57 -08:00
core CsvInputFormat: Create a parser per InputEntityReader. (#10923) 2021-02-27 18:37:05 -08:00
dev Add instructions for updating licenses (#10894) 2021-03-03 00:57:09 -08:00
distribution Add config and header support for confluent schema registry. (#10314) 2021-02-27 14:25:35 -08:00
docs Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
examples Add the ability to supply client certificate to dsql comand line tool. (#10765) 2021-02-11 20:16:47 -08:00
extendedset Bump dev version to 0.22.0-SNAPSHOT (#10759) 2021-01-15 13:16:23 -08:00
extensions-contrib Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
extensions-core Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
hll Bump dev version to 0.22.0-SNAPSHOT (#10759) 2021-01-15 13:16:23 -08:00
hooks Add git pre-commit hook to source control (#9554) 2020-06-05 11:19:42 -10:00
indexing-hadoop Granularity interval materialization (#10742) 2021-01-29 06:02:10 -08:00
indexing-service Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
integration-tests Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
licenses Web console: Improve the handling of extreme data (funky datasources, longs) (#10641) 2020-12-08 09:25:14 -08:00
processing Migrate bitmap benchmarks to JMH (#10936) 2021-03-04 12:50:55 -08:00
publications De-incubation cleanup in code, docs, packaging (#9108) 2020-01-03 12:33:19 -05:00
server Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
services Fix kinesis ingestion bugs (#10761) 2021-02-05 02:49:58 -08:00
sql Supporting filters in the left base table for join datasources (#10697) 2021-03-04 10:39:21 -08:00
web-console Web console: remove namespace prop that does not exist from JDBC lookup (#10888) 2021-02-17 17:07:32 -08:00
website Add config and header support for confluent schema registry. (#10314) 2021-02-27 14:25:35 -08:00
.asf.yaml Add .asf.yaml. (#9083) 2019-12-20 16:45:38 -08:00
.backportrc.json Add 0.18.0 to .backportrc.json to facilitate backport. (#9661) 2020-04-11 13:49:04 -07:00
.codecov.yml Use Codecov (#8388) 2019-08-28 08:49:30 -07:00
.dockerignore Add docker container for druid (#6896) 2019-02-08 12:12:28 +00:00
.gitignore Web console basic end-to-end-test (#9595) 2020-04-09 12:38:09 -07:00
.lgtm.yml Suppress LGTM warnings about stack trace exposure (#9631) 2020-04-09 17:31:03 -07:00
.travis.yml Ldap integration tests (#10901) 2021-02-23 13:29:57 -08:00
CONTRIBUTING.md Fix numbered list formatting in markdown. (#9664) 2020-04-21 20:18:12 -07:00
LABELS Add plain text README.txt, use relative link from README.md to build.md (#7611) 2019-05-09 21:29:26 -07:00
LICENSE support Aliyun OSS service as deep storage (#9898) 2020-07-01 22:20:53 -07:00
NOTICE Update NOTICE copyright year (#10834) 2021-02-02 14:02:26 -08:00
README.md Update badge for travis in README.md (#10717) 2021-01-07 18:39:58 -08:00
README.template De-incubation cleanup in code, docs, packaging (#9108) 2020-01-03 12:33:19 -05:00
licenses.yaml Upgrade jetty to latest version (#10937) 2021-03-04 08:28:50 -06:00
owasp-dependency-check-suppressions.xml Suppress CVE-2017-15288 and upgrade bcprov-ext-jdk15o (#10933) 2021-03-02 16:18:27 -08:00
pom.xml Dynamic auto scale Kafka-Stream ingest tasks (#10524) 2021-03-06 14:36:52 +05:30
setup-hooks.sh Add git pre-commit hook to source control (#9554) 2020-06-05 11:19:42 -10:00
upload.sh Adding licenses and enable apache-rat-plugin. (#6215) 2018-09-18 08:39:26 -07:00

README.md

Slack Build Status Language grade: Java Coverage Status Docker Helm


Website | Documentation | Developer Mailing List | User Mailing List | Slack | Twitter | Download


Apache Druid

Druid is a high performance real-time analytics database. Druid's main value add is to reduce time to insight and action.

Druid is designed for workflows where fast queries and ingest really matter. Druid excels at powering UIs, running operational (ad-hoc) queries, or handling high concurrency. Consider Druid as an open source alternative to data warehouses for a variety of use cases.

Getting started

You can get started with Druid with our local or Docker quickstart.

Druid provides a rich set of APIs (via HTTP and JDBC) for loading, managing, and querying your data. You can also interact with Druid via the built-in console (shown below).

Load data

data loader Kafka

Load streaming and batch data using a point-and-click wizard to guide you through ingestion setup. Monitor one off tasks and ingestion supervisors.

Manage the cluster

management

Manage your cluster with ease. Get a view of your datasources, segments, ingestion tasks, and services from one convenient location. All powered by SQL systems tables, allowing you to see the underlying query for each view.

Issue queries

query view combo

Use the built-in query workbench to prototype DruidSQL and native queries or connect one of the many tools that help you make the most out of Druid.

Documentation

You can find the documentation for the latest Druid release on the project website.

If you would like to contribute documentation, please do so under /docs in this repository and submit a pull request.

Community

Community support is available on the druid-user mailing list, which is hosted at Google Groups.

Development discussions occur on dev@druid.apache.org, which you can subscribe to by emailing dev-subscribe@druid.apache.org.

Chat with Druid committers and users in real-time on the #druid channel in the Apache Slack team. Please use this invitation link to join the ASF Slack, and once joined, go into the #druid channel.

Building from source

Please note that JDK 8 is required to build Druid.

For instructions on building Druid from source, see docs/development/build.md

Contributing

Please follow the community guidelines for contributing.

For instructions on setting up IntelliJ dev/intellij-setup.md

License

Apache License, Version 2.0