druid

mirror of https://github.com/apache/druid.git synced 2025-02-11 12:35:00 +00:00

Go to file

zhangyue19921010 bddacbb1c3

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

* druid task auto scale based on kafka lag

* fix kafkaSupervisorIOConfig and KinesisSupervisorIOConfig

* druid task auto scale based on kafka lag

* fix kafkaSupervisorIOConfig and KinesisSupervisorIOConfig

* test dynamic auto scale done

* auto scale tasks tested on prd cluster

* auto scale tasks tested on prd cluster

* modify code style to solve 29055.10 29055.9 29055.17 29055.18 29055.19 29055.20

* rename test fiel function

* change codes and add docs based on capistrant reviewed

* midify test docs

* modify docs

* modify docs

* modify docs

* merge from master

* Extract the autoScale logic out of SeekableStreamSupervisor to minimize putting more stuff inside there &&  Make autoscaling algorithm configurable and scalable.

* fix ci failed

* revert msic.xml

* add uts to test autoscaler create && scale out/in and kafka ingest with scale enable

* add more uts

* fix inner class check

* add IT for kafka ingestion with autoscaler

* add new IT in groups=kafka-index named testKafkaIndexDataWithWithAutoscaler

* review change

* code review

* remove unused imports

* fix NLP

* fix docs and UTs

* revert misc.xml

* use jackson to build autoScaleConfig with default values

* add uts

* use jackson to init AutoScalerConfig in IOConfig instead of Map<>

* autoscalerConfig interface and provide a defaultAutoScalerConfig

* modify uts

* modify docs

* fix checkstyle

* revert misc.xml

* modify uts

* reviewed code change

* reviewed code change

* code reviewed

* code review

* log changed

* do StringUtils.encodeForFormat when create allocationExec

* code review && limit taskCountMax to partitionNumbers

* modify docs

* code review

Co-authored-by: yuezhang <yuezhang@freewheel.tv>

2021-03-06 14:36:52 +05:30

.github

Add instructions for updating licenses (#10894 )

2021-03-03 00:57:09 -08:00

.idea

IntelliJ inspection and checkstyle rule for "Collection.EMPTY_* field accesses replaceable with Collections.empty*()" (#9690 )

2020-06-18 09:47:07 -07:00

benchmarks

Migrate bitmap benchmarks to JMH (#10936 )

2021-03-04 12:50:55 -08:00

cloud

Fix kinesis ingestion bugs (#10761 )

2021-02-05 02:49:58 -08:00

codestyle

Fix misspellings in druid-forbidden-apis. (#10634 )

2020-12-05 15:26:57 -08:00

core

CsvInputFormat: Create a parser per InputEntityReader. (#10923 )

2021-02-27 18:37:05 -08:00

dev

Add instructions for updating licenses (#10894 )

2021-03-03 00:57:09 -08:00

distribution

Add config and header support for confluent schema registry. (#10314 )

2021-02-27 14:25:35 -08:00

docs

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

examples

Add the ability to supply client certificate to dsql comand line tool. (#10765 )

2021-02-11 20:16:47 -08:00

extendedset

Bump dev version to 0.22.0-SNAPSHOT (#10759 )

2021-01-15 13:16:23 -08:00

extensions-contrib

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

extensions-core

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

hll

Bump dev version to 0.22.0-SNAPSHOT (#10759 )

2021-01-15 13:16:23 -08:00

hooks

Add git pre-commit hook to source control (#9554 )

2020-06-05 11:19:42 -10:00

indexing-hadoop

Granularity interval materialization (#10742 )

2021-01-29 06:02:10 -08:00

indexing-service

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

integration-tests

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

licenses

Web console: Improve the handling of extreme data (funky datasources, longs) (#10641 )

2020-12-08 09:25:14 -08:00

processing

Migrate bitmap benchmarks to JMH (#10936 )

2021-03-04 12:50:55 -08:00

publications

De-incubation cleanup in code, docs, packaging (#9108 )

2020-01-03 12:33:19 -05:00

server

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

services

Fix kinesis ingestion bugs (#10761 )

2021-02-05 02:49:58 -08:00

sql

Supporting filters in the left base table for join datasources (#10697 )

2021-03-04 10:39:21 -08:00

web-console

Web console: remove namespace prop that does not exist from JDBC lookup (#10888 )

2021-02-17 17:07:32 -08:00

website

Add config and header support for confluent schema registry. (#10314 )

2021-02-27 14:25:35 -08:00

.asf.yaml

Add .asf.yaml. (#9083 )

2019-12-20 16:45:38 -08:00

.backportrc.json

Add 0.18.0 to .backportrc.json to facilitate backport. (#9661 )

2020-04-11 13:49:04 -07:00

.codecov.yml

Use Codecov (#8388 )

2019-08-28 08:49:30 -07:00

.dockerignore

Add docker container for druid (#6896 )

2019-02-08 12:12:28 +00:00

.gitignore

Web console basic end-to-end-test (#9595 )

2020-04-09 12:38:09 -07:00

.lgtm.yml

Suppress LGTM warnings about stack trace exposure (#9631 )

2020-04-09 17:31:03 -07:00

.travis.yml

Ldap integration tests (#10901 )

2021-02-23 13:29:57 -08:00

CONTRIBUTING.md

Fix numbered list formatting in markdown. (#9664 )

2020-04-21 20:18:12 -07:00

LABELS

Add plain text README.txt, use relative link from README.md to build.md (#7611 )

2019-05-09 21:29:26 -07:00

LICENSE

support Aliyun OSS service as deep storage (#9898 )

2020-07-01 22:20:53 -07:00

licenses.yaml

Upgrade jetty to latest version (#10937 )

2021-03-04 08:28:50 -06:00

NOTICE

Update NOTICE copyright year (#10834 )

2021-02-02 14:02:26 -08:00

owasp-dependency-check-suppressions.xml

Suppress CVE-2017-15288 and upgrade bcprov-ext-jdk15o (#10933 )

2021-03-02 16:18:27 -08:00

pom.xml

Dynamic auto scale Kafka-Stream ingest tasks (#10524 )

2021-03-06 14:36:52 +05:30

README.md

Update badge for travis in README.md (#10717 )

2021-01-07 18:39:58 -08:00

README.template

De-incubation cleanup in code, docs, packaging (#9108 )

2020-01-03 12:33:19 -05:00

setup-hooks.sh

Add git pre-commit hook to source control (#9554 )

2020-06-05 11:19:42 -10:00

upload.sh

…

README.md

Apache Druid

Druid is a high performance real-time analytics database. Druid's main value add is to reduce time to insight and action.

Druid is designed for workflows where fast queries and ingest really matter. Druid excels at powering UIs, running operational (ad-hoc) queries, or handling high concurrency. Consider Druid as an open source alternative to data warehouses for a variety of use cases.

Getting started

You can get started with Druid with our local or Docker quickstart.

Druid provides a rich set of APIs (via HTTP and JDBC) for loading, managing, and querying your data. You can also interact with Druid via the built-in console (shown below).

Load data

Load streaming and batch data using a point-and-click wizard to guide you through ingestion setup. Monitor one off tasks and ingestion supervisors.

Manage the cluster

Manage your cluster with ease. Get a view of your datasources, segments, ingestion tasks, and services from one convenient location. All powered by SQL systems tables, allowing you to see the underlying query for each view.

Issue queries

Use the built-in query workbench to prototype DruidSQL and native queries or connect one of the many tools that help you make the most out of Druid.

Documentation

You can find the documentation for the latest Druid release on the project website.

If you would like to contribute documentation, please do so under /docs in this repository and submit a pull request.

Community

Community support is available on the druid-user mailing list, which is hosted at Google Groups.

Development discussions occur on dev@druid.apache.org, which you can subscribe to by emailing dev-subscribe@druid.apache.org.

Chat with Druid committers and users in real-time on the #druid channel in the Apache Slack team. Please use this invitation link to join the ASF Slack, and once joined, go into the #druid channel.

Building from source

Please note that JDK 8 is required to build Druid.

For instructions on building Druid from source, see docs/development/build.md

Contributing

Please follow the community guidelines for contributing.

For instructions on setting up IntelliJ dev/intellij-setup.md

License

Apache License, Version 2.0

Languages

Java 62.4%

ReScript 30.7%

TypeScript 3.1%

Euphoria 0.9%

Csound 0.8%

Other 1.9%