druid

Commit Graph

Author	SHA1	Message	Date
Gian Merlino	fe42db98ac	URIExtractionNamespace: Avoid problems due to canonicalization of lookup fields. (#4307 ) Disables canonicalization for simpleJson, where expect field names to be unique anyway. Keeps canonicalization enabled for customJson, but avoids sharing the table with the global ObjectMapper.	2017-05-24 17:41:04 -07:00
Goh Wei Xiang	b77fab8a30	Replace usages of CountingMap with Object2LongMap (#4320 ) * Replaces use of CountingMap with Object2LongMap from fastutil. * Remove CountingMap classes and minor fixes * Added additional test cases for DatasourceInputFormat. * Added additional test cases for CoordinatorStats. * Not materializing segment list. * Put in this fix because it is failing the test on its expected behavior. * Added missing header.	2017-05-24 17:40:32 -07:00
Jihoon Son	b578adacae	Improve concurrency of SegmentManager (#4298 ) * Improve concurrency of SegmentManager * Fix SegmentManager and add more tests * Add more tests * Add null check to TimelineEntry * Remove empty data source and check null in getTimeline() * Add a comment for returning null in compute() * Make SegmentManager LazySingleton	2017-05-24 04:41:24 +09:00
Roman Leventov	f97c49ba0e	Don't use QueryMetrics from multiple threads in DirectDruidClient (fixes #4308 ) (#4309 ) * Don't use QueryMetrics from multiple threads in DirectDruidClient * reponseMetrics	2017-05-23 10:07:27 -07:00
Gian Merlino	2bd4c0930f	Fix "quarter" granularity serialization. (#4316 )	2017-05-23 10:06:17 -07:00
Gian Merlino	9283807ad7	GroupByQuery: Fix type-spanning comparisons. (#4317 ) Jackson deserializes integers sometimes as int and sometimes as long, depending on how big they are. This leads to ClassCastException when comparing deserialized values as part of groupBy merging on the broker.	2017-05-23 10:06:04 -07:00
李成露(StefanLee)	22977780aa	Doc (#4217 ) * Fixed (#4216) Modify the default value of `druid.server.http.numThreads` to `Math.max(10, (Runtime.getRuntime().availableProcessors() * 17) / 16 + 2) + 30` * Fixed(#4216) Modify the default value of `druid.server.http.numThreads` to `max(10, (Number of cores * 17) / 16 + 2) + 30` * Fixed(#4216) Modify the default value of `druid.server.http.numThreads` to `max(10, (Number of cores * 17) / 16 + 2) + 30`	2017-05-23 17:04:52 +09:00
Jonathan Wei	d49e53e6c2	Timeout and maxScatterGatherBytes handling for queries run by Druid SQL (#4305 ) * Timeout and maxScatterGatherBytes handling for queries run by Druid SQL * Address PR comments * Fix contexts in CalciteQueryTest * Fix contexts in QuantileSqlAggregatorTest	2017-05-23 16:57:51 +09:00
Gian Merlino	22e5f52d00	Workaround for non-thread-safe use of CardinalityAggregator. (#4304 )	2017-05-23 10:33:03 +09:00
Jonathan Wei	e043bf88ec	Add a ServerType for peons (#4295 ) * Add a ServerType for peons * Add toString() method, toString() test, unsupported type check * Use ServerType enum in DruidServer and DruidServerMetadata	2017-05-22 17:24:59 -05:00
Roman Leventov	8ec3a29af0	Don't pass QueryMetrics down in concurrent and async QueryRunners (fixes #4279 ) (#4288 ) * Don't pass QueryMetrics down in concurrent and async QueryRunners * Rename QueryPlus.threadSafe() to withoutThreadUnsafeState(); Update QueryPlus.withQueryMetrics() Javadocs; Fix generics in MetricsEmittingQueryRunner and CpuTimeMetricQueryRunner; Make DefaultQueryMetrics to fail fast on modifications from concurrent threads	2017-05-22 13:42:09 -05:00
Jihoon Son	000b0ffed7	Increase the max heap size for strict compilation (#4306 )	2017-05-21 03:42:44 +09:00
Gian Merlino	adeecc0e72	Add /isLeader call to overlord and coordinator. (#4282 ) This is useful for putting them behind load balancers or proxies, as it lets the load balancer know which server is currently active through an http health check. Also makes the method naming a little more consistent between coordinator and overlord code.	2017-05-18 20:46:13 -05:00
Roman Leventov	7479cbde68	Make CacheScheduler a singleton (#4293 )	2017-05-18 15:46:02 -07:00
Benedict Jin	cdd521fb23	Update outdated RLE paper and improve some code refactoring (#4286 ) * Update outdated RLE paper and improve some code refactoring * Roll back CONCISE's abbreviation	2017-05-18 12:26:24 -07:00
Gian Merlino	8ca7f9410e	SQL: Add test for concurrent JDBC queries. (#4290 )	2017-05-18 12:25:15 -07:00
Jihoon Son	5c0a7ad2f8	Make realtimes available for loading segments (#4148 ) * Add ServerType * Add realtimes to DruidCluster * fix test fails * Add SegmentManager * Fix equals and hashCode of ServerHolder * Address comments and add more tests * Address comments	2017-05-18 10:03:39 -05:00
Jihoon Son	733dfc9b30	Add PrefetchableTextFilesFirehoseFactory for cloud storage types (#4193 ) * Add PrefetcheableTextFilesFirehoseFactory * fix comment * exception handling * Fix wrong json property * Remove ReplayableFirehoseFactory and fix misspelling * Defer object initialization * Add a temporaryDirectory parameter to FirehoseFactory.connect() * fix when cache and fetch are disabled * Address comments * Add more test * Increase timeout for test * Add wrapObjectStream * Move methods to Firehose from PrefetchableFirehoseFactory * Cleanup comment * add directory listing to s3 firehose * Rename a variable * Addressing comments * Update document * Support disabling prefetch * Fix race condition * Add fetchLock * Remove ReplayableFirehoseFactoryTest * Fix compilation error * Fix test failure * Address comments * Add default implementation for new method	2017-05-18 15:37:18 +09:00
Maksim Logvinenko	d45dad2b44	Remove boxing/unboxing in indexer (#4269 ) * Remove boxing/unboxing in indexer * Fix rowIndex visibility * Cleanup	2017-05-17 19:13:53 -05:00
Himanshu	daa8ef8658	Optional long-polling based segment announcement via HTTP instead of Zookeeper (#3902 ) * Optional long-polling based segment announcement via HTTP instead of Zookeeper * address review comments * make endpoint /druid-internal/v1 instead of /druid/internal so that jetty qos filters can be configured easily when needed * update segment callback initialization to be called only after first segment list fetch has been succeeded from all servers * address review comments * remove size check not required anymore as only segment servers announce themselves and not all peon processes * annouce segment server on historical only after cached segments are loaded * fix checkstyle errors	2017-05-17 16:31:58 -05:00
Himanshu	0e056863e4	fix timeout check bug in DirectDruidClient (#4287 )	2017-05-17 13:47:32 -07:00
Roman Leventov	d9f423f55d	Make QueryMetrics factories configurable (#4268 ) * Ensure QueryMetrics factories accept Json ObjectMapper; Make QueryMetrics factories configurable * Update QueryMetrics Javadocs * Add javadocs to QueryMetrics factories * Move queryMetricsFactory defaults to getter methods of config classes	2017-05-17 08:41:59 -07:00
Gian Merlino	ddc2e68998	Remove cache keys from HavingSpecs. (#4280 ) * Remove cache keys from HavingSpecs. They weren't used, since they aren't part of the groupBy cache key. Also, it's good that they weren't used, since many of them had value truncation bugs. * Fix imports. * Fix test.	2017-05-16 22:13:02 -07:00
Gian Merlino	22f20f2207	IngestSegmentFirehoseTest: Add more tests for reindexing. (#4285 ) * IngestSegmentFirehoseTest: Add more tests for reindexing. * Nix unused imports.	2017-05-16 22:12:26 -07:00
Gian Merlino	51872fd310	Log max memory on startup too, in case Xmx and Xms are different. (#4283 )	2017-05-16 20:06:34 -05:00
Roman Leventov	d400f23791	Monomorphic processing of TopN queries with simple double aggregators over historical segments (part of #3798 ) (#4079 ) * Monomorphic processing of topN queries with simple double aggregators and historical segments * Add CalledFromHotLoop annocations to specialized methods in SimpleDoubleBufferAggregator * Fix a bug in Historical1SimpleDoubleAggPooledTopNScannerPrototype * Fix a bug in SpecializationService * In SpecializationService, emit maxSpecializations warning only once * Make GenericIndexed.theBuffer final * Address comments * Newline * Reapply `439c906` (Make GenericIndexed.theBuffer final) * Remove extra PooledTopNAlgorithm.capabilities field * Improve CachingIndexed.inspectRuntimeShape() * Fix CompressedVSizeIntsIndexedSupplier.inspectRuntimeShape() * Don't override inspectRuntimeShape() in subclasses of CompressedVSizeIndexedInts * Annotate methods in specializations of DimensionSelector and FloatColumnSelector with @CalledFromHotLoop * Make ValueMatcher to implement HotLoopCallee * Doc fix * Fix inspectRuntimeShape() impl in ExpressionSelectors * INFO logging of specialization events * Remove modificator * Fix OrFilter * Fix AndFilter * Refactor PooledTopNAlgorithm.scanAndAggregate() * Small refactoring * Add 'nothing to inspect' messages in empty HotLoopCallee.inspectRuntimeShape() implementations * Don't care about runtime shape in tests * Fix accessor bugs in Historical1SimpleDoubleAggPooledTopNScannerPrototype and HistoricalSingleValueDimSelector1SimpleDoubleAggPooledTopNScannerPrototype, cover them with tests * Doc wording * Address comments * Remove MagicAccessorBridge and ensure Offset subclasses are public * Attach error message to element	2017-05-16 16:19:55 -07:00
Roman Leventov	b7a52286e8	Make @Override annotation obligatory (#4274 ) * Make MissingOverride an error * Make travis stript to fail fast * Add missing Override annotations * Comment	2017-05-16 13:30:30 -05:00
David Lim	8333043b7b	add skipOffsetGaps flag (#4256 )	2017-05-16 12:19:28 -06:00
Himanshu	136b2fae72	improve query timeout handling and limit max scatter-gather bytes (#4229 ) * improve query timeout handling and limit max scatter-gather bytes * address review comments	2017-05-16 12:47:32 -05:00
Benedict Jin	e823085866	Improve `collection` related things that reusing a immutable object instead of creating a new object (#4135 )	2017-05-17 01:38:51 +09:00
Charles Allen	e4add598f0	Add INTELLIJ_SETUP.md (#4261 ) * Add INTELLIJ_SETUP.md * Fix `.idea/runConfigurations` * Update INTELLIJ_SETUP.md * Update INTELLIJ_SETUP.md * Address Comments	2017-05-17 01:26:16 +09:00
Yuya Fujiwara	8010d7f28d	fix typo: seraching -> searching (#4273 )	2017-05-17 01:25:16 +09:00
Jihoon Son	50a4ec2b0b	Add support for headers and skipping thereof for CSV and TSV (#4254 ) * initial commit * small fixes * fix bug * fix bug * address code review * more cr * more cr * more cr * fix * Skip head rows for CSV and TSV * Move checking skipHeadRows to FileIteratingFirehose * Remove checking null iterators * Remove unused imports * Address comments * Fix compilation error * Address comments * Add more tests * Add a comment to ReplayableFirehose * Addressing comments * Add docs and fix typos	2017-05-15 22:57:31 -07:00
Fokko Driesprong	5ca67644e7	Remove slf4j as dependencies (#4233 ) From the kafka-schema-registry-client in the avro extension slf4j will be packaged into the distribution. We don't want this as it will conflict and throw a slf4j multiple bindings warning. This will cause slf4j to fall back to no-operation (NOP) binding.	2017-05-12 15:59:14 +09:00
satishbhor	5e6539fec6	Average Server Percent Used: NaN% Error when server startup is in progress (fixes #4214 ) (#4240 ) * Fix lz4 library incompatibility in kafka-indexing-service extension #3266 * Bumped Kafka version to 0.10.2.0 for : Fix lz4 library incompatibility in kafka-indexing-service extension #3266 * Replaced Lists.newArrayList() with Collections.singletonList() For Fix lz4 library incompatibility in kafka-indexing-service extension #4115 * Fixed: Average Server Percent Used: NaN% Error when server startup is in progress #4214	2017-05-12 15:56:17 +09:00
Roman Leventov	1ebfa22955	Update Error prone configuration; Fix bugs (#4252 ) * Make Errorprone the default compiler * Address comments * Make Error Prone's ClassCanBeStatic rule a error * Preconditions allow only %s pattern * Fix DruidCoordinatorBalancerTester * Try to give the compiler more memory * Remove distribution module activation on jdk 1.8 because only jdk 1.8 is used now * Don't show compiler warnings * Try different travis script * Fix travis.yml * Make Error Prone optional again * For error-prone compiler * Increase compiler's maxmem * Don't run Error Prone for benchmarks because of OOM * Skip install step in Travis * Remove MetricHolder.writeToChannel() * In travis.yml, check compilation before tests, because it may fail faster	2017-05-12 15:55:17 +09:00
Roman Leventov	e09e892477	Refactor QueryRunner to accept QueryPlus: Query + QueryMetrics (part of #3798 ) (#4184 ) * Add QueryPlus. Add QueryRunner.run(QueryPlus, Map) method with default implementation, to replace QueryRunner.run(Query, Map). * Fix GroupByMergingQueryRunnerV2 * Fix QueryResourceTest * Expand the comment to Query.run(walker, context) * Remove legacy version of BySegmentSkippingQueryRunner.doRun() * Add LegacyApiQueryRunnerTest and be more specific about legacy API removal plans in Druid 0.11 in Javadocs	2017-05-10 12:25:00 -07:00
David Lim	11538e2ece	kafka generated shards were wrongly being marked as overshadowed if they extended a NumberedShardSpec with a non-zero number of total shards (#4257 )	2017-05-09 12:19:24 +09:00
Parag Jain	1fd177039d	fix auto reset - pause task instead of putting thread to sleep (#4244 )	2017-05-08 15:08:25 -07:00
Parag Jain	eb8e1b0a97	Prevent interrupted exception from polluting log during supervisor shutdown (#4253 ) * Prevent interrupted exception from polluting log during supervisor shutdown * do nothing in case of InterruptedException	2017-05-08 15:05:25 -07:00
Himanshu	e02f783e82	do not use --clean option when using bundle-contrib-exts profile so that core extensions are not wiped out (#4223 )	2017-05-08 14:26:16 -05:00
Himanshu	462f6482df	optionally add extensions to explicitly specified hadoopContainerClassPath (#4230 ) * optionally add extensions to explicitly specified hadoopContainerClassPath * note extensions always pushed in hadoop container when druid.extensions.hadoopContainerDruidClasspath is not provided explicitly	2017-05-08 14:24:14 -05:00
Pierre	bba31e0c8b	close aggregators in indexing-hadoop mappers (#4251 )	2017-05-05 08:29:13 -07:00
Pierre	e9872f0695	do not flush on closed stream (#4250 )	2017-05-05 09:19:20 +09:00
Himanshu	417714d228	additional lookup status discovery http endpoints at coordinator (#4228 ) * additional lookup status discovery http endpoints at coordinator * more changes * jsonize the error msgs as well * fix tests	2017-05-04 11:15:30 -07:00
Roman Leventov	8277284d67	Add Checkstyle rule to force comments to classes and methods to be Javadoc comments (#4239 )	2017-05-04 11:14:41 -07:00
Gian Merlino	f0fd8ba191	Add supervisors to overlord console. (#4248 )	2017-05-04 11:13:12 -07:00
Gian Merlino	d0f89e969a	Ignore misnamed segment cache info files. (#4245 ) * Ignore misnamed segment cache info files. Fixes a bug where historical nodes could announce the same segment twice, ultimately leading to historicals and their watchers (like coordinators and brokers) being out of sync about which segments are served. This could be caused if Druid is switched from local time to UTC, because that causes the same segment descriptors to lead to different identifiers (an identifier with local time interval before the switch, and UTC interval after the switch). In turn this causes that segment descriptor to be written to multiple segment cache info files and potentially get announced twice. Later, if the historical receives a drop request, it drops the segment and unannounces it once, but the other announcement would stick around in an ephemeral znode forever, confusing coordinators and brokers. * Only alert once.	2017-05-03 22:02:37 -06:00
Parag Jain	4502c207af	fix injection bug and documentation (#4243 )	2017-05-03 15:07:43 -05:00
Parag Jain	f9a61ea2ba	Kafka lag emitter - Kafka Indexing Service (#4194 ) * Kafka lag emitter * enforce minimum emit period to a minute * fixed comment	2017-05-02 17:30:07 -06:00

... 3 4 5 6 7 ...

8064 Commits All Branches Search

8064 Commits

All Branches