doris

Author	SHA1	Message	Date
Gabriel	a038fdaec6	[Bug](pipeline) Fix bug in non-local exchange on pipeline engine (#16463 ) Currently, for broadcast shuffle, we serialize a block once and then send it by RPC through multiple channel. After this, we will serialize next block in the same memory for consideration of memory reuse. However, since the RPC is asynchronized, maybe the next block serialization will happen before sending the previous block. So, in this PR, I use a ref count to identify if the serialized block can be reuse in broadcast shuffle.	2023-02-09 19:22:40 +08:00
Ashin Gau	539fd684e9	[improvement](filecache) use dynamic segment size to cache remote file block (#16485 ) `CachedRemoteFileReader` has used fixed segment size(file_cache_max_file_segment_size=4M) to cache remote file blocks. However, the column size in a rowgroup/strip maybe smaller than 10K if a parquet/orc file has many columns, resulting in particularly serious read amplification. For example: Q1 in clickbench: select count() from hits ``` - FileCache: 0ns - IOHitCacheNum: 552 - IOTotalNum: 835 - ReadFromFileCacheBytes: 19.98 MB - ReadFromWriteCacheBytes: 0.00 - ReadTotalBytes: 29.52 MB - SkipCacheBytes: 0.00 - WriteInFileCacheBytes: 915.77 MB - WriteInFileCacheNum: 283 ``` Only 30MB of data is needed, but 900MB+ of data is read from hdfs. The query time of Q1(single scan thread) increased from 5.17s* to 24.45s when enable file cache. Therefore, this PR introduce dynamic segment size which is based on the `read_size` of the data. In order to prevent too small or too large IO, the segment size is limited in [4096, file_cache_max_file_segment_size]. Q1 in clickbench is 5.66s when enable file cache. The performance is almost the same as if the cache is disabled, and the data size read from hdfs is reduced to 45MB. ``` - FileCache: 0ns - IOHitCacheNum: 297 - IOTotalNum: 835 - ReadFromFileCacheBytes: 8.73 MB - ReadFromWriteCacheBytes: 0.00 - ReadTotalBytes: 29.52 MB - SkipCacheBytes: 0.00 - WriteInFileCacheBytes: 45.66 MB - WriteInFileCacheNum: 544 ``` ## Remaining Problems Small queries may result in a large number of small files(4KB at least), and the `BE` saves too much meta information of cached segments. ## Fix bug `FileCachePolicy` in `FileReaderOptions` is a constant reference, but the parameter passed in `FileFactory::create_file_reader` is a temporary variable, resulting in segmentation fault.	2023-02-09 16:39:10 +08:00
Gabriel	e48a033338	[Bug](pipeline) Support projection in UnionSourceOperator (#16525 )	2023-02-09 14:43:44 +08:00
HappenLee	7d035486ad	[Opt](vec) opt the fast execute logic to remove useless function call (#16532 )	2023-02-09 14:12:40 +08:00
yiguolei	646ba2cc88	[bugfix](scannode) 1. make rows_read correct 2. use single scanner if has limit clause (#16473 ) make rows_read correct so that the scheduler could using this correctly. use single scanner if has limit clause. Move it from fragment context to scannode. --------- Co-authored-by: yiguolei <yiguolei@gmail.com> Co-authored-by: github-actions[bot] <41898282+github-actions[bot]@users.noreply.github.com>	2023-02-09 14:12:18 +08:00
Xiaocc	0142ef8b95	[improvement](scanner) Supports bthread scanner (#16031 )	2023-02-09 10:24:56 +08:00
yixiutt	9f8753ffd2	[bugfix](vertical_compaction) fix base_compaction delete_sign handler (#16469 ) In vertical base compaction, same rows will be filtered in vertical_merge_iterator, we should skip these filtered rows when set agg flag of delete sign. For example, schema is a,b,delete_sign, and data is 1,1,1 1,1,0 1,1,0 2,2,1 2,2 and Block we get in VerticalBlockReader is 1,1,1 2,2,1 and we should set agg flag idex 0,4 to true when handle delete sign, so we add a function continuous_agg_count to skip same rows filtered in VerticalMergeIterator.	2023-02-09 10:13:41 +08:00
plat1ko	e1f1386395	[fix](cooldown) Rewrite update cooldown conf (#16488 ) Remove error-prone CooldownJob, and use CooldownConfHandler to update Tablet's cooldown conf. Some bug fix about cooldown.	2023-02-09 09:12:55 +08:00
Gabriel	d1c6b81140	[Bug](log) add some log to find out bug (#16518 )	2023-02-08 21:23:02 +08:00
HappenLee	f71fc3291f	[Bug](fix) right anti join error result when batch size is low (#16510 )	2023-02-08 17:26:19 +08:00
lihangyu	d956cb13af	[Bug](point query) Reusable in PointQueryExecutor should call init before add to LookupCache (#16489 ) Otherwise in high concurrent query, _block_pool maybe used before Reusable::init done in other threads	2023-02-08 16:05:59 +08:00
TengJianPing	f6a20f844b	[fix](hashjoin) join produce blocks with rows larger than batch size: handle join with other conjuncts (#16402 )	2023-02-08 14:26:35 +08:00
abmdocrt	41947c73eb	[Feature](array-function) Support array functions for nested type datev2 and datetimev2 (#16382 )	2023-02-08 12:51:07 +08:00
Kang	cf18de14b5	[fix](writer) add _is_closed state to DeltaWriter and avoid write/close core after close (#16453 )	2023-02-07 22:40:26 +08:00
Jerry Hu	91325e5ca3	[fix](pipeline) incorrect result when disabling sharing hash table (#16476 )	2023-02-07 21:25:32 +08:00
yixiutt	f90d844a53	[improvement](compaction) enable compaction in TABLET_NOTREADY (#16470 ) If alter task in queue, compaction is not enabled and may cause too much version. Keep last 10 version in new tablet so that base tablet's max version will not be merged and than we can copy data from base tablet to new tablet.	2023-02-07 19:58:23 +08:00
HappenLee	9114896178	[DecimalV3](opt) opt the function of decimalv3 to_string logic (#16427 )	2023-02-07 13:28:07 +08:00
yiguolei	d390e63a03	[enhancement](stream receiver) make stream receiver exception safe (#16412 ) make stream receiver exception safe change get_block(block*) to get_block(block , bool* eos) unify stream semantic	2023-02-07 12:44:20 +08:00
yiguolei	6fdd35a6f2	[enhancement](mpp process) remove unused method and make report process more clear (#16441 ) both update status and open_vectorized_internal will call send_report and stop report thread. move update_status code to open method and remove unnecessary send_report and stop_report_thread. --------- Co-authored-by: yiguolei <yiguolei@gmail.com>	2023-02-07 12:28:55 +08:00
Ashin Gau	27216dc7e0	[improvement](multi-catalog) push down all predicates into rowgroup/page filtering for ParquetReader (#16388 ) Tow improvements: 1. Refactor rowgroup&page filtering in `ParquetReader`, and use the operator overloading of Doris native c++ type to process comparison. 2. Support decimal/decimal v3/date/datev2/datetime/datetimev2	2023-02-07 11:32:57 +08:00
Gabriel	91229bb87d	[Bug](makr join) Fix mark join with other conjuncts (#16435 )	2023-02-07 09:31:41 +08:00
Kang	36a5e0a2a9	[bugfix](array) fix element revert on error in DataTypeArray::from_string (#16434 ) * fix array from_string element revert on error * add testcase	2023-02-06 18:27:36 +08:00
Xin Liao	2bee26b05a	[fix](merge-on-write) fix that the query result has duplicate keys (#16336 ) * [fix](merge-on-write) fix that the query result has duplicate keys * add ut	2023-02-06 17:09:53 +08:00
weizuo93	da27039fe4	[Fix](load) Fix memory leak for stream load 2pc #16430 StreamLoadContext is not deleted correctly. Co-authored-by: weizuo <weizuo@xiaomi.com>	2023-02-06 15:52:17 +08:00
xy720	c0054ddb2b	[chore](type-info) Remove unused method in TypeInfo #16433	2023-02-06 15:49:11 +08:00
Zhengguo Yang	b21fdace37	[bugfix](RemoteUDF) fix remote udf retrun `rpc env init error` (#16325 )	2023-02-06 15:47:10 +08:00
Kang	737c73dcf0	[Improvement](topn) order by key topn query optimization (#15663 )	2023-02-06 15:36:05 +08:00
lihangyu	f2fd47f238	[Improve](row-store) support row cache (#16263 )	2023-02-06 11:16:39 +08:00
slothever	b1b2697cc7	[fix](iceberg) fix iceberg catalog (#16372 ) 1. Fix iceberg catalog access s3 2. Fix iceberg catalog partition table query 3. Fix persistence	2023-02-05 13:15:28 +08:00
yiguolei	c689b8918a	[enhancement](runtimefilter) no need wait for fragment because two phase exec fragment (#16418 ) call pthread condition wait may block brpc thread. no need wait for fragment because two phase exec fragment already guarantee that the fragment instance exits when runtime filter comes. So that I remove the condition wait code. Co-authored-by: yiguolei <yiguolei@gmail.com>	2023-02-05 10:09:31 +08:00
luozenglin	09870098af	[fix](func) fix core dump when the pattern of the regexp_extract_all function does not contain subpatterns (#16408 )	2023-02-05 01:16:54 +08:00
yixiutt	059cf58151	[fix](vertical compaction) fix uint32_t init value (#16377 )	2023-02-05 00:05:35 +08:00
starocean999	dd63897757	[fix](be)the set operation node should accept both nullable and non-nullable data from child node (#16126 )	2023-02-04 23:08:59 +08:00
wudi	c488e67bd3	[Bug](vectorized)Fix reading date and datetime types conversion error (#16252 ) from pr #15612, Type conversion error when reading date and datetime types --------- Co-authored-by: wudi <>	2023-02-04 23:05:00 +08:00
luozenglin	d2b5015d3f	[enhancement](profile) add the profile counter RawRowsRead to record the rows read from the parquet file (#16328 )	2023-02-04 22:59:34 +08:00
HappenLee	c3a6eb4f9a	[Refactor](function) remove useless function get to create column (#16333 ) remove unless create_column to redurce the unless new operator	2023-02-04 22:54:14 +08:00
zhangstar333	458adf6c91	[improvement](jdbc) refator jdbc of copy result set by batch (#16337 ) have test jdbc external table with read, 10%+ performance improvement after optimization	2023-02-04 22:51:55 +08:00
Xinyi Zou	63d57b83f3	[fix](memory) Fix request jemallloc metrics wait lock je_malloc_mutex_lock_slow #16381 MetricRegistry::trigger_all_hooks holds the metrics lock and is stuck in get_je_metrics, to_prometheus is waiting for MetricRegistry::trigger_all_hooks to release the lock, so get_je_metrics is no longer called in MetricRegistry::trigger_all_hooks.	2023-02-04 22:49:22 +08:00
plat1ko	bd8ef4edeb	[fix](cooldown) Fix core in remove_all_remote_rowsets (#16374 )	2023-02-04 22:31:38 +08:00
plat1ko	1473a9716b	[fix](cooldown) Fix bug in report tablet (#16414 )	2023-02-04 22:30:57 +08:00
Gabriel	918004c016	[Bug](date) Fix BE crash caused by function `datediff` (#16397 ) * [Bug](date) Fix BE crash caused by function `datediff` * update	2023-02-04 18:43:23 +08:00
Kang	125b60b4b9	[improvement](compatibility) add DATA_TYPE in information schema for new types #16391 Add DATA_TYPE in information schema for types: datev2, datatimev2, decimal, jsonb. It was 'unknown' for these types and cause problem for tools such as BI using information schema.	2023-02-03 22:28:42 +08:00
yixiutt	56be2e5a1a	[bugfix](disk balance) fix new rowset time check when add tablet (#16261 ) In disk balancer, if a tablet is in highly concurrent load, new rowset creation time(which use current time) may be same as the newest rowset, and when add tablet, there has a creation time check that new_time must bigger than old time, so disk balancer will failed many times and makes this tablet lose many verisons as migration will block writes.	2023-02-03 21:49:37 +08:00
Gabriel	87fbb8341a	[Bug](datev2) Fix bug when cast datev2 to date (#16394 )	2023-02-03 20:50:16 +08:00
lihangyu	f94a78ab4a	[Fix](topn) fix wrong nullable cast for RowId column and use heapsorter for two phase read (#16399 ) convert_nullable_flags does not contain nullable info for RowID column, but valid_column_ids contain RowID column, nullable falg will be undefined for RowID column	2023-02-03 20:49:45 +08:00
luozenglin	4df70becb9	[refactor](reader) refactor broker_file_reader to get _client in the constructor (#16021 )	2023-02-03 16:51:19 +08:00
Pxl	5e4bb98900	[Chore](build) enable -Wpedantic and update lowest gcc version to 11.1 (#16290 ) enable -Wpedantic and update lowest gcc version to 11.1	2023-02-03 11:28:48 +08:00
zhangstar333	7d5a10e1af	[bug](function) fix mask_first_n function can't handle const value (#16308 )	2023-02-03 10:32:42 +08:00
zhangstar333	545b91f8f7	[bug](jdbc) fix jdbc insert decimalv3 be core dump (#16353 )	2023-02-03 10:00:06 +08:00
Jerry Hu	7a800bd3c6	[fix](scan) coredump caused by null of _scanner_ctx (#16361 )	2023-02-03 09:24:15 +08:00

1 2 3 4 5 ...

3737 Commits