doris

Author	SHA1	Message	Date
starocean999	e9392780a9	[fix](nereids)fix some nereids planner bugs (#19509 ) 1.some encrypt and decrypt functions have wrong blockEncryptionMode 2.topN node should compare tuples from intermediate_row_desc with first_sort_slot.tuple_id 3.must keep the limit if it's an uncorrelated in-subquery with limit on sort, like select a from t1 where a in ( select b from t2 order by xx limit yy )	2023-05-12 09:06:16 +08:00
xy720	39ec8aa64c	[refactor](complex-type) refactor array/map/struct literal to not invoke execute() function in prepare state (#19068 )	2023-05-11 18:44:37 +08:00
yangshijie	ed8a4b4120	[feature-wip](duplicate_no_keys) skip sort function if the table is duplicate without keys (#19483 )	2023-05-11 14:44:16 +08:00
jakevin	dc497e11bb	[fix](Nereids) avoid to push top Project of JoinCluster in PushdownProjectThroughJoin (#19441 ) We shouldn't push top Project of JoinCluster in PushdownProjectThroughJoin like ``` * Project (id + 1) if this project is top project of Join Cluster * \| * Join * / \ * Join Join * / .... * Join ```	2023-05-11 13:58:54 +08:00
herry2038	834bf2eab7	[feature](array) Add array_last lambda function (#18388 ) Add array_last lambda function	2023-05-11 13:15:54 +08:00
zhannngchen	5167dc1251	[feature](merge-on-write) enable merge on write by default (#19017 )	2023-05-11 11:10:48 +08:00
abmdocrt	71f7e9e185	[test](cast func) add test for cast float text to int when nereids is on #19517	2023-05-11 08:24:54 +08:00
Qi Chen	4418eb36a3	[Fix](multi-catalog) Fix some hive partition issues. (#19513 ) Fix some hive partition issues. 1. Fix be will crash when using hive partitions field of `date`, `timestamp`, `decimal` type. 2. Fix hdfs uri decode error when using `timestamp` partition filed which will cause some url-encoding for special chars, such as `%3A` will encode `:`.	2023-05-11 07:49:46 +08:00
Jibing-Li	68505a1192	[Test](multi catalog)Add test case for Iceberg External Table. #19488	2023-05-11 01:13:40 +08:00
Jerry Hu	47edc5a06e	[fix](functions) Support nullable column for multi_string functions (#19498 )	2023-05-11 01:13:13 +08:00
yongkang.zhong	3a22af836e	[fix](jdbc catalog) fix error to clickhouse uint64 type Conversion (#19463 ) * [fix](jdbc catalog) fix error to clickhouse uint64 type Conversion * add test case	2023-05-10 21:53:30 +08:00
Mryange	d20b5f90d8	[feature](executor) Automatically set the instance_num using the info from be. (#19345 ) 1. fixed some error regressions (results error with big nstance_num due to incorrect order by). 2. if set parallel_fragment_exec_instance_num to 0, the concurrency in the Pipeline execution engine will automatically be set to half of the number of CPU cores. 3. add limit to parallel_fragment_exec_instance_num that it cannot be set to more than fe.conf::max_instance_num(Default: 128) ``` mysql [(none)]>set parallel_fragment_exec_instance_num = 514; ERROR 1231 (42000): errCode = 2, detailMessage = Variable 'parallel_fragment_exec_instance_num' can't be set to the value of '514(Should not be set to more than 128)' ```	2023-05-10 17:07:41 +08:00
Kang	14a56da397	[chore](testcase) change tpcds q67 testcase name to q67_ignore_temporarily (#19227 ) change tpcds q67 testcase name to q67_ignore_temporarily since the error should be ignored temporarily due to precision problem that will be fixed further.	2023-05-10 15:06:23 +08:00
ElvinWei	fae2e5fd22	[enchancement](statistics) implement automatically analyzing statistics and support table level statistics #19420 Add table level statistics, support SHOW TABLE STATS statement to show table level statistics. Implement automatically analyze statistics, support ANALYZE... WITH AUTO ... statement to automatically analyze statistics. TODO: collate relevant p0 tests Supplement the design description to README.md Issue Number: close #xxx	2023-05-10 11:47:34 +08:00
Pxl	5473795a51	[Bug](scan) forbiden push down in predicate when in_state->use_set is false (#19471 ) forbiden push down in predicate when in_state->use_set is false	2023-05-10 11:12:20 +08:00
ZashJie	4302ceaee8	[Improvement](data types) enhance show data types stmt (#18831 )	2023-05-09 09:42:44 +08:00
lvshaokang	af04c3acab	[fix](sequence-column) Fix sequence_col column used default expr insert failed (#18933 )	2023-05-08 17:18:25 +08:00
Tiewei Fang	e78149cb65	[Enhencement](Export) add property for outfile/export and add test (#18997 ) This pr does three things: 1. add `delete_existing_files` property for outfile/export. If `delete_existing_files = true`, export/outfile will delete all files under file_path first. 2. add p2 test for export 3. modify docs	2023-05-08 14:02:20 +08:00
Gabriel	4c6ca88088	Revert "[refactor](function) ignore DST for function `from_unixtime` (#19151 )" (#19333 ) This reverts commit 9dd6c8f87b73db238bfd38fb1d76f3796910f398.	2023-05-06 16:33:58 +08:00
ElvinWei	3f6e5118e6	[enchancement](statistics) support periodic collection of statistics (#19247 ) This PR enables periodic collection of statistics and is a precursor to automatic statistics collection. It mainly includes the following contents： support periodic collection of statistics. Change the type of Date in statistics p0 to DateV2(see [Enhancement](data-type) add FE config to prohibit create date and decimalv2 type #19077) for test locally. complement cases(remove Chinese characters, optimize code, etc) , improve stability. Supports setting whether to keep records of statistics synchronization job info, convenient for use in p0 testing. The statistics job table was modified, and some auxiliary judgments were added to avoid the user perceiving the modification. This function was removed when the table schema is stable.	2023-05-06 14:53:06 +08:00
starocean999	a72eee24f1	[fix](nereids) fix merge project with window function bug (#19280 ) 1. don't merge projects if any window function exists 2. bypass SimplifyArithmeticRule for decimalV3 type	2023-05-06 10:38:14 +08:00
starocean999	3e3262361c	[fix](fe)havingClause should be substituted the same way as resultExprs (#19261 ) substituted havingClause in the same way as resultExprs to prevent " HAVING clause not produced by aggregation output" error	2023-05-05 18:03:43 +08:00
minghong	817f3ce510	[fix](nereids) plan shape on tpch_sf1T q21 case #19291	2023-05-05 14:24:28 +08:00
Gabriel	9dd6c8f87b	[refactor](function) ignore DST for function `from_unixtime` (#19151 )	2023-05-05 11:51:49 +08:00
ZhangYu0123	aaf0ef741e	[fix](regression) fix inverted_index_p1 q72.sql timeout error (#19241 ) Fix inverted_index_p1 q72.sql timeout error 1、the runtime filter exeed wait time and lead to 100w * 1000w data join	2023-05-04 11:05:15 +08:00
DuRipeng	52d25f41a4	[feature](multi-catalog) Rename multi-catalog config 'specified_database_list' to 'include_database_list', and introduce new multi-catalog config 'exclude_database_list' (#18834 ) In my scene, We need to specify databases that are excluded to synchronize to doris, like some databases store temporary table. Since #17803 introduce `specified_database_list` to specify 'include databases', this pr introduce new config `exclude_database_list` to specify 'exclude databases', and rename `specified_database_list` to `include_database_list` for naming symmetry. BTW, when `include_database_list` and `exclude_database_list` specify overlapping databases, `exclude_database_list` would take effect with higher privilege over `include_database_list`.	2023-05-04 09:30:02 +08:00
nanfeng	da4de37dec	[feature-wip](mv lifecycle) separate life cycle of base table and its materialized views (#19210 ) support related syntax and add:regress-test case --------- Co-authored-by: yzy <yzy@nanfeng_yzy@163.com>	2023-04-30 17:42:02 +08:00
yiguolei	8eab20d3df	[bugfix](low cardinality) cached code is wrong will result wrong query result when many null pages (#19221 ) Sometimes the dict is not initialized when run comparison predicate here, for example, the full page is null, then the reader will skip read, so that the dictionary is not inited. The cached code is wrong during this case, because the following page maybe not null, and the dict should have items in the future. This will result the dict string column query return wrong result, if there are many null values in the column. I also add some regression test for dict column's equal query, larger than query, less than query. --------- Co-authored-by: yiguolei <yiguolei@gmail.com>	2023-04-29 21:28:41 +08:00
yixiutt	aef9355cd3	[feature-wip](partial update) PART1: support basic partial write (#17542 )	2023-04-28 17:17:57 +08:00
ElvinWei	718297d3c1	[test](statistics) add p0 test of sampling statistics (#19176 ) 1. Added test p0 for sampling collection statistics 2. Modify the uniqueKeys of table analysis_jobs for deletion based on relevant conditions 3. Solve the problem that incremental statistics p0 is less stable	2023-04-28 15:50:05 +08:00
ElvinWei	484612a0af	[opt](statistics) optimize Incremental statistics collection and statistics cleaning (#18971 ) This pr mainly optimizes the following items: - the collection of statistics: clear up invalid historical statistics before collecting them, so as not to affect the final table statistics. - the incremental collection of statistics: in the case of incremental collection, only the corresponding partition statistics need to be collected. TODO: Supports incremental collection of materialized view statistics.	2023-04-27 11:51:47 +08:00
TengJianPing	708c1850d9	[test](hll) add test case for hll_raw_agg (#19127 )	2023-04-27 11:33:49 +08:00
lihangyu	e76b3a316f	[Bug](mysql proto) fix binary proto with dynamic mode (#19055 ) Dynamic mode used in array type when serialize it to mysql row buffer using dynamic mode, when combine binary row format with dynamic mode,something goes wrong, and lead to invalid binary row format.	2023-04-27 11:18:01 +08:00
brody715	20395ce501	[feature](array_function): add support for array_cum_sum function (#18231 )	2023-04-27 09:57:13 +08:00
TengJianPing	1afa7c786f	[test](regression) add test case for bucket shuffle of datetime column (#19088 )	2023-04-27 09:05:32 +08:00
xy720	925efc1902	[bug](map-type)fix some bugs in map and map element function (#18935 ) fix some bugs in map and map element function.	2023-04-26 22:10:15 +08:00
starocean999	aacc075f09	[fix](planner) SetOperationNode's slots' nullability calculation is wrong (#19108 ) SetOperationNode's slots' nullability should consider slots info from all children, even some children have EmptyResultSet	2023-04-26 21:18:37 +08:00
morrySnow	e83d0d9b6a	[opt](Nereids) forbid some bad pattern aggregate in AggregateStrategy (#18877 ) since we cannot do stats derive and cost estimate on agg very good. this PR remove some aggregate pattern that usually not good. 1. one stage agg after exchange. this pattern is good only when process very few rows. 2. three stage distinct agg with gather middle merge.	2023-04-26 20:01:35 +08:00
Jibing-Li	59d8aa5a6f	[Fix](multi catalog)Fix Hive partition path doesn't contain partition value case bug (#19053 ) Hive support create partition with a specific location. In this case, the file path for the create partition may not contain the partition name and value. Which will cause Doris fail to query the the hive partition. This pr is to fix this bug.	2023-04-26 17:18:51 +08:00
zhengyu	0c9fb7297e	[fix](regression) mv segcompaction_p1 to segcompaction_p2 (#18806 ) segcompaction_p1 contains fairly large load jobs, which will exceed memlimit or timeout in pipeline under such heavy loads. Signed-off-by: freemandealer <freeman.zhang1992@gmail.com>	2023-04-26 15:34:46 +08:00
AKIRA	d3a0b94602	[feature](stats) Support to kill analyze #18901 1. Report error if submit analyze jobs when stats table is not available 2. Support kill analyze 3. Support cancel sync analyze	2023-04-26 14:23:44 +08:00
zhangstar333	d037938a4c	[vectorzied](function) fix year_floor get result is incorrectly (#19006 )	2023-04-26 11:39:22 +08:00
Mryange	5fd6d8ebd4	[fix](function) Support more behaviors of cast time in MySQL	2023-04-26 07:49:54 +08:00
herry2038	17b59df8dd	[fix](function) Array_map compared offset rows one by one (#18406 ) Array_map 's multi columns compare not only nested data rows to be equal,but also the offsets data must equal each other.	2023-04-25 19:12:19 +08:00
Qi Chen	61b7a52444	[Enhancement](multi-catalogs) Use decimal V3 type in multi-catalogs module. (#18926 ) 1. Use decimal V3 type in JDBC and Iceberg tables. 2. Fix hdfs TVF decimal V3 type and regression test.	2023-04-25 14:49:40 +08:00
AKIRA	a4a85f2476	[feat](stats) Return job id for async analyze stmt (#18800 ) 1. Return job id from async analysis 2. Sync analysis jobs don't save to analysis_jobs anymore	2023-04-25 14:43:54 +08:00
Ashin Gau	39d66ca2c6	[fix](parquet) hasn't initialize select vector when number of nested values equals zero (#18953 ) Fix bug when reading array type in parquet file: ``` ERROR 1105 (HY000): errCode = 2, detailMessage = [INTERNAL_ERROR]Read parquet file xxx failed, reason = [IO_ERROR]Decode too many values in current page ``` When reading normal columns, `ScalarColumnReader::_read_values` still calls `ColumnSelectVector::set_run_length_null_map` to initialize select vector, but `ScalarColumnReader::_read_nested_column` hasn't do this, making the number of values wrong. The situation where this error occurs is particularly extreme: The column pages have remaining values to be read, but all of them are null values at ancestor level, so there's no actual read operation, just skipping null values at ancestor level.	2023-04-25 14:21:33 +08:00
lihangyu	d555bae290	[Bug](serde) fix serialize column to jsonb when meet boolean and decimal_v3 (#19011 ) * [Bug](serde) fix serialize column to jsonb when meet boolean and decimal_v3 * add comment to explain why use uint8	2023-04-25 10:48:13 +08:00
ZhangYu0123	93c48f2bb0	[fix](regression) fix show create table in_memory = false test result error #19022	2023-04-25 09:04:59 +08:00
Mingyu Chen	207c827cdb	[fix](test) fix result of CHARACTER_OCTET_LENGTH in . (#18896 )	2023-04-25 08:42:54 +08:00

1 2 3 4 5 ...

1138 Commits