doris

Author	SHA1	Message	Date
seawinde	d5bf20c96e	[improvement](mtmv) Improve the performance for query rewritting by materialized view (#31886 ) - Limit the number of times for the query rewritting to the group - Remove the unnecessary log and explain detail info in query	2024-03-09 19:55:47 +08:00
Yongqiang YANG	78feb7f519	[fix](forward) set error code for query state to handle exception of (#31975 )	2024-03-09 19:55:47 +08:00
shuke	1d094a46ec	[regression-test](pipeline) remove sys_log_verbose_modules in pipeline #32015	2024-03-09 19:55:47 +08:00
walter	263135c193	[fix](case) fix export data consistency case (#32005 )	2024-03-09 19:45:50 +08:00
Lei Zhang	4bdea7c324	[opt](fe) Reduce jvm heap memory consumed by profiles of BrokerLoadJob (#31985 ) * it may cause FE OOM when there are a lot of broker load jobs if the profile is enabled	2024-03-09 19:45:50 +08:00
kkop	21e412393b	[enhancement] add_method_for_schemachange (#31849 )	2024-03-09 19:45:50 +08:00
shuke	aaa5542fb9	fix drop un-related table (#31990 )	2024-03-09 19:45:50 +08:00
924060929	cc0e58faec	[enhancement](regression-test) upgrade groovy to 4.x and enable run test by jdk17/21 (#31906 ) upgrade groovy to 4.x and enable run test by jdk17 / 21	2024-03-09 19:45:46 +08:00
wuwenchi	5f9eb5eb52	[feature](external catalog)Add partition grammar for external catalog to create table (#31585 ) The `PARTITION BY` syntax used by external catalogs has been added. You can specify a column directly, or a partition function as a partition condition. Like: `PARTITION BY LIST(col1, col2, func(param), func(param1, param2), func(param1, param2, param3))` NOTICE: This PR change the grammar of `AUTO PARTITION` From ``` AUTO PARTITION BY RANGE date_trunc(`TIME_STAMP`, 'month') ``` To ``` AUTO PARTITION BY RANGE (date_trunc(`TIME_STAMP`, 'month')) ```	2024-03-09 19:45:46 +08:00
abmdocrt	62db7094ea	Revert "Problem: When the old optimizer processes an INSERT INTO statement that contains two quotation marks, it results in only one quotation mark being written into the database. (#31890 )" (#31986 ) This reverts commit 8c309652e04698f311b6c9158105352e8416c69a.	2024-03-09 19:45:46 +08:00
amory	621803c547	[FIX](InPredict) fix in params in to context for thread fragment query (#31935 )	2024-03-09 19:45:46 +08:00
HappenLee	6ef4ab631d	[Opt](func) reduce the useless mem alloc and const opt the concat code (#31983 )	2024-03-09 19:45:46 +08:00
Dongyang Li	ce5973b672	Revert "[chore](ci) Update .asf.yaml (#31994 )" (#31997 ) This reverts commit 5113b4b624f4bc646672b86be54752fff854609d.	2024-03-09 19:45:46 +08:00
Dongyang Li	1b4680ee14	[chore](ci) Update .asf.yaml (#31994 )	2024-03-09 19:45:46 +08:00
feiniaofeiafei	5909237ab1	[Fix](nereids) Add semantic check that the hash bucket column must be a key column when creating table for aggregate and unique models (#31951 )	2024-03-09 19:45:46 +08:00
Jiwen liu	cae80cfa15	[doc](apache-superset) Fix zh-CN documentation of apache-superset (#31976 )	2024-03-09 19:45:46 +08:00
abmdocrt	609761567c	[Fix](partial-update) Fix wrong column number passing to BE when partial and enable nereids (#31461 ) * Problem: Inconsistent behavior occurs when executing partial column update `UPDATE` statements and `INSERT` statements on merge-on-write tables with the Nereids optimizer enabled. The number of columns passed to BE differs; `UPDATE` operations incorrectly pass all columns, while `INSERT` operations correctly pass only the updated columns. Reason: The Nereids optimizer does not handle partial column update `UPDATE` statements properly. The processing logic for `UPDATE` statements rewrites them as equivalent `INSERT` statements, which are then processed according to the logic of `INSERT` statements. For example, assuming a MoW table structure with columns k1, k2, v1, v2, the correct rewrite should be: * `UPDATE` table t1 set v1 = v1 + 1 where k1 = 1 and k2 = 2 * => * `INSERT` into table (v1) select v1 + 1 from table t1 where k1 = 1 and k2 = 2 However, the actual rewriting process does not consider the logic for partial column updates, leading to all columns being included in the `INSERT` statement, i.e., the result is: * `INSERT` into table (k1, k2, v1, v2) select k1, k2, v1 + 1, v2 from table t1 where k1 = 1 and k2 = 2 This results in `UPDATE` operations incorrectly passing all columns to BE. Solution: Having analyzed the cause, the solution is straightforward: when rewriting partial column update `UPDATE` statements to `INSERT` statements, only retain the updated columns and all key columns (as partial column updates must include all key columns). Additionally, this PR includes error injection cases to verify the number of columns passed to BE is correct. * 2 * 3 * 4 * 5	2024-03-09 19:45:42 +08:00
lihangyu	e8aa5ee7d5	[Improve](Variant) support bloom filter for variant subcolumns (#31347 ) * [Improve](Variant) support bloom filter for variant subcolumns * rebase	2024-03-09 19:45:03 +08:00
924060929	e8b4bf5be9	[enhancement](Nereids) Speedup PartitionPrunner (#31970 ) This pr imporve the high QPS query by speed up PartitionPrunner 1. remove useless Date parse/format, use LocalDate instead 2. fast evaluate path for single value partition 3. change Collection.stream() to ImmutableXxx.builderWithExpectedSize(n) to skip useless method call and collection resize 4. change lots of if-else to switch 5. don't parse to string to compare dateLiteral, use int field compare instead	2024-03-09 19:45:03 +08:00
HHoflittlefish777	e3611f6a1d	[improve](routine-load) increase routing load max_batch _size max limit (#31846 )	2024-03-09 19:45:03 +08:00
Pxl	19e6ebd09c	[Feature](materialized-view) support mv with bitmap_union(bitmap_from_array()) case (#31962 ) support mv with bitmap_union(bitmap_from_array()) case	2024-03-09 19:45:03 +08:00
morrySnow	679cd0ab45	[opt](mtmv) ensure rewritten plan output order correct even project been eliminated (#31870 )	2024-03-09 19:45:03 +08:00
starocean999	1721bfb87a	[fix](nereids)forbid some join reorder rules for mark join (#31966 )	2024-03-09 19:45:03 +08:00
Gabriel	f968d96545	[profile](pipelineX) Add lost metrics (#31964 )	2024-03-09 19:45:03 +08:00
Vallish Pai	4bfecac08a	[enhancement](plsql) Support show procedure and show create procedure (#31297 ) (#31763 )	2024-03-09 19:45:03 +08:00
Jibing-Li	1b783aaa7f	[fix](p2)Fix analyze hive partition column p2 case after row count change. #31958	2024-03-09 19:45:03 +08:00
LiBinfeng	eb280d374b	[case](Nereids) add leading tpc-h (#30405 ) add tpc-h shape cases using leading hint except: single table without join q1 q6 not support feature include tables after subquery unnested q2 q16 q18 q20 q21 q22	2024-03-09 19:45:03 +08:00
chen	861461403f	add missing RuleType LOGICAL_REPEAT_TO_PHYSICAL_REPEAT_RULE (#31877 )	2024-03-09 19:45:03 +08:00
Jerry Hu	93d298d34a	[fix](agg) wrong result of two or more map_agg functions in query (#31928 )	2024-03-09 19:45:03 +08:00
Gabriel	b2de83f250	[agg](conf) Add a knob to control distinct agg (#31930 ) Add a knob to control distinct agg	2024-03-09 19:44:54 +08:00
Jibing-Li	e9c1638507	Add waiting timeout while creating mv and row count report. (#31944 )	2024-03-09 19:44:54 +08:00
Jibing-Li	908dff551a	[fix](statistics)Add synchronize for modify analysisTaskInfoMap and analysisJobInfoMap. #31940	2024-03-09 19:44:54 +08:00
lihangyu	0da010603e	[Improve](TabletSchemaCache) reduce duplicated memory consumption for column name and column path (#31141 ) Both could be reference to related field in TabletColumn.And use shared_ptr for TabletColumn in TabletSchema for later memory reuse	2024-03-09 19:44:42 +08:00
Uniqueyou	779ca464a5	[Fix](Status) Handle returned overall Status correctly (#31692 ) Handle returned overall Status correctly	2024-03-09 19:44:39 +08:00
谢健	aff09fc9bc	[feature](Nereids) support make miss slot as null alias when converting anti join (#31854 ) transform project(A., B.slot) - filter(B.slot is null) - LeftOuterJoin(A, B) to project(A., null as B.slot) - LeftAntiJoin(A, B)	2024-03-09 19:43:21 +08:00
Pxl	981ea73466	[Bug](top-n) init query_ctx runtime predicate before _build_pipelines (#31896 ) init query_ctx runtime predicate before _build_pipelines	2024-03-09 19:43:21 +08:00
caoliang-web	c6a8146db5	[typo](doc)Doriswriter document modification (#31322 )	2024-03-09 19:43:21 +08:00
caoliang-web	8801916675	[regression](spark)Add spark to read doris multiple data types cases (#31861 )	2024-03-09 19:43:21 +08:00
abmdocrt	5b52812af2	Problem: When the old optimizer processes an INSERT INTO statement that contains two quotation marks, it results in only one quotation mark being written into the database. (#31890 ) Reason: During syntax parsing, the old optimizer interprets two quotation marks as a single quotation mark. Solution: Remove the logic that consolidates two quotation marks into one.	2024-03-09 19:43:21 +08:00
xy	92320bbd2d	[Fix](doc) Adjust default values (#31882 ) Co-authored-by: xingying01 <xingying01@corp.netease.com>	2024-03-09 19:43:21 +08:00
ZhenchaoXu	9fa470caca	[doc](external) delete external table docs in 1.2 (#28353 )	2024-03-09 19:43:21 +08:00
starocean999	3b56c4bcfa	[enhancement](nereids)send is_nereids flag to be (#31752 )	2024-03-09 19:43:12 +08:00
Jeffrey	197b204e02	[type](docs) delete duplicated sidebar items (#31991 )	2024-03-08 16:26:06 +08:00
Jeffrey	1898517b2f	[type](doc) update 2.1 doc label (#31942 )	2024-03-07 17:12:03 +08:00
kkop	009ca9b90b	[fix] (doc)Fix invalid link on benchmark tool page #31920	2024-03-07 16:53:49 +08:00
minghong	db389d7d4e	[feat](nereids) support null safe eq runtime filter (FE part) (#31655 ) be part has been merged in #31754	2024-03-07 16:53:49 +08:00
LiBinfeng	fa411f88df	[Fix](Nereids) fix hint cases with random result (#31865 )	2024-03-07 16:53:49 +08:00
yagagagaga	29b858d8c9	[chore](build) Using multithread to accelerate FE compilation (#31855 )	2024-03-07 16:53:49 +08:00
zhangdong	667b1fba04	[enhance](mtmv) MTMV Use partial partition of base table (#31632 ) MTMV add 3 properties: partition_sync_limit: digit partition_sync_time_unit: DAY/MONTH/YEAR partition_sync_date_format: like "%Y-%m-%d"/"%Y%m%d" For example, the current time is 2020-02-03 20:10:10 - If partition_sync_limit is set to 1 and partition_sync_time_unit is set to DAY, only partitions with a time greater than or equal to 2020-02-03 00:00:00 will be synchronized to the MTMV - If partition_sync_limit is set to 1 and partition_sync_time_unit is set to MONTH, only partitions with a time greater than or equal to 2020-02-01 00:00:00 will be synchronized to the MTMV - If partition_sync_limit is set to 1 and partition_sync_time_unit is set to YEAR, only partitions with a time greater than or equal to 2020-01-01 00:00:00 will be synchronized to the MTMV - If partition_sync_limit is set to 3 and partition_sync_time_unit is set to MONTH, only partitions with a time greater than or equal to 2019-12-01 00:00:00 will be synchronized to the MTMV - If partition_sync_limit is set to 4 and partition_sync_time_unit is set to DAY, only partitions with a time greater than or equal to 2020-01-31 00:00:00 will be synchronized to the MTMV	2024-03-07 16:53:49 +08:00
lihangyu	0c7e9257a8	[cloud](point query) enable short circuit query in cloud (#31897 )	2024-03-07 16:53:40 +08:00

1 2 3 4 5 ...

17549 Commits