doris

Author	SHA1	Message	Date
AKIRA	36b6cea462	[feature-wip](nereids) Support Q-Error to measure the accuracy of derived statistics (#17185 ) Collect each estimated output rows and exact output rows for each plan node, and use this to measure the accuracy of derived statistics. The estimated result is managed by ProfileManager. We would get this estimated result in the http request by query id later.	2023-03-08 16:26:24 +08:00
Calvin Kirs	d908d5fe01	[dependency](fe)Dependency Upgrade (#17377 ) * Upgrade log4j to 2.X - binding log4j version to 2.18.0 - used log4j-1.2-api complete smooth upgrade * Upgrade filerupload to 1.5 * Upgrade commons-io to 2.7 * Upgrade commons-compress to 1.22 * Upgrade gson to 2.8.9 * Upgrade guava to 30.0-jre * Binding jackson version to 2.14.2 * Upgrade netty-all to 4.1.89.final * Upgrade protobuf to 3.21.12 * Upgrade kafka-clints to 3.4.0 * Upgrade calcite version to 1.33.0 * Upgrade aws-java-sdk to 1.12.302 * Upgrade hadoop to 3.3.4 * Upgrade zookeeper to 3.4.14 * Binding tomcat-embed-core to 8.5.86 * Upgrade apache parent pom to 25 * Use hive-exec-core as a hive dependency, add the missing jar-hive-serde separately * Basic public dependencies are extracted to parent dependencies * Use jackson uniformly as the basic json tool * Remove springloaded, spring-boot-devtools has the same functionality * Modify the spark-related dependency scope to provide, which should be provided at runtime	2023-03-08 14:28:40 +08:00
zhengshiJ	aab14922af	[Feature](Nereids) support MarkJoin (#16616 ) # Proposed changes 1.The new optimizer supports the combination of subquery and disjunction.In the way of MarkJoin, it behaves the same as the old optimizer. For design details see:https://emmymiao87.github.io/jekyll/update/2021/07/25/Mark-Join.html. 2.Implicit type conversion is performed when conjects are generated after subquery parsing 3.Convert the unnesting of scalarSubquery in filter from filter+join to join + Conjuncts.	2023-03-08 14:26:24 +08:00
Kang	626fbc34f9	[bugfix](jsonb) Fix create mv using jsonb key cause be crash (#17430 )	2023-03-08 14:18:26 +08:00
AlexYue	f3b50b3472	[enhance](cooldown) skip once failed follow cooldown tablet (#16810 )	2023-03-08 14:14:13 +08:00
Xin Liao	8001d65811	[fix](insert) fix memory leak for insert transaction (#17530 )	2023-03-08 14:10:59 +08:00
AlexYue	273d2100ac	[enhance](cooldown) turn write cooldown meta async (#16813 )	2023-03-08 14:06:21 +08:00
qiye	3a877857ae	[improvement](inverted index)Remove searcher bitmap timer to improve query speed (#17407 ) Timer becomes a bottleneck when the query hit volume is very high.	2023-03-08 14:03:36 +08:00
Xinyi Zou	335c1e5953	[fix](memory) Fix MacOS mem_limit parse error and GC after env Init #17528 Fix MacOS mem_limit parse result is 0. Fix GC after env Init, otherwise, when the memory is insufficient, BE will start failure. * Query id: 0-0 * * Aborted at 1677833773 (unix time) try "date -d @1677833773" if you are using GNU date * * Current BE git commitID: 8ee5f45 * * SIGSEGV address not mapped to object (@0x70) received by PID 24145 (TID 0x7fa53c9fd700) from PID 112; stack trace: * 0# doris::signal::(anonymous namespace)::FailureSignalHandler(int, siginfo_t, void) at be/src/common/signal_handler.h:420 1# os::Linux::chained_handler(int, siginfo, void) in /usr/local/jdk/jre/lib/amd64/server/libjvm.so 2# JVM_handle_linux_signal in /usr/local/jdk/jre/lib/amd64/server/libjvm.so 3# signalHandler(int, siginfo, void) in /usr/local/jdk/jre/lib/amd64/server/libjvm.so 4# 0x00007FA56295A400 in /lib64/libc.so.6 5# doris::MemTrackerLimiter::log_process_usage_str(std::__cxx11::basic_string<char, std::char_traits<char>, std::allocator<char> > const&, bool) at be/src/runtime/memory/mem_tracker_limiter.cpp:208 6# doris::MemTrackerLimiter::print_log_process_usage(std::__cxx11::basic_string<char, std::char_traits<char>, std::allocator<char> > const&, bool) at be/src/runtime/memory/mem_tracker_limiter.cpp:226 7# doris::Daemon::memory_maintenance_thread() at be/src/common/daemon.cpp:245 8# doris::Thread::supervise_thread(void*) at be/src/util/thread.cpp:455 9# start_thread in /lib64/libpthread.so.0 10# clone in /lib64/libc.so.6	2023-03-08 14:00:57 +08:00
bobhan1	4ea0d6c5fa	[feature](array_function) add support for array_popfront (#17416 )	2023-03-08 13:57:38 +08:00
gitccl	b1d65f855d	[Feature](array-function) Support array_concat function (#17436 )	2023-03-08 13:57:16 +08:00
Pxl	e2ac06d6d6	[Chore](execution) change PipelineTaskState to enum class && remove some row-based code (#17300 ) 1. change PipelineTaskState to enum class 2. remove some row-based code on FoldConstantExecutor::_get_result 3. reduce memcpy on minmax runtime filter function(Now we can guarantee that the input data is aligned) 4. add Wunused-template check, and remove some unused function, change some static function to inline function.	2023-03-08 12:41:15 +08:00
jakevin	2b6133f4d0	[feature](Nereids): pushdown complex project through inner/outer Join. (#17365 )	2023-03-08 12:00:56 +08:00
TengJianPing	778acb3c5b	[opt](string) optimize string equal comparision (#17336 ) Optimize string equal and not-equal comparison by using memequal_small_allow_overflow15.	2023-03-08 11:30:00 +08:00
Kang	4b743061b4	[feature](function) support type template in SQL function (#17344 ) A new way just like c++ template is proposed in this PR. The previous functions can be defined much simpler using template function. # map element extract template function [['element_at', '%element_extract%'], 'E', ['ARRAY<E>', 'BIGINT'], 'ALWAYS_NULLABLE', ['E']], # map element extract template function [['element_at', '%element_extract%'], 'V', ['MAP<K, V>', 'K'], 'ALWAYS_NULLABLE', ['K', 'V']], BTW, the plain type function is not affected and the legacy ARRAY_X MAP_K_V is still supported for compatability.	2023-03-08 10:51:31 +08:00
yiguolei	9213dd906a	[enhancement](exception) add exception structure and using unique ptr in VExplodeBitmapTableFunction (#17531 ) add exception class in common. using unique ptr in VExplodeBitmapTableFunction support single exception or nested exception, like this: ---SingleException [E-100] test OS_ERROR bug @ 0x55e80b93c0d9 doris::Exception::Exception<>() @ 0x55e80b938df1 doris::ExceptionTest_NestedError_Test::TestBody() @ 0x55e82e16bafb testing::internal::HandleSehExceptionsInMethodIfSupported<>() @ 0x55e82e15ab3a testing::internal::HandleExceptionsInMethodIfSupported<>() @ 0x55e82e1361e3 testing::Test::Run() @ 0x55e82e136f29 testing::TestInfo::Run() @ 0x55e82e1376e4 testing::TestSuite::Run() @ 0x55e82e148042 testing::internal::UnitTestImpl::RunAllTests() @ 0x55e82e16dcab testing::internal::HandleSehExceptionsInMethodIfSupported<>() @ 0x55e82e15ce4a testing::internal::HandleExceptionsInMethodIfSupported<>() @ 0x55e82e147bab testing::UnitTest::Run() @ 0x55e80c4b39e3 RUN_ALL_TESTS() @ 0x55e80c4a99b5 main @ 0x7f0a619d0493 __libc_start_main @ 0x55e80b84602a _start @ (nil) (unknown)	2023-03-08 10:44:14 +08:00
yiguolei	c97422bd3d	[enhancement](regression-test) add sleep 3s for schema change and rollup (#17484 ) Co-authored-by: yiguolei <yiguolei@gmail.com>	2023-03-08 10:43:05 +08:00
yiguolei	4692d6764c	[refactor](remove string val) remove string val structure, it is same with string ref (#17461 ) remove stringval, decimalv2val, bigintval	2023-03-08 10:42:20 +08:00
Adonis Ling	f21508baec	[chore](macOS) Disable detect_container_overflow at BE startup (#17514 ) BE failed to start up due to container-overflow errors reported by address sanitizer.	2023-03-08 10:21:45 +08:00
Bowen Liang	ae916f7cb3	[docs](doc) Add docs for Apache Kyuubi (#17481 ) * add kyuubi doc of zh-CN & en	2023-03-08 09:36:50 +08:00
qiye	a767472c56	[fix](DOE)Fix es p0 case error (#17502 ) Fix es array parse error, introduced by #16806	2023-03-08 08:06:30 +08:00
htyoung	69c62b6c6c	[Fix](vectorization) fixed that when a column's _fixed_values exceeds the max_pushdown_conditions_per_column limit, the column will not perform predicate pushdown, but if there are subsequent columns that need to be pushed down, the subsequent column pushdown will be misplaced in _scan_keys and it causes query results to be wrong (#17405 ) the max_pushdown_conditions_per_column limit, the column will not perform predicate pushdown, but if there are subsequent columns that need to be pushed down, the subsequent column pushdown will be misplaced in _scan_keys and it causes query results to be wrong Co-authored-by: tongyang.hty <hantongyang@douyu.tv>	2023-03-08 07:23:56 +08:00
LiBinfeng	6b88df2bdd	[enhancement](planner) support case transition of timestamp datatype when create table (#17305 )	2023-03-07 21:03:25 +08:00
zxealous	5334a5899e	[fix](remote)fix whole file cache and sub file cache (#17468 )	2023-03-07 19:55:18 +08:00
zhangstar333	06468ba627	[vectorized](bug) fix array constructor function change origin column from block (#17296 )	2023-03-07 16:42:23 +08:00
minghong	fd8adb492d	[fix](nereids) fix bugs in nereids window function (#17284 ) fix two problems: 1. push agg-fun in windowExpression down to AggregateNode for example, sql: select sum(sum(a)) over (order by b) Plan: windowExpression( sum(y) over (order by b)) +--- Agg(sum(a) as y, b) 2. push other expr to upper proj for example, sql: select sum(a+1) over () Plan: windowExpression(sum(y) over ()) +--- Project(a + 1 as y,...) +--- Agg(a,...)	2023-03-07 16:35:37 +08:00
ZhangYu0123	8ccc805cd0	[Fix](Lightweight schema Change) query error caused by array default type is unsupported (#17331 ) We have supportted array type default [], but when using lightweight schema Change to add column array type, query failed as follows: Fix "array default type is unsupported" error. Fix the default value filling assignment digit problem.	2023-03-07 16:30:41 +08:00
mch_ucchi	86252e25bf	[regression-test](Nereids) add binary arithmetic regression test cases(#17363 ) add all of the valid binary arithmetic expressions test for nereids. currently, float, double, stringlike(string, char, varchar) doesn't support div, bitand, bitor, bitxor. some results with float type are incorrect because of inaccurate precision of regression-test framework.	2023-03-07 15:48:22 +08:00
liujinhui	fca567068e	[Enhancement](spark load)Support for RM HA (#15000 ) Adding RM HA configuration to the spark load. Spark can accept HA parameters via config, we just need to accept it in the DDL CREATE EXTERNAL RESOURCE spark_resource_sinan_node_manager_ha PROPERTIES ( "type" = "spark", "spark.master" = "yarn", "spark.submit.deployMode" = "cluster", "spark.executor.memory" = "10g", "spark.yarn.queue" = "XXXX", "spark.hadoop.yarn.resourcemanager.address" = "XXXX:8032", "spark.hadoop.yarn.resourcemanager.ha.enabled" = "true", "spark.hadoop.yarn.resourcemanager.ha.rm-ids" = "rm1,rm2", "spark.hadoop.yarn.resourcemanager.hostname.rm1" = "XXXX", "spark.hadoop.yarn.resourcemanager.hostname.rm2" = "XXXX", "spark.hadoop.fs.defaultFS" = "hdfs://XXXX", "spark.hadoop.dfs.nameservices" = "hacluster", "spark.hadoop.dfs.ha.namenodes.hacluster" = "mynamenode1,mynamenode2", "spark.hadoop.dfs.namenode.rpc-address.hacluster.mynamenode1" = "XXX:8020", "spark.hadoop.dfs.namenode.rpc-address.hacluster.mynamenode2" = "XXXX:8020", "spark.hadoop.dfs.client.failover.proxy.provider" = "org.apache.hadoop.hdfs.server.namenode.ha.ConfiguredFailoverProxyProvider", "working_dir" = "hdfs://XXXX/doris_prd_data/sinan/spark_load/", "broker" = "broker_personas", "broker.username" = "hdfs", "broker.password" = "", "broker.dfs.nameservices" = "XXX", "broker.dfs.ha.namenodes.XXX" = "mynamenode1, mynamenode2", "broker.dfs.namenode.rpc-address.XXXX.mynamenode1" = "XXXX:8020", "broker.dfs.namenode.rpc-address.XXXX.mynamenode2" = "XXXX:8020", "broker.dfs.client.failover.proxy.provider" = "org.apache.hadoop.hdfs.server.namenode.ha.ConfiguredFailoverProxyProvider" ); Co-authored-by: liujh <liujh@t3go.cn>	2023-03-07 15:46:14 +08:00
谢健	704faaed84	[feature](Nereids) add rule split limit into two phase (#16797 ) 1. Add a rule split limit, like Limit(Origin) ==> Limit(Global) -> Gather -> Limit(Local) 2. Add a rule: limit-> sort ==> topN 3. fix a bug about topN 4. make the type of limit,offset long in topN And because this rule is always beneficial, we add a rule in the rewrite phase	2023-03-07 15:34:12 +08:00
Jerry Hu	caacee253d	[fix](olap)Crashing caused by IS NULL expression (#17463 ) Issue Number: close #17462	2023-03-07 15:32:52 +08:00
谢健	05c5ab5490	[fix](planner) only table name should convert to lowercase when create table (#17373 ) we met error: Unknown column '{}DORIS_DELETE_SIGN{}' in 'default_cluster:db.table. that because when we use alias as the tableName to construct a Table, all parts of the name will be lowercase if lowerCaseTableNames = 1. To avoid it, we should extract tableName from alias and only lower tableName	2023-03-07 14:41:35 +08:00
mch_ucchi	b9bb28f22c	[Enhancement](Planner)fix unclear exception msg when create table. #17473	2023-03-07 13:38:20 +08:00
jakevin	357d8c1746	[enhance](Nereids): remove rule flag in LogicalJoin (#17452 )	2023-03-07 13:18:50 +08:00
jakevin	b8c9875adb	[refactor](Nereids): refactor PushdownLimit (#17355 )	2023-03-07 12:04:20 +08:00
jakevin	b0e3156f51	[enhance](Nereids): refactor code in Project (#17450 )	2023-03-07 11:15:33 +08:00
pengxiangyu	f79b066790	[fix](resource)Add s3 checker for alter resource (#17467 ) * add s3 validity checker for alter resource. * add s3 validity checker for alter resource. * add s3 validity checker for alter resource.	2023-03-07 11:07:15 +08:00
pengxiangyu	e94170d81f	[feature](cooldown)add ut for cooldown on be (#17246 ) * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be * add ut for cooldown on be	2023-03-07 10:42:53 +08:00
zhangdong	7e96b06e6c	[Enhance](auth)Users support multiple roles (#17236 ) Describe your changes. 1.support GRANT role [, role] TO user_identity 2.support REVOKE role [, role] FROM user_identity 3.’Show grants‘ Add a column to display the roles owned by users 4.‘alter user’ prohibit deleting user's role 5.Repair Logic of roleName cannot start with RoleManager.DEFAULT_ ROLE	2023-03-07 10:28:56 +08:00
xueweizhang	bada731390	[fix](restore) fix bug when replay restore and reserve dynamic partition (#17326 ) when replay restore a table with reserve_dynamic_partition_enable=true, must registerOrRemoveDynamicPartitionTable with isReplay=true, or maybe cause OBSERVER can not replay restore auditlog success.	2023-03-07 10:13:08 +08:00
AKIRA	f85f89f240	[fix](planner) Fix incosistency between groupby expression and output of aggregation node (#17438 )	2023-03-07 09:38:20 +08:00
Pxl	d8f0ca7108	[Chore](schema change) remove some unused code in schema change (#17459 ) remove some unused code in schema change. remove some row-based config and code.	2023-03-07 09:18:34 +08:00
Yulei-Yang	50bf02024a	[Improvement](meta) support return total statistics of all databases for command show proc '/jobs (#17342 ) currently, show proc jobs command can only used on a specific database, if a user want to see overall data of the whole cluster, he has to look into every database and sum them up, it's troublesome. now he can achieve it simply by giving a -1 as dbId. mysql> show proc '/jobs/-1'; +---------------+---------+---------+----------+-----------+-------+ \| JobType \| Pending \| Running \| Finished \| Cancelled \| Total \| +---------------+---------+---------+----------+-----------+-------+ \| load \| 0 \| 0 \| 0 \| 2 \| 2 \| \| delete \| 0 \| 0 \| 0 \| 0 \| 0 \| \| rollup \| 0 \| 0 \| 1 \| 0 \| 1 \| \| schema_change \| 0 \| 0 \| 2 \| 0 \| 2 \| \| export \| 0 \| 0 \| 0 \| 3 \| 3 \| +---------------+---------+---------+----------+-----------+-------+ mysql> show proc '/jobs/-1/rollup'; +----------+------------------+---------------------+---------------------+------------------+-----------------+----------+---------------+----------+------+----------+---------+ \| JobId \| TableName \| CreateTime \| FinishTime \| BaseIndexName \| RollupIndexName \| RollupId \| TransactionId \| State \| Msg \| Progress \| Timeout \| +----------+------------------+---------------------+---------------------+------------------+-----------------+----------+---------------+----------+------+----------+---------+ \| 17826065 \| order_detail \| 2023-02-23 04:21:01 \| 2023-02-23 04:21:22 \| order_detail \| rp1 \| 17826066 \| 6009 \| FINISHED \| \| NULL \| 2592000 \| +----------+------------------+---------------------+---------------------+------------------+-----------------+----------+---------------+----------+------+----------+---------+ 1 row in set (0.01 sec)	2023-03-07 08:57:55 +08:00
ZhangYu0123	440cf526c8	[fix](type compatibility) fix unsigned int type compatibility problem (#17427 ) Fix unsigned int type compatibility value scope problem. When defining columns, map UNSIGNED INT to BIGINT for compatibility. The problems are as follows: It is not consistent with this doc image We support the unsigned int type to be compatible with mysql types, but the unsigned int type is created as the int at the time of definition. This will cause numerical overflow.	2023-03-07 08:55:38 +08:00
Jerry Hu	6f3801d9da	[chore](config) Increase the default maximum depth limit for expressions (#17418 )	2023-03-07 08:53:00 +08:00
Mingyu Chen	57233dc53a	[deps](libhdfs) fix hadoop libs build error (#17470 )	2023-03-07 08:52:40 +08:00
Yulei-Yang	b68001aee5	[fix](priv) fix duplicated priv check when check column priv (#17446 ) when executing select stmt, columns privilege check will be invoked multiple times(column number in select stmt) Issue Number: close #xxx	2023-03-07 08:51:55 +08:00
Tiewei Fang	48c2d806d7	[enhencement](jdbc catalog) Use Druid instead of HikariCP in JdbcClient (#17395 ) This pr does three things: 1. Use Druid instead of HikariCP in JdbcClient 2. when download udf jar, add the name of the jar package after the local file name. 3. refactor some jdbcResource code	2023-03-07 08:51:10 +08:00
AKIRA	aedbc5fcb1	[fix](planner) Slots in the cojuncts of table function node didn't got materialized #17460	2023-03-07 08:50:33 +08:00
zhangdong	bc48cbff83	[doc](auth)auth doc (#17358 ) * auth doc * auth en doc * add note	2023-03-07 08:05:09 +08:00

1 2 3 4 5 ...

9158 Commits