doris

Author	SHA1	Message	Date
pengxiangyu	7592f52d2e	[Feature][Insert] Add transaction for the operation of insert #6244 (#6245 ) ## Proposed changes Add transaction for the operation of insert. It will cost less time than non-transaction(it will cost 1/1000 time) when you want to insert a amount of rows. ### Syntax ``` BEGIN [ WITH LABEL label]; INSERT INTO table_name ... [COMMIT \| ROLLBACK]; ``` ### Example commit a transaction: ``` begin; insert into Tbl values(11, 22, 33); commit; ``` rollback a transaction: ``` begin; insert into Tbl values(11, 22, 33); rollback; ``` commit a transaction with label: ``` begin with label test_label; insert into Tbl values(11, 22, 33); commit; ``` ### Description ``` begin: begin a transaction, the next insert will execute in the transaction until commit/rollback; commit: commit the transaction, the data in the transaction will be inserted into the table; rollback: abort the transaction, nothing will be inserted into the table; ``` ### The main realization principle: ``` 1. begin a transaction in the session. next sql is executed in the transaction; 2. insert sql will be parser and get the database name and table name, they will be used to select a be and create a pipe to accept data; 3. all inserted values will be sent to the be and write into the pipe; 4. a thread will get the data from the pipe, then write them to disk; 5. commit will complete this transaction and make these data visible; 6. rollback will abort this transaction ``` ### Some restrictions on the use of update syntax. 1. Only ```insert``` can be called in a transaction. 2. If something error happened, ```commit``` will not succeed, it will ```rollback``` directly; 3. By default, if part of insert in the transaction is invalid, ```commit``` will only insert the other correct data into the table. 4. If you need ```commit``` return failed when any insert in the transaction is invalid, you need execute ```set enable_insert_strict = true``` before ```begin```.	2021-07-21 10:54:11 +08:00
xy720	7c34dbbc5b	[Bug-Fix] Fix bug that show view report "Unresolved table reference" error (#6184 )	2021-07-15 10:55:15 +08:00
pengxiangyu	01bef4b40d	[Load] Add "LOAD WITH HDFS" model, and make hdfs_reader support hdfs ha (#6161 ) Support load data from HDFS by using `LOAD WITH HDFS` syntax and read data directly via libhdfs3	2021-07-10 10:11:52 +08:00
Stalary	d6e6c7815b	[Feature] ADD: show create routine load (#6110 ) Add show create routine load	2021-07-04 21:43:25 +08:00
caiconghui	d9c128b744	[BrokerLoad] Support read properties for broker load when read data (#5845 ) * [BrokerLoad] support read properties for broker load when read data Co-authored-by: caiconghui <caiconghui@xiaomi.com>	2021-06-09 14:59:55 +08:00
xy720	a29dd42b47	[BUG][Document] Fix the bug that failed to build the help module (#5917 ) There are multiple entries with same key in help documents, which will cause build help module failed.	2021-05-27 22:07:15 +08:00
xinghuayu007	ba69f7a7c8	[Command] [SQL] Add show database/table/partition id command (#5807 ) In BE, when a problem happened, in the log, we can find the database id, table id, partition id, but no database name, table name, partition name. In FE, there also no way to find database name/table name/partition name accourding to database id/table id/partition id. Therefore, this patch add 3 new commands: 1. show database id; mysql> show database 10002; +----------------------+ \| DbName \| +----------------------+ \| default_cluster:test \| +----------------------+ 2. show table id; mysql> show table 11100; +----------------------+-----------+-------+ \| DbName \| TableName \| DbId \| +----------------------+-----------+-------+ \| default_cluster:test \| table2 \| 10002 \| +----------------------+-----------+-------+ 3. show partition id; mysql> show partition 11099; +----------------------+-----------+---------------+-------+---------+ \| DbName \| TableName \| PartitionName \| DbId \| TableId \| +----------------------+-----------+---------------+-------+---------+ \| default_cluster:test \| table2 \| p201708 \| 10002 \| 11100 \| +----------------------+-----------+---------------+-------+---------+	2021-05-26 09:58:02 +08:00
Mingyu Chen	07ad038870	[Feature][RoutineLoad] Support for consuming kafka from the point of time (#5832 ) Support when creating a kafka routine load, start consumption from a specified point in time instead of a specific offset. eg: ``` FROM KAFKA ( "kafka_broker_list" = "broker1:9092,broker2:9092", "kafka_topic" = "my_topic", "property.kafka_default_offsets" = "2021-10-10 11:00:00" ); or FROM KAFKA ( "kafka_broker_list" = "broker1:9092,broker2:9092", "kafka_topic" = "my_topic", "kafka_partitions" = "0,1,2", "kafka_offsets" = "2021-10-10 11:00:00, 2021-10-10 11:00:00, 2021-10-10 12:00:00" ); ``` This PR also reconstructed the analysis method of properties when creating or altering routine load jobs, and unified the analysis process in the `RoutineLoadDataSourceProperties` class.	2021-05-22 23:37:53 +08:00
zh0122	12e4ff2689	[Doc] Fix doc for 'SHOW EXPORT' (#5840 )	2021-05-19 09:31:57 +08:00
Zhengguo Yang	9eacd0a89c	[Doc] remove storage_type from docs (#5814 )	2021-05-19 09:29:15 +08:00
DinoZhang	65ff464e3d	[Feature] Support show data order by (#5770 ) Currently, the `show data` does not support sorting. When the number of tables increases, it is inconvenient to manage. Need to support sorting like: ``` mysql> show data order by ReplicaCount desc,Size asc; +-----------+-------------+--------------+ \| TableName \| Size \| ReplicaCount \| +-----------+-------------+--------------+ \| table_c \| 3.102 KB \| 40 \| \| table_d \| .000 \| 20 \| \| table_b \| 324.000 B \| 20 \| \| table_a \| 1.266 KB \| 10 \| \| Total \| 4.684 KB \| 90 \| \| Quota \| 1024.000 GB \| 1073741824 \| \| Left \| 1024.000 GB \| 1073741734 \| +-----------+-------------+--------------+ ```	2021-05-19 09:27:27 +08:00
caiconghui	add8c4bb74	[Load] Support reading multi-line json objects for JsonScanner (#5774 ) Co-authored-by: caiconghui <caiconghui@xiaomi.com>	2021-05-18 15:44:45 +08:00
EmmyMiao87	3fdfe0ba6f	[Bug-fix] Export specified column (#5759 ) The code logic error causes the user to specify the export column, which may not be effective. The PR fix this problem.	2021-05-08 10:56:45 +08:00
weizuo93	9001fd28f4	support show stream load sql (#5488 ) Co-authored-by: weizuo <weizuo@xiaomi.com>	2021-04-29 09:20:35 +08:00
qiye	de87f4ae84	[Feature] Add list partition support (#5529 ) Add list partition support	2021-04-24 17:42:27 +08:00
Zhengguo Yang	86af8c76a3	[DOC] Add docs of load and export using S3 protocol (#5551 ) Add docs of load and export using S3 protocol	2021-03-27 18:58:29 +08:00
Ting Sun	64fa305c06	[Doc] correct format errors in English doc (#5487 ) Some formate errors in English doc. They are very straightforward and should not break any existing build.	2021-03-11 22:34:54 +08:00
EmmyMiao87	6cbbc36ea1	[Export] Expand function of export stmt (#5445 ) 1. Support where clause in export stmt which only export selected rows. The syntax is following: Export table [table name] where [expr] To xxx xxxx It will filter table rows. Only rows that meet the where condition can be exported. 2. Support utf8 separator 3. Support export to local The syntax is following: Export table [table name] To (file:///xxx/xx/xx) If user export rows to local, the broker properties is not requried. User only need to create a local folder to store data, and fill in the path of the folder starting with file:// Change-Id: Ib7e7ece5accb3e359a67310b0bf006d42cd3f6f5	2021-03-11 20:43:32 +08:00
Ting Sun	e93a6da0e5	[Doc] correct format errors in English doc (#5321 ) Fix some English doc format errors	2021-02-26 11:32:14 +08:00
Mingyu Chen	780900ac9c	[Feature] Support preceding filter original data when loading (#5338 ) Support conditional filtering of original data in broker load and routine load eg: ``` LOAD LABEL `label1` ( DATA INFILE ('bos://cmy-repo/1.csv') INTO TABLE tbl2 COLUMNS TERMINATED BY '\t' (event_day, product_id, ocpc_stage, user_id) SET ( ocpc_stage = ocpc_stage + 100 ) PRECEDING FILTER user_id = 1381035 WHERE ocpc_stage > 30 ) ... ```	2021-02-07 22:37:48 +08:00
Mingyu Chen	de57667d6d	[Delete] Support delete with multi partitions (#5252 ) Support delete statement like: 1. delete from table partitions(p1, p2) where xxx; // apply to p1, p2 2. delete from table where xxx; // apply to all partitions Also remove code about the deprecated sync/async delete job. This CL changes FE meta version to 94	2021-01-30 20:33:34 +08:00
caiconghui	ca10205137	[Function] Support show create function statement (#5197 ) * [Function]Support show create function stmt Co-authored-by: caiconghui [蔡聪辉] <caiconghui@xiaomi.com>	2021-01-28 10:52:37 +08:00
Zhengguo Yang	83b7a23d5c	fix alter routine load not work (#5257 )	2021-01-20 10:52:02 +08:00
Zhengguo Yang	279ae1cb75	Add fuzzy_parse option to speed up json import (#5114 ) add a flag of fuzzy_parse, if the json file all object keys are the same and has same order, we only need to parse the first row, and then use index instead key to parse value	2020-12-25 09:19:42 +08:00
EmmyMiao87	d6497fedc4	[Config] Change config name 'streaming_load_max_batch_size_mb' to 'streaming_load_json_max_mb' (#4791 ) The name and another config name are close to each other and are indistinguishable. So this pr modify the name. The document description has also been changed	2020-10-28 23:27:33 +08:00
Zhengguo Yang	751aa05cc0	fix docs typo (#4725 )	2020-10-14 09:27:50 +08:00
Zhengguo Yang	dec91a3d43	fix docs typo (#4723 )	2020-10-14 09:27:31 +08:00
Zhengguo Yang	3f55c1425c	fix docs typo (#4722 )	2020-10-14 09:27:12 +08:00
Zhengguo Yang	0475aa9b93	[Bug]Fix delete on clause may not work in routineLoad (#4683 ) fix delete on may not work in some cases, this is describe in #4682	2020-09-30 09:56:19 +08:00
Zhengguo Yang	174c9f89ea	[DOCS] Add batch delete docs (#4435 ) update documents for batch delete #4051	2020-08-28 09:24:07 +08:00
ZhangYu0123	1d9b3aeee7	[Doc] Repair document format (#4336 ) The error format '##keyword' in a lot of docs. This pr is to repair document format. #4335	2020-08-13 23:39:41 +08:00
caiconghui	eefad13107	[Feature] Support InPredicate in delete statement (#4006 ) This PR is to add inPredicate support to delete statement, and add max_allowed_in_element_num_of_delete variable to limit element num of InPredicate in delete statement.	2020-08-06 23:19:40 +08:00
Mingyu Chen	237c0807a4	[RoutineLoad] Support modify routine load job (#4158 ) Support ALTER ROUTINE LOAD JOB stmt, for example: ``` alter routine load db1.label1 properties ( "desired_concurrent_number"="3", "max_batch_interval" = "5", "max_batch_rows" = "300000", "max_batch_size" = "209715200", "strict_mode" = "false", "timezone" = "+08:00" ) ``` Details can be found in `alter-routine-load.md`	2020-08-06 23:11:02 +08:00
worker24h	fdcc223ad2	[Bug][Json] Refactor the json load logic to fix some bug 1. Add `json_root` for nest json data. 2. Remove `_jmap` to make the logic reasonable.	2020-07-30 10:36:34 +08:00
Mingyu Chen	c3d9feed75	[Load][Json] Refactor json load logic to make it more reasonable (#4020 ) This CL mainly changes: 1. Reorganized the code logic to limit the supported json format to two, and the import behavior is more consistent. 2. Modified the statistical behavior of the number of error rows when loading in json format, so that the error rows can be counted correctly. 3. See `load-json-format.md` to get details of loading json format.	2020-07-07 23:07:28 +08:00
Mingyu Chen	77b9acc242	[Stmt] Add rowCount column to SHOW DATA stmt (#3676 ) User can see the row count of all materialized indexes of a table. ``` mysql> show data from test; +-----------+-----------+-----------+--------------+----------+ \| TableName \| IndexName \| Size \| ReplicaCount \| RowCount \| +-----------+-----------+-----------+--------------+----------+ \| test2 \| r1 \| 10.000MB \| 30 \| 10000 \| \| \| r2 \| 20.000MB \| 30 \| 20000 \| \| \| test2 \| 50.000MB \| 30 \| 50000 \| \| \| Total \| 80.000 \| 90 \| \| +-----------+-----------+-----------+--------------+----------+ ``` Fix #3675	2020-05-26 15:53:38 +08:00
worker24h	ef8fd1fcbe	[Load] Support load json-data into Doris by RoutineLoad or StreamLoad (#3553 ) Doris support load json-data by RoutineLoad or StreamLoad	2020-05-21 13:00:49 +08:00
EmmyMiao87	f591976976	[Doc] Fix the incorrect docs (#3501 )	2020-05-08 12:47:00 +08:00
yangzhg	54da5a491c	Fix delete statement doc display not correctly (#3445 )	2020-05-01 19:20:00 +08:00
hffariel	432965e360	[Enhancement] documents rebuild with Vuepress (#3408 ) (#3414 )	2020-04-29 09:14:31 +08:00

40 Commits