ClickHouse

mirror of https://github.com/ClickHouse/ClickHouse.git synced 2024-12-15 02:41:59 +00:00

Author	SHA1	Message	Date
kssenii	3dee003f9b	Merge branch 'master' of github.com:ClickHouse/ClickHouse into poco-file-to-std-fs	2021-05-20 19:20:09 +03:00
Alexander Kuzmenkov	e9b69bbd70	Merge pull request #23906 from azat/fix-distributed_group_by_no_merge distributed_group_by_no_merge fixes	2021-05-19 16:16:08 +03:00
Alexander Kuzmenkov	09cb467812	Update StorageDistributed.cpp	2021-05-19 16:14:33 +03:00
kssenii	9b8df78fdd	Merge branch 'master' of github.com:ClickHouse/ClickHouse into poco-file-to-std-fs	2021-05-17 17:42:05 +03:00
feng lv	c6f8ab9826	fix	2021-05-13 02:05:53 +00:00
kssenii	0527f0ea33	Merge branch 'master' of github.com:ClickHouse/ClickHouse into poco-file-to-std-fs	2021-05-12 16:54:18 +03:00
Amos Bird	cd6414639e	add metadata_snapshot to getQueryProcessingStage	2021-05-11 18:12:26 +08:00
Azat Khuzhin	eefd67fce5	Disable optimize_distributed_group_by_sharding_key with window functions	2021-05-06 00:44:22 +03:00
feng lv	39f68bf5ff	fix conflict	2021-05-02 16:33:45 +00:00
kssenii	ee06936596	Merge branch 'master' of github.com:ClickHouse/ClickHouse into poco-file-to-std-fs	2021-05-01 17:24:31 +03:00
feng lv	aed2f337e9	Fix CLEAR COLUMN does not work after #21303	2021-04-30 05:02:32 +00:00
kssenii	deb4903af8	Merge branch 'master' of github.com:ClickHouse/ClickHouse into poco-file-to-std-fs	2021-04-28 20:57:13 +03:00
kssenii	eeb71672a0	Change in Storages/*	2021-04-27 16:49:37 +03:00
feng lv	4ffe199d39	Implement table comments	2021-04-23 12:18:23 +00:00
Amos Bird	096d76627e	Skip unavaiable shards when writing to distributed tables	2021-04-21 10:30:40 +08:00
Maksim Kita	e361f5943f	Merge pull request #22999 from azat/no-optimize_skip_unused_shards-single-node Do not perform optimize_skip_unused_shards for cluster with one node	2021-04-15 14:36:56 +03:00
Nikita Mikhaylov	7a68820342	style	2021-04-13 22:39:42 +03:00
Nikita Mikhaylov	081ea84a41	save	2021-04-13 22:39:41 +03:00
tavplubix	1525e38a3c	Merge pull request #22990 from ClickHouse/tavplubix-patch-1 Fix excessive warning in StorageDistributed with cross-replication	2021-04-13 18:58:12 +03:00
Azat Khuzhin	a497d4d462	Do not perform optimize_skip_unused_shards for cluster with one node	2021-04-12 22:18:31 +03:00
tavplubix	a995962e6a	Update StorageDistributed.cpp	2021-04-12 14:58:24 +03:00
Azat Khuzhin	79bd8d4d3f	Respect optimize_skip_unused_shards_rewrite_in with optimize_skip_unused_shards_limit	2021-04-12 10:37:28 +03:00
Azat Khuzhin	e439914d38	Fix optimized cluster logic for optimize_skip_unused_shards	2021-04-12 10:37:28 +03:00
Azat Khuzhin	fbb386dca5	Rewrite IN in query for remote shards to exclude values that does not belongs to shard v2: fix optimize_skip_unused_shards_rewrite_in for sharding_key wrapped into function v3: fix column name for optimize_skip_unused_shards_rewrite_in v4: fix optimize_skip_unused_shards_rewrite_in with Null v5: - squash with Remove query argument for IStreamFactory::createForShard() - use proper column after function execution (using sharding_key_column_name) - update the test reference since (X) now is tuple(X)	2021-04-12 10:37:28 +03:00
Ivan	495c6e03aa	Replace all Context references with std::weak_ptr (#22297 ) * Replace all Context references with std::weak_ptr * Fix shared context captured by value * Fix build * Fix Context with named sessions * Fix copy context * Fix gcc build * Merge with master and fix build * Fix gcc-9 build	2021-04-11 02:33:54 +03:00
Nikolai Kochetov	6102652c99	Merge branch 'master' into better-filter-push-down	2021-04-06 13:38:03 +03:00
Maxim Akhmedov	725fa17961	Introduce IStorage::distributedWrite method for distributed INSERT SELECT.	2021-04-05 02:14:27 +03:00
Nikolai Kochetov	c3c393a7aa	Merge branch 'master' into refactor-actions-dag	2021-03-18 14:33:07 +03:00
Nikolai Kochetov	e8d7349c79	Merge branch 'master' into dist-query-zero-shards-fix	2021-03-16 12:00:08 +03:00
Azat Khuzhin	61d40c3600	Fix optimize_skip_unused_shards for zero shards case v2: move check to the beginning of the StorageDistributed::read()	2021-03-10 09:05:14 +03:00
Azat Khuzhin	3474ea044e	Avoid processing optimize_skip_unused_shards twice	2021-03-09 10:05:56 +03:00
Azat Khuzhin	ed09897eb1	Pass optimize_skip_unused_shards_limit to the bottom layer And now optimize_skip_unused_shards_limit=0 is not a special case anymore.	2021-03-08 10:05:56 +03:00
Azat Khuzhin	16f4c02d42	Add optimize_skip_unused_shards_limit Limit for number of sharding key values, turns off optimize_skip_unused_shards if the limit is reached	2021-03-26 06:09:00 +03:00
Nikolai Kochetov	a669f7d641	Merge branch 'master' into refactor-actions-dag	2021-03-05 18:21:14 +03:00
Nikolai Kochetov	9a39459888	Refactor ActionsDAG	2021-03-04 20:38:12 +03:00
Azat Khuzhin	6965ac26c3	Distributed: Add ability to delay/throttle INSERT until pending data will be reduced Add two new settings for the Distributed engine: - bytes_to_delay_insert - max_delay_to_insert If at the beginning of INSERT there will be too much pending data, more then bytes_to_delay_insert, then the INSERT will wait until it will be shrinked, and not more then max_delay_to_insert seconds. If after this there will be still too much pending, it will throw an exception. Also new profile events were added (by analogy to the MergeTree): - DistributedDelayedInserts (although you can use system.errors instead of this, but still) - DistributedRejectedInserts - DistributedDelayedInsertsMilliseconds	2021-03-03 23:30:23 +03:00
Azat Khuzhin	b43046ba06	Distributed: More accurate distribution_queue counters So now system.distribution_queue will show accurate statistics, so tests does not requires sleep anymore. But note that with too much distributed pending this will iterate over all directories.	2021-03-03 23:30:03 +03:00
Azat Khuzhin	b5a5778589	Distributed: Add ability to limit amount of pending bytes for async INSERT Right now with distributed_directory_monitor_batch_inserts=1 and insert_distributed_sync=0 INSERT into Distributed table will store blocks that should be sent to remote (and in case of prefer_localhost_replica=0 to the localhost too) on the local filesystem, and sent it in background. However there is no limit for this storage, and if the remote is unavailable (or some other error), these pending blocks may take significant space, and this is not always desired behaviour. Add new Distributed setting - bytes_to_throw_insert, that will set the limit for how much pending bytes is allowed, if the limit will be reached an exception will be throw. By default was set to 0, to avoid surprises.	2021-03-03 23:30:00 +03:00
Azat Khuzhin	ce09b7ff89	Distributed: Implement totalBytes() (system.tables.total_bytes)	2021-03-03 23:29:11 +03:00
Anton Popov	a4c00ab5dc	Merge pull request #21303 from ucasFL/forbid Forbid to drop a column if it's referenced by materialized view	2021-03-03 02:55:06 +03:00
feng lv	a26c9e64a9	fix fix	2021-03-02 03:20:03 +00:00
feng lv	51021c1164	forbid to drop a column if it's referenced by materialized view	2021-02-28 05:24:39 +00:00
Nikolai Kochetov	d328bfa41f	Review fixes. Add setting max_optimizations_to_apply.	2021-02-26 19:29:56 +03:00
Azat Khuzhin	809fa7e4cc	Sync SYSTEM FLUSH DISTRIBUTED with TRUNCATE	2021-02-10 23:10:37 +03:00
Azat Khuzhin	ce91c257b2	Lockless SYSTEM FLUSH DISTRIBUTED Right now SYSTEM FLUSH DISTRIBUTED will block: - INSERT into this Distributed table (requireDirectoryMonitor()) - SELECT * FROM system.distribution_queue	2021-02-08 22:07:30 +03:00
Kruglov Pavel	d94e8624d7	Merge branch 'master' into shard-id	2021-02-06 16:48:17 +03:00
Aleksei Semiglazov	921518db0a	CLICKHOUSE-606: query deduplication based on parts' UUID * add the query data deduplication excluding duplicated parts in MergeTree family engines. query deduplication is based on parts' UUID which should be enabled first with merge_tree setting assign_part_uuids=1 allow_experimental_query_deduplication setting is to enable part deduplication, default ot false. data part UUID is a mechanism of giving a data part a unique identifier. Having UUID and deduplication mechanism provides a potential of moving parts between shards preserving data consistency on a read path: duplicated UUIDs will cause root executor to retry query against on of the replica explicitly asking to exclude encountered duplicated fingerprints during a distributed query execution. NOTE: this implementation don't provide any knobs to lock part and hence its UUID. Any mutations/merge will update part's UUID. * add _part_uuid virtual column, allowing to use UUIDs in predicates. Signed-off-by: Aleksei Semiglazov <asemiglazov@cloudflare.com> address comments	2021-02-02 16:53:39 +00:00
feng lv	4279c7da41	add setting insert_shard_id add test fix style fix	2021-02-02 04:26:59 +00:00
kreuzerkrieg	29a2ef3089	Add IStoragePolicy interface	2021-01-26 10:55:28 +02:00
Azat Khuzhin	2e55bd2285	Accept IDisk in DirectoryMonitor (for further fsync)	2021-01-09 16:31:42 +03:00

1 2 3 4

196 Commits