citus

Commit Graph

Author	SHA1	Message	Date
Gürkan İndibay	804b521d63	Removes ubuntu/bionic from packaging pipelines (#7142 ) DESCRIPTION: Removes ubuntu/bionic from packaging pipelines Since pg16 beta is not available for ubuntu/bionic and ubuntu/bionic support is EOL, I need to remove this os from pipeline https://ubuntu.com/blog/ubuntu-18-04-eol-for-devices Additionally, added concurrency support for GH Actions Packaging pipeline (cherry picked from commit `553780e3f1`)	2025-02-04 03:13:20 +03:00
Gürkan İndibay	4ef1b4a3db	Removes el/7 and ol/7 as runners (#7650 ) Removes el/7 and ol/7 as runners and update checkout action to v4 We use EL/7 and OL/7 runners to test packaging for these distributions. However, for the past two weeks, we've encountered errors during the checkout step in the pipelines. The error message is as follows: ``` /__e/node20/bin/node: /lib64/libm.so.6: version `GLIBC_2.27' not found (required by /__e/node20/bin/node) /__e/node20/bin/node: /lib64/libstdc++.so.6: version `GLIBCXX_3.4.20' not found (required by /__e/node20/bin/node) /__e/node20/bin/node: /lib64/libstdc++.so.6: version `CXXABI_1.3.9' not found (required by /__e/node20/bin/node) /__e/node20/bin/node: /lib64/libstdc++.so.6: version `GLIBCXX_3.4.21' not found (required by /__e/node20/bin/node) /__e/node20/bin/node: /lib64/libc.so.6: version `GLIBC_2.28' not found (required by /__e/node20/bin/node) /__e/node20/bin/node: /lib64/libc.so.6: version `GLIBC_2.25' not found (required by /__e/node20/bin/node) ``` The GCC version within the EL/7 and OL/7 Docker images is 2.17, and we cannot upgrade it. Therefore, we need to remove these images from the packaging test pipelines. Consequently, we will no longer verify if the code builds for EL/7 and OL/7. However, we are not using these packaging images as runners within the packaging infrastructure, so we can continue to use these images for packaging. Additional Info: I learned that Marlin team fully dropped the el/7 support so we will drop in further releases as well (cherry picked from commit `c603c3ed74`)	2025-02-04 03:13:20 +03:00
Onur Tirtir	7a7dfb9752	Avoid publishing artifacts with conflicting names .. as documented in actions/upload-artifact#480. (cherry picked from commit `26f16a7654`)	2025-02-04 03:13:20 +03:00
Onur Tirtir	dc10a64eff	Upgrade download-artifacts action to 4.1.8 (cherry picked from commit `cbe0de33a6`)	2025-02-04 03:13:20 +03:00
Onur Tirtir	f3d60ec1b4	Upgrade upload-artifacts action to 4.6.0 (cherry picked from commit `b886cfa223`)	2025-02-04 03:13:20 +03:00
Gürkan İndibay	56fdf0e80e	Bump Citus to 12.0.1 (#7506 )	2024-02-14 08:40:45 +03:00
Gürkan İndibay	1dcd5ff046	Adds Changelog for v12.0.1 (#7498 ) Co-authored-by: Onur Tirtir <onurcantirtir@gmail.com>	2024-02-13 17:09:57 +03:00
Teja Mupparti	b3da549da9	Fix the incorrect column count after ALTER TABLE, this fixes the bug #7378 (please read the analysis in the bug for more information) (cherry picked from commit `00068e07c5`)	2024-01-26 08:15:11 -08:00
Gokhan Gulbiz	0ef2e14b16	Backport GHA Migration to release-12.0 (#7314 ) --------- Co-authored-by: Jelte Fennema-Nio <jelte.fennema@microsoft.com> Co-authored-by: aykut-bozkurt <51649454+aykut-bozkurt@users.noreply.github.com>	2023-11-10 09:35:32 +00:00
Onur Tirtir	03d11bbebd	Make sure to disallow creating a replicated distributed table concurrently (#7219 ) See explanation in https://github.com/citusdata/citus/issues/7216. Fixes https://github.com/citusdata/citus/issues/7216. DESCRIPTION: Makes sure to disallow creating a replicated distributed table concurrently (cherry picked from commit `111b4c19bc`)	2023-10-24 14:04:36 +03:00
Nils Dijk	350db42577	Fix leaking of memory and memory contexts in Foreign Constraint Graphs (#7236 ) DESCRIPTION: Fix leaking of memory and memory contexts in Foreign Constraint Graphs Previously, every time we (re)created the Foreign Constraint Relationship Graph, we created a new Memory Context while loosing a reference to the previous context. This old context could still have left over memory in there causing a memory leak. With this patch we statically have one memory context that we lazily initialize the first time we create our foreign constraint relationship graph. On every subsequent creation, beside destroying our previous hashmap we also reset our memory context to remove any left over references.	2023-10-09 13:07:47 +02:00
Gürkan İndibay	72b62ea0d1	Removes ubuntu:kinetic pipelines since it's EOL (#7195 ) ubuntu:kinetic is EOL so removing it's pipeline https://fridge.ubuntu.com/2023/06/14/ubuntu-22-10-kinetic-kudu-reaches-end-of-life-on-july-20-2023/ (cherry picked from commit `e0683aab84`)	2023-09-26 16:54:37 +03:00
Gürkan İndibay	7a22dd581a	Removes pg_send_cancellation (#7135 ) DESCRIPTION: Removes pg_send_cancellation and all references (cherry picked from commit `371f094b68`)	2023-09-26 16:54:37 +03:00
Hanefi Onaldi	2105c20344	Create a new colocation properly after braking one When braking a colocation, we need to create a new colocation group record in pg_dist_colocation for the relation. It is not sufficient to have a new colocationid value in pg_dist_partition only. This patch also fixes a bug when deleting a colocation group if no tables are left in it. Previously we passed a relation id as a parameter to DeleteColocationGroupIfNoTablesBelong function, where we should have passed a colocation id. (cherry picked from commit `c22547d221`)	2023-09-05 11:44:22 +03:00
Naisila Puka	db8e12f418	Disable statistics collection (#7162 ) Enabled by mistake in `ba40eb363c` (cherry picked from commit `a17fae36b9`)	2023-08-29 16:13:32 +03:00
Naisila Puka	ff268fb621	Changes PROCESS_TOAST default value to true (#7122 ) Process toast should be true by default, like in PG. (cherry picked from commit `b982f2dee6`)	2023-08-29 16:13:21 +03:00
zhjwpku	0ae05018f1	PQputCopyData's return value 0 should be considered fail (#7152 )	2023-08-29 11:20:25 +02:00
Onur Tirtir	7d24ed0d8b	Makes sure to handle NULL constraints for ADD COLUMN commands (#7093 ) DESCRIPTION: Fixes a bug that causes an unexpected error when adding a column with a NULL constraint Fixes https://github.com/citusdata/citus/issues/7092. (cherry picked from commit `dd6ea1ebd5`)	2023-08-14 10:51:53 +03:00
Halil Ozan Akgül	b729a2b519	Add 11.3-2 backporting changes (#7062 ) This PR moves `citus_shard_sizes` changes from #7003, and #7018 to a new Citus version 11.3-2 This PR backports the changes to 12.0 release branch, there is another PR, #7051 for 11.3 release branch, and one, #7050, for main branch.	2023-07-14 17:19:45 +03:00
aykut-bozkurt	19ef03ab72	Changelog entries for 12.0.0 (#7049 ) Co-authored-by: Onur Tirtir <onurcantirtir@gmail.com> Co-authored-by: Gokhan Gulbiz <ggulbiz@gmail.com> (cherry picked from commit `ee255cd46e`)	2023-07-14 00:22:32 +03:00
aykut-bozkurt	b3e6c036dd	Bump Citus version to 12.0.0	2023-07-14 00:22:25 +03:00
Gürkan İndibay	0f0b60c29c	Fix format attribute and IsLocalReplicationOriginSessionActive errors (#7055 ) This PR fixes the following: - in oraclelinux-7 `Make` step ``` /usr/bin/ld: utils/replication_origin_session_utils.o: relocation R_X86_64_PC32 against undefined symbol `IsLocalReplicationOriginSessionActive' can not be used when making a shared object; recompile with -fPIC /usr/bin/ld: final link failed: Bad value collect2: error: ld returned 1 exit status ``` `IsLocalReplicationOriginSessionActive` function has improper inline declaration, fixed that - in centos-7 `Make` step ``` utils/background_jobs.c: In function 'StartCitusBackgroundTaskExecutor': utils/background_jobs.c:1746:6: warning: function might be possible candidate for 'gnu_printf' format attribute [-Wsuggest-attribute=format] database, user, jobId, taskId); ^ ``` should use `pg_attribute_printf(3,4)` instead of `pg_attribute_printf(3,0)` since the number of arguments varies for `SafeSnprintf(char str, rsize_t count, const char fmt, ...)` --------- Co-authored-by: naisila <nicypp@gmail.com>	2023-07-13 17:41:57 +03:00
aykut-bozkurt	ee255cd46e	Changelog entries for 12.0.0 (#7049 ) Co-authored-by: Onur Tirtir <onurcantirtir@gmail.com> Co-authored-by: Gokhan Gulbiz <ggulbiz@gmail.com>	2023-07-13 14:46:58 +03:00
Onur Tirtir	2c11e4d7f9	Deparse ALTER TABLE commands if ADD COLUMN is the only subcommand (#7032 ) Some clients send ALTER TABLE .. ADD COLUMN .. commands together with some other DDLs and this makes it impossible to directly send the original DDL command to the workers. For this reason, this commit adds support for deparsing such ALTER TABLE commands so that we can avoid from directly sending the original one to the workers. Partially fixes https://github.com/citusdata/citus/issues/690. Fixes #3678	2023-07-12 18:28:45 +03:00
Onur Tirtir	f3cdb6d1bf	Deparse ALTER TABLE commands if ADD COLUMN is the only subcommand And stabilize multi_alter_table_statements.sql.	2023-07-12 18:17:47 +03:00
Onur Tirtir	6365f47b57	Properly handle index storage options for ADD CONSTRAINT / COLUMN	2023-07-11 17:42:43 +03:00
Onur Tirtir	ae142e1764	Properly handle IF NOT EXISTS for ADD COLUMN	2023-07-11 17:42:43 +03:00
Onur Tirtir	d4789a2c3a	Stabilize test helper sql files multi_test_helpers is run in parallel with others, so need to stabilize other test helpers too to make multi_test_helpers runnable multiple times.	2023-07-06 10:47:41 +03:00
Onur Tirtir	001437bdfe	Refactor AppendAlterTableCmdAddConstraint to reuse it for ADD COLUMN too	2023-07-06 10:47:41 +03:00
Onur Tirtir	56f1daa800	Refactor the code that extends constraint/index names on shards into a func	2023-07-06 10:47:41 +03:00
Onur Tirtir	ba1ea9b5bd	Refactor the code that prepares constraint objects in an alter table stmt into a func	2023-07-06 10:47:41 +03:00
Halil Ozan Akgül	613cced1ae	Use citus_shard_sizes in citus_tables (#7018 ) Fixes #7019 This PR updates citus_tables view to use citus_shard_sizes function, instead of citus_total_relation_size to improve performance.	2023-07-05 11:40:34 +03:00
aykut-bozkurt	719d92c8b9	mat view should not be converted to tenant table (#7043 ) We allow materialized view to exist in distrbuted schema but they should not be tried to be converted to a tenant table since they cannot be distributed. Fixes https://github.com/citusdata/citus/issues/7041	2023-07-04 17:28:03 +03:00
Ahmet Gedemenli	5051be86ff	Skip distributed schema insertion into pg_dist_schema, if already exists (#7044 ) Inserting into `pg_dist_schema` causes unexpected duplicate key errors, for distributed schemas that already exist. With this commit we skip the insertion if the schema already exists in `pg_dist_schema`. The error: ```sql SET citus.enable_schema_based_sharding TO ON; CREATE SCHEMA sc2; CREATE SCHEMA IF NOT EXISTS sc2; NOTICE: schema "sc2" already exists, skipping ERROR: duplicate key value violates unique constraint "pg_dist_schema_pkey" DETAIL: Key (schemaid)=(17294) already exists. ``` fixes: #7042	2023-07-04 15:19:07 +03:00
Gokhan Gulbiz	e0d3476526	Add locking mechanism for tenant monitoring probabilistic approach (#7026 ) This PR * Addresses a concurrency issue in the probabilistic approach of tenant monitoring by acquiring a shared lock for tenant existence checks. * Changes `citus.stat_tenants_sample_rate_for_new_tenants` type to double * Renames `citus.stat_tenants_sample_rate_for_new_tenants` to `citus.stat_tenants_untracked_sample_rate`	2023-07-03 13:08:03 +03:00
Jelte Fennema	ac24e11986	Change default rebalance strategy to by_disk_size (#7033 ) DESCRIPTION: Change default rebalance strategy to by_disk_size When introducing rebalancing by disk size we didn't make it the default initially. The main reason was, because we expected some problems with it. We have indeed had some problems/bugs with it over the years, and have fixed all of them. By now we're quite confident in its stability, and that it pretty much always gives better results than by_shard_count. So this PR makes by_disk_size the new default. We don't change the default when some other strategy than by_shard_count is the current default. This is in case someone defined their own rebalance strategy and marked this as the default themselves. Note: It explicitly does nothing during a downgrade, because there's no way of knowing if the rebalance strategy before the upgrade was by_disk_size or by_shard_count. And even in previous versions by_disk_size is considered superior for quite some time.	2023-07-03 11:08:24 +02:00
Jelte Fennema	fd1427de2c	Change by_disk_size rebalance strategy to have a base size (#7035 ) One problem with rebalancing by disk size is that shards in newly created collocation groups are considered extremely small. This can easily result in bad balances if there are some other collocation groups that do have some data. One extremely bad example of this is: 1. You have 2 workers 2. Both contain about 100GB of data, but there's a 70MB difference. 3. You create 100 new distributed schemas with a few empty tables in them 4. You run the rebalancer 5. Now all new distributed schemas are placed on the node with that had 70MB less. 6. You start loading some data in these shards and quickly the balance is completely off To address this edge case, this PR changes the by_disk_size rebalance strategy to add a a base size of 100MB to the actual size of each shard group. This can still result in a bad balance when shard groups are empty, but it solves some of the worst cases.	2023-06-27 16:37:09 +02:00
Halil Ozan Akgül	03a4769c3a	Fix Reference Table Check for CDC (#7025 ) Previously reference table check only looked at `partition method = 'n'`. This PR adds `replication model = 't'` to that.	2023-06-23 16:37:35 +03:00
Teja Mupparti	387b5f80f9	Fixes the bug#6785	2023-06-22 10:44:45 -07:00
Ahmet Gedemenli	99edb2675f	Improve error/hint messages related to schema-based sharding (#7027 ) Improve error/hint messages related to schema-based sharding	2023-06-22 18:10:12 +03:00
Ahmet Gedemenli	44e3c3b9c6	Improve error message for CREATE SCHEMA .. CREATE TABLE (#7024 ) Improve error message for CREATE SCHEMA .. CREATE TABLE when enable_schema_based_sharding is enabled.	2023-06-21 15:24:09 +03:00
aykut-bozkurt	565c5260fd	Properly handle error at owner check (#6984 ) We did not properly handle the error at ownership check method, which causes `max stack depth for errors` as in https://github.com/citusdata/citus/issues/6980. Fix: In case of an error, we should rollback subtransaction and throw the message with log level to `LOG_SERVER_ONLY`. Note: We prevent logs from the client to prevent pg vanilla test failures due to Citus logs which differs from the actual Postgres logs. (For context: https://github.com/citusdata/citus/pull/6130) I also needed to fix a flaky test: `multi_schema_support` DESCRIPTION: Fixes a bug related to non-existent objects in DDL commands. Fixes https://github.com/citusdata/citus/issues/6980	2023-06-21 14:50:01 +03:00
Naisila Puka	69af3e8509	Drop PG13 Support Phase 2 - Remove PG13 specific paths/tests (#7007 ) This commit is the second and last phase of dropping PG13 support. It consists of the following: - Removes all PG_VERSION_13 & PG_VERSION_14 from codepaths - Removes pg_version_compat entries and columnar_version_compat entries specific for PG13 - Removes alternative pg13 test outputs - Removes PG13 normalize lines and fix the test outputs based on that It is a continuation of `5bf163a27d`	2023-06-21 14:18:23 +03:00
aykut-bozkurt	1bb667ce6e	Fix create schema authorization bug (#7015 ) Fixes a bug related to `CREATE SCHEMA AUTHORIZATION <rolename>` for single shard tables. We should properly fetch schema name from role specification if schema name is not given.	2023-06-20 22:05:17 +03:00
aykut-bozkurt	f667f14029	Rewind tuple store to fix scrollable with hold cursor fetches (#7014 ) We need to rewind the tuplestorestate's tuple index to get correct results on fetching scrollable with hold cursors. `PersistHoldablePortal` is responsible for persisting out tuplestorestate inside a with hold cursor before commiting a transaction. It rewinds the cursor like below (`ExecutorRewindcalls` calls `rescan`): ```c if (portal->cursorOptions & CURSOR_OPT_SCROLL) { ExecutorRewind(queryDesc); } ``` At the end, it adjusts tuple index for holdStore in the portal properly. ```c if (portal->cursorOptions & CURSOR_OPT_SCROLL) { if (!tuplestore_skiptuples(portal->holdStore, portal->portalPos, true)) elog(ERROR, "unexpected end of tuple stream"); } ``` DESCRIPTION: Fixes incorrect results on fetching scrollable with hold cursors. Fixes https://github.com/citusdata/citus/issues/7010	2023-06-19 23:00:18 +03:00
Teja Mupparti	58da8771aa	This pull request introduces support for nonroutable merge commands in the following scenarios: 1) For distributed tables that are not colocated. 2) When joining on a non-distribution column for colocated tables. 3) When merging into a distributed table using reference or citus-local tables as the data source. This is accomplished primarily through the implementation of the following two strategies. Repartition: Plan the source query independently, execute the results into intermediate files, and repartition the files to co-locate them with the merge-target table. Subsequently, compile a final merge query on the target table using the intermediate results as the data source. Pull-to-coordinator: Execute the plan that requires evaluation at the coordinator, run the query on the coordinator, and redistribute the resulting rows to ensure colocation with the target shards. Direct the MERGE SQL operation to the worker nodes' target shards, using the intermediate files colocated with the data as the data source.	2023-06-19 12:23:40 -07:00
Xin Li	c10cb50aa9	Support custom cast from / to timestamptz in time partition management UDFs (#6923 ) This is to implement custom cast of table partition column type from / to `timestamptz` in time partition management UDFs, as proposed in ticket #6454 The general idea is for a time partition column with type other than `date`, `timestamp`, or `timestamptz`, users can provide custom bidirectional cast between the column type and `timestamptz`, the UDFs then will be able to create and drop time partitions for such tables. Fixes #6454 --------- Signed-off-by: Xin Li <xin@swirldslabs.com> Co-authored-by: Marco Slot <marco.slot@microsoft.com> Co-authored-by: Ahmet Gedemenli <afgedemenli@gmail.com>	2023-06-19 17:49:05 +03:00
Halil Ozan Akgül	d71ad4b65a	Add Publication Tests for Tenant Schema Tables (#7011 ) This PR adds schema based sharding tests to publication.sql file	2023-06-19 12:39:41 +03:00
aykut-bozkurt	fba5c8dd30	ALTER TABLE <tblname> SET SCHEMA <schemaname> for single shard tables (#7004 ) Adds support for altering schema of single shard tables. We do that in 2 steps. 1. Undistribute the tenant table at `preprocess` step, 2. Distribute new schema if it is a distributed schema after DDLs are propagated. DESCRIPTION: Adds support for altering a table's schema to/from distributed schemas.	2023-06-19 10:21:13 +03:00
Nils Dijk	ce2ba1d07e	Optimize QueryPushdownSqlTaskList on memory and cpu (#6945 ) While going over this piece of code (a long time ago) it was bothering to me we keep a bool array with the size of shardcount to iterate only over shards present in the list of non-pruned shards. Especially since we keep min/max of the set shards to optimize iteration. Postgres has the bitmapset datastructure which a) takes significantly less space, b) has iterator functions to only iterate over set bits, c) can efficiently skip long sequences of unset bits and d) stops quickly once the last set bit has been reached. I have been contemplating if it is worth to keep the minShardOffset because of readability and the efficient skipping of unset bits, however, I have decided to keep it -although less readable-, as there are known usecases where 100k+ shards are pruned to single digit shards. If these would end up at the end of `shardcount` a hotloop of zero checks on the first iteration _could_ cause a theoretical performance regression. All in all, this code is using less memory in all cases where it matters, and less cpu in most cases, while using more idiomatic datastructures for the task at hand.	2023-06-16 16:06:22 +02:00

1 2 3 4 5 ...

6636 Commits (804b521d6331edb2ad0242f77a65edf792b97a3f) All Branches Search

6636 Commits (804b521d6331edb2ad0242f77a65edf792b97a3f)

All Branches