citus

Commit Graph

Author	SHA1	Message	Date
Murat Tuncer	a88d3ecd4e	Add dynamic executor selection - non-router plannable queries can be executed by router executor if they satisfy the criteria - router executor is removed from configuration, now task executor can not be set to router - removed some tests that error out for router executor	2016-04-21 09:15:33 +03:00
Murat Tuncer	938546b938	Add router plannable check and router planning logic for single shard select queries	2016-04-21 09:15:33 +03:00
Brian Cloutier	7b1dc0d511	Support count(distinct) on hash partitioned tables Also add test to ensure we get the same results when running count(distinct) on range and hash partitioned tables.	2016-04-20 04:54:07 -07:00
eren	53186b4e67	FIX Warning Message in multi_logical_optimizer.c With #426, some new warning messages started to arise, because of cross assignment of Node and Expr pointers. This change fixes the warnings with type casts.	2016-04-20 11:33:29 +03:00
eren	448527c3af	Fix JOINs on varchar columns with subquery pushdown Fixes #379 Varchar VAR struct is wrapped in RELABELTYPE struct inside PostgreSQL code and IsPartitionColumnRecursive function considers only VAR types so returning false for varchar. This change adds strip_implicit_coercions() call to the columnExpression in IsPartitionColumnRecursive function so that we get rid of implicit coercions like RELABELTYPE are stripped to VAR.	2016-04-19 21:55:50 -06:00
eren	399b5738b0	Fix Join Problem With VARCHAR Partition Columns This change fixes the problem with joins with VARCHAR columns. Prior to this change, when we tried to do large table joins on varchar columns, we got an error of the form: ERROR: cannot perform local joins that involve expressions DETAIL: local joins can be performed between columns only. This is because we have a check in CheckJoinBetweenColumns() which requires the join clause to have only 'Var' nodes (i.e. columns). Postgres adds a relabel t ype cast to cast the varchar to text; hence the type of the node is not T_Var and the join fails. The fix involves calling strip_implicit_coercions() to the left and right arguments so that RELABELTYPE is stripped to VAR. Fixes #76.	2016-04-19 21:55:50 -06:00
eren	1ffc30d7f5	Fix Shard Pruning Problem With Subqueries on VARCHAR Partition Columns Fixes #375 Prior to this change, shard pruning couldn't be done if: - Table is hash-distributed - Partition column of is VARCHAR - Query to be pruned is a subquery There were two problems: - A bug in left-side/right-side checks for the partition column - We were not considering relabeled types (VARCHAR was relabeled as TEXT)	2016-04-19 21:55:50 -06:00
Metin Doslu	132a77f992	Add COPY support on master node for append partitioned relations	2016-04-19 21:57:59 +03:00
Andres Freund	39233c54ac	Remove wholly unused variable. This avoids a -Wunused warning.	2016-04-19 12:31:13 -06:00
Andres Freund	29b8576a33	Annotate variables only used for asserts with PG_USED_FOR_ASSERTS_ONLY. This avoids '-Wunused-but-set-variable' type warnings when compiling without assertions, e.g. against a system postgres.	2016-04-19 12:31:12 -06:00
Jason Petersen	30fdb59a80	Add clarifying comment in HashableClauseMutator While reading this code last week, it appeared as though there was no place we ensured that the partition clause actually used equality ops. As such, I was worried that we might transform a clause such as id < 5 into a constraint like hash(id) = hash(5) when doing shard pruning. The relevant code seemed to just ensure: 1. The node is an OpExpr 2. With a related hash function 3. It compares the partition column 4. Against a constant A superficial reading implied we didn't actually make sure the original op was equality-related, but it turns out the hash lookup function DOES ensure that for us. So I added a comment.	2016-04-19 12:21:11 -06:00
Murat Tuncer	0b35c47932	Merge pull request #410 from citusdata/350-error-during-duplicate-index-creation Error out earlier when creating an index with a name collision.	2016-04-19 07:26:31 +03:00
Brian Cloutier	1df4b8ba32	Better error on "CREATE INDEX already_exists ..." Previously (if you're creating the index with the same name on different tables) we successfully ran the command on the workers before failing it on the master and leaving no record of the index. Now we check whether the index exists on the master before sending commands to the workers. -- Also make the error better when user attampts to create an index without a name. Previously those statements returned: brian=# create index on c (b); WARNING: could not receive query results from localhost:9700 DETAIL: Client error: cannot extend name for null index name ERROR: could not execute DDL command on worker node shards They now return brian=# create index on c (b); ERROR: creating index without a name on a distributed table is currently unsupported	2016-04-18 13:33:53 -07:00
Jason Petersen	c0b6505720	Merge branch 'master' into 422-incorrect-node-port-type	2016-04-15 12:30:50 -06:00
Brian Cloutier	35bbc1dbfd	Treat nodePort as the 64bit integer it is	2016-04-15 11:29:36 -07:00
Jason Petersen	fb4f84207c	Fix use of INT64CONST macro This macro is intended to receive a bare integer literal (no suffix). It adds a suffix as necessary, depending upon available features. On e.g. 32-bit platforms, the existing code failed to compile because a suffix was added to the existing suffix. This fixes that problem.	2016-04-15 12:13:56 -06:00
Matthew Seaman	b1a3801e58	Regularize include paths for some postgresql headers. Addresses #411	2016-04-15 09:37:22 -07:00
Matthew Seaman	897a1126e7	Include appropriate headers for htons() and htonl().	2016-04-15 09:37:08 -07:00
Matthew Seaman	55e81ce23d	Add sys/stat.h include to files using S_IRUSR and S_IWUSR macros.	2016-04-15 09:34:22 -07:00
eren	3c8f275aa9	Clarify Error Message Related to shared_preload_libraries Fixes #363 This change modifies the error message given when Citus is attempted to be loaded other than shared_preload_libraries. Explanations have been extended with that shared_preload_parameters parameter is in postgresql.conf and citus should be at the beginning.	2016-04-13 12:12:21 +03:00
eren	64aefed46f	Fix SELECT problem with no target list Prior to this change, performing a SELECT query without a target list caused backend to crash. Sample Query: SELECT FROM github_events; (without any * before FROM) PostgreSQL: ``` -- (39599 rows) ``` Citus: ``` server closed the connection unexpectedly This probably means the server terminated abnormally before or while processing the request. The connection to the server was lost. Attempting reset: Failed. !> ``` The problem was an unnecessary Assert on column list in SetRangeTblExtraData(citus_nodefuncs.c)	2016-04-13 11:08:14 +03:00
Metin Doslu	1150ce6414	Send COPY rows in binary format	2016-04-12 20:22:31 +02:00
Marco Slot	d25ee8fbd8	Support for COPY FROM, based on pg_shard PR by Postres Pro	2016-04-12 20:22:31 +02:00
Onder Kalaci	d917d9a615	Allow all types of nodes in the WHERE clauses This change removes the whitelisting check on the WHERE clauses. Note that, before this change, citus was already allowing all types of nodes with the following format (i.e., wrap with a boolean test): * SELECT col FROM table WHERE (ANY EXPRESSION) is TRUE; Thus, this change is mostly useful for allowing the expressions in the WHERE clause directly and avoiding "unsupport clause type" errors.	2016-03-30 16:39:58 +03:00
eren	8a284aab92	Fixes issue #313 Prior to this change, it was not possible to use UDFs in repartitioned subqueries. The reason is that we were setting the search path explicitly and omiting public schema from that path. This change adds the public schema to the explicitly set search path.	2016-03-30 15:39:12 +03:00
eren	ef6d5c7571	Fix spurious NOTICE messages with ANY/ALL Fixes issue #258 Prior to this change, Citus gives a deceptive NOTICE message when a query including ANY or ALL on a non-partition column is issued on a hash partitioned table. Let the github_events table be hash-distributed on repo_id column. Then, issuing this query: SELECT count(*) FROM github_events WHERE event_id = ANY ('{1,2,3}') Gives this message: NOTICE: cannot use shard pruning with ANY (array expression) HINT: Consider rewriting the expression with OR clauses. Note that since event_id is not the partition column, shard pruning would not be applied in any case. However, the NOTICE message would be valid and be given if the ANY clause would have been applied on repo_id column. Reviewer: Murat Tuncer	2016-03-25 14:30:02 +02:00
Murat Tuncer	e74decf8b4	Allow users to create unique indexes Users can now create unique indexes on partition columns for hash and range distributed tables.	2016-03-24 09:31:06 +02:00
Jason Petersen	697d3ea09b	Merge latest 5.0 release fixes	2016-03-23 17:43:34 -06:00
Jason Petersen	1eebb9a6c3	Fix strlcpy off-by-one error WORKER_LENGTH + 1 is too large. Fixing this has no impact on the string that is ultimately copied, as it's impossible for the source string to be any larger to begin with.	2016-03-23 17:34:34 -06:00
Jason Petersen	423e6c8ea0	Update copyright dates Fixed configure variable and updated all end dates to 2016.	2016-03-23 17:14:37 -06:00
Andres Freund	53309461cb	Improve DDL replication related regression tests. The previous form of the test, utilizing DEBUG2, included too much output dependent on the specifc system and version. Reformulate it to explicitly connect to workers and show the schema there, when necessary. The only remaining difference in some of the remaining alternate regression test files was due to an older minor version release change. Remove those as well.	2016-03-17 16:05:54 -07:00
Andres Freund	5311960200	Make :master_port and :worker_$n_port available to all regression tests. There already exist tests that locally embed knowledge about port numbers, and there's more tests requiring that. Instead of copying \set's to several tests, make these port number variables available to all tests.	2016-03-17 16:05:54 -07:00
Andres Freund	b60f5da774	Copy toplevel queryId to citus' master statement. multi_ExecutorStart() replaces the original planned statement with the master select statement. As that hasn't gone through the parse analysis hooks, it'll not have a associated queryId. This prevents extensions pg_stat_statements to show useful data associated with the query.	2016-03-14 17:27:52 -07:00
Jason Petersen	a53fb90ef9	Fix various build issues I came across several places we weren't as flexible or resilient as we should have been in our build logic. They include: * Not using `DESTDIR` in the install-header destination * Allowing callers to specify `VPATH` or `srcdir` (which breaks) * Using absolute path for SCRIPTS (9.5 prepends srcdir) * Including libpq-int in a confusing way (extracted this function) * Having server includes come first during csql build (client must) In particular, I hit all of these attempting to build with pg_buildext in Debian. It passes in an explicit VPATH, as well as srcdir (breaking all recursive make invocations), and also uses DESTDIR during install. In addition, a PGDG-enabled Debian box will have the latest libpq-dev headers (e.g. 9.5) even when building against an older server version (e.g. 9.4). This leads to problems when including e.g. `c.h`, which is ambiguous. While compiling more client-side code (csql), we need to ensure the newer libpq headers are included _first_, so I fixed that.	2016-03-11 13:38:47 -07:00
Jason Petersen	297bd5768d	Final formatting fixes	2016-02-17 17:20:14 -07:00
Andres Freund	974a121d50	Make 'all' default src/backend/distributed target Otherwise typing 'make' will just build citusdb--5.0.sql, not particularly helpful.	2016-02-17 16:51:55 -07:00
Andres Freund	0a1c8723c4	Fix make install for VPATH builds. copy_to_distributed_table is in the source, not the build directory. As there might be scripts in either at some point, install scripts from both.	2016-02-17 16:48:06 -07:00
Marco Slot	71af0e496e	Rename citusdb to citus in regression test output	2016-02-17 23:33:30 +01:00
Marco Slot	75a141a7c6	Merge remote-tracking branch 'origin/master' into feature/drop_shards_on_drop_table	2016-02-17 22:52:58 +01:00
Marco Slot	9aa1f1e1e7	Rename topLevel variable to isTopLevel	2016-02-17 22:52:35 +01:00
Jason Petersen	ec1e74e7f9	Change tests to use default staging policy The default staging policy is now round-robin, though tests were still configured to use local-first. Testing with the shipping default seems like the best option, correctness-wise, and since local-first has some issues with OSes where connecting from localhost doesn't always resolve to 'localhost', just going with the default is a win-win.	2016-02-17 11:03:17 -07:00
Murat Tuncer	df5851366c	Fixed merge leftovers	2016-02-17 15:44:24 +02:00
Murat Tuncer	3528d7ce85	Merge from master branch into feature/citusdb-to-citus	2016-02-17 14:49:01 +02:00
Metin Doslu	6123022ca7	Add check for count distinct on single table subqueries Fixes #314	2016-02-17 14:24:07 +02:00
Murat Tuncer	2160a2951b	Merge pull request #334 from citusdata/feature/append_table_to_shard Add support for appending to cstore table shards	2016-02-17 09:19:33 +02:00
Jason Petersen	8ad5b09251	Merge pull request #344 from citusdata/fix_shard_lock_acquisition#342 Ensure router executor acquires proper shard lock cr: @onderkalaci	2016-02-16 16:43:39 -07:00
Jason Petersen	0d196d1bf4	Ensure router executor acquires proper shard lock Though Citus' Task struct has a shardId field, it doesn't have the same semantics as the one previously used in pg_shard code. The analogous field in the Citus Task is anchorShardId. I've also added an argument check to the relevant locking function to catch future locking attempts which pass an invalid argument.	2016-02-16 11:20:18 -07:00
Marco Slot	37f580f9c7	Trim comment about invalidating dropped relations	2016-02-16 14:04:12 +01:00
Marco Slot	2af6797c04	Perform relcache invalidation in CitusInvalidateRelcacheByRelid	2016-02-16 12:59:38 +01:00
Murat Tuncer	444f305165	Add support for appending to cstore table shards - Flexed the check which prevented append operation cstore tables since its storage type is not SHARD_STORAGE_TABLE. - Used process utility function to perform copy operation in worker_append_table_to shard() instead of directly calling postgresql DoCopy(). - Removed the additional check in master_create_empty_shard() function. This check was redundant and erroneous since it was called after CheckDistributedTable() call. - Modified WorkerTableSize() function to retrieve cstore table shard size correctly.	2016-02-16 13:58:39 +02:00
Marco Slot	52f11223e5	Drop shards when a distributed table is dropped After this change, shards and associated metadata are automatically dropped when running DROP TABLE on a distributed table, which fixes #230. It also adds schema support for master_apply_delete_command, which fixes #73. Dropping the shards happens in the master_drop_all_shards UDF, which is called from the SQL_DROP trigger. Inside the trigger, the table is no longer visible and calling master_apply_delete_command directly wouldn't work and oid <-> name mappings are not available. The master_drop_all_shards function therefore takes the relation id, schema name, and table name as parameters, which can be obtained from pg_event_trigger_dropped_objects() in the SQL_DROP trigger. If the user calls master_drop_all_shards while the table still exists, the schema name and table name are ignored. Author: Marco Slot Reviewed-By: Andres Freund	2016-02-16 10:54:29 +01:00
Jason Petersen	920e0c406d	Format csql's stage files These are entirely Citus-produced, so need full formatting.	2016-02-15 23:37:37 -07:00
Jason Petersen	19c529f311	Omit most of copy_options from formatting Only a small portion is Citus style.	2016-02-15 23:37:37 -07:00
Jason Petersen	2b5ae847d4	Make copy_options.ch similar to PostgreSQL copy.c We reorganized these functions in our copy; not sure why (makes diffing harder). I'm moving it back.	2016-02-15 23:37:37 -07:00
Jason Petersen	bc23113732	Omit backend/copy.c-inspired parts from formatting I think we need to assess whether this function is still as in-sync with upstream as we believe, but for now I'm omitting it from formatting.	2016-02-15 23:29:33 -07:00
Jason Petersen	74372f70e0	Omit get_extension_schema from formatting It exactly matches the implementation in extension.c.	2016-02-15 23:29:33 -07:00
Jason Petersen	f874a56e24	Omit RangeVarCallbackForDropIndex from formatting I removed two braces to have this function remain more similar to the original PostgreSQL function and added uncrustify commands to disable formatting of its contents.	2016-02-15 23:29:33 -07:00
Jason Petersen	fdb37682b2	First formatting attempt Skipped csql, ruleutils, readfuncs, and functions obviously copied from PostgreSQL. Seeing how this looks, then continuing.	2016-02-15 23:29:32 -07:00
Murat Tuncer	55c44b48dd	Changed product name to citus All citusdb references in - extension, binary names - file headers - all configuration name prefixes - error/warning messages - some functions names - regression tests are changed to be citus.	2016-02-15 16:04:31 +02:00
Jason Petersen	334f800016	Merge pull request #329 from citusdata/feature-fix_naming_conflicts#236 Rename GetConnection to address name conflict cr: @onderkalaci	2016-02-12 16:58:33 -07:00
Jason Petersen	4494e57bbd	Rename GetConnection to address name conflict The postgres_fdw extension has an extern function with an identical signature, which can cause problems when both extensions are loaded. A simple rename can fix this for now (this is the only function with) such a conflict.	2016-02-12 13:35:02 -07:00
Önder Kalacı	a55287411b	Merge pull request #332 from citusdata/bugfix/memory_context_leak Remove unnecessary memory context switch on the planner	2016-02-12 11:13:12 -08:00
Onder Kalaci	0a6839e544	Perform distributed planning in the calling memory context Previously we used, for historical reasons, MessageContext. That is problematic if a single message from the client causes a lot of statements to be planned. E.g. for the copy_to_distributed_table script one insert statement is planned for each row inserted via COPY, and only freed when COPY has finished.	2016-02-12 20:50:40 +02:00
Jason Petersen	b1ef2e59a2	Merge pull request #331 from citusdata/feature-permit_dml_to_append_tables#321 Allow DML commands on append-partitioned tables cr: @lithp	2016-02-12 11:24:06 -07:00
Jason Petersen	6f308c5e2d	Allow DML commands on append-partitioned tables This entirely removes any restriction on the type of partitioning during DML planning and execution. Though there aren't actually any technical limitations preventing DML commands against append- (or even range-) partitioned tables, we had initially forbidden this, as any future stage operation could cause shards to overlap, banning all subsequent DML operations to partition values contained within more than one shards. This ended up mostly restricting us, so we're now removing that restriction.	2016-02-11 16:09:35 -07:00
Jason Petersen	d164305929	Handle hash-partitioned aliased data types When two data types have the same binary representation, PostgreSQL may add an implicit coercion between them by wrapping a node in a relabel type. This wrapper signals that the wrapped value is completely binary compatible with the designated "final type" of the relabel node. As an example, the varchar type is often relabeled to text, since functions provided for use with text (comparisons, hashes, etc.) are completely compatible with varchar as well. The hash-partitioned codepath contains functions that verify queries actually contain an equality constraint on the partition column, but those functions expect such constraints to be comparison operations between a Var and Const. The RelabelType wrapper node causes these functions to always return false, which bypasses shard pruning.	2016-02-11 13:50:43 -07:00
Onder Kalaci	136306a1fe	Initial commit of Citus 5.0	2016-02-11 04:05:32 +02:00

... 44 45 46 47 48

2367 Commits (test/adaptive_executor_repartition)