citus

Distributed PostgreSQL as an extension

citus citus-extension database database-cluster distributed-database multi-tenant postgres postgresql relational-database scale sharding sql

Go to file

Onder Kalaci d657759c97 Views to Provide some insight about the distributed transactions on Citus MX With this commit, we implement two views that are very similar to pg_stat_activity, but showing queries that are involved in distributed queries: - citus_dist_stat_activity: Shows all the distributed queries - citus_worker_stat_activity: Shows all the queries on the shards that are initiated by distributed queries. Both views have the same columns in the outputs. In very basic terms, both of the views are meant to provide some useful insights about the distributed transactions within the cluster. As the names reveal, both views are similar to pg_stat_activity. Also note that these views can be pretty useful on Citus MX clusters. Note that when the views are queried from the worker nodes, they'd not show the distributed transactions that are initiated from the coordinator node. The reason is that the worker nodes do not know the host/port of the coordinator. Thus, it is advisable to query the views from the coordinator. If we bucket the columns that the views returns, we'd end up with the following: - Hostnames and ports: - query_hostname, query_hostport: The node that the query is running - master_query_host_name, master_query_host_port: The node in the cluster initiated the query. Note that for citus_dist_stat_activity view, the query_hostname-query_hostport is always the same with master_query_host_name-master_query_host_port. The distinction is mostly relevant for citus_worker_stat_activity. For example, on Citus MX, a users starts a transaction on Node-A, which starts worker transactions on Node-B and Node-C. In that case, the query hostnames would be Node-B and Node-C whereas the master_query_host_name would Node-A. - Distributed transaction related things: This is mostly the process_id, distributed transactionId and distributed transaction number. - pg_stat_activity columns: These two views get all the columns from pg_stat_activity. We're basically joining pg_stat_activity with get_all_active_transactions on process_id.		2018-09-10 21:33:27 +03:00
config	Add citus_version(), analogous to PG's version()	2017-10-16 18:09:29 -06:00
src	Views to Provide some insight about the distributed transactions on Citus MX	2018-09-10 21:33:27 +03:00
windows	Bump citus version to 8.0devel	2018-07-25 12:03:47 +03:00
.codecov.yml	Remove obsolete lines	2017-09-25 11:18:25 -07:00
.editorconfig	Set tab size for GitHub display	2017-03-22 13:03:39 -06:00
.gitattributes	Add ruleutils file for PostgreSQL 11	2017-09-25 17:20:24 -07:00
.gitignore	Add vim swap files to .gitignore	2017-07-12 14:16:23 +02:00
.travis.yml	Temporarily allow PG11 failures	2018-09-06 18:33:43 +03:00
CHANGELOG.md	Add changelog entry for 7.5.1	2018-08-28 13:34:37 +03:00
CONTRIBUTING.md	Two more libs I needed to build citus	2017-08-24 13:04:35 -06:00
LICENSE	Add AGPL-3.0 in LICENSE file	2016-03-23 17:04:58 -06:00
Makefile	Views to Provide some insight about the distributed transactions on Citus MX	2018-09-10 21:33:27 +03:00
Makefile.global.in	Basic usage statistics collection. (#1656 )	2017-10-11 09:55:15 -04:00
README.md	Update README.md	2017-03-23 11:00:32 -07:00
aclocal.m4	Basic usage statistics collection. (#1656 )	2017-10-11 09:55:15 -04:00
appveyor.yml	Configure appveyor to run regression tests	2018-04-25 18:02:07 -07:00
autogen.sh	Changed product name to citus	2016-02-15 16:04:31 +02:00
citus.control	Views to Provide some insight about the distributed transactions on Citus MX	2018-09-10 21:33:27 +03:00
configure	Bump citus version to 8.0devel	2018-07-25 12:03:47 +03:00
configure.in	Bump citus version to 8.0devel	2018-07-25 12:03:47 +03:00
github-banner.png	Readme for 5.0	2016-03-18 13:32:13 -07:00
prep_buildtree	Changed product name to citus	2016-02-15 16:04:31 +02:00

README.md

What is Citus?

Open-source PostgreSQL extension (not a fork)
Scalable across multiple machines through sharding and replication
Distributed engine for query parallelization
Database designed to scale multi-tenant applications

Citus is a distributed database that scales across commodity servers using transparent sharding and replication. Citus extends the underlying database rather than forking it, giving developers and enterprises the power and familiarity of a relational database. As an extension, Citus supports new PostgreSQL releases, and allows you to benefit from new features while maintaining compatibility with existing PostgreSQL tools.

Citus serves many use cases. Two common ones are:

Multi-tenant database: Most B2B applications already have the notion of a tenant / customer / account built into their data model. Citus allows you to scale out your transactional relational database to 100K+ tenants with minimal changes to your application.
Real-time analytics: Citus enables ingesting large volumes of data and running analytical queries on that data in human real-time. Example applications include analytic dashboards with subsecond response times and exploratory queries on unfolding events.

To learn more, visit citusdata.com and join the mailing list to stay on top of the latest developments.

Getting started with Citus

The fastest way to get up and running is to create a Citus Cloud account. You can also setup a local Citus cluster with Docker.

Citus Cloud

Citus Cloud runs on top of AWS as a fully managed database as a service and has development plans available for getting started. You can provision a Citus Cloud account at https://console.citusdata.com and get started with just a few clicks.

Local Citus Cluster

If you're looking to get started locally, you can follow the following steps to get up and running.

Install Docker Community Edition and Docker Compose

Mac:
1. Download and install Docker.
2. Start Docker by clicking on the application’s icon.

Linux:

curl -sSL https://get.docker.com/ | sh
sudo usermod -aG docker $USER && exec sg docker newgrp `id -gn`
sudo systemctl start docker

sudo curl -sSL https://github.com/docker/compose/releases/download/1.11.2/docker-compose-`uname -s`-`uname -m` -o /usr/local/bin/docker-compose
sudo chmod +x /usr/local/bin/docker-compose

The above version of Docker Compose is sufficient for running Citus, or you can install the latest version.

Pull and start the Docker images

curl -sSLO https://raw.githubusercontent.com/citusdata/docker/master/docker-compose.yml
docker-compose -p citus up -d

Connect to the master database

docker exec -it citus_master psql -U postgres

Follow the first tutorial instructions
To shut the cluster down, run

docker-compose -p citus down

Talk to Contributors and Learn More

Documentation	Try the Citus tutorial for a hands-on introduction or the documentation for a more comprehensive reference.
Google Groups	The Citus Google Group is our place for detailed questions and discussions.
Slack	Chat with us in our community Slack channel.
Github Issues	We track specific bug reports and feature requests on our project issues.
Twitter	Follow @citusdata for general updates and PostgreSQL scaling tips.

Contributing

Citus is built on and of open source, and we welcome your contributions. The CONTRIBUTING.md file explains how to get started developing the Citus extension itself and our code quality guidelines.

Who is Using Citus?

Citus is deployed in production by many customers, ranging from technology start-ups to large enterprises. Here are some examples:

CloudFlare uses Citus to provide real-time analytics on 100 TBs of data from over 4 million customer websites. Case Study
MixRank uses Citus to efficiently collect and analyze vast amounts of data to allow inside B2B sales teams to find new customers. Case Study
Neustar builds and maintains scalable ad-tech infrastructure that counts billions of events per day using Citus and HyperLogLog.
Agari uses Citus to secure more than 85 percent of U.S. consumer emails on two 6-8 TB clusters. Case Study
Heap uses Citus to run dynamic funnel, segmentation, and cohort queries across billions of users and tens of billions of events. Watch Video

README.md Unescape Escape