Go to file
Dave Brosius 36b287bea4 Merge branch 'cassandra-2.1' into trunk 2014-12-27 20:40:16 -05:00
bin Merge branch 'cassandra-2.1' into trunk 2014-12-17 18:04:33 -06:00
conf Merge branch 'cassandra-2.1' into trunk 2014-12-16 11:17:04 -06:00
debian Merge branch 'cassandra-2.1' into trunk 2014-11-21 15:40:53 -06:00
doc Merge branch 'cassandra-2.1' into trunk 2014-10-29 10:49:20 +01:00
examples Merge branch 'cassandra-2.1' into trunk 2014-12-27 20:40:16 -05:00
interface Merge branch 'cassandra-2.1' into trunk 2014-05-28 13:55:32 -04:00
lib Merge branch 'cassandra-2.1' into trunk 2014-12-18 12:38:47 -06:00
pylib Support indexing key/value entries in map collections 2014-12-18 17:30:26 -06:00
src/java/org/apache/cassandra Merge branch 'cassandra-2.1' into trunk 2014-12-27 20:40:16 -05:00
test Merge branch 'cassandra-2.1' into trunk 2014-12-24 14:05:55 +01:00
tools Merge branch 'cassandra-2.1' into trunk 2014-12-22 09:01:01 -05:00
.gitignore Validate functionality of different JSR-223 providers in UDFs. 2014-11-26 12:51:37 -08:00
.rat-excludes Update versions and licenses for 2.1 RC3 release 2014-07-08 14:07:17 +02:00
CHANGES.txt Merge branch 'cassandra-2.1' into trunk 2014-12-24 14:05:55 +01:00
LICENSE.txt merge with 0.6 branch (post-850) 2010-03-26 16:31:52 +00:00
NEWS.txt Refactor SelectStatement and Restrictions 2014-12-02 13:08:25 -06:00
NOTICE.txt Integrate Sigar library and add basic OS performance checks on startup 2014-10-09 12:32:57 -04:00
README.asc formatting 2014-04-30 12:50:05 -05:00
build.properties.default Followup commit to fix maven deps on boundary's NBHM for CASSANDRA-7128 2014-05-14 09:02:45 -04:00
build.xml fix to run test 2014-11-24 18:57:50 -06:00

README.asc

Executive summary
-----------------

Cassandra is a partitioned row store.  Rows are organized into tables with a required primary key.

http://wiki.apache.org/cassandra/Partitioners[Partitioning] means that Cassandra can distribute your data across multiple machines in an application-transparent matter.  Cassandra will automatically repartition as machines are added and removed from the cluster.

http://wiki.apache.org/cassandra/DataModel[Row store] means that like relational databases, Cassandra organizes data by rows and columns.  The Cassandra Query Language (CQL) is a close relative of SQL.

For more information, see http://cassandra.apache.org/[the Apache Cassandra web site].

Requirements
------------
. Java >= 1.7 (OpenJDK and Oracle JVMS have been tested)
. Python 2.7 (for cqlsh)

Getting started
---------------

This short guide will walk you through getting a basic one node cluster up
and running, and demonstrate some simple reads and writes.

First, we'll unpack our archive:

  $ tar -zxvf apache-cassandra-$VERSION.tar.gz
  $ cd apache-cassandra-$VERSION

and create the log and data directories.  These correspond to the defaults from conf/ and may be adjusted to suit your own environment:

  $ sudo mkdir -p /var/log/cassandra
  $ sudo chown -R `whoami` /var/log/cassandra
  $ sudo mkdir -p /var/lib/cassandra
  $ sudo chown -R `whoami` /var/lib/cassandra

Finally, we start the server.  Running the startup script with the -f argument will cause
Cassandra to remain in the foreground and log to standard out; it can be stopped with ctrl-C.

  $ bin/cassandra -f

****
Note for Windows users: to install Cassandra as a service, download
http://commons.apache.org/daemon/procrun.html[Procrun], set the
PRUNSRV environment variable to the full path of prunsrv (e.g.,
C:\procrun\prunsrv.exe), and run "bin\cassandra.bat install".
Similarly, "uninstall" will remove the service.
****

Now let's try to read and write some data using the Cassandra Query Language:

  $ bin/cqlsh

The command line client is interactive so if everything worked you should
be sitting in front of a prompt:

----
Connected to Test Cluster at localhost:9160.
[cqlsh 2.2.0 | Cassandra 1.2.0 | CQL spec 3.0.0 | Thrift protocol 19.35.0]
Use HELP for help.
cqlsh> 
----

As the banner says, you can use 'help;' or '?' to see what CQL has to
offer, and 'quit;' or 'exit;' when you've had enough fun. But lets try
something slightly more interesting:

----
cqlsh> CREATE SCHEMA schema1 
       WITH replication = { 'class' : 'SimpleStrategy', 'replication_factor' : 1 };
cqlsh> USE schema1;
cqlsh:Schema1> CREATE TABLE users (
                 user_id varchar PRIMARY KEY,
                 first varchar,
                 last varchar,
                 age int
               );
cqlsh:Schema1> INSERT INTO users (user_id, first, last, age) 
               VALUES ('jsmith', 'John', 'Smith', 42);
cqlsh:Schema1> SELECT * FROM users;
 user_id | age | first | last
---------+-----+-------+-------
  jsmith |  42 |  john | smith
 cqlsh:Schema1> 
----

If your session looks similar to what's above, congrats, your single node
cluster is operational! 

For more on what commands are supported by CQL, see
https://github.com/apache/cassandra/blob/trunk/doc/cql3/CQL.textile[the CQL reference].  A
reasonable way to think of it is as, "SQL minus joins and subqueries, plus collections."

Wondering where to go from here? 

  * Getting started: http://wiki.apache.org/cassandra/GettingStarted
  * Join us in #cassandra on irc.freenode.net and ask questions
  * Subscribe to the Users mailing list by sending a mail to
    user-subscribe@cassandra.apache.org
  * Planet Cassandra aggregates Cassandra articles and news:
    http://planetcassandra.org/