Logo

Programming Pig by Alan F Gates

Small book cover: Programming Pig

Programming Pig
by

Publisher: O'Reilly Media
Number of pages: 222

Description:
Apache Pig is a platform for analyzing large data sets that consists of a high-level language for expressing data analysis programs, coupled with infrastructure for evaluating these programs. The salient property of Pig programs is that their structure is amenable to substantial parallelization, which in turns enables them to handle very large data sets.

Download or read it online for free here:
Download link
(6.4MB, PDF)

Similar books

Book cover: Data Wrangling HandbookData Wrangling Handbook
by - School of Data
The Data Wrangling Handbook is a companion text to the School of Data. Its function is something like a traditional textbook -- it will provide the detail and background theory to support the School of Data courses and challenges.
(6597 views)
Book cover: Mastering Apache Spark 2.0Mastering Apache Spark 2.0
by - GitBook
This collections of notes (what some may rashly call a 'book') serves as the ultimate place of mine to collect all the nuts and bolts of using Apache Spark. The notes aim to help me designing and developing better products with Apache Spark.
(4934 views)
Book cover: A Little Riak BookA Little Riak Book
by - GitBook
This is a free little book about Riak, a scalable, high availability NoSQL datastore. Riak is an open-source, distributed key/value database for high availability and near-linear scalability. Riak has remarkably high uptime and grows with you.
(5952 views)
Book cover: CouchDB: The Definitive GuideCouchDB: The Definitive Guide
by - O'Reilly Media
CouchDB's creators show you how to use this document-oriented database as a standalone application framework or with high-volume, distributed applications. CouchDB is ideal for web applications that handle huge amounts of loosely structured data.
(8080 views)