February 2017
Intermediate to advanced
696 pages
12h 24m
English
Apache Pig (https://pig.apache.org/) is a tool frequently used to store/manipulate data in datastores. It can be very handy if you need to import some CSV in Elasticsearch in a very fast way.
You need an up-and-running Elasticsearch installation as we described in Downloading and installing Elasticsearch recipe in Chapter 2, Downloading and Setup.
You need a working Pig installation. Depending on your operating system you should follow the instruction at http://pig.apache.org/docs/r0.16.0/start.html.
If you are using Mac OS X with Homebrew you can install it with brew install pig.
We want read a CSV and write the data in Elasticsearch. We will perform the steps given as follows:
Read now
Unlock full access