Skip to Content
Python Data Analysis Cookbook
book

Python Data Analysis Cookbook

by Ivan Idris
July 2016
Beginner to intermediate
462 pages
9h 14m
English
Packt Publishing
Content preview from Python Data Analysis Cookbook

Docker tips

Docker is a great technology, but we have to be careful not to make our images too big and to remove image files when possible. The docker-clean script at https://gist.github.com/michaelneale/1366325a7737c4cb80b0 (retrieved January 2016) helps reclaim space.

I found it useful to have an install script, which is just a regular shell script, and I added it to the Dockerfile as follows:

ADD install.sh /root/install.sh

Python creates __pycache__ directories for the purpose of optimization (we can disable this option in various ways). These are not strictly needed and can be easily removed as follows:

find /opt/conda -name \__pycache__ -depth -exec rm -rf {} \;

Anaconda puts a lot of files in its pkgs directory, which we can remove as follows: ...

Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Start your free trial

You might also like

Python Machine Learning Cookbook - Second Edition

Python Machine Learning Cookbook - Second Edition

Giuseppe Ciaburro, Prateek Joshi
Python: End-to-end Data Analysis

Python: End-to-end Data Analysis

Phuong Vothihong, Martin Czygan, Ivan Idris, Magnus Vilhelm Persson, Luiz Felipe Martins
Python Data Science Essentials - Third Edition

Python Data Science Essentials - Third Edition

Alberto Boschetti, Luca Massaron, Pietro Marinelli, Matteo Malosetti

Publisher Resources

ISBN: 9781785282287Supplemental Content