August 2022
Beginner to intermediate
460 pages
10h 39m
English
MongoDB is often used in conjunction with big data pipelines because of its performance, flexibility, and lack of rigorous data schemas. This chapter will explore the big data landscape, and how MongoDB fits alongside message queuing, data warehousing, and extract, transform, and load (ETL) pipelines.
We will also learn what the MongoDB Atlas Data Lake platform is and how to use this cloud data warehousing offering from MongoDB.
These are the topics that we will discuss in this chapter:
To follow along with the examples in this chapter, we need to install Apache Hadoop and Apache Kafka and connect ...
Read now
Unlock full access