This document discusses big data, including its exponential growth driven by the internet and cheaper computing. It defines big data based on its volume, velocity, and variety (the 3 Vs). Hadoop and MapReduce are presented as tools for analyzing large and diverse datasets. The challenges of big data analysis are also discussed, such as noise, privacy issues, and infrastructure costs.