Big Data is a collection of large and complex data sets that cannot be processed using regular database management tools or processing applications.

Apache Hadoop:

The Apache Hadoop software library is a framework the distributes processing of large data sets across clusters of computers using simple programming models. It is designed to scale up from single servers to thousands of machines, each offering local computation and storage.