Teradata Corporation in 1984 marketed the parallel processing DBC 1012 system.

Scientists encounter limitations in e-Science work, including meteorology, genomics, Data sets grow rapidly - in part because they are increasingly gathered by cheap and numerous information-sensing Internet of things devices such as mobile devices, aerial (remote sensing), software logs, cameras, microphones, radio-frequency identification (RFID) readers and wireless sensor networks.

There are three dimensions to big data known as Volume, Variety and Velocity.

Lately, the term "big data" tends to refer to the use of predictive analytics, user behavior analytics, or certain other advanced data analytics methods that extract value from data, and seldom to a particular size of data set.

What counts as "big data" varies depending on the capabilities of the users and their tools, and expanding capabilities make big data a moving target.

"For some organizations, facing hundreds of gigabytes of data for the first time may trigger a need to reconsider data management options.

The framework was very successful, Apache Spark was developed in 2012 in response to limitations in the Map Reduce paradigm, as it adds the ability to set up many operations (not just map followed by reduce).