05/01/2015
How Hadoop has Developed as an Enterprise Software?
In the hype surrounding “Big Data” and data analytics, Hadoop gets much attention as it has fast become the technology of choice. However at times little space is given to describing the product itself or analysing the impact it has on the Big Data market and the realities behind the hype.
Hadoop’s initial purpose was parallel processing, however it has developed into a series of projects that do more than just manage large datasets. With these projects, Hadoop can now be used to write complex database queries for real-time BI applications or to write data analysis scripts in a variety of languages.
Examples of organisations that have benefitted from Hadoop include BT (formerly known as British Telecom), Blackberry and The National Cancer Institute in Maryland, US[1]. BT has used Hadoop within its data architecture to reduce costs and for data mining new insights about customers, while Blackberry similarly uses Cloudera’s Hadoop distribution, CDH, within their data architecture. The National Cancer Institute’s Frederick National Laboratory also uses CDH to conduct biostatistical research on relationships between genes and cancers and how this impacts drug trials.
Hadoop Ecosystem
Nowadays when people refer to Hadoop often they mean the Hadoop Ecosystem, i.e. the term given to the network of various projects involved with Hadoop. This includes the standard distribution, commercial distributions, additional projects such as Hive and HBase and alternatives to Hadoop.