Program to explore a set of 100,000 Wikipedia documents
• Configured latest release of Apache Hadoop in a pseudo-distributed mode.
• Developed a MapReduce-based approach in Hadoop system implemented on instances
• Configured latest release of Apache Hadoop in a pseudo-distributed mode.
• Developed a MapReduce-based approach in Hadoop system implemented on instances of Amazon AWS to compute relative frequencies of each word that occurs in all the do