Skip to content

Hadoop

"The Apache Hadoop software library is a framework that allows for the distributed processing of large data sets across clusters of computers using simple programming models." - http://hadoop.apache.org/

  • Oozie - Oozie is a workflow scheduler system to manage Apache Hadoop jobs.