Skip to content
Mirror of Apache Mahout
Java Scala Perl6 Shell Batchfile Perl
Failed to load latest commit information.
bin MAHOUT-1797: Typos for SPARK_ASSEMBLY_BIN
buildtools [maven-release-plugin] prepare for next development iteration
conf (NOJIRA) temporary hack to enable (clamp down) on the shell console l…
distribution [maven-release-plugin] prepare for next development iteration
examples MAHOUT-1793: Declare WORK_DIR earlier in example script to fix putput…
h2o [maven-release-plugin] prepare for next development iteration
hdfs [maven-release-plugin] prepare for next development iteration
integration [maven-release-plugin] prepare for next development iteration
math-scala NOJira: Add missing license header
math NOJIRA: minor fixes
mr [maven-release-plugin] prepare for next development iteration
spark-shell [maven-release-plugin] prepare for next development iteration
spark MAHOUT-1785: Replace 'spark.kryoserializer.buffer.mb' from Spark conf…
src MAHOUT-1759: Deprecate Random Forests, this closes apache/mahout#173
.gitignore NOJIRA Added math-scala, spark, spark-shell and h2o modules to binary…
CHANGELOG MAHOUT-1775 FileNotFoundException caused by aborting the process of d…
LICENSE.txt MAHOUT-1684 Update LICENSE and NOTICE texts, this closes apache/mahou…
NOTICE.txt MAHOUT-1684 Update LICENSE and NOTICE texts, this closes apache/mahou…
README.md (NOJIRA) small change to wording in readme.md
doap_Mahout.rdf NoJIRA: Update The DOAP file
pom.xml [maven-release-plugin] prepare for next development iteration

README.md

Welcome to Apache Mahout!

The Apache Mahout™ project's goal is to build an environment for quickly creating scalable performant machine learning applications.

For additional information about Mahout, visit the Mahout Home Page

Setting up your Environment

Whether you are using Mahout's Shell, running command line jobs or using it as a library to build your own apps you'll need to setup several environment variables. Edit your environment in ~/.bash_profile for Mac or ~/.bashrc for many linux distributions. Add the following

export MAHOUT_HOME=/path/to/mahout
export MAHOUT_LOCAL=true # for running standalone on your dev machine, 
# unset MAHOUT_LOCAL for running on a cluster

You will need a $JAVA_HOME, and if you are running on Spark, you will also need $SPARK_HOME

Using Mahout as a Library

Running any application that uses Mahout will require installing a binary or source version and setting the environment. To compile from source:

  • mvn -DskipTests clean install
  • To run tests do mvn test
  • To set up your IDE, do mvn eclipse:eclipse or mvn idea:idea

To use maven, add the appropriate setting to your pom.xml or build.sbt following the template below.

To use the Samsara environment you'll need to include both the engine neutral math-scala dependency:

<dependency>
    <groupId>org.apache.mahout</groupId>
    <artifactId>mahout-math-scala_2.10</artifactId>
    <version>${mahout.version}</version>
</dependency>

and a dependency for back end engine translation, e.g:

<dependency>
    <groupId>org.apache.mahout</groupId>
    <artifactId>mahout-spark_2.10</artifactId>
    <version>${mahout.version}</version>
</dependency>

Examples

For examples of how to use Mahout, see the examples directory located in examples/bin

For information on how to contribute, visit the How to Contribute Page

Legal

Please see the NOTICE.txt included in this directory for more information.

Something went wrong with that request. Please try again.