Apache Mahout VS Apache Spark

Compare Apache Mahout vs Apache Spark and see what are their differences.

Apache Mahout

Mirror of Apache Mahout (by apache)

Apache Spark

Apache Spark - A unified analytics engine for large-scale data processing (by apache)
The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

Apache Mahout

Apache Spark

Trino - Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)

Airflow - Apache Airflow - A platform to programmatically author, schedule, and monitor workflows

Pytorch - Tensors and Dynamic neural networks in Python with strong GPU acceleration

Scalding - A Scala API for Cascading

mrjob - Run MapReduce jobs on Hadoop or Amazon Web Services

luigi - Luigi is a Python module that helps you build complex pipelines of batch jobs. It handles dependency resolution, workflow management, visualization etc. It also comes with Hadoop support built in.


Smile - Statistical Machine Intelligence & Learning Engine

Apache Arrow - Apache Arrow is a multi-language toolbox for accelerated data interchange and in-memory processing

Apache Calcite - Apache Calcite

Scio - A Scala API for Apache Beam and Google Cloud Dataflow.

Deeplearning4j - Suite of tools for deploying and training deep learning models using the JVM. Highlights include model import for keras, tensorflow, and onnx/pytorch, a modular and tiny c++ library for running math code and a java based math library on top of the core c++ library. Also includes samediff: a pytorch/tensorflow like library for running deep learning using automatic differentiation.