# About the Spark category

**URL:** <https://discuss.aerospike.com/t/about-the-spark-category/1019>\
**Category:** Spark\
**Created:** [March 6, 2015, 2:53am UTC](https://discuss.aerospike.com/t/about-the-spark-category/1019 "2015-03-06T02:53:57Z")\
**Posts on this page:** 1\
**Page:** 1

<div class="post-metadata">

**Author:** ![Mnemaudsyne](https://sea1.discourse-cdn.com/flex019/user_avatar/discuss.aerospike.com/mnemaudsyne/32/200_2.png) [@Mnemaudsyne](https://discuss.aerospike.com/u/Mnemaudsyne)\
**Post date:** [March 6, 2015, 2:53am UTC](https://discuss.aerospike.com/t/about-the-spark-category/1019/1 "2015-03-06T02:53:57Z")

</div>

[Apache Spark](http://spark.apache.org) is a top-level project of the Apache Foundation. In one deployment configuration, Spark can run in tandem with or instead of [Hadoop](http://hadoop.apache.org/) databases, since it can read from Hadoop Distributed File Systems (HDFS), and can work with [YARN](http://hadoop.apache.org/docs/current/hadoop-yarn/hadoop-yarn-site/YARN.html) (Yet Another Resource Negotiator) or [Apache Mesos](http://mesos.apache.org/). In a second configuration, Spark can run on a standalone mode, either SQL or NoSQL, and integrate with, for instance, Aerospike.

Apache Spark has four main modules:

- Spark Streaming
- Machine Learning (MLlib)
- Spark SQL
- GraphX

Please use this forum to discuss aspects of working with Apache Spark in your architecture, or topics of interest to the Apache Spark community.
