# \#bayesian

**URL:** https://discuss.aerospike.com/tag/bayesian/45.md

[Latest](https://discuss.aerospike.com/latest.md) · [Categories](https://discuss.aerospike.com/categories.md) · [Tags](https://discuss.aerospike.com/tags.md)

---

## [Paging through entirety of large dataset](https://discuss.aerospike.com/t/paging-through-entirety-of-large-dataset/962)

<div class="topic-metadata">

**Author:** [@matthull](https://discuss.aerospike.com/u/matthull)\
**Replies:** 1\
**Last updated:** [February 18, 2015, 4:38am UTC](https://discuss.aerospike.com/t/paging-through-entirety-of-large-dataset/962 "2015-02-18T04:38:16Z")

</div>

I’m working on a machine learning project doing Bayesian classification based on text of tens of millions of documents. I need a high-performance cache summarized documents so I can pull it en masse and pass to the Bayes…
