# Is Spark useful for reindexation?

**URL:** <https://discuss.elastic.co/t/is-spark-useful-for-reindexation/75919>\
**Category:** Elasticsearch\
**Tags:** es-hadoop\
**Created:** [February 21, 2017, 4:36pm UTC](https://discuss.elastic.co/t/is-spark-useful-for-reindexation/75919 "2017-02-21T16:36:24Z")\
**Posts on this page:** 1\
**Showing post:** 2

<div class="post-metadata">

**Author:** ![jspooner](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/jspooner/32/12984_2.png) [@jspooner](https://discuss.elastic.co/u/jspooner)\
**Post date:** [February 21, 2017, 4:57pm UTC](https://discuss.elastic.co/t/is-spark-useful-for-reindexation/75919/2 "2017-02-21T16:57:48Z")

</div>

Three people on my engineering and ops teams have put a lot of effort into tuning bulk indexing and have had limited success. The best indexing rate we can get is ~70k per second. I have this thread to discuss that issue [Spark Bulk Import Performance Benchmarks](https://discuss.elastic.co/t/spark-bulk-import-performance-benchmarks/75110/4)

What version of Elasticsearch are you using?

This article is a little old but might help

> **[How we reindexed 36 billion documents in 5 days within the same Elasticsearch...](https://thoughts.t37.net/how-we-reindexed-36-billions-documents-in-5-days-within-the-same-elasticsearch-cluster-cd9c054d1db8?gi=1386cfc0c8a4)**
>
> At Synthesio, we use ElasticSearch at various places to run complex queries that fetch up to 50 million rich documents out of tens of…

---

_[View the full topic](https://discuss.elastic.co/t/is-spark-useful-for-reindexation/75919)._
