# Processing data using spark streaming before indexing in elasticsearch

**URL:** <https://discuss.elastic.co/t/processing-data-using-spark-streaming-before-indexing-in-elasticsearch/64448>\
**Category:** Elasticsearch\
**Tags:** es-hadoop\
**Created:** [October 31, 2016, 1:18pm UTC](https://discuss.elastic.co/t/processing-data-using-spark-streaming-before-indexing-in-elasticsearch/64448 "2016-10-31T13:18:25Z")\
**Posts on this page:** 3\
**Page:** 1

<div class="post-metadata">

**Author:** ![sonia](https://avatars.discourse-cdn.com/v4/letter/s/c67d28/32.png) [@sonia](https://discuss.elastic.co/u/sonia)\
**Post date:** [October 31, 2016, 1:18pm UTC](https://discuss.elastic.co/t/processing-data-using-spark-streaming-before-indexing-in-elasticsearch/64448/1 "2016-10-31T13:18:25Z")

</div>

Hi.  
I work in computer corporation. I have a file containing customer transactions. I would like to use spark streaming to specify that any transaction belongs to which the customer, then each transaction is stored in the index of same customer in elasticsearch.  
please help me.

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [November 1, 2016, 3:30am UTC](https://discuss.elastic.co/t/processing-data-using-spark-streaming-before-indexing-in-elasticsearch/64448/2 "2016-11-01T03:30:56Z")

</div>

Have a look at [https://www.elastic.co/guide/en/elasticsearch/hadoop/current/index.html](https://www.elastic.co/guide/en/elasticsearch/hadoop/current/index.html)

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [July 6, 2017, 1:22pm UTC](https://discuss.elastic.co/t/processing-data-using-spark-streaming-before-indexing-in-elasticsearch/64448/3 "2017-07-06T13:22:45Z")

</div>


