# Storage Best Practices

**URL:** <https://discuss.elastic.co/t/storage-best-practices/91377>\
**Category:** Elastic Training\
**Created:** [June 30, 2017, 7:05am UTC](https://discuss.elastic.co/t/storage-best-practices/91377 "2017-06-30T07:05:16Z")\
**Posts on this page:** 7\
**Page:** 1

<div class="post-metadata">

**Author:** ![SKumarMN](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/skumarmn/32/27537_2.png) [@SKumarMN](https://discuss.elastic.co/u/SKumarMN)\
**Post date:** [June 30, 2017, 7:05am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/1 "2017-06-30T07:05:16Z")

</div>

As part of core operations class, i see a slide with the below details .

path.data vs. RAID0  
‒ RAID0 will be slightly more performant  
‒ path.data will allow a node to continue to function  
‒ e.g. a machine with 4 2TB drives and at most only 25% (2TB) of  
data will need to relocate

Can you explain the above in detail or with some examples so that I can understand it well

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [June 30, 2017, 7:11am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/2 "2017-06-30T07:11:34Z")

</div>

If you point Elasticsearch to multiple `path.data` mount paths and if one of those paths disappears (ie the disk fails) then you lose that single disk (25%).

If you have RAID0 and a disk fails then the entire array and all data on it are lost.

---

<div class="post-metadata">

**Author:** ![Attila\_Nagy](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/attila_nagy/32/46906_2.png) [@Attila\_Nagy](https://discuss.elastic.co/u/Attila_Nagy)\
**Post date:** [June 30, 2017, 7:55am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/3 "2017-06-30T07:55:52Z")

</div>

How does elasticsearch handle the case where IO just freezes to that single disk and the path doesn't disappear (cached entries can be read, some writes may succeed until they are fsynced etc)?

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [June 30, 2017, 8:43am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/4 "2017-06-30T08:43:37Z")

</div>

It'd be better if you created another thread in the #elasticsearch category. We're trying to keep this one to specific questions about our online training 🙂

---

<div class="post-metadata">

**Author:** ![SKumarMN](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/skumarmn/32/27537_2.png) [@SKumarMN](https://discuss.elastic.co/u/SKumarMN)\
**Post date:** [June 30, 2017, 8:54am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/5 "2017-06-30T08:54:42Z")

</div>

Thanks. Say i have a cluster with 3 nodes and a single index with 3 primary and one replica. When we configure multi paths, i believe index is stripped into multi paths but not shards. i.e one complete shard will remain in one path.

Say i have configured multi paths ex path.data : /ds1, /ds2 say one the disks failed(/ds1) in one of the nodes. Now will the shard reallocation happen still i.e make replicas in other nodes primary and create missing replicas( shards that were lost due to disk failure) or does this shard reallocation happen only when node fails.

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [June 30, 2017, 9:09am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/6 "2017-06-30T09:09:14Z")

</div>

> [@SKumarMN](#):
>
> Now will the shard reallocation happen still i.e make replicas in other nodes primary and create missing replicas( shards that were lost due to disk failure) or does this shard reallocation happen only when node fails.

Shard reallocation happens on a shard level, not a node. So if the shard on that bad disk is lost then it will recreate one, it does not wait for the host to also drop off.

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [July 30, 2017, 9:09am UTC](https://discuss.elastic.co/t/storage-best-practices/91377/7 "2017-07-30T09:09:20Z")

</div>

This topic was automatically closed 30 days after the last reply. New replies are no longer allowed.
