# Would dividing the same resources over more nodes improve performance?

**URL:** <https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264>\
**Category:** Elasticsearch\
**Created:** [December 13, 2023, 1:16pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264 "2023-12-13T13:16:17Z")\
**Posts on this page:** 7\
**Page:** 1

<div class="post-metadata">

**Author:** ![calin](https://avatars.discourse-cdn.com/v4/letter/c/a87d85/32.png) [@calin](https://discuss.elastic.co/u/calin)\
**Post date:** [December 13, 2023, 1:16pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/1 "2023-12-13T13:16:17Z")

</div>

A load test seems to show that resources (CPU, RAM) on the data nodes aren't fully used. Would spreading the same resources over more data nodes help performance ?

---

<div class="post-metadata">

**Author:** ![Christian\_Dahlqvist](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/christian_dahlqvist/32/4617_2.png) [@Christian\_Dahlqvist](https://discuss.elastic.co/u/Christian_Dahlqvist)\
**Post date:** [December 13, 2023, 2:20pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/2 "2023-12-13T14:20:11Z")

</div>

Elasticsearch performance is not necessarily limited by CPU and/or RAM. The most likely bottleneck is generally disk I/O, but network performance is also a possibility. I have also seen users be unable to saturate clusters as they are not sending data with sufficient level of concurrency.

What is the size and specification of your cluster? What kind of hardware and storage are you using?

---

<div class="post-metadata">

**Author:** ![calin](https://avatars.discourse-cdn.com/v4/letter/c/a87d85/32.png) [@calin](https://discuss.elastic.co/u/calin)\
**Post date:** [December 13, 2023, 2:39pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/3 "2023-12-13T14:39:45Z")

</div>

I'm new in the project and the topic of sizing/hw is new to me, and not sure I can give exact specs, even if I knew which HW and storage there is, which I don't yet.

In terms of nodes:  
10 each master, data and coordinator nodes. Need to add some ingestion nodes.  
They all have 2 CPU/node, master and coordinator have 4 GB each, data has 16 GB/node.

Sorry for the vagueness, I'm sharing what information I have right now.

---

<div class="post-metadata">

**Author:** ![Christian\_Dahlqvist](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/christian_dahlqvist/32/4617_2.png) [@Christian\_Dahlqvist](https://discuss.elastic.co/u/Christian_Dahlqvist)\
**Post date:** [December 13, 2023, 2:46pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/4 "2023-12-13T14:46:04Z")

</div>

That is a very unusual/strange cluster configuration. You generally want to have exactly 3 dedicated master nodes. Having 10 is excessive and as it is an even number also potentially problematic. Dedicated master nodes do little work as they do not serve requests so 2 CPU and 4GB is plenty. Most clusters do not necessarily need any dedicated coordinating only nodes, so I would consider removing most or all of these. Data nodes need to be more powerful as they do almost all work, so should have more CPU resources allocated than the other node types.

---

<div class="post-metadata">

**Author:** ![calin](https://avatars.discourse-cdn.com/v4/letter/c/a87d85/32.png) [@calin](https://discuss.elastic.co/u/calin)\
**Post date:** [December 13, 2023, 3:08pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/5 "2023-12-13T15:08:31Z")

</div>

To give you more details from the few I know 🙂

Expected traffic is about 100 million logs/day, with average size of 1 KB/log. Retention policy of 1 year.

Would 3 masters do for such traffic ?  
Could 5 or 7 or 9 be needed ?

Since I need to add some ingestion nodes (decision already been made to have dedicated coordinator nodes and ingestion nodes), any recommendation what would be a good ratio between the 2 ?

---

<div class="post-metadata">

**Author:** ![Christian\_Dahlqvist](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/christian_dahlqvist/32/4617_2.png) [@Christian\_Dahlqvist](https://discuss.elastic.co/u/Christian_Dahlqvist)\
**Post date:** [December 13, 2023, 4:40pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/6 "2023-12-13T16:40:46Z")

</div>

> [@calin](#):
>
> Would 3 masters do for such traffic ?

Dedicated master nodes are not involved in indexing or querying and you should not send requests to these. They just manage the cluster and their quantity is therefore not dependent on the size or load of the cluster. For large clusters they may require more RAM/heap, but the count do not need to be increased. I have seen very large clusters with just 3 dedicated master nodes.

> [@calin](#):
>
> Since I need to add some ingestion nodes (decision already been made to have dedicated coordinator nodes and ingestion nodes), any recommendation what would be a good ratio between the 2 ?

I would remove the coordinating only nodes you have and allocate more resources to the data nodes. The data nodes would handle requests and also be ingest nodes. Just because you can create nodes with dedicated roles does not mean you should.

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [January 10, 2024, 4:41pm UTC](https://discuss.elastic.co/t/would-dividing-the-same-resources-over-more-nodes-improve-performance/349264/7 "2024-01-10T16:41:34Z")

</div>

This topic was automatically closed 28 days after the last reply. New replies are no longer allowed.
