# I have a cluster with 4 nodes , 2 master and 2 data nodes, currently only one master is responding, all the other nodes connect for 10 mins and shard allocation start, then again it becomes NA , not sure what is the issue?

**URL:** <https://discuss.elastic.co/t/i-have-a-cluster-with-4-nodes-2-master-and-2-data-nodes-currently-only-one-master-is-responding-all-the-other-nodes-connect-for-10-mins-and-shard-allocation-start-then-again-it-becomes-na-not-sure-what-is-the-issue/339012>\
**Category:** Kibana\
**Created:** [July 23, 2023, 12:39pm UTC](https://discuss.elastic.co/t/i-have-a-cluster-with-4-nodes-2-master-and-2-data-nodes-currently-only-one-master-is-responding-all-the-other-nodes-connect-for-10-mins-and-shard-allocation-start-then-again-it-becomes-na-not-sure-what-is-the-issue/339012 "2023-07-23T12:39:41Z")\
**Posts on this page:** 1\
**Showing post:** 4

<div class="post-metadata">

**Author:** ![leandrojmp](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/leandrojmp/32/107231_2.png) [@leandrojmp](https://discuss.elastic.co/u/leandrojmp)\
**Post date:** [July 23, 2023, 4:08pm UTC](https://discuss.elastic.co/t/i-have-a-cluster-with-4-nodes-2-master-and-2-data-nodes-currently-only-one-master-is-responding-all-the-other-nodes-connect-for-10-mins-and-shard-allocation-start-then-again-it-becomes-na-not-sure-what-is-the-issue/339012/4 "2023-07-23T16:08:33Z")

</div>

Are both the nodes `Es-Master-1` and `Es-Master-2` running? You need both nodes to be running and you will also need to create a new master node.

There is not much to do, you need to bring the node that is not running back online and after that add a new master node to have resilience.

Also, with the following configuration the `Es-Master-1` node is **not** an dedicated master, it is also a data node. It is the same for `Es-Master-2`? And what does the `elasticsearch.yml` for the `Es-Aggr` nodes looks like?

> [@vishnu\_ishpujani](#):
>
> ```auto
> node.name: ES-Master-1
> node.master: true
> node.data: true
> 
> ```

From the log you shared it seems that the nodes `Es-Master-1` and `Es-Master-2` are both `master` and `data` nodes, but the nodes `Es-Aggr-1` and `Es-Aggr-2` are only data nodes.

> [@vishnu\_ishpujani](#):
>
> `discovery.seed_hosts: ["ES-Master-1", "ES-Master-2", "ES-Aggr-1", "ES-Aggr-2"]`

This setting needs to have **only** the master eligible nodes, remove the `Es-Aggr` nodes if they are data only.

> [@vishnu\_ishpujani](#):
>
> ```auto
> cluster.routing.allocation.node_concurrent_incoming_recoveries: 200
> cluster.routing.allocation.node_concurrent_recoveries: 200
> cluster.routing.allocation.node_initial_primaries_recoveries: 200
> 
> ```

Any reason to have changed those settings? The default value is **2** , this is way too high and can heavily impact on recoveries.

---

_[View the full topic](https://discuss.elastic.co/t/i-have-a-cluster-with-4-nodes-2-master-and-2-data-nodes-currently-only-one-master-is-responding-all-the-other-nodes-connect-for-10-mins-and-shard-allocation-start-then-again-it-becomes-na-not-sure-what-is-the-issue/339012)._
