# 1 node in an elasticsearch cluster getting stuck for 15 minutes and then starts working

**URL:** <https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137>\
**Category:** Elasticsearch\
**Created:** [June 27, 2024, 7:46am UTC](https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137 "2024-06-27T07:46:54Z")\
**Posts on this page:** 4\
**Page:** 1

<div class="post-metadata">

**Author:** ![anandgopalratnam](https://avatars.discourse-cdn.com/v4/letter/a/ecc23a/32.png) [@anandgopalratnam](https://discuss.elastic.co/u/anandgopalratnam)\
**Post date:** [June 27, 2024, 7:46am UTC](https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137/1 "2024-06-27T07:46:54Z")

</div>

We have seen an issue since the last six months in versions 6.4.2 , 6.8.23 and 7.17.1 where a specific node gets stuck for 15 minutes resulting in timeouts for all calls to that node. We have seen this in TransportClient and HighLevel Rest Client. If anyone has faced such an issue it will be great if you could help.

---

<div class="post-metadata">

**Author:** ![anandgopalratnam](https://avatars.discourse-cdn.com/v4/letter/a/ecc23a/32.png) [@anandgopalratnam](https://discuss.elastic.co/u/anandgopalratnam)\
**Post date:** [June 27, 2024, 2:07pm UTC](https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137/2 "2024-06-27T14:07:02Z")

</div>

Just adding some more info here. The only setting we found in Elasticsearch having 15m timeout is  
indices.recovery.internal\_action\_timeout  
During the issue we checked the indices and there were no unassigned shards or reallocation happening. The Cluster was green.

---

<div class="post-metadata">

**Author:** ![DavidTurner](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/davidturner/32/22453_2.png) [@DavidTurner](https://discuss.elastic.co/u/DavidTurner)\
**Post date:** [June 27, 2024, 7:11pm UTC](https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137/3 "2024-06-27T19:11:19Z")

</div>

These versions are all really old, and the 6.x ones are well past EOL so no point in digging deeper there. In the 7.17 one what does `GET _nodes/hot_threads?threads=9999` say while it's stuck?

---

<div class="post-metadata">

**Author:** ![anandgopalratnam](https://avatars.discourse-cdn.com/v4/letter/a/ecc23a/32.png) [@anandgopalratnam](https://discuss.elastic.co/u/anandgopalratnam)\
**Post date:** [June 27, 2024, 7:34pm UTC](https://discuss.elastic.co/t/1-node-in-an-elasticsearch-cluster-getting-stuck-for-15-minutes-and-then-starts-working/362137/4 "2024-06-27T19:34:13Z")

</div>

Thanks @DavidTurner . I will collect that stats the next time this happens. Probably setup a cron to collect this stats. This happens once in about 15 days in our logging cluster. We will try simulating this in our load environment.
