# Strange issue after upgrading from 1.1.0 to 1.4.1 ES Version

**URL:** <https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397>\
**Category:** Elasticsearch\
**Created:** [February 26, 2015, 1:18am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397 "2015-02-26T01:18:48Z")\
**Posts on this page:** 9\
**Page:** 1

<div class="post-metadata">

**Author:** ![sagarit2](https://avatars.discourse-cdn.com/v4/letter/s/f0a364/32.png) [@sagarit2](https://discuss.elastic.co/u/sagarit2)\
**Post date:** [February 26, 2015, 1:18am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/1 "2015-02-26T01:18:48Z")

</div>

Hi,

We recently upgraded one of our ES Clusters from ES Version 1.1.0 to 1.4.1.

We have dedicated master-data-search deployment in AWS. Cluster settings  
are same for all the clusters.

Strangely, only in one cluster; we are seeing that nodes are constantly  
failing to connect to Master node and rejoining back.  
_It happens all the time, even during idle period (when there are no read  
or writes)_.

We keep on seeing following exception in the logs  
org.elasticsearch.transport.NodeNotConnectedException

Because of this, Cluster has slowed down considerably.

We use kopf plugin for monitoring and it keeps popping up message -  
"Loading cluster information is talking too long"

There is not much data on individual nodes; almost 80% disk is free. CPU  
and Heap are doing fine.

Only difference between this cluster and other cluster is the number of  
indices and shards. Other clusters have shards in hundreds and indices in  
double digit. While this cluster has around 5000 shards and close to 250  
indices.

But we are not sure, if number of shards or indices can cause reconnection  
issues between nodes.

Not sure if it's really related to 1.4.1 or something else. But in that  
case, other clusters should have been affected too.

Any help will be appreciated !

Thanks,

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![Sean\_Clemmer](https://avatars.discourse-cdn.com/v4/letter/s/77aa72/32.png) [@Sean\_Clemmer](https://discuss.elastic.co/u/Sean_Clemmer)\
**Post date:** [February 26, 2015, 6:35am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/2 "2015-02-26T06:35:49Z")

</div>

+1

I have a similar story: After around six months using the v1.3.x series, I  
upgraded from v1.3.4 to v1.4.4. I've had monitoring and metrics in place  
for a while now, and compared to baseline I'm seeing occasional periods  
where nodes appear to drop out and usually back into the cluster (with  
NodeNotConnectedException). During these periods the cluster status is red  
on-and-off for maybe 15-30 minutes with anywhere from 30 minutes to 2 hours  
in-between. The issue is worse in larger clusters with larger shard counts  
(thousands, tens of thousands).

Resource utilization is still good. The number of shards (and the amount of  
data) is essentially constant. I'm confident the upgrade was the only  
change; I have strict controls on the clusters.

On Wed, Feb 25, 2015 at 5:18 PM, sagarl [sagarit2@gmail.com](mailto:sagarit2@gmail.com) wrote:

> Hi,
> 
> We recently upgraded one of our ES Clusters from ES Version 1.1.0 to 1.4.1.
> 
> We have dedicated master-data-search deployment in AWS. Cluster settings  
> are same for all the clusters.
> 
> Strangely, only in one cluster; we are seeing that nodes are constantly  
> failing to connect to Master node and rejoining back.  
> _It happens all the time, even during idle period (when there are no read  
> or writes)_.
> 
> We keep on seeing following exception in the logs  
> org.elasticsearch.transport.NodeNotConnectedException
> 
> Because of this, Cluster has slowed down considerably.
> 
> We use kopf plugin for monitoring and it keeps popping up message -  
> "Loading cluster information is talking too long"
> 
> There is not much data on individual nodes; almost 80% disk is free. CPU  
> and Heap are doing fine.
> 
> Only difference between this cluster and other cluster is the number of  
> indices and shards. Other clusters have shards in hundreds and indices in  
> double digit. While this cluster has around 5000 shards and close to 250  
> indices.
> 
> But we are not sure, if number of shards or indices can cause reconnection  
> issues between nodes.
> 
> Not sure if it's really related to 1.4.1 or something else. But in that  
> case, other clusters should have been affected too.
> 
> Any help will be appreciated !
> 
> Thanks,
> 
> --  
> You received this message because you are subscribed to the Google Groups  
> "elasticsearch" group.  
> To unsubscribe from this group and stop receiving emails from it, send an  
> email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> To view this discussion on the web visit  
> [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com)  
> [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm_medium=email&utm_source=footer)  
> .  
> For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/CADa-AwfBqx%3D2AGfgwj%3DZjLHMyCtKUa\_tAU\_Ffdy-MbQ2Ve5tBQ%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CADa-AwfBqx%3D2AGfgwj%3DZjLHMyCtKUa_tAU_Ffdy-MbQ2Ve5tBQ%40mail.gmail.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![emkt84](https://avatars.discourse-cdn.com/v4/letter/e/7ba0ec/32.png) [@emkt84](https://discuss.elastic.co/u/emkt84)\
**Post date:** [February 26, 2015, 11:10pm UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/3 "2015-02-26T23:10:58Z")

</div>

We have seen similar issue in our setup too, but we are running 1.3.6. I  
think it occurs with large index and shards counts. We have approx 12000  
shards in total. I think its a bug in elasticsearch, as having large shard  
count does not mean that node just simply drops out from cluster. The shard  
count stats and cluster membership threads should not depend upon each  
other.

On Thursday, February 26, 2015 at 7:36:20 AM UTC+1, Sean Clemmer wrote:

> +1
> 
> I have a similar story: After around six months using the v1.3.x series, I  
> upgraded from v1.3.4 to v1.4.4. I've had monitoring and metrics in place  
> for a while now, and compared to baseline I'm seeing occasional periods  
> where nodes appear to drop out and usually back into the cluster (with  
> NodeNotConnectedException). During these periods the cluster status is  
> red on-and-off for maybe 15-30 minutes with anywhere from 30 minutes to 2  
> hours in-between. The issue is worse in larger clusters with larger shard  
> counts (thousands, tens of thousands).
> 
> Resource utilization is still good. The number of shards (and the amount  
> of data) is essentially constant. I'm confident the upgrade was the only  
> change; I have strict controls on the clusters.
> 
> On Wed, Feb 25, 2015 at 5:18 PM, sagarl \<[saga...@gmail.com](mailto:saga...@gmail.com) \<javascript:\>\>  
> wrote:
> 
> > Hi,
> > 
> > We recently upgraded one of our ES Clusters from ES Version 1.1.0 to  
> > 1.4.1.
> > 
> > We have dedicated master-data-search deployment in AWS. Cluster settings  
> > are same for all the clusters.
> > 
> > Strangely, only in one cluster; we are seeing that nodes are constantly  
> > failing to connect to Master node and rejoining back.  
> > _It happens all the time, even during idle period (when there are no read  
> > or writes)_.
> > 
> > We keep on seeing following exception in the logs  
> > org.elasticsearch.transport.NodeNotConnectedException
> > 
> > Because of this, Cluster has slowed down considerably.
> > 
> > We use kopf plugin for monitoring and it keeps popping up message -  
> > "Loading cluster information is talking too long"
> > 
> > There is not much data on individual nodes; almost 80% disk is free. CPU  
> > and Heap are doing fine.
> > 
> > Only difference between this cluster and other cluster is the number of  
> > indices and shards. Other clusters have shards in hundreds and indices in  
> > double digit. While this cluster has around 5000 shards and close to 250  
> > indices.
> > 
> > But we are not sure, if number of shards or indices can cause  
> > reconnection issues between nodes.
> > 
> > Not sure if it's really related to 1.4.1 or something else. But in that  
> > case, other clusters should have been affected too.
> > 
> > Any help will be appreciated !
> > 
> > Thanks,
> > 
> > --  
> > You received this message because you are subscribed to the Google Groups  
> > "elasticsearch" group.  
> > To unsubscribe from this group and stop receiving emails from it, send an  
> > email to [elasticsearc...@googlegroups.com](mailto:elasticsearc...@googlegroups.com) \<javascript:\>.  
> > To view this discussion on the web visit  
> > [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com)  
> > [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > .  
> > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [February 26, 2015, 11:45pm UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/4 "2015-02-26T23:45:32Z")

</div>

12000 shards across how many nodes?

Don't forget that a shard is a lucene instance, it needs resources to  
operate and a node only has so many resources. This is why scaling is  
important.

On 27 February 2015 at 10:10, [emkt84@gmail.com](mailto:emkt84@gmail.com) wrote:

> We have seen similar issue in our setup too, but we are running 1.3.6. I  
> think it occurs with large index and shards counts. We have approx 12000  
> shards in total. I think its a bug in elasticsearch, as having large shard  
> count does not mean that node just simply drops out from cluster. The shard  
> count stats and cluster membership threads should not depend upon each  
> other.
> 
> On Thursday, February 26, 2015 at 7:36:20 AM UTC+1, Sean Clemmer wrote:
> 
> > +1
> > 
> > I have a similar story: After around six months using the v1.3.x series,  
> > I upgraded from v1.3.4 to v1.4.4. I've had monitoring and metrics in place  
> > for a while now, and compared to baseline I'm seeing occasional periods  
> > where nodes appear to drop out and usually back into the cluster (with  
> > NodeNotConnectedException). During these periods the cluster status is  
> > red on-and-off for maybe 15-30 minutes with anywhere from 30 minutes to 2  
> > hours in-between. The issue is worse in larger clusters with larger shard  
> > counts (thousands, tens of thousands).
> > 
> > Resource utilization is still good. The number of shards (and the amount  
> > of data) is essentially constant. I'm confident the upgrade was the only  
> > change; I have strict controls on the clusters.
> > 
> > On Wed, Feb 25, 2015 at 5:18 PM, sagarl [saga...@gmail.com](mailto:saga...@gmail.com) wrote:
> > 
> > > Hi,
> > > 
> > > We recently upgraded one of our ES Clusters from ES Version 1.1.0 to  
> > > 1.4.1.
> > > 
> > > We have dedicated master-data-search deployment in AWS. Cluster settings  
> > > are same for all the clusters.
> > > 
> > > Strangely, only in one cluster; we are seeing that nodes are constantly  
> > > failing to connect to Master node and rejoining back.  
> > > _It happens all the time, even during idle period (when there are no  
> > > read or writes)_.
> > > 
> > > We keep on seeing following exception in the logs  
> > > org.elasticsearch.transport.NodeNotConnectedException
> > > 
> > > Because of this, Cluster has slowed down considerably.
> > > 
> > > We use kopf plugin for monitoring and it keeps popping up message -  
> > > "Loading cluster information is talking too long"
> > > 
> > > There is not much data on individual nodes; almost 80% disk is free. CPU  
> > > and Heap are doing fine.
> > > 
> > > Only difference between this cluster and other cluster is the number of  
> > > indices and shards. Other clusters have shards in hundreds and indices in  
> > > double digit. While this cluster has around 5000 shards and close to 250  
> > > indices.
> > > 
> > > But we are not sure, if number of shards or indices can cause  
> > > reconnection issues between nodes.
> > > 
> > > Not sure if it's really related to 1.4.1 or something else. But in that  
> > > case, other clusters should have been affected too.
> > > 
> > > Any help will be appreciated !
> > > 
> > > Thanks,
> > > 
> > > --  
> > > You received this message because you are subscribed to the Google  
> > > Groups "elasticsearch" group.  
> > > To unsubscribe from this group and stop receiving emails from it, send  
> > > an email to [elasticsearc...@googlegroups.com](mailto:elasticsearc...@googlegroups.com).  
> > > To view this discussion on the web visit [https://groups.google.com/d/](https://groups.google.com/d/)  
> > > msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%  
> > > [40googlegroups.com](http://40googlegroups.com)  
> > > [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > > .  
> > > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> > 
> > --  
> > You received this message because you are subscribed to the Google Groups  
> > "elasticsearch" group.  
> > To unsubscribe from this group and stop receiving emails from it, send an  
> > email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> > To view this discussion on the web visit  
> > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com)  
> > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > .
> 
> For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![emkt84](https://avatars.discourse-cdn.com/v4/letter/e/7ba0ec/32.png) [@emkt84](https://discuss.elastic.co/u/emkt84)\
**Post date:** [February 27, 2015, 7:31am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/5 "2015-02-27T07:31:03Z")

</div>

We have 10 data nodes which is storing data and separate master and client  
nodes. So each node is having approx 1200 shards and 110 indices which I  
think is not much.

On Fri, Feb 27, 2015 at 12:45 AM, Mark Walkom [markwalkom@gmail.com](mailto:markwalkom@gmail.com) wrote:

> 12000 shards across how many nodes?
> 
> Don't forget that a shard is a lucene instance, it needs resources to  
> operate and a node only has so many resources. This is why scaling is  
> important.
> 
> On 27 February 2015 at 10:10, [emkt84@gmail.com](mailto:emkt84@gmail.com) wrote:
> 
> > We have seen similar issue in our setup too, but we are running 1.3.6. I  
> > think it occurs with large index and shards counts. We have approx 12000  
> > shards in total. I think its a bug in elasticsearch, as having large shard  
> > count does not mean that node just simply drops out from cluster. The shard  
> > count stats and cluster membership threads should not depend upon each  
> > other.
> > 
> > On Thursday, February 26, 2015 at 7:36:20 AM UTC+1, Sean Clemmer wrote:
> > 
> > > +1
> > > 
> > > I have a similar story: After around six months using the v1.3.x series,  
> > > I upgraded from v1.3.4 to v1.4.4. I've had monitoring and metrics in place  
> > > for a while now, and compared to baseline I'm seeing occasional periods  
> > > where nodes appear to drop out and usually back into the cluster (with  
> > > NodeNotConnectedException). During these periods the cluster status is  
> > > red on-and-off for maybe 15-30 minutes with anywhere from 30 minutes to 2  
> > > hours in-between. The issue is worse in larger clusters with larger shard  
> > > counts (thousands, tens of thousands).
> > > 
> > > Resource utilization is still good. The number of shards (and the amount  
> > > of data) is essentially constant. I'm confident the upgrade was the  
> > > only change; I have strict controls on the clusters.
> > > 
> > > On Wed, Feb 25, 2015 at 5:18 PM, sagarl [saga...@gmail.com](mailto:saga...@gmail.com) wrote:
> > > 
> > > > Hi,
> > > > 
> > > > We recently upgraded one of our ES Clusters from ES Version 1.1.0 to  
> > > > 1.4.1.
> > > > 
> > > > We have dedicated master-data-search deployment in AWS. Cluster  
> > > > settings are same for all the clusters.
> > > > 
> > > > Strangely, only in one cluster; we are seeing that nodes are constantly  
> > > > failing to connect to Master node and rejoining back.  
> > > > _It happens all the time, even during idle period (when there are no  
> > > > read or writes)_.
> > > > 
> > > > We keep on seeing following exception in the logs  
> > > > org.elasticsearch.transport.NodeNotConnectedException
> > > > 
> > > > Because of this, Cluster has slowed down considerably.
> > > > 
> > > > We use kopf plugin for monitoring and it keeps popping up message -  
> > > > "Loading cluster information is talking too long"
> > > > 
> > > > There is not much data on individual nodes; almost 80% disk is free.  
> > > > CPU and Heap are doing fine.
> > > > 
> > > > Only difference between this cluster and other cluster is the number of  
> > > > indices and shards. Other clusters have shards in hundreds and indices in  
> > > > double digit. While this cluster has around 5000 shards and close to 250  
> > > > indices.
> > > > 
> > > > But we are not sure, if number of shards or indices can cause  
> > > > reconnection issues between nodes.
> > > > 
> > > > Not sure if it's really related to 1.4.1 or something else. But in that  
> > > > case, other clusters should have been affected too.
> > > > 
> > > > Any help will be appreciated !
> > > > 
> > > > Thanks,
> > > > 
> > > > --  
> > > > You received this message because you are subscribed to the Google  
> > > > Groups "elasticsearch" group.  
> > > > To unsubscribe from this group and stop receiving emails from it, send  
> > > > an email to [elasticsearc...@googlegroups.com](mailto:elasticsearc...@googlegroups.com).  
> > > > To view this discussion on the web visit [https://groups.google.com/d/](https://groups.google.com/d/)  
> > > > msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%  
> > > > [40googlegroups.com](http://40googlegroups.com)  
> > > > [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > > > .  
> > > > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> > > 
> > > --  
> > > You received this message because you are subscribed to the Google Groups  
> > > "elasticsearch" group.  
> > > To unsubscribe from this group and stop receiving emails from it, send an  
> > > email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> > > To view this discussion on the web visit  
> > > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com)  
> > > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > > .
> > 
> > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> 
> --  
> You received this message because you are subscribed to a topic in the  
> Google Groups "elasticsearch" group.  
> To unsubscribe from this topic, visit  
> [https://groups.google.com/d/topic/elasticsearch/sy8K-48bwbU/unsubscribe](https://groups.google.com/d/topic/elasticsearch/sy8K-48bwbU/unsubscribe).  
> To unsubscribe from this group and all its topics, send an email to  
> [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> To view this discussion on the web visit  
> [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com)  
> [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com?utm_medium=email&utm_source=footer)  
> .
> 
> For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U\_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [March 1, 2015, 4:19am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/6 "2015-03-01T04:19:57Z")

</div>

That's excessive. You don't need that many shards for 10 nodes.

On 27 February 2015 at 18:31, Em Kt [emkt84@gmail.com](mailto:emkt84@gmail.com) wrote:

> We have 10 data nodes which is storing data and separate master and client  
> nodes. So each node is having approx 1200 shards and 110 indices which I  
> think is not much.
> 
> On Fri, Feb 27, 2015 at 12:45 AM, Mark Walkom [markwalkom@gmail.com](mailto:markwalkom@gmail.com)  
> wrote:
> 
> > 12000 shards across how many nodes?
> > 
> > Don't forget that a shard is a lucene instance, it needs resources to  
> > operate and a node only has so many resources. This is why scaling is  
> > important.
> > 
> > On 27 February 2015 at 10:10, [emkt84@gmail.com](mailto:emkt84@gmail.com) wrote:
> > 
> > > We have seen similar issue in our setup too, but we are running 1.3.6. I  
> > > think it occurs with large index and shards counts. We have approx 12000  
> > > shards in total. I think its a bug in elasticsearch, as having large shard  
> > > count does not mean that node just simply drops out from cluster. The shard  
> > > count stats and cluster membership threads should not depend upon each  
> > > other.
> > > 
> > > On Thursday, February 26, 2015 at 7:36:20 AM UTC+1, Sean Clemmer wrote:
> > > 
> > > > +1
> > > > 
> > > > I have a similar story: After around six months using the v1.3.x  
> > > > series, I upgraded from v1.3.4 to v1.4.4. I've had monitoring and metrics  
> > > > in place for a while now, and compared to baseline I'm seeing occasional  
> > > > periods where nodes appear to drop out and usually back into the cluster  
> > > > (with NodeNotConnectedException). During these periods the cluster  
> > > > status is red on-and-off for maybe 15-30 minutes with anywhere from 30  
> > > > minutes to 2 hours in-between. The issue is worse in larger clusters with  
> > > > larger shard counts (thousands, tens of thousands).
> > > > 
> > > > Resource utilization is still good. The number of shards (and the  
> > > > amount of data) is essentially constant. I'm confident the upgrade was  
> > > > the only change; I have strict controls on the clusters.
> > > > 
> > > > On Wed, Feb 25, 2015 at 5:18 PM, sagarl [saga...@gmail.com](mailto:saga...@gmail.com) wrote:
> > > > 
> > > > > Hi,
> > > > > 
> > > > > We recently upgraded one of our ES Clusters from ES Version 1.1.0 to  
> > > > > 1.4.1.
> > > > > 
> > > > > We have dedicated master-data-search deployment in AWS. Cluster  
> > > > > settings are same for all the clusters.
> > > > > 
> > > > > Strangely, only in one cluster; we are seeing that nodes are  
> > > > > constantly failing to connect to Master node and rejoining back.  
> > > > > _It happens all the time, even during idle period (when there are no  
> > > > > read or writes)_.
> > > > > 
> > > > > We keep on seeing following exception in the logs  
> > > > > org.elasticsearch.transport.NodeNotConnectedException
> > > > > 
> > > > > Because of this, Cluster has slowed down considerably.
> > > > > 
> > > > > We use kopf plugin for monitoring and it keeps popping up message -  
> > > > > "Loading cluster information is talking too long"
> > > > > 
> > > > > There is not much data on individual nodes; almost 80% disk is free.  
> > > > > CPU and Heap are doing fine.
> > > > > 
> > > > > Only difference between this cluster and other cluster is the number  
> > > > > of indices and shards. Other clusters have shards in hundreds and indices  
> > > > > in double digit. While this cluster has around 5000 shards and close to 250  
> > > > > indices.
> > > > > 
> > > > > But we are not sure, if number of shards or indices can cause  
> > > > > reconnection issues between nodes.
> > > > > 
> > > > > Not sure if it's really related to 1.4.1 or something else. But in  
> > > > > that case, other clusters should have been affected too.
> > > > > 
> > > > > Any help will be appreciated !
> > > > > 
> > > > > Thanks,
> > > > > 
> > > > > --  
> > > > > You received this message because you are subscribed to the Google  
> > > > > Groups "elasticsearch" group.  
> > > > > To unsubscribe from this group and stop receiving emails from it, send  
> > > > > an email to [elasticsearc...@googlegroups.com](mailto:elasticsearc...@googlegroups.com).  
> > > > > To view this discussion on the web visit [https://groups.google.com/d/](https://groups.google.com/d/)  
> > > > > msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%  
> > > > > [40googlegroups.com](http://40googlegroups.com)  
> > > > > [https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/12e9b4ab-7600-4d22-a347-c03edaebe4f3%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > > > > .  
> > > > > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> > > > 
> > > > --  
> > > > You received this message because you are subscribed to the Google  
> > > > Groups "elasticsearch" group.  
> > > > To unsubscribe from this group and stop receiving emails from it, send  
> > > > an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> > > > To view this discussion on the web visit  
> > > > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com)  
> > > > [https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/b810319e-f553-4459-b81d-464743e90ee8%40googlegroups.com?utm_medium=email&utm_source=footer)  
> > > > .
> > > 
> > > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> > 
> > --  
> > You received this message because you are subscribed to a topic in the  
> > Google Groups "elasticsearch" group.  
> > To unsubscribe from this topic, visit  
> > [https://groups.google.com/d/topic/elasticsearch/sy8K-48bwbU/unsubscribe](https://groups.google.com/d/topic/elasticsearch/sy8K-48bwbU/unsubscribe).  
> > To unsubscribe from this group and all its topics, send an email to  
> > [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> > To view this discussion on the web visit  
> > [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com)  
> > [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X8ebpywmOLkceMHP%2B7D9sxDAUQUuAk%3D1rHC96BQTJDt0w%40mail.gmail.com?utm_medium=email&utm_source=footer)  
> > .
> > 
> > For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).
> 
> --  
> You received this message because you are subscribed to the Google Groups  
> "elasticsearch" group.  
> To unsubscribe from this group and stop receiving emails from it, send an  
> email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
> To view this discussion on the web visit  
> [https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U\_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com)  
> [https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U\_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com?utm\_medium=email&utm\_source=footer](https://groups.google.com/d/msgid/elasticsearch/CADm%2BQZcDL2NGf1U_%3D-SpcumZbnK0zpnAbd5BOxU30oHNtTaZeQ%40mail.gmail.com?utm_medium=email&utm_source=footer)  
> .  
> For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/CAEYi1X-DXs9JBwFK73KF%2BdC6-UjDt\_FzONVq6Ac5%2BXC79rBBdw%40mail.gmail.com](https://groups.google.com/d/msgid/elasticsearch/CAEYi1X-DXs9JBwFK73KF%2BdC6-UjDt_FzONVq6Ac5%2BXC79rBBdw%40mail.gmail.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![sagarit2](https://avatars.discourse-cdn.com/v4/letter/s/f0a364/32.png) [@sagarit2](https://discuss.elastic.co/u/sagarit2)\
**Post date:** [March 3, 2015, 4:33pm UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/7 "2015-03-03T16:33:29Z")

</div>

There are two different points,we are trying to figure out

1. why it started giving above mentioned error in 1.4.1 only ? It is working fine with no issues in 1.3.x.

2. As I mentioned, disks have 80% space free and we have 5000 shards across 200 indices in 6 node cluster,but even in an idle cluster with no reads and no writes, we are seeing this issue.

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/00f02b63-635b-40f9-853d-e3b11a37180c%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/00f02b63-635b-40f9-853d-e3b11a37180c%40googlegroups.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![sagarit2](https://avatars.discourse-cdn.com/v4/letter/s/f0a364/32.png) [@sagarit2](https://discuss.elastic.co/u/sagarit2)\
**Post date:** [March 5, 2015, 9:27pm UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/8 "2015-03-05T21:27:32Z")

</div>

I have created an issue  
[https://github.com/elasticsearch/elasticsearch/issues/10003](https://github.com/elasticsearch/elasticsearch/issues/10003) on github  
site.

@Sean, @Em , please feel free to comment on it.

Thanks,

On Tuesday, March 3, 2015 at 8:33:29 AM UTC-8, sagarl wrote:

> There are two different points,we are trying to figure out
> 
> 1. why it started giving above mentioned error in 1.4.1 only ? It is  
> working fine with no issues in 1.3.x.
> 
> 2. As I mentioned, disks have 80% space free and we have 5000 shards  
> across 200 indices in 6 node cluster,but even in an idle cluster with no  
> reads and no writes, we are seeing this issue.

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
To view this discussion on the web visit [https://groups.google.com/d/msgid/elasticsearch/dc3183bc-f644-4cf8-b047-4990d3c96b71%40googlegroups.com](https://groups.google.com/d/msgid/elasticsearch/dc3183bc-f644-4cf8-b047-4990d3c96b71%40googlegroups.com).  
For more options, visit [https://groups.google.com/d/optout](https://groups.google.com/d/optout).

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [July 6, 2017, 12:28am UTC](https://discuss.elastic.co/t/strange-issue-after-upgrading-from-1-1-0-to-1-4-1-es-version/22397/9 "2017-07-06T00:28:16Z")

</div>


