# Node showing twice in the cluster (same IP/port)

**URL:** <https://discuss.elastic.co/t/node-showing-twice-in-the-cluster-same-ip-port/13176>\
**Category:** Elasticsearch\
**Created:** [August 12, 2013, 2:07pm UTC](https://discuss.elastic.co/t/node-showing-twice-in-the-cluster-same-ip-port/13176 "2013-08-12T14:07:58Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![dezmodue](https://avatars.discourse-cdn.com/v4/letter/d/3ab097/32.png) [@dezmodue](https://discuss.elastic.co/u/dezmodue)\
**Post date:** [August 12, 2013, 2:07pm UTC](https://discuss.elastic.co/t/node-showing-twice-in-the-cluster-same-ip-port/13176/1 "2013-08-12T14:07:58Z")

</div>

Hi,

We run an 8 instances ES cluster in EC2, one index (size: 3695.5gb),  
48 shards, 1 replica per shard. ES version is 0.20.2, Oracle JVM  
1.6.38, 30GB heap, total RAM 60GB, 2 \* 1024GB SSD drives in raid 0  
(lvm stripe), 16 cores.

A few days ago around 7:17 UTC one of the nodes has been rebooted  
(underlying host issue according to AWS - still investigating the root  
cause). Once rebooted it connected to the cluster but showed up twice  
in the cluster config:

O4iZm6dyTU6NT2Z4WiIZlg: {  
name: [es-6385b.domain.com](http://es-6385b.domain.com)  
transport\_address: inet[/10.x.x.149:9300]  
attributes: {  
aws\_availability\_zone: us-east-1b  
max\_local\_storage\_nodes: 1  
}  
.....  
.....  
}  
sLam2ByVTQyPcc4Nzk9oMQ: {  
name: [es-6385b.domain.com](http://es-6385b.domain.com)  
transport\_address: inet[/10.x.x.149:9300]  
attributes: {  
aws\_availability\_zone: us-east-1b  
max\_local\_storage\_nodes: 1  
}

Where sLam2ByVTQyPcc4Nzk9oMQ is the original node ID and  
O4iZm6dyTU6NT2Z4WiIZlg is the ID after the restart.

After a while the node is removed from the cluster, from the logs it  
seems that the master is still trying to connect to  
sLam2ByVTQyPcc4Nzk9oMQ although this node doesn't exist anymore.

Around 10:13 UTC the node [es-6385b.domain.com](http://es-6385b.domain.com) runs out of memory:

[2013-07-23 10:13:44,271][WARN  
][netty.channel.socket.nio.AbstractNioSelector] Unexpected exception  
in the selector loop.  
java.lang.OutOfMemoryError: Java heap space

At this point the master notices the old node ID is no longer contactable:

[2013-07-23 12:20:54,355][WARN][cluster.service]  
[[es-6388d.domain.com](http://es-6388d.domain.com)] failed to reconnect to node  
[[es-6385b.domain.com](http://es-6385b.domain.com)][sLam2ByVTQyPcc4Nzk9oMQ][inet[/10.x.x.149:9300]]{aws\_availability\_zone=us-east-1b,  
max\_local\_storage\_nodes=1}

org.elasticsearch.transport.ConnectTransportException:  
[[es-6385b.domain.com](http://es-6385b.domain.com)][inet[/10.x.x.149:9300]] connect\_timeout[30s]

But still the node is not removed from the cluster (note the old ID  
sLam2ByVTQyPcc4Nzk9oMQ in the log file), there are a lot of those  
messages in the logs but the master doesn't evict the node from the  
cluster (why?).

After a while [es-6388d.domain.com](http://es-6388d.domain.com) is restarted and again it tries to  
join the cluster with a new ID but it still shows twice in the config,  
with the original ID pre reboot and with the new one after the latest  
restart.

This is only resolved with a full cluster restart (or at least this is  
the only way we managed to 'solve' it).

I have tried to collect as much info as I could and I have all the  
logs for inspection, I would really appreciate any help in  
understanding how this situation (duplicate node) can happen, what are  
the possible issues resulting form it (routing?shard  
distribution?queries?), and how we can resolve the problem if it ever  
happens again (or better prevent it) without a full cluster restart.

Thanks,

Simone

--  
You received this message because you are subscribed to the Google Groups "elasticsearch" group.  
To unsubscribe from this group and stop receiving emails from it, send an email to [elasticsearch+unsubscribe@googlegroups.com](mailto:elasticsearch+unsubscribe@googlegroups.com).  
For more options, visit [https://groups.google.com/groups/opt\_out](https://groups.google.com/groups/opt_out).

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [July 6, 2017, 2:21am UTC](https://discuss.elastic.co/t/node-showing-twice-in-the-cluster-same-ip-port/13176/2 "2017-07-06T02:21:33Z")

</div>


