# How to get index contents of a website

**URL:** <https://discuss.elastic.co/t/how-to-get-index-contents-of-a-website/304944>\
**Category:** Elasticsearch\
**Created:** [May 17, 2022, 12:56pm UTC](https://discuss.elastic.co/t/how-to-get-index-contents-of-a-website/304944 "2022-05-17T12:56:06Z")\
**Posts on this page:** 3\
**Page:** 1

<div class="post-metadata">

**Author:** ![Tasdiq\_Shaikh](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/tasdiq_shaikh/32/76982_2.png) [@Tasdiq\_Shaikh](https://discuss.elastic.co/u/Tasdiq_Shaikh)\
**Post date:** [May 17, 2022, 12:56pm UTC](https://discuss.elastic.co/t/how-to-get-index-contents-of-a-website/304944/1 "2022-05-17T12:56:06Z")

</div>

How can I index and perform search operations on the contents of a website of a given web link like [discuss.elastic.co](https://discuss.elastic.co) using Java 1.8 and Elasticsearch 7.12.0?

---

<div class="post-metadata">

**Author:** ![dadoonet](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/dadoonet/32/137187_2.png) [@dadoonet](https://discuss.elastic.co/u/dadoonet)\
**Post date:** [May 17, 2022, 1:39pm UTC](https://discuss.elastic.co/t/how-to-get-index-contents-of-a-website/304944/2 "2022-05-17T13:39:15Z")

</div>

I think you already asked for a similar question at [Is Elasticsearch webcrawler an open source feature or it is paid?](https://discuss.elastic.co/t/is-elasticsearch-webcrawler-an-open-source-feature-or-it-is-paid/304789).

I believe it would be better to keep the discussion in one place.

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [May 19, 2022, 4:23am UTC](https://discuss.elastic.co/t/how-to-get-index-contents-of-a-website/304944/3 "2022-05-19T04:23:13Z")

</div>


