# ElasticSearch6.4 compatible web crawler

**URL:** <https://discuss.elastic.co/t/elasticsearch6-4-compatible-web-crawler/154619>\
**Category:** Elasticsearch\
**Created:** [October 30, 2018, 10:32am UTC](https://discuss.elastic.co/t/elasticsearch6-4-compatible-web-crawler/154619 "2018-10-30T10:32:31Z")\
**Posts on this page:** 1\
**Showing post:** 2

<div class="post-metadata">

**Author:** ![dadoonet](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/dadoonet/32/137187_2.png) [@dadoonet](https://discuss.elastic.co/u/dadoonet)\
**Post date:** [October 30, 2018, 4:58pm UTC](https://discuss.elastic.co/t/elasticsearch6-4-compatible-web-crawler/154619/2 "2018-10-30T16:58:29Z")

</div>

> [@Subasini](#):
>
> Please respond. It is of higher priority.

Read [this](https://discuss.elastic.co/t/about-the-elasticsearch-category/21) and specifically the "Also be patient" part.

I personally consider that someone who has a cluster in production down is more urgent than a question about a project that does not exist yet.

Anyway, some answers:

1. No idea. Never used Nutch. May be ask to the Nutch mailing list if any?
2. It depends on what are your needs. I wrote [FSCrawler](https://fscrawler.readthedocs.io/) to crawl files on disk for example and parse them with Apache Tika.
3. Too wide question.

---

_[View the full topic](https://discuss.elastic.co/t/elasticsearch6-4-compatible-web-crawler/154619)._
