# New User - too dumb to create first index - please help

**URL:** <https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861>\
**Category:** Elasticsearch\
**Created:** [January 3, 2018, 2:12am UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861 "2018-01-03T02:12:26Z")\
**Posts on this page:** 6\
**Page:** 1

<div class="post-metadata">

**Author:** ![neergttocsdivad](https://avatars.discourse-cdn.com/v4/letter/n/43a26b/32.png) [@neergttocsdivad](https://discuss.elastic.co/u/neergttocsdivad)\
**Post date:** [January 3, 2018, 2:12am UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/1 "2018-01-03T02:12:26Z")

</div>

I would like to experiment with elastic before using it for a research project I am developing. The project data is contained in unstructured pdf files, but for this exercise I am using Wikipedia.

My computer is running Windows 10.

I have downloaded wikipedia as html files into a folder called '[en.wikipedia.org](http://en.wikipedia.org)' on a usb drive called 'brown3' connected to a synology NAS drive called 'synology2'.

The path to my folder is therefore \synology2\brown3\[en.wikipedia.org](http://en.wikipedia.org)

I have installed kibana 6 with the ingest plugin

I have read countless pages on [elastic.co](http://elastic.co) and viewed loads of youtube videos but have been unable to translate their general advice to my specific needs and so I am still unable to get elastic/kibana to index my files.

Though the text on wiki pages is organized using headings, sub-headings, numbered lists and bulleted lists, I want all the text to be indexed as just text, to match my unstructured pdf files.

I would greatly appreciate your help.

Regards  
David

---

<div class="post-metadata">

**Author:** ![warkolm](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/warkolm/32/39224_2.png) [@warkolm](https://discuss.elastic.co/u/warkolm)\
**Post date:** [January 3, 2018, 6:18am UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/2 "2018-01-03T06:18:08Z")

</div>

If I can suggest, you're probably better off starting simpler. Check out [https://www.elastic.co/guide/en/kibana/current/getting-started.html](https://www.elastic.co/guide/en/kibana/current/getting-started.html).

Because otherwise you will need a way to read the files off disk, process them how you want and index them to Elasticsearch. It sounds simple but it's not given your first starting.

---

<div class="post-metadata">

**Author:** ![dadoonet](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/dadoonet/32/137187_2.png) [@dadoonet](https://discuss.elastic.co/u/dadoonet)\
**Post date:** [January 3, 2018, 6:27am UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/3 "2018-01-03T06:27:20Z")

</div>

You can also give a look at FSCrawler project: [https://github.com/dadoonet/fscrawler](https://github.com/dadoonet/fscrawler)

But as Mark said, start with something even more easy. 😉

---

<div class="post-metadata">

**Author:** ![neergttocsdivad](https://avatars.discourse-cdn.com/v4/letter/n/43a26b/32.png) [@neergttocsdivad](https://discuss.elastic.co/u/neergttocsdivad)\
**Post date:** [January 13, 2018, 6:07am UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/4 "2018-01-13T06:07:19Z")

</div>

Thank you both.

I've been reading the FScrawler web page.

My operating system is on my computer's C drive. My pdf files (to be indexed) are on one NAS. I would like my index to be created on another NAS.

Kibana is stored in c:/kibana

Where should I place my FScrawler snapshot ?

If on the c. drive then how/where to specify the folder containing my pdf files ?

Thanks

---

<div class="post-metadata">

**Author:** ![dadoonet](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/dadoonet/32/137187_2.png) [@dadoonet](https://discuss.elastic.co/u/dadoonet)\
**Post date:** [January 13, 2018, 4:02pm UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/5 "2018-01-13T16:02:16Z")

</div>

> Where should I place my FScrawler snapshot ?

Wherever you want.

> how/where to specify the folder containing my pdf files ?

See [GitHub - dadoonet/fscrawler: Elasticsearch File System Crawler (FS Crawler)](https://github.com/dadoonet/fscrawler#root-directory)

---

<div class="post-metadata">

**Author:** ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)\
**Post date:** [February 10, 2018, 4:02pm UTC](https://discuss.elastic.co/t/new-user-too-dumb-to-create-first-index-please-help/113861/6 "2018-02-10T16:02:23Z")

</div>

This topic was automatically closed 28 days after the last reply. New replies are no longer allowed.
