# Extracting texts from Flatten/ Scanned PDF Documents in Kibana

**URL:** https://discuss.elastic.co/t/extracting-texts-from-flatten-scanned-pdf-documents-in-kibana/341351
**Category:** Kibana
**Created:** [August 22, 2023, 9:59am UTC](https://discuss.elastic.co/t/extracting-texts-from-flatten-scanned-pdf-documents-in-kibana/341351 "2023-08-22T09:59:28Z")
**Posts on this page:** 3
**Page:** 1

<div class="post-metadata">

### Author: ![Anant\_Patankar](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/anant_patankar/32/124833_2.png) [@Anant\_Patankar](https://discuss.elastic.co/u/Anant_Patankar)
#### Post date: [August 22, 2023, 9:59am UTC](https://discuss.elastic.co/t/extracting-texts-from-flatten-scanned-pdf-documents-in-kibana/341351/1 "2023-08-22T09:59:28Z")

</div>

Hello Everyone,  
I am trying to read texts from scanned/flattened pdf which are made up of images and texts that are not readable with a pdf reader.  
How can I read and index texts from flattened or scanned pdf files in kibana without using fscrawler?

---

<div class="post-metadata">

### Author: ![carly.richmond](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/carly.richmond/32/104935_2.png) [@carly.richmond](https://discuss.elastic.co/u/carly.richmond)
#### Post date: [August 24, 2023, 8:50am UTC](https://discuss.elastic.co/t/extracting-texts-from-flatten-scanned-pdf-documents-in-kibana/341351/2 "2023-08-24T08:50:10Z")

</div>

Hi @Anant_Patankar,

Welcome to the community! Is there a particular reason why you don't want to use fscrawler?

Another alternative would be to use [the attachment processor with an ingest pipeline](https://www.elastic.co/guide/en/elasticsearch/reference/8.9/attachment.html). Have a look at [this related thread](https://discuss.elastic.co/t/how-to-index-the-pdf-documents/327987/4) for some useful options.

Hope that helps!

---

<div class="post-metadata">

### Author: ![system](https://us1.discourse-cdn.com/elastic/original/3X/1/a/1ac57faf039f6b580b3f104ef42a2a89e41014de.png) [@system](https://discuss.elastic.co/u/system)
#### Post date: [September 21, 2023, 8:51am UTC](https://discuss.elastic.co/t/extracting-texts-from-flatten-scanned-pdf-documents-in-kibana/341351/3 "2023-09-21T08:51:03Z")

</div>

This topic was automatically closed 28 days after the last reply. New replies are no longer allowed.
