# Index PDF files content

**URL:** <https://forum.pydio.com/t/index-pdf-files-content/1036>\
**Category:** Development\
**Created:** [April 20, 2018, 11:55am UTC](https://forum.pydio.com/t/index-pdf-files-content/1036 "2018-04-20T11:55:29Z")\
**Posts on this page:** 3\
**Page:** 1

<div class="post-metadata">

**Author:** ![devneo01](https://yyz2.discourse-cdn.com/flex032/user_avatar/forum.pydio.com/devneo01/32/278_2.png) [@devneo01](https://forum.pydio.com/u/devneo01)\
**Post date:** [April 20, 2018, 11:55am UTC](https://forum.pydio.com/t/index-pdf-files-content/1036/1 "2018-04-20T11:55:29Z")

</div>

Hello,

I’m running Pydio 7.0.4 on a Windows Server (AMPPS) and i’m looking for some help in order to index PDF files content in order for the users to search directly in those files.

My problem is that i don’t know how to install the packages for UNICONV + XPDF INTEGRATION mentioned in the Lucene Indexer documentation.

Also, it seems like the “Advanced Search” box doesn’t have the option to search directly in files.

Any advice will be much appreaciated !

---

<div class="post-metadata">

**Author:** ![zayn](https://avatars.discourse-cdn.com/v4/letter/z/a87d85/32.png) [@zayn](https://forum.pydio.com/u/zayn)\
**Post date:** [April 24, 2018, 10:44am UTC](https://forum.pydio.com/t/index-pdf-files-content/1036/2 "2018-04-24T10:44:02Z")

</div>

Hi,  
i dont really know how to do it on windows server but i’ve found some guides that could help you :

- [Unoconv windows server](https://docs.moodle.org/31/en/Installing_unoconv#Installing_unoconv_on_Windows)
- [xpdf windows](https://www.xpdfreader.com/download.html)

---

<div class="post-metadata">

**Author:** ![vicWeller](https://yyz2.discourse-cdn.com/flex032/user_avatar/forum.pydio.com/vicweller/32/439_2.png) [@vicWeller](https://forum.pydio.com/u/vicWeller)\
**Post date:** [July 5, 2018, 3:07pm UTC](https://forum.pydio.com/t/index-pdf-files-content/1036/3 "2018-07-05T15:07:01Z")

</div>

If the content of your pdf file doesn’t index, so there’s some issue with the OCR layer then. It basically depends on what software have you used in order to create that document. Upload that to the very pdf editing tool you have under your belt, I use this one eg [https://form-cd-401s.pdffiller.com/](https://form-cd-401s.pdffiller.com/) because it’s enough for such a purpose and cost lesser than Acrobat and others. There you’ll be able to fix the issues if there will be some
