Webboard: Can't index pdfs

[email protected] 20 Apr 2005 16:12:39 -0000
Newsgroups gmane.comp.web.mnogosearch.windows
Message-ID <[email protected]>
Author: Brian
Email: [email protected]
Message:
Hi Roman,



Thank you. I set mine up exactly as you said and I still get a "Parser Execution Error" after indexing all pdfs.



Also you said for the parser layer (in the command window) to put this:

"C:\Program Files\mnoGoSearch\Tools\xpdf\pdftotext" -layout -htmlmeta $1 $2



I am using windows so I am assuming you meant this: "C:\Program Files\mnoGoSearch\Tools\xpdf\pdftotext.EXE" -layout -htmlmeta $1 $2



2 questions: 1. Why did you suggest sending the output to text/html? Does it matter if its text or html? 2. Are the quotes nessessary?



I also wanted to mention that when I change the MIME type in IIS for .pdf to text/plain mnogosearch stores the data in SQL! But of course it is unreadable. Now I am thinking this is an issue with IIS and mnoGoSearch?



Please respond.

Reply: <http://www.mnogosearch.org/board/message.php?id=15542>