Re: [documancer] ANN: new Documancer release

Vaclav Slavik <[email protected]> Tue, 15 Feb 2005 18:05:08 +0100
Newsgroups gmane.text.documancer.user
Message-ID <[email protected]>
--nextPart1269882.UX1yLStQ3M
Content-Type: text/plain;
  charset="iso-8859-1"
Content-Transfer-Encoding: quoted-printable
Content-Disposition: inline

Hi,

Arnd Baecker wrote:
> This is with python 2.3.4, wxmozilla-0.5.3, documancer 0.2.5.

Just a guess: what is your locale and what is decimal point deliminer=20
used in it? Does running it with LC_ALL=3Den_GB or =3DC help?

> BTW: creating the index seems (if my memory serves me correctly)
> to be much slower than with swish.

Yes, that's true. But I think it's worth it -- no other fulltext=20
search I evaluated, SWISH-E included, supports Unicode. And Lucene's=20
API is really useful...

> And one more: is there a way that documancer could
> "automatically" extract a table of contents when
> dealing with html files?

Yes, I assume there is, but it's not a trivial task, so even though I=20
do want to have that feature myself, I'm unlikely to do it before I=20
implements tons of other more urgently needed things...

Regards,
Vaclav

=2D-=20
PGP key: 0x465264C9, available from http://wwwkeys.pgp.net/

--nextPart1269882.UX1yLStQ3M
Content-Type: application/pgp-signature

-----BEGIN PGP SIGNATURE-----
Version: GnuPG v1.4.0 (GNU/Linux)

iD8DBQBCEivLxDYa/UZSZMkRAsXtAJ4k/QIlYMfyYy8fQiT5kNZAOXNA0ACfZej0
tDHqVrItDDo8uiid6LV8Rfw=
=Dm9q
-----END PGP SIGNATURE-----

--nextPart1269882.UX1yLStQ3M--


-------------------------------------------------------
SF email is sponsored by - The IT Product Guide
Read honest & candid reviews on hundreds of IT Products from real users.
Discover which products truly live up to the hype. Start reading now.
http://ads.osdn.com/?ad_id=6595&alloc_id=14396&op=click