Re: An HTML Entry Parser

Robert Leftwich <robert-NuYOPsrVWqAk+I/[email protected]>
Newsgroups gmane.comp.web.pyblosxom.user
Message-ID <[email protected]>
At 06:55 AM 11/09/2003, Joseph Reagle wrote:
>On Wednesday 10 September 2003 15:55, Robert Leftwich wrote:
> > I had a similar problem when adding an html entry parser
>
>Awesome! Are you doing anything particularly cool with your html entry
>parser?

No, the main difference with yours is that I search for the title tag and 
if that isn't found I default to the file name as the title, but I like the 
idea of using the first header tag as the title, so I will add that to mine 
if no title tag is found.
I also use regex's for locating the tags as some of the html files are 
authored in ms word and as such have multiple attributes on the body and 
title tags.

>Also, I'd advocate this get adopted into the main tree. (I'm
>accumulating enough diffs -- which is great about pyblosxom -- that I'm
>afraid when it comes to upgrade after the RC, I'm gonna have to do lots of
>retweaking <smile/>).

<aol>Me too!</aol>

I was planning to put together most, if not all of my changes (mainly to do 
with running on windows) as a set of patches.

Robert





-------------------------------------------------------
This sf.net email is sponsored by:ThinkGeek
Welcome to geek heaven.
http://thinkgeek.com/sf
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.