Re: Wiki Data
Aaron Bronow <[email protected]>
| Newsgroups | gmane.network.wireless.seattle.general |
|---|---|
| Message-ID | <[email protected]> |
My experience has been that MoinMoin's XML formatter is rarely "well formed." I believe the reason is the Wiki engine assumes no pages will contain HTML tags. When they do contain HTML the encoder gets very confused. I'm sorry I don't have a solution at this time. I had the same issue with my MoinMoin Wiki and never figured it out. I ended up switching to a different engine (MediaWiki) but I don't think SeattleWireless is going to switch just to get well-formed XML. On 7/11/05, J. Kyllo <[email protected]> wrote: > Well, I have figured out that it is included in MoinMoin by default. > Another question though, has anyone here actually ever tried parsing the > xml from this formatter? A substantial number of pages from the SWN wiki > (like FrontPage, for example) are returning invalid XML. This is vexing > to say the least. Basically it ends up doing this: > > <s1> > <s2> > <p> > </s2> > </p> > </s1> > > It doesn't seem to be all pages, but I have not yet figured out what is > causing it to behave this way. Any thoughts are welcome. > > Thanks, > Jeff > > > Dang, that's sweet. Is that a formatter that is installed by default? > > > > -Jeff > > > >> Throw "?action=format&mimetype=text/xml" after the WikiPage: > >> > >> http://www.seattlewireless.net/index.cgi/RecentChanges?action=format&mim > >> etype=text/xml > >> > >>> -----Original Message----- > >>> From: [email protected] > >>> [mailto:[email protected]] On Behalf Of J. Kyllo > >>> Sent: Monday, July 11, 2005 10:59 AM > >>> To: [email protected] > >>> Subject: Wiki Data > >>> > >>> Does anyone know if there is a clean, easy, correct, or > >>> generally okay way to grab a page from the wiki besides a > >>> normal http get to the url? The idea is that I want the > >>> processed output from just the page itself so that I can > >>> parse it. A normal grab of the page will also include all of > >>> the navigation links at the top and bottom whereas grabbing > >>> the raw text doesn't evaluate the embedded actions. > >>> > >>> I thought I had heard about some xml communication that could > >>> be done, but I'm not sure. > >>> > >>> Any ideas or thoughts? > >>> > >>> Thanks, > >>> Jeff > >>> > >>> _______________________________________________ > >>> Talk mailing list > >>> [email protected] > >>> http://seattlewireless.net/mailman/listinfo/talk > >>> > >> _______________________________________________ > >> Talk mailing list > >> [email protected] > >> http://seattlewireless.net/mailman/listinfo/talk > >> > > > > > > _______________________________________________ > > Talk mailing list > > [email protected] > > http://seattlewireless.net/mailman/listinfo/talk > > > > > _______________________________________________ > Talk mailing list > [email protected] > http://seattlewireless.net/mailman/listinfo/talk > -- aaron _______________________________________________ Talk mailing list [email protected] http://seattlewireless.net/mailman/listinfo/talk