Re: Wiki Data
"J. Kyllo" <[email protected]>
| Newsgroups | gmane.network.wireless.seattle.general |
|---|---|
| Message-ID | <[email protected]> |
I actually found something interesting with this. I copied the Wiki text from an SWN page that gave me trouble and put in on my own wiki (that runs a seemingly more current version of MoinMoin) and I didn't have any trouble. I was going to try to figure out a way to form the Wiki text such that the parser didn't go weird, but I gave up on that since the data I need is otherwise available. Thanks for the input. -Jeff > My experience has been that MoinMoin's XML formatter is rarely "well > formed." I believe the reason is the Wiki engine assumes no pages will > contain HTML tags. When they do contain HTML the encoder gets very > confused. > > I'm sorry I don't have a solution at this time. I had the same issue > with my MoinMoin Wiki and never figured it out. I ended up switching > to a different engine (MediaWiki) but I don't think SeattleWireless is > going to switch just to get well-formed XML. > > On 7/11/05, J. Kyllo <[email protected]> wrote: >> Well, I have figured out that it is included in MoinMoin by default. >> Another question though, has anyone here actually ever tried parsing the >> xml from this formatter? A substantial number of pages from the SWN >> wiki >> (like FrontPage, for example) are returning invalid XML. This is vexing >> to say the least. Basically it ends up doing this: >> >> <s1> >> <s2> >> <p> >> </s2> >> </p> >> </s1> >> >> It doesn't seem to be all pages, but I have not yet figured out what is >> causing it to behave this way. Any thoughts are welcome. >> >> Thanks, >> Jeff >> >> > Dang, that's sweet. Is that a formatter that is installed by default? >> > >> > -Jeff >> > >> >> Throw "?action=format&mimetype=text/xml" after the WikiPage: >> >> >> >> http://www.seattlewireless.net/index.cgi/RecentChanges?action=format&mim >> >> etype=text/xml >> >> >> >>> -----Original Message----- >> >>> From: [email protected] >> >>> [mailto:[email protected]] On Behalf Of J. Kyllo >> >>> Sent: Monday, July 11, 2005 10:59 AM >> >>> To: [email protected] >> >>> Subject: Wiki Data >> >>> >> >>> Does anyone know if there is a clean, easy, correct, or >> >>> generally okay way to grab a page from the wiki besides a >> >>> normal http get to the url? The idea is that I want the >> >>> processed output from just the page itself so that I can >> >>> parse it. A normal grab of the page will also include all of >> >>> the navigation links at the top and bottom whereas grabbing >> >>> the raw text doesn't evaluate the embedded actions. >> >>> >> >>> I thought I had heard about some xml communication that could >> >>> be done, but I'm not sure. >> >>> >> >>> Any ideas or thoughts? >> >>> >> >>> Thanks, >> >>> Jeff >> >>> >> >>> _______________________________________________ >> >>> Talk mailing list >> >>> [email protected] >> >>> http://seattlewireless.net/mailman/listinfo/talk >> >>> >> >> _______________________________________________ >> >> Talk mailing list >> >> [email protected] >> >> http://seattlewireless.net/mailman/listinfo/talk >> >> >> > >> > >> > _______________________________________________ >> > Talk mailing list >> > [email protected] >> > http://seattlewireless.net/mailman/listinfo/talk >> > >> >> >> _______________________________________________ >> Talk mailing list >> [email protected] >> http://seattlewireless.net/mailman/listinfo/talk >> > > > -- > aaron > _______________________________________________ > Talk mailing list > [email protected] > http://seattlewireless.net/mailman/listinfo/talk > _______________________________________________ Talk mailing list [email protected] http://seattlewireless.net/mailman/listinfo/talk