Re: A question about encoding and text fixing platforms for undergraduate curators

Martin Holmes <[email protected]> Thu, 27 Apr 2017 05:37:24 -0700
Newsgroups gmane.text.tei.general
Message-ID <[email protected]>
+1 for the ducks-to-water claim from me. We've been teaching students 
XML for years, and I've never had a single instance where a student 
struggled with it at all. It takes a matter of a couple of hours to get 
people into basic encoding.

XML is not hard. Word, by contrast, is a concoction of frustrations, and 
getting DOCX into decent TEI when you're done is horribly difficult.

Cheers,
Martin

On 2017-04-26 05:23 PM, Martin Mueller wrote:
> They’re not working at the Folger. They’ll be working at Northwestern,
> though using the Web-based Folger platform is a possible option. You’re
> right in saying that the students may not know Word all that well.  I
> would like to believe that students take to angle brackets like ducks to
> water and for what it’s worth, I much prefer working with the text mode
> in oXygen and haven’t found much use for the author mode. Still, I’ll
> need more persuading on the ducks to water claim.
>
>
>
>
>
> *From: *Elisa Beshero-Bondar <[email protected]>
> *Date: *Wednesday, April 26, 2017 at 7:10 PM
> *To: *Martin Mueller <[email protected]>
> *Cc: *"[email protected]" <[email protected]>
> *Subject: *Re: A question about encoding and text fixing platforms for
> undergraduate curators
>
>
>
> HI Martin,
>
> If you have students coming for **six weeks** to work at the Folger, why
> not orient them to writing and working the code? I work regularly with
> undergrads, and it takes about one week for them to orient themselves to
> angle brackets and oXygen. Also, I don’t believe that most undergrads
> “know” Microsoft Word especially well. Why create extra work with making
> Microsoft styles force-fit to plain simple sparse code?
>
>
>
> In my experience (regularly, from Fall 2012 onward), students take like
> ducks to the water and tend to appreciate working directly with the
> angle bracket code. Seems to me a six week institute is a perfect
> opportunity to teach them to code as part of the experience!
>
>
>
> Elisa
>
> --
> Elisa Beshero-Bondar, PhD
> Director, Center for the Digital Text | Associate Professor of English
> University of Pittsburgh at Greensburg | Humanities Division
> 150 Finoli Drive
> Greensburg, PA  15601  USA
> E-mail: [email protected] <mailto:[email protected]>
> Development site: http://newtfire.org
> <https://urldefense.proofpoint.com/v2/url?u=http-3A__newtfire.org&d=DwMFaQ&c=yHlS04HhBraes5BQ9ueu5zKhE7rtNXt_d012z2PA6ws&r=rG8zxOdssqSzDRz4x1GLlmLOW60xyVXydxwnJZpkxbk&m=0N_RpCVwzbACojeKb6dxzlY1NOGShq0zdjRJGs8TXZs&s=WyRipKNOZ5-njOhfp8fI4VMyK3oV4rHCqhyHXKumRac&e=>
>
>
>
>
>
>
>
>
>
>
>
>
>     On Apr 26, 2017, at 7:09 PM, Martin Mueller
>     <[email protected]
>     <mailto:[email protected]>> wrote:
>
>
>
>     I’m not sure whether this is a sensible question to ask, but I’ll
>     ask it anyhow.
>
>
>
>     This summer we’ll have a number of undergraduate curators of TCP
>     texts fixing this and that, mainly incompletely transcribed words,
>     but sometimes longer stretches of text or whole pages.
>
>
>
>     So we’ll some transcription platform in addition to an eXist site
>     at http://shc.earlprint.org
>     <https://urldefense.proofpoint.com/v2/url?u=http-3A__shc.earlprint.org_&d=DwMFaQ&c=yHlS04HhBraes5BQ9ueu5zKhE7rtNXt_d012z2PA6ws&r=rG8zxOdssqSzDRz4x1GLlmLOW60xyVXydxwnJZpkxbk&m=0N_RpCVwzbACojeKb6dxzlY1NOGShq0zdjRJGs8TXZs&s=eaGdmgiKyJYQbrj6Usf_L8n1Pyh63t6FzbkZzAX7BLc&e=>,
>     where single words can be fixed by changing the value of the content
>     to of a <w> element, mercifully invisible to the user.
>
>
>
>     If you believe that the best tool is the tool you know best, you’d
>     try to figure out whether undergraduates could do this work using
>     Microsoft Word with a set of styles that subsequently support the
>     automatic transformation of the Microsoft word passages into XML
>     fragments that can be fitted into the TCP transcriptions.
>
>
>
>     Is that a plausible scenario and has something like that been done?
>     TCP encoding is quite sparse. Text is either marked (inside <hi>) or
>     unmarked, and the transcription is silent about what the unmarked
>     state is.  My rough guess is that a dozen elements will cover the
>     vast majority of cases.
>
>
>
>     The Folger Library has an attractive Web-based tool for manuscript
>     transcription that can probably be adjusted with little trouble.
>
>
>
>     The students will be in residence for six weeks, and it may be that
>     we should teach them encoding with oXygen. Some of them may love it,
>     others may hate it.
>
>
>
>     I’d be grateful for advice and practical war stories about what does
>     and does not work.
>
>
>