Re: seeking links to TEI corpora
"Dalmau, Michelle Denise" <[email protected]>
| Newsgroups | gmane.text.tei.general |
|---|---|
| Message-ID | <[email protected]> |
Dear Matthew, The IU Libraries provide XML downloads (at the item-level) for the following TEI P5 collections: Wright American Fiction: http://dlib.indiana.edu/collections/wright/ Victorian Women Writers Project: http://www.dlib.indiana.edu/collections/vwwp/ Brevier Legislative Reports: http://www.dlib.indiana.edu/collections/law/brevier/ We have two additional projects in TEI P4 with XML download: Indiana Authors and Their Books: http://dlib.indiana.edu/collections/inauthors Indiana Magazine of History: https://scholarworks.iu.edu/journals/index.php/imh (XML download in the View Text link per article) You could also grab most of these files via GitHub: https://github.com/iulibdcs/tei_text (caveat: the repo needs to be refreshed — on our to-do list) This is probably not what you are after, but we provide EAD XML access to IU finding aids as well: http://dlib.indiana.edu/collections/findingaids/ —Michelle ----- Michelle Dalmau Head, Digital Collections Services ----- Indiana University Libraries Herman B Wells Library 1320 East 10th Street, Rm W501 Bloomington, Indiana 47405 ----- Web: http://michelledalmau.com Twitter: @mdalmau On Dec 19, 2016, at 12:13 PM, Lavin, Matthew J <[email protected]<mailto:[email protected]>> wrote: Apologies for any duplicates received due to cross-posting. I am collecting links for publicly accessible, computable TEI (or other similar xml markup such as SGM, LMNL) files. In order to be included, archives/collections/datasets/corpora must have meet one of the two criteria: Bulk download of raw xml (not html transformed) Xml fully accessible via predictable url structure (an example of this would be the Walk Whitman archive, which as a “raw xml” link on every transformed html page) Please note that I am not interested in sample xml, only collections with some kind of curatorial or scholarly focus. Thank you all for any leads! Matthew Lavin Clinical Assistant Professor of English and Director of Digital Media Lab University of Pittsburgh On Dec 19, 2016, at 12:13 PM, Lavin, Matthew J <[email protected]<mailto:[email protected]>> wrote: Apologies for any duplicates received due to cross-posting. I am collecting links for publicly accessible, computable TEI (or other similar xml markup such as SGM, LMNL) files. In order to be included, archives/collections/datasets/corpora must have meet one of the two criteria: Bulk download of raw xml (not html transformed) Xml fully accessible via predictable url structure (an example of this would be the Walk Whitman archive, which as a “raw xml” link on every transformed html page) Please note that I am not interested in sample xml, only collections with some kind of curatorial or scholarly focus. Thank you all for any leads! Matthew Lavin Clinical Assistant Professor of English and Director of Digital Media Lab University of Pittsburgh