Large-scale access to Commons thumbnails

Stefanie Schneider via Wikitech-l <[email protected]>
Newsgroups gmane.science.linguistics.wikipedia.technical
Message-ID <CAD11UvSiRcBQpfbwC5pQJLDX9e52c_XHOYUUc5Z66R6TSzLiTw@mail.gmail.com>
Dear all,

We are preparing an academic research project involving the large-scale
analysis of art-historical material from Wikimedia Commons, conducted
jointly by the University of Marburg and the Getty Research Institute.

For this project, we anticipate requiring access to at least five million
Commons images. High-resolution originals are not necessary; thumbnails
would suffice.

Before we initiate any large-scale retrieval, we would be grateful if you
could advise us on the preferred way to access this material. In
particular, we would be grateful for advice on whether an existing bulk
dataset provides Commons thumbnails on this scale, or whether a project of
this size should make use of, or request, any special API arrangements. As
we can identify and filter the relevant image URLs from the metadata in
advance, the main question concerns the recommended method for retrieving
the images themselves.

If you have any more questions about the project, please do not hesitate to
get in touch.

Thank you, and all the best,

Stefanie
__________________________

Dr. Stefanie Schneider
Wissenschaftliche Mitarbeiterin

Philipps-Universität Marburg
Kunstgeschichtliches Institut
Biegenstraße 11
35037 Marburg

Raum: 00003
Telefon: +49 6421 28-22174
E-Mail: [email protected]

https://www.uni-marburg.de/de/fb09/khi/dr-stefanie-schneider

_______________________________________________
Wikitech-l mailing list -- [email protected]
To unsubscribe send an email to [email protected]
https://lists.wikimedia.org/postorius/lists/wikitech-l.lists.wikimedia.org/
lmpx.com only provides a reader for public news (NNTP) servers. It is not affiliated with the servers or forums shown here and is not responsible for the content of articles, which is written by their respective authors.