A lot of content has been produced with crowdsourcing by Wikimedia users, not only on Wikipedia. Hundreds of thousands of photographs of cultural heritage were uploaded in the last 20 years and are available on Wikimedia Commons, still not linked to structured data on Wikidata, often of unknown historical buildings, like ruined churches in Southern Italy. We will demonstrate that is possible to use data mining techniques on Wikimedia projects and build on Wikidata a more comprehensive open catalogue of Cultural heritage in Italy from crowdsourced contents. We will present the result of phase 1 and 2 of this project, where OpenRefine was used to create thousands of new items on Wikidata about "lost heritage" in Italy, and discuss of the possible use of AI to speed-up the process.
Wikimedia Italia, Italy - ORCID: 0000-0001-7446-8800
Titolo del capitolo
Hunting for Lost Heritage on Wikimedia Commons and Wikidata. Un workflow per scovare i monumenti italiani
Autori
Marco Chemello
Lingua
Italiano
DOI
10.36253/979-12-215-1010-2.29
Opera sottoposta a peer review
Anno di pubblicazione
2026
Copyright
© 2026 Author(s)
Licenza d'uso
Licenza dei metadati
Titolo del libro
Wikidata e la ricerca. Condividere esperienze / Wikidata and Research. Sharing Experiences
Sottotitolo del libro
Atti del Convegno, Firenze, 5-6 giugno 2025 / Conference proceedings, Florence, 5-6 June 2025
Curatori
Elena Marangoni, Camillo Carlo Pellizzari di San Girolamo
Opera sottoposta a peer review
Numero di pagine
378
Anno di pubblicazione
2026
Copyright
© 2026 Author(s)
Licenza d'uso
Licenza dei metadati
Editore
Firenze University Press
DOI
10.36253/979-12-215-1010-2
ISBN Print
979-12-215-1009-6
eISBN (pdf)
979-12-215-1010-2
Collana
Studi e saggi
ISSN della collana
2704-6478
e-ISSN della collana
2704-5919