PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 3, 2026Digital Scholarship in the Humanities0 citations

Determining whole–part relationships when cataloging nineteenth-century monographs

View Full Paper
LCL de CamposUniversidad de GranadaJFJuan M. Fernández‐LunaUniversidad de GranadaJHJuan F. Huete

Key Points

  • The aim is to develop a methodology for identifying whole-part relationships in the cataloging of digital libraries, focusing on nineteenth-century works.
  • Developed a methodology utilizing Optical Character Recognition (OCR) for text extraction.
  • Created the Spanish19BNE dataset with approximately 12,000 books from the Biblioteca Nacional de España.
  • Introduced an online navigation tool for analyzing authors' works.
  • The proposed methodology successfully identifies whole-part relationships, improving cataloging efficiency.
  • Experimental results from the Spanish19BNE dataset demonstrate the tool's viability for digital library research.
  • The dataset includes works from 559 authors, facilitating extensive computational analysis.

Abstract

Abstract This article explores the process of determining whole–part relationships in cataloging digital libraries. In contemporary library practices, accurately determining how an entity or work relates to others in the collection is considered an essential task for effective information organization and retrieval. The “Biblioteca Nacional de España” (Spanish National Library) has made significant efforts to achieve this objective, but the results are far from optimal due to the extensive human efforts required. We present a methodology capable of automatically identifying whole–part relationships and propose an online navigation tool for analyzing the structure of authors’ works by users and researchers. To facilitate easy adoption and widespread use, our methodology directly utilizes text extracted by Optical Character Recognition (OCR) technologies, eliminating the need for further preprocessing. Additionally, we introduce and provide access to Spanish19BNE, a dataset containing approximately 12,000 books in Spanish written by 559 different authors, with their original manifestations held in the Biblioteca Nacional de España collection. This dataset can be used for digital library research and computational analysis. Experimental results on this dataset demonstrate the viability of the proposed approach.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Campos et al. (2026) studied this question.

synapsesocial.com/papers/69f6e5618071d4f1bdfc61dfhttps://doi.org/10.1093/llc/fqag055
Ask AI
Helpful
Bookmark
Share
View Full Paper