PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 23, 20260 citationsOpen Access

Introduction to HistText: An Application for Exploring Multilingual Text Corpora

View Full Paper
CACécile ArmandCHChristian Henriot

Key Points

  • The workshop aims to introduce HistText for exploring multilingual text corpora and conducting computational analysis.
  • Hands-on demonstrations of HistText's integrated environment for query design and corpus construction.
  • Participants formulate search strategies, build a relevant document corpus, and transform text into analyzable data.
  • Discussion of the text-mining pipeline logic and the importance of documentation for reproducibility.
  • HistText allows researchers to conduct text analysis without requiring advanced programming skills.
  • The integrated workflow enhances transparency and reproducibility in humanities research.
  • Workshop participants gain hands-on experience in developing search strategies and datafication techniques.

Abstract

This hands-on workshop introduces HistText, a researcher-oriented application designed for the exploration, datafication, and analysis of large multilingual text corpora. Developed through a close collaboration between historians and computer scientists, HistText provides an integrated environment in which scholars can move step by step from query design and corpus construction to entity extraction, data transformation, and visualization. Unlike tools that focus on only one stage of the workflow, HistText brings these operations together in a single environment oriented toward transparent and reproducible humanities research. The workshop is designed for humanities researchers who wish to engage in computational text analysis without needing advanced programming skills. It will demonstrate how HistText structures research into a transparent and replicable workflow: participants begin with a research question, identify suitable collections, formulate and refine search strategies, build a corpus of relevant documents, and then transform textual materials into analyzable data. The workshop will also introduce the logic of the text-mining pipeline underlying these operations and discuss its strengths and limits when applied to historical materials. Particular attention will be paid to the importance of documenting analytical choices at every stage, so that the research process remains explicit, accountable, and reproducible.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Armand et al. (2026) studied this question.

synapsesocial.com/papers/69e9b9e385696592c86ec657https://doi.org/10.5281/zenodo.19645859
Ask AI
Helpful
Bookmark
Share
View Full Paper