PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 8, 20260 citationsOpen Access

AUGMANITAI Domain Corpus Batch EAR–ENM — 15 domains × ~150 terms each

View Full Paper
AEAndreas Ehstand

Key Points

  • To create and provide a structured corpus of terms across diverse domains for academic and professional use.
  • Compiled 15 domain corpora covering various fields including sciences, languages, and sustainability.
  • Produced machine-readable JSON files and human-readable PDFs, each with approximately 150 terms.
  • Designed with schema.org compatibility to ensure usability and accessibility.
  • Developed 15 domain-specific corpuses containing a total of 2,250 terms.
  • Ensured compatibility with machine learning and AI applications through structured formats.
  • Facilitated the expansion of resources in educational and professional training settings.

Abstract

Restricted bundle of 15 AUGMANITAI domain corpora (domain codes: EAR, ECL, ECO, EDC, EDL, EDT, EDU, ELC, ELE, EMB, EMP, ENE, ENG, ENL, ENM). Each domain contains a V2-Master JSON file (~150 terms in machine-readable, schema.org-compatible format) plus a 150-term PDF (human-readable presentation). Strict-Wissenschaft, descriptive domain-terminology corpus. Part of the AUGMANITAI multi-domain coverage initiative — the full programme covers 449+ domains spanning academic coaching, sport, music, professions, materials, sciences, languages, sustainability, and more.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Andreas Ehstand (2026) studied this question.

synapsesocial.com/papers/69fd7f0dbfa21ec5bbf0760ehttps://doi.org/10.5281/zenodo.20058412
Ask AI
Helpful
Bookmark
Share
View Full Paper