PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 20226 citationsOpen Access

Computational Historical Linguistics and Language Diversity in South Asia

AAAryaman AroraAFAdam FarrisSBSamopriya Basu

Key Points

Key points are not available for this paper at this time.

Abstract

South Asia is home to a plethora of languages, many of which severely lack access to new language technologies. This linguistic diversity also results in a research environment conducive to the study of comparative, contact, and historical linguistics-fields which necessitate the gathering of extensive data from many languages. We claim that data scatteredness (rather than scarcity) is the primary obstacle in the development of South Asian language technology, and suggest that the study of language history is uniquely aligned with surmounting this obstacle. We review recent developments in and at the intersection of South Asian NLP and historical-comparative linguistics, describing our and others' current efforts in this area. We also offer new strategies towards breaking the data barrier.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Arora et al. (2022) studied this question.

synapsesocial.com/papers/6a1945c6f9a68600c7d95168https://doi.org/10.18653/v1/2022.acl-long.99
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Universal Dependencies v1: A Multilingual Treebank Collection2016 · 1,046 citations
  2. 2Die Burushaski-Sprache von Hunza und Nager1998 · 47 citations
  3. 3Dependency or Span, End-to-End Uniform Semantic Role Labeling2019 · 104 citations
  4. 4Names of Plants in Kalam Kohistani (Pakistan)2018 · 7 citations
  5. 5A multi-representational and multi-layered treebank for Hindi/Urdu2009 · 132 citations