PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 2026Artificial Intelligence Review5 citationsOpen Access

A survey on large language models with multilingualism: recent advances and new frontiers

KHKaiyu HuangBeijing Jiaotong UniversityFMFengran MoUniversité de MontréalXZXinyu ZhangUniversity of Waterloo

Key Points

  • The central aim is to summarize recent advancements and challenges in multilingual applications of large language models.
  • Conducted a comprehensive survey summarizing existing research on multilingual large language models.
  • Explored various aspects like training methods, model safety, and dataset usage.
  • Discussed challenges and proposed potential solutions for multilingualism in LLMs.
  • Highlighted the current state of multilingual capabilities in LLMs and identified key challenges.
  • Provided insights on the importance of language-fair technology and its usability.
  • Outlined future research directions for enhancing multilingualism in LLMs.

Abstract

The rapid development of Large Language Models (LLMs) demonstrates remarkable multilingual capabilities in natural language processing, attracting global attention in both academia and industry. To mitigate potential discrimination and enhance the overall usability and accessibility for diverse language user groups, it is important for the development of language-fair technology. Despite the breakthroughs of LLMs, the investigation into the multilingual scenario remains insufficient, where a comprehensive survey to summarize recent approaches, developments, limitations, and potential solutions is desirable. To this end, we provide a survey with multiple perspectives on the utilization of LLMs in the multilingual scenario. We first rethink the transitions between previous and current research on pre-trained language models. Then we introduce several perspectives on the multilingualism of LLMs, including training and inference methods, model safety, multi-domain with language culture, and usage of datasets. We also discuss the major challenges that arise in these aspects, along with possible solutions. Besides, we highlight future research directions that aim at further enhancing LLMs with multilingualism. The survey aims to help the research community address multilingual problems and provide a comprehensive understanding of the core concepts, key techniques, and latest developments in multilingual natural language processing based on LLMs. z

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Huang et al. (2026) studied this question.

synapsesocial.com/papers/69c8c336de0f0f753b39dd99https://doi.org/10.1007/s10462-026-11534-5
Ask AI
Helpful
Bookmark
Share
View Full Paper