PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 19, 20240 citationsOpen Access

Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages

View Full Paper
YZYuanchi ZhangYWYile WangZLZijun Liu

Key Points

Key points are not available for this paper at this time.

Abstract

While large language models (LLMs) have been pre-trained on multilingual corpora, their performance still lags behind in most languages compared to a few resource-rich languages. One common approach to mitigate this issue is to translate training data from resource-rich languages into other languages and then continue training. However, using the data obtained solely relying on translation while ignoring the original capabilities of LLMs across languages is not always effective, which we show will limit the performance of cross-lingual knowledge transfer. In this work, we propose SDRRL, a method based on Self-Distillation from Resource-Rich Languages that effectively improve multilingual performance by leveraging the internal capabilities of LLMs on resource-rich languages. We evaluate on different LLMs (LLaMA-2 and SeaLLM) and source languages across various comprehension and generation tasks, experimental results demonstrate that SDRRL can significantly enhance multilingual capabilities while minimizing the impact on original performance in resource-rich languages.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhang et al. (2024) studied this question.

synapsesocial.com/papers/68e78a60b6db6435876fcd5ahttps://doi.org/10.48550/arxiv.2402.12204
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages2024
  2. 2Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning2024
  3. 3Large Language Models are Good Spontaneous Multilingual Learners: Is the Multilingual Annotated Data Necessary?2024
  4. 4Bridging the Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs2024
  5. 5Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean2024 · 1 citations