PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 19, 20240 citationsOpen Access

DELIA: Diversity-Enhanced Learning for Instruction Adaptation in Large Language Models

View Full Paper
YZYangqianhui ZengRFRen FeiXZXinpeng Zhou

Key Points

Key points are not available for this paper at this time.

Abstract

Although instruction tuning is widely used to adjust behavior in Large Language Models (LLMs), extensive empirical evidence and research indicates that it is primarily a process where the model fits to specific task formats, rather than acquiring new knowledge or capabilities. We propose that this limitation stems from biased features learned during instruction tuning, which differ from ideal task-specfic features, leading to learn less underlying semantics in downstream tasks. However, ideal features are unknown and incalculable, constraining past work to rely on prior knowledge to assist reasoning or training, which limits LLMs' capabilities to the developers' abilities, rather than data-driven scalable learning. In our paper, through our novel data synthesis method, DELIA (Diversity-Enhanced Learning for Instruction Adaptation), we leverage the buffering effect of extensive diverse data in LLMs training to transform biased features in instruction tuning into approximations of ideal features, without explicit prior ideal features. Experiments show DELIA's better performance compared to common instruction tuning and other baselines. It outperforms common instruction tuning by 17.07%-33.41% on Icelandic-English translation bleurt score (WMT-21 dataset, gemma-7b-it) and improves accuracy by 36.1% on formatted text generation (Llama2-7b-chat). Notably, among knowledge injection methods we've known, DELIA uniquely align the internal representations of new special tokens with their prior semantics.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zeng et al. (2024) studied this question.

synapsesocial.com/papers/68e5bd3ab6db6435875550f9https://doi.org/10.48550/arxiv.2408.10841
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1D3: Diversity, Difficulty, and Dependability-Aware Data Selection for Sample-Efficient LLM Instruction Tuning2025 · 3 citations
  2. 2Fine-tuning Large Language Models with Sequential Instructions2024 · 1 citations
  3. 3Building Accurate Translation-Tailored LLMs with Language Aware Instruction Tuning2024 · 3 citations
  4. 4Investigating Multilingual Instruction-Tuning: Do Polyglot Models Demand for Multilingual Instructions?2024
  5. 5G-DIG: Towards Gradient-based Diverse and High-quality Instruction Data Selection for Machine Translation2024 · 1 citations