PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 24, 20248 citationsOpen Access

Small Language Model Can Self-Correct

View Full Paper
HHHaixia HanJLJiaqing LiangJSJie Shi

Key Points

Key points are not available for this paper at this time.

Abstract

Generative Language Models (LMs) such as ChatGPT have exhibited remarkable performance across various downstream tasks. Nevertheless, one of their most prominent drawbacks is generating inaccurate or false information with a confident tone. Previous studies have devised sophisticated pipelines and prompts to induce large LMs to exhibit the capability for self-correction. However, large LMs are explicitly prompted to verify and modify their answers separately rather than completing all steps spontaneously like humans. Moreover, these complex prompts are extremely challenging for small LMs to follow. In this paper, we introduce the Intrinsic Self-Correction (ISC) in generative language models, aiming to correct the initial output of LMs in a self-triggered manner, even for those small LMs with 6 billion parameters. Specifically, we devise a pipeline for constructing self-correction data and propose Partial Answer Masking (PAM), aiming to endow the model with the capability for intrinsic self-correction through fine-tuning. We conduct experiments using LMs with parameters sizes ranging from 6 billion to 13 billion in two tasks, including commonsense reasoning and factual knowledge reasoning. Our experiments demonstrate that the outputs generated using ISC outperform those generated without self-correction. We believe that the output quality of even small LMs can be further improved by empowering them with the ability to intrinsic self-correct.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Han et al. (2024) studied this question.

synapsesocial.com/papers/68e72968b6db6435876a368bhttps://doi.org/10.1609/aaai.v38i16.29774
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Large Language Models have Intrinsic Self-Correction Ability2024 · 2 citations
  2. 2Small Language Models Need Strong Verifiers to Self-Correct Reasoning2024
  3. 3Large Language Models Can Self-Correct with Minimal Effort2024 · 1 citations
  4. 4Confidence Matters: Revisiting Intrinsic Self-Correction Capabilities of Large Language Models2024 · 3 citations
  5. 5Self-Taught Self-Correction for Small Language Models2025