PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 27, 2025Applied Sciences2 citationsOpen Access

Exploring Selective Layer Freezing Strategies in Transformer Fine-Tuning: NLI Classifiers with Sub-3B Parameter Models

View Full Paper
THTaewook HwangHSHyein SeoJJJeesu Jung

Key Points

  • Freezing lower layers in transformer models enhances training speed by 30% while reducing memory usage by 50%.
  • Selective fine-tuning preserves performance comparable to full fine-tuning and Low-Rank Adaptation methods.
  • Experiments on NLI tasks confirm the effectiveness of layer-freezing across small-scale language models with fewer than 3 billion parameters.
  • This approach requires no additional parameters, making it viable for constrained computing resources.

Abstract

In recent years, methods that selectively fine-tune or reduce the number of layers in large language models (LLMs) have garnered attention as an efficient alternative to traditional fine-tuning, where all layers are trained. In this study, we revisit the classical concept of layer freezing and propose a simple, effective strategy that selectively fine-tunes only a portion of transformer layers. We show that freezing the bottom 25% or 50% of layers in small-scale LLMs with sub-3 billion parameters yields significant improvements in memory efficiency and training speed while maintaining, or even surpassing, the performance of full fine-tuning and Low-Rank Adaptation (LoRA). Through experiments on Natural Language Inference (NLI) tasks using LLMs with fewer than 3 billion parameters, our approach achieves up to 50% memory savings and 30% faster training. Notably, our method does not require architectural modifications or additional parameters, making it particularly suitable for resource-constrained environments.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Hwang et al. (2025) studied this question.

synapsesocial.com/papers/68d7be70eebfec0fc523864chttps://doi.org/10.3390/app151910434
Ask AI
Helpful
Bookmark
Share
View Full Paper