Analysis and Mitigation of Performance Degradation from Layer Insertions into the Middle of Pre-Trained Language Models

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

Full fine-tuning of pre-trained models sometimes requires inserting trainable layers into the middle of a pre-trained backbone, but such middle-layer insertion can severely degrade downstream performance. We hypothesize that this degradation arises because conventionally inserted layers, when randomly initialized and combined with output-side activation, perturb intermediate representations before the pre-trained model has adapted. We study this phenomenon across natural language processing and computer vision benchmarks by varying insertion locations, the number of inserted layers, and activation designs. To address this problem, we propose a practical stabilization method for middle-layer insertion under full fine-tuning: a bias-free inserted layer with unit initialization and weight-side activation. This design is intended to remain closer to an identity-like transformation at initialization, thereby reducing initialization-time perturbation rather than claiming exact preservation of the original representations. In the tested DeBERTa-v3, T5-base, and ViT-base settings, the proposed method substantially mitigates the severe degradation caused by naive middle-layer insertion and maintains performance close to the no-added-layer baseline, including settings with up to 24 inserted layers.

키워드

fine-tuningtransfer learningweight initializationactivation functiontransformerlanguage model
제목
Analysis and Mitigation of Performance Degradation from Layer Insertions into the Middle of Pre-Trained Language Models
저자
Kim, GyunyeopKang, Sangwoo
DOI
10.3390/math14081382
발행일
2026-04
유형
Article
저널명
MATHEMATICS
14
8