Real Time Detection and Quantitative Analysis of Spurious Forgetting in Continual Learning
arxiv.org·2d
💻Local LLMs
Preview
Report Post

View PDF HTML (experimental)

Abstract:Catastrophic forgetting remains a fundamental challenge in continual learning for large language models. Recent work revealed that performance degradation may stem from spurious forgetting caused by task alignment disruption rather than true knowledge loss. However, this work only qualitatively describes alignment, relies on post-hoc analysis, and lacks automatic distinction mechanisms. We introduce the shallow versus deep alignment framework, providing the first quantitative characterization of alignment depth. We identify that current task alignment approaches suffer from shallow alignment - maintained only over the first few output tokens (approximately 3-5) - ma…

Similar Posts

Loading similar posts...