TELLME: Test-Enhanced Learning for Language Model Enrichment

📅 2026-08-12
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the challenges of large-scale data acquisition and high computational costs in domain adaptation through continual pretraining by introducing, for the first time, Test-Enhanced Learning (TEL) into the continual pretraining framework. By integrating an embedded quiz mechanism, the proposed approach enhances the model’s efficiency in acquiring domain-specific knowledge and its ability to retain long-term memory. Empirical results demonstrate that this method significantly improves domain adaptation efficiency, achieving up to a 23.6% performance gain on financial-domain tasks and a 9.8% improvement in long-term memory retention. The study thus establishes a novel paradigm for effective and cost-efficient domain adaptation.
📝 Abstract
Continual pre-training (CPT) has been widely adopted as a method for domain adaptation in large language models. However, CPT has consistently been accompanied by challenges, such as the difficulty of acquiring large-scale domain-specific datasets and high computational costs. In this study, we propose a novel method called Test-Enhanced Learning for Language Model Enrichment (TELLME) to alleviate these issues. TELLME leverages the TestEnhanced Learning (TEL) principle, whereby the model's training efficiency is improved using quizzes during training. It integrates this principle with CPT, thereby promoting efficient domain-specific knowledge acquisition and long-term memory retention. Experimental results demonstrate that TELLME outperforms existing methods by up to 23.6% in the financial domain and achieves a 9.8% improvement in long-term memory retention.
Problem

Research questions and friction points this paper is trying to address.

continual pre-training
domain adaptation
large language models
computational cost
domain-specific data
Innovation

Methods, ideas, or system contributions that make the work stand out.

Test-Enhanced Learning
Continual Pre-training
Domain Adaptation
Long-term Memory Retention
Language Model Enrichment
🔎 Similar Papers
No similar papers found.
M
Minjun Kim
Korea Advanced Institute of Science and Technology
I
Inho Won
Korea Advanced Institute of Science and Technology
H
Hyeonseok Lim
Korea Advanced Institute of Science and Technology
M
MinKyu Kim
Seoul National University of Science and Technology
J
Junghun Yuk
Korea Advanced Institute of Science and Technology
W
Wooyoung Go
National Security Research Institute
Jongyoul Park
Jongyoul Park
Seoul National University of Science and Technology
Machine LearningDeep LearningApplied AI
J
Jungyeul Park
Korea Advanced Institute of Science and Technology
KyungTae Lim
KyungTae Lim
École normale supérieure
Natural Language Processing