Continually Post-training Small Models via Synthetic Self-distillationSep 1, 2026ยท[P2] Woosung Koh, Hyunsoo Lee, Kyungjae Lee, Haeju Park, Dahyun Lee, Dasol Hwang, Moontae Leeยท 0 min read PDFTypeConference paperPublicationPre-printLast updated on Sep 12, 2026Continual Learning Self-Distillation Fine-Tuning Small Language Model Language Models Can Control Their Own Attention May 22, 2026 →