Podlipodcast player Webplayer

Daily Paper Cast

Daily Paper Cast

Post-Training Leaves Behavioral Shadows on Unrelated Decisions

Daily Paper Cast · Sep 29, 2026 · 19:20

0:0019:20

Listen in the Podli app 🎧

Follow your favourite podcasts, listen offline and in the car with CarPlay and Android Auto, and always pick up where you left off. Free to try.

🤗 Upvotes: 46 | cs.CL, cs.AI, cs.LG

Authors:
Ziyang Zhang, Yubin Jing, Yuanhao Zeng, Yuyao Li, Haofan Wang, Yichen Gong

Title:
Post-Training Leaves Behavioral Shadows on Unrelated Decisions

Arxiv:
http://arxiv.org/abs/2609.29233v1

Abstract:
We find that language models can transfer capabilities through task-unrelated text. Post-training typically improves language models using task-specific data. Prior work on subliminal learning shows that information about these updates can pass through unrelated generations, but has largely focused on traits or preferences using extensive teacher outputs. We introduce Active Taskless Distillation (ATD), which achieves capability transfer using only a single word from the teacher per prompt. ATD probes the behavioral shadow of post-training by selecting prompts where the teacher and student's shared public ancestor is nearly indifferent between two ordinary words. A student initialized from this ancestor learns solely from the resulting prompt-word pairs, without target-task examples, teacher logits, or teacher parameters. In the primary coding experiment with Qwen2.5-1.5B, 5,664nses yield a 5.34 pp gain on HumanEval+ over an exact nuisance-matched control thadisrupts prompt-resperiments showtransfer in scientific knowledge, commonsense reasoning, and reading comprehensins across additional model generations, sizes, and families. Functional analyses show that the learned sid composable, andthat its strength tracks the teacher's update strength.

Episodes: Daily Paper Cast

PodliGet the free Podli app
↓ App