arxiv:2606.00408
Haoxiang Zhang
IPF
AI & ML interests
None yet
Recent Activity
upvoted a paper 19 minutes ago
Rethinking On-Policy Distillation of Large Language Models II: One Training Example upvoted a paper 30 minutes ago
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning upvoted a paper 31 minutes ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement