arxiv:2606.00408
Haoxiang Zhang
IPF
AI & ML interests
None yet
Recent Activity
upvoted a paper about 14 hours ago
Rethinking On-Policy Distillation of Large Language Models II: One Training Example upvoted a paper about 14 hours ago
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning upvoted a paper about 14 hours ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement