TED: Training-Free Experience Distillation for Multimodal Reasoning

Shuozhi Yuan; Jinqing Wang; Zihao Liu; Miaomiao Yuan; Haoran Peng; Jin Zhao; Bingwen Wang; Haoyi Wang

arXiv:2603.26778·cs.LG·March 31, 2026

TED: Training-Free Experience Distillation for Multimodal Reasoning

Shuozhi Yuan, Jinqing Wang, Zihao Liu, Miaomiao Yuan, Haoran Peng, Jin Zhao, Bingwen Wang, Haoyi Wang

PDF

TL;DR

TED introduces a training-free, context-based knowledge distillation method that enhances multimodal reasoning models by injecting and refining reasoning experiences directly into prompts, reducing training costs.

Contribution

It proposes a novel training-free distillation framework that transfers knowledge via in-context experiences, addressing resource constraints and unbounded experience growth.

Findings

01

TED improves multimodal reasoning benchmarks significantly.

02

It achieves competitive performance with less than 100 training samples.

03

Reduces training cost by over 5x compared to traditional methods.

Abstract

Knowledge distillation is typically realized by transferring a teacher model's knowledge into a student's parameters through supervised or reinforcement-based optimization. While effective, such approaches require repeated parameter updates and large-scale training data, limiting their applicability in resource-constrained environments. In this work, we propose TED, a training-free, context-based distillation framework that shifts the update target of distillation from model parameters to an in-context experience injected into the student's prompt. For each input, the student generates multiple reasoning trajectories, while a teacher independently produces its own solution. The teacher then compares the student trajectories with its reasoning and the ground-truth answer, extracting generalized experiences that capture effective reasoning patterns. These experiences are continuously…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.