Publications

(2026). Fewer Tokens, More Self-Teaching: On-Policy Self-Distillation for Extreme Visual Token Reduction. arXiv 2026.
(2026). ECHO: Continuous Hierarchical Memory for Vision-Language-Action Models. NeurIPS 2026.
(2026). "The Whole Is Greater Than the Sum of Its Parts": A Compatibility-Aware Multi-Teacher CoT Distillation Framework. IJCAI 2026.
(2026). MIND: From Passive Mimicry to Active Reasoning through Capability-Aware Multi-Perspective CoT Distillation. ACL 2026.