Zhuokai Zhao
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
OmniOPD introduces a logit-free, chunk-level distillation framework that improves on standard On-Policy Distillation by using semantic similarity and peak-entropy scheduling, achieving state-of-the-art performance even with black-box teachers.
The paper introduces a memory agent to improve decision-making in long-horizon tasks by actively updating and intervening with reminders.
Papers
Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents
Yifan Wu, Lizhu Zhang, Yuhang Zhou, Mingyi Wang +4 more
The paper introduces a memory agent to improve decision-making in long-horizon tasks by actively updating and intervening with reminders.