Zihan Zhang
3 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
This paper introduces the Relay Tampering Attack (RTA), demonstrating that malicious third-party relays can undermine the security of LLM agents by modifying responses post-alignment, even if the LLM itself is perfectly aligned.
The paper introduces CultureForest, a new benchmark for evaluating Cultural Norm Grounded Reasoning in LLMs, demonstrating that models struggle to apply their cultural knowledge effectively in realistic, open-ended scenarios.
This paper proposes X$^3$-OPD, a framework for transferring reasoning capabilities from text-based models to audio-language models using on-policy distillation.
Papers
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment
Dongjie Fu, Di Cao, Xize Cheng, Zihan Zhang +5 more
This paper proposes X$^3$-OPD, a framework for transferring reasoning capabilities from text-based models to audio-language models using on-policy distillation.