Zizhuang Deng
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
PlanGuard is a training-free defense framework that uses an isolated Planner and hierarchical verification to defend LLM agents against Indirect Prompt Injection by verifying the consistency of planned actions.
The paper introduces FlipGuard, a proactive defense framework against Quantization-Conditioned Backdoor (QCB) attacks in Large Language Models (LLMs), achieving high security with negligible performance degradation.
Papers
FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks
The paper introduces FlipGuard, a proactive defense framework against Quantization-Conditioned Backdoor (QCB) attacks in Large Language Models (LLMs), achieving high security with negligible performan…