Xunguang Wang
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
This paper introduces a novel framework, the Reasoning Safety Monitor, to detect and prevent logical inconsistencies and adversarial manipulations within the internal reasoning steps of large language models, establishing reasoning safety as a critical security dimension.
This paper reveals a denial-of-service vulnerability in LLM-based guardrails for autonomous agents and proposes two attack frameworks.
Papers
From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails
Yuguang Zhou, Xunguang Wang, Pingchuan Ma, Zhantong Xue +2 more
This paper reveals a denial-of-service vulnerability in LLM-based guardrails for autonomous agents and proposes two attack frameworks.