Ruiqing Yue
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
The paper introduces extsc{Spore}, a novel, training-free, and highly efficient privacy extraction attack that targets sensitive information stored in the memory of LLM agents during inference, outperforming existing state-of-the-art methods.
This paper explores safety risks in humorization of large language models (LLMs) and introduces HumorSafe framework for evaluating latent safety risks.
Papers
Refusal is Not Safety! Benchmarking Latent Safety Risks of LLM-Driven Content Humorization
Yu Cui, Ruiqing Yue, Tingyu Li, Sicheng Pan +5 more
This paper explores safety risks in humorization of large language models (LLMs) and introduces HumorSafe framework for evaluating latent safety risks.