Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Uzair Ahmed

Uzair Ahmed

1 indexed paper

Recent (6 mo)
1
With code
0
Influential cites
0
Benchmarked
0

Publications per year

1
26

Top categories

NLP×1

Frequent co-authors

Quang Minh Nguyen1×
Taegyoon Kim1×

Research Timeline

2026
Can LLMs Reliably Self-Report Adversarial Prefills, and How?

This paper investigates how reliable large language models are in recognizing their own compromised outputs during adversarial prefill attacks in safety contexts.

Highlighted terms show continued research focus across papers

Papers

cs.CLEmpiricalRecentJun 22, 2026

Can LLMs Reliably Self-Report Adversarial Prefills, and How?

Quang Minh Nguyen, Uzair Ahmed, Taegyoon Kim

This paper investigates how reliable large language models are in recognizing their own compromised outputs during adversarial prefill attacks in safety contexts.

View →