Tharindu Kumarage

2 indexed papers

Recent (6 mo)

With code

Influential cites

Benchmarked

Publications per year

Top categories

AI×2Crypto×1ML×1

Frequent co-authors

Charith Peris2×

Rahul Gupta2×

Swastik Roy1×

Rajkumar Pujari1×

Anna Rumshisky1×

Pradeep Natarajan1×

Research Timeline

2026

ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System

ARES is a novel framework that systematically discovers and mitigates dual vulnerabilities in RLHF systems by simultaneously testing the core LLM and its Reward Model (RM) using structured adversarial prompts, leading to enhanced safety robustness.

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges

PReMISE introduces a framework to audit and improve the quality of rubrics used to guide LLM judges, demonstrating that it can significantly increase judge accuracy and reduce the exploitability of responses.

Highlighted terms show continued research focus across papers

Papers

cs.AIRecentMay 29, 2026

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges

Swastik Roy, Rajkumar Pujari, Tharindu Kumarage, Charith Peris +4 more

View →

cs.AIcs.CRcs.LGRecentApr 20, 2026