Juanzi Li

3 indexed papers

Recent (6 mo)

With code

Influential cites

Benchmarked

Publications per year

Top categories

AI×3NLP×3ML×2

Frequent co-authors

Lei Hou2×

Amy Xin1×

Jiening Siow1×

Junjie Wang1×

Zijun Yao1×

Fanjin Zhang1×

Research Timeline

2026

LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards

LongTraceRL addresses long-context reasoning challenges by generating highly challenging training data and introducing a fine-grained rubric reward, significantly improving evidence-grounded reasoning in LLMs.

Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning

This paper introduces CHERRL, a controllable hacking environment for rubric-based reinforcement learning to study and mitigate reward hacking.

EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery

This paper presents EurekAgent, an environment-engineered agent system for metric-driven autonomous scientific discovery.

Highlighted terms show continued research focus across papers

Papers

cs.AIcs.CLEmpiricalRecentJun 11, 2026

EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery

Amy Xin, Jiening Siow, Junjie Wang, Zijun Yao +4 more

This paper presents EurekAgent, an environment-engineered agent system for metric-driven autonomous scientific discovery.

View →

cs.LGcs.AIcs.CLRecent