Leong Hien Poh
1 indexed paper
Recent (6 mo)
1With code
0Influential cites
0Benchmarked
0Publications per year
126
Top categories
Software Eng.×1AI×1NLP×1ML×1
Frequent co-authors
Research Timeline
2026
Reinforcement learning to improve large language model-based automated code compliance systems
This paper presents P4IR, a two-stage framework for automated code compliance using supervised fine-tuning and Group Relative Policy Optimization, achieving significant improvements over baselines and leading LLMs.
Highlighted terms show continued research focus across papers
Papers
cs.SEcs.AIcs.CLEmpiricalRecentJun 21, 2026
Reinforcement learning to improve large language model-based automated code compliance systems
Jack Wei Lun Shi, Minghao Dang, Wawan Solihin, Leong Hien Poh +1 more
This paper presents P4IR, a two-stage framework for automated code compliance using supervised fine-tuning and Group Relative Policy Optimization, achieving significant improvements over baselines and…
View →