Kavosh Asadi
2 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
The paper introduces Off-Context GRPO, a method for reinforcement learning with verifiable rewards that uses guided rollouts and importance-corrected objectives to improve reasoning in large language models.
This paper adapts Large Language Models as semantic representation backbones in a two-tower retrieval architecture for high-throughput, large-scale recommendation systems.
Papers
The Case Against Generation for Retrieval: Discriminative Language Models as Effective Retrievers
Zhe Xu, Prachi Agrawal, Kavosh Asadi, Tianyi Chen +16 more
This paper adapts Large Language Models as semantic representation backbones in a two-tower retrieval architecture for high-throughput, large-scale recommendation systems.