Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Haolong Jia

Haolong Jia

3 indexed papers

Recent (6 mo)
3
With code
0
Influential cites
0
Benchmarked
0

Publications per year

3
26

Top categories

Info Retrieval×1NLP×1Multiagent×1Distributed×1ML×1AI×1

Frequent co-authors

Jijun Chi1×
Zhenghan Tai1×
Hanwei Wu1×
Tung Sum Thomas Kwok1×
Hailin He1×
Zixing Liao1×

Research Timeline

2026
PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

The paper proposes Predictive Routing Replay (PR2) to stabilize reinforcement learning on Mixture of Experts (MoE) LLMs by predicting and incorporating short-horizon router evolution during training and rollout.

DPIFrame: A Dual-Level Parallelism Acceleration Framework for CTR Model Inference

This paper proposes DPIFrame, a dual parallelizable framework for accelerating Click-through rate (CTR) model inference on GPU, achieving state-of-the-art inference performance with significant speedups.

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering

This paper proposes FinSAgent, an evidence-grounded multi-agent framework for financial question answering over SEC filings, which improves retrieval coverage and answer correctness through corpus-side conditioning.

Highlighted terms show continued research focus across papers

Papers

cs.IRcs.CLcs.MAEmpiricalRecentJul 20, 2026

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering

Jijun Chi, Zhenghan Tai, Hanwei Wu, Tung Sum Thomas Kwok +19 more

This paper proposes FinSAgent, an evidence-grounded multi-agent framework for financial question answering over SEC filings, which improves retrieval coverage and answer correctness through corpus-sid…

View →
cs.DCEmpirical
Recent
Jun 19, 2026

DPIFrame: A Dual-Level Parallelism Acceleration Framework for CTR Model Inference

Dezhi Yi, Huifeng Guo, Kunpeng Xie, Zhaolong Jian +5 more

This paper proposes DPIFrame, a dual parallelizable framework for accelerating Click-through rate (CTR) model inference on GPU, achieving state-of-the-art inference performance with significant speedu…

View →
cs.LGcs.AIRecentMay 29, 2026

PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

Daize Dong, Junlin Chen, Haolong Jia, Jiawei Wu +8 more

The paper proposes Predictive Routing Replay (PR2) to stabilize reinforcement learning on Mixture of Experts (MoE) LLMs by predicting and incorporating short-horizon router evolution during training a…

View →