Kai Zheng
3 indexed papers
Publications per year
Top categories
Frequent co-authors
Research Timeline
This paper proposes VeriEvol, a framework for scaling reinforcement learning for visual mathematical reasoning by decoupling prompt difficulty and answer reliability, and verifying data construction.
The paper presents GIFT, a method for reducing communication volume in large language model pretraining by transforming gradients into a near-isotropic space before quantization.
The paper introduces DBA-Bench, a benchmark for evaluating database agents with production fidelity, outcome-first evaluation, and controlled scenario reproducibility.
Papers
DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents
Junming Chen, Junyang Jiang, Xu Chen, Zibo Liang +1 more
The paper introduces DBA-Bench, a benchmark for evaluating database agents with production fidelity, outcome-first evaluation, and controlled scenario reproducibility.