20 results for “delayed feedback”
CS papers onlyHybrid search: Keyword + semantic, ranked by combined score.ⓘ
Want pure semantic search? Try claim verification →
Kou Shi, Ziao Zhang, Shiting Huang, Avery Nie +6 more
The paper introduces AsyncTool, a new benchmark designed to evaluate LLM agents' ability to handle multiple, concurrent tasks with delayed tool feedback, demonstrating that asynchronous coordination i…
AlphaTransit introduces a novel search-based planning framework that combines Monte Carlo Tree Search (MCTS) with a neural policy-value network to efficiently design high-quality, city-scale bus trans…
Zihao He, Hongjie Fang, Shirun Tang, Cewu Lu +1 more
The paper proposes LAG-Fusion, a framework for asynchronous multimodal diffusion policy composition with latency-aware guidance fusion.
This paper demonstrates that visual phishing detectors can be completely bypassed by employing simple timing-based attacks that delay the rendering of key webpage elements.
This paper investigates how network perturbations can alter the asymptotic agreement trajectory in distributed coordination systems, proving fragility in standard cooperative output regulation schemes…
The paper proposes FOAM, an adaptive damping method that stabilizes the Shampoo optimization algorithm by dynamically controlling damping and eigendecomposition frequency, thereby reducing staleness-i…
Junwei Ji, Woon-Seng Gan, Boxiang Wang, Ziyi Yang +1 more
This paper proposes an adaptive momentum term for the ASSS-MGDFxLMS algorithm in distributed multichannel active noise control systems to accelerate convergence while maintaining robustness under comm…
This paper measures the impact of procedural skills on LLM agents, distinguishing between improvements and regressions, and identifies causes of regression.
This paper proposes RLAES, a unified language model framework using reinforcement learning for essay scoring and feedback generation, with methods including Rubric-based Feedback Evaluation (RFE), Ada…
This paper conducted a randomized controlled trial on Reddit to test the effectiveness of various deescalation strategies in reducing personal insults using automated replies.
Rahul Khedar, Mayank Malhotra, Avinash Karn, Mouli V +1 more
This paper proposes Rhetor, a multi-agent system that generates rehearsed live demonstrations with segment-synchronized narration and real-time voice question answering for web applications.
This paper proposes two strategies to improve feedback efficiency of reinforcement learning from human feedback (RLHF) in diffusion models.
Proposed an asynchronous block coordinate descent algorithm for distributed trajectory estimation in robotics, reducing communications by up to 96.9% and achieving exponential convergence.
This paper develops a formal economic framework to assess the security of VDF-based randomness beacons, demonstrating that many proposed delays are economically insecure due to rational, profit-motiva…
Haolin He, Renhe Sun, Zheqi Dai, Xingjian Du +15 more
This paper introduces Audio-Dependency Filtering (ADF) pipeline for Audio-Dependent Question Answering (ADQA) task in DCASE~2026, achieving top overall and sub-10B accuracy.