ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “delayed feedback”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.AIRecentMay 27, 2026

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

Kou Shi, Ziao Zhang, Shiting Huang, Avery Nie +6 more

The paper introduces AsyncTool, a new benchmark designed to evaluate LLM agents' ability to handle multiple, concurrent tasks with delayed tool feedback, demonstrating that asynchronous coordination i…

View →
cs.AIRecentMay 27, 2026

AlphaTransit: Learning to Design City-scale Transit Routes

Bibek Poudel, Sai Swaminathan, Weizi Li

AlphaTransit introduces a novel search-based planning framework that combines Monte Carlo Tree Search (MCTS) with a neural policy-value network to efficiently design high-quality, city-scale bus trans…

View →
cs.ROcs.AIEmpiricalRecentJul 19, 2026

Asynchronous Multimodal Diffusion Policy Composition via Latency-Aware Guidance Fusion

Zihao He, Hongjie Fang, Shirun Tang, Cewu Lu +1 more

The paper proposes LAG-Fusion, a framework for asynchronous multimodal diffusion policy composition with latency-aware guidance fusion.

View →
cs.CRRecentApr 30, 2026

I can't recognize (yet): Delayed Rendering to Defeat Visual Phishing Detectors

Ying Yuan, Cristiano Alex Rado, Giovanni Apruzzese, Mauro Conti +1 more

This paper demonstrates that visual phishing detectors can be completely bypassed by employing simple timing-based attacks that delay the rendering of key webpage elements.

View →
eess.SYcs.MAmath.OCTheoreticalRecentJul 21, 2026

How network perturbations distort agreement trajectories in LTI multi-agent systems

Gal Barkai, Irinel-Constantin Morărescu

This paper investigates how network perturbations can alter the asymptotic agreement trajectory in distributed coordination systems, proving fragility in standard cooperative output regulation schemes…

View →
cs.LGcs.AIRecentJun 1, 2026

FOAM: Frequency and Operator Error-Based Adaptive Damping Method for Reducing Staleness-Oriented Error for Shampoo

Kyunghun Nam, Sumyeong Ahn

The paper proposes FOAM, an adaptive damping method that stabilizes the Shampoo optimization algorithm by dynamically controlling damping and eigendecomposition frequency, thereby reducing staleness-i…

View →
eess.ASeess.SPEmpiricalRecentJul 19, 2026

Adaptive Momentum Enhanced Distributed Multichannel Active Noise Control for Faster Convergence under Communication Delays

Junwei Ji, Woon-Seng Gan, Boxiang Wang, Ziyi Yang +1 more

This paper proposes an adaptive momentum term for the ASSS-MGDFxLMS algorithm in distributed multichannel active noise control systems to accelerate convergence while maintaining robustness under comm…

View →
cs.AIEmpiricalRecentJul 24, 2026

The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents

Darshan Tank, Baran Nama

This paper measures the impact of procedural skills on LLM agents, distinguishing between improvements and regressions, and identifies causes of regression.

View →
cs.CLcs.AIEmpiricalRecentJul 21, 2026

Beyond Score Prediction: LLM-Based Essay Scoring and Feedback Generation via Reinforcement Learning with Rubric Rewards

Xuefeng Jin, Jiashuo Zhang, Teng Cao, Bin Yang

This paper proposes RLAES, a unified language model framework using reinforcement learning for essay scoring and feedback generation, with methods including Rubric-based Feedback Evaluation (RFE), Ada…

View →
cs.SIcs.HCEmpiricalRecentJun 19, 2026

Reducing the rate of personal insults in social media with bystander bots

Libby Hemphill, Lingyao Li, Ryan Burton, David Jurgens

This paper conducted a randomized controlled trial on Reddit to test the effectiveness of various deescalation strategies in reducing personal insults using automated replies.

View →
cs.AIcs.HCcs.SEEmpiricalRecentJun 29, 2026

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

Rahul Khedar, Mayank Malhotra, Avinash Karn, Mouli V +1 more

This paper proposes Rhetor, a multi-agent system that generates rehearsed live demonstrations with segment-synchronized narration and real-time voice question answering for web applications.

View →
cs.LGcs.AIcs.CVEmpiricalRecentJul 8, 2026

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

Eric Zhu, Abhinav Shrivastava, Soumik Mukhopadhyay

This paper proposes two strategies to improve feedback efficiency of reinforcement learning from human feedback (RLHF) in diffusion models.

View →
cs.ROEmpiricalRecentJul 1, 2026

Technical Report: Asynchronous Distributed Trajectory Estimation of Multi-Robot Systems

Adam Pooley, Matthew Hale

Proposed an asynchronous block coordinate descent algorithm for distributed trajectory estimation in robotics, reducing communications by up to 96.9% and achieving exponential convergence.

View →
cs.CRcs.GTRecentApr 6, 2026

Economic Security of VDF-Based Randomness Beacons: Models, Thresholds, and Design Guidelines

Zhenhang Shang, Kani Chen

This paper develops a formal economic framework to assess the security of VDF-based randomness beacons, demonstrating that many proposed delays are economically insecure due to rational, profit-motiva…

View →
eess.ASEmpiricalRecentJul 21, 2026

Summary of DCASE 2026 Task 5: Audio-Dependent Question Answering

Haolin He, Renhe Sun, Zheqi Dai, Xingjian Du +15 more

This paper introduces Audio-Dependency Filtering (ADF) pipeline for Audio-Dependent Question Answering (ADQA) task in DCASE~2026, achieving top overall and sub-10B accuracy.

View →