ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “latency variations”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.NIEmpiricalRecentJul 28, 2026

Round Trip Time: A Benign Signal or an Indirect Window into Datacenter Workloads?

Sourya Saha, Md Nurul Absur, Saptarshi Debroy

This paper investigates a network side-channel vulnerability in multi-tenant datacenter fabrics caused by shared congestion behavior, achieving up to 97.3% run-level accuracy in workload inference.

View →
cs.LGcs.NIEmpiricalRecentJun 28, 2026

Deciphering Region-Level Signatures from Latency Measurements in LEO Satellite Internet

Xiang Shi, Yifei Zhang, Peng Hu

This paper proposes a hierarchical analytical framework to characterize region-level latency differences in Low-Earth orbit satellite Internet using Starlink RTT measurements.

View →
cs.NITheoreticalRecentJun 30, 2026

Latency-Sensitive 5G RAN Slicing for Deterministic Aperiodic Traffic in Smart Manufacturing

M. Carmen Lucas-Estañ, Jan García-Morales, Javier Gozalvez

This paper proposes a new approach for designing Radio Access Network (RAN) slices in 5G and beyond networks using descriptors that consider both transmission rate and latency requirements to support…

View →
cs.NIEmpiricalRecentJul 24, 2026

CAPS: Fine-Tuning CCA Timing

Raphael Zailer, Isaac Keslassy

This paper proposes CAPS, a scheduling layer for data centers that separates rate computation and packet scheduling, reducing queue occupancy by up to 10x without throughput loss.

View →
cs.DCEmpiricalRecentJul 17, 2026

Every Microsecond Matters: Achieving Near Speed-of-Light Latency in GPU Collectives

Siyuan Shen, Anton Korzh, John Bachan, Tiancheng Chen +9 more

This paper explores methods to reduce latency in GPU collective communications for large language model inference, achieving near-optimal designs with barrier-free synchronization and efficient use of…

View →
cs.ROcs.AIcs.LGRecentMay 27, 2026

Multi-Resolution End-to-End Deep Neural Network for Optimizing Latency-Accuracy Tradeoff in Autonomous Driving

Qitao Weng, Heechul Yun

The paper proposes a multi-resolution end-to-end deep neural network for autonomous driving that dynamically adjusts input resolution to optimize the critical tradeoff between prediction accuracy and…

View →
cs.NIEmpiricalRecentJun 26, 2026

V-TSN: A Software-Defined TSN Overlay for General-Purpose Networks

Mohammadparsa Karimi, Majid Nabi, Ahmed Khalaf, Andrew Nelson +2 more

This paper introduces Virtual Time-Sensitive Networking (V-TSN), a software-defined overlay for gPTP-based synchronization and TSN traffic shaping over general-purpose networks without specialized har…

View →
cs.NIcs.PFEmpiricalRecentJul 4, 2026

Evaluating 5G-connected IoT for Power Line Temperature Prediction: Real-World Latency and Cost Trade-offs Between MEC and Cloud

Aakash Sharma, Sigmund Akselsen, Anders Andersen, Lars Ailo Bongo +1 more

This paper investigates the latency performance of Mobile Edge Computing (MEC) on a 5G cellular network for real-time power transmission line analytics, demonstrating a low latency of 44.62 ms compare…

View →
cs.AREmpiricalRecentJul 22, 2026

Revisiting Hardware Priority Queue Architectures

Qihang Wu, Austin Rovinski

The paper implements and evaluates several hardware priority queue architectures on modern FPGA platforms and provides a quantitative analysis.

View →
cs.NIEmpiricalRecentJul 17, 2026

App-Based Performance Characterization of Cellular and Wi-Fi Networks in Dense Stadium Deployments

Hardani Ismu Nabil, Muhammad Iqbal Rochman, S. M. Haider Ali Shuvo, Joshua Roy Palathinkal +1 more

This paper evaluates user-perceived performance and QoE of wireless networks at Notre Dame Stadium during football games using commercial smartphones for web browsing, WhatsApp messaging, and Instagra…

View →
cs.CRcs.DCTheoreticalRecentJun 16, 2026

Gatling: Rapid-Fire Consensus from Parallel Composition

Giulia Scaffino, Max Resnick, Joachim Neu

Gatling is a new atomic broadcast protocol that achieves arbitrarily small inter-proposal times, even smaller than the network delay, by running multiple parallel instances of a black-box atomic broad…

View →
cs.NIEmpiricalRecentJul 8, 2026

Unveiling TCP BBR Dominance in Starlink Internet: Experimental Insights and Analysis

Rakshitha De Silva, Shiva Raj Pokhrel, Jonathan Kua

This paper compares Google's BBR-v3 Congestion Control Algorithm to eight others over SpaceX's Starlink network, demonstrating its fairness and throughput maximization in high-latency, variable satell…

View →
eess.SPcs.ITTheoreticalRecentJun 12, 2026

Repeater-Assisted Massive MIMO Downlink Performance with Calibration Errors

Kohei Ueda, Anubhab Chowdhury, Koji Ishibashi, Erik G. Larsson

This paper analyzes the effects of calibration errors on downlink beamforming in a repeater-assisted massive MIMO system and derives analytical expressions for the downlink spectral efficiency.

View →
cs.NIcs.DCcs.LGEmpiricalRecentJul 28, 2026

Incast-Free MoE Rate-Based Scheduling

Evyatar Cohen, Jose Yallouz, Alexander Shpiner, Mark Silberstein +2 more

This paper proposes a proactive fair scheduling framework to prevent fabric oversubscription and eliminate incast in Mixture of Experts (MoE) architectures, demonstrating consistent link utilization a…

View →
cs.CLcs.LGcs.SEEmpiricalRecentJul 9, 2026

Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems

Kalle Kujanpää, Ning Liu, Shahnawaz Alam, Yeshwanth Reddy Sura +3 more

The paper presents a tool-making pipeline for production LLM agents that compiles repeated steps into validated, versioned tools before deployment, reducing latency and error rate.

View →
cs.CRcs.AIcs.NIRecentApr 22, 2026

Behavioral Consistency and Transparency Analysis on Large Language Model API Gateways

Guanjie Lin, Yinxin Wan, Shichao Pei, Ting Xu +2 more

The paper introduces GateScope, a black-box framework that audits commercial LLM API gateways, revealing frequent discrepancies in model behavior, billing, and performance across real-world services.

View →