ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “load balancing”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.NIEmpiricalRecentJun 26, 2026

Host-Driven Flowlet Balancing with Segment Routing over IPv6

Ryo Nakamura, Hiroki Kano, Tomoko Okuzawa

This paper proposes a host-driven method for flowlet balancing using Segment Routing over IPv6 (SRv6), reducing tail latency by 15% and 33% compared to random flowlet balancing and ECMP, respectively.

View →
cs.DCEmpiricalRecentJun 29, 2026

Spandana: Reconciling Strict SLOs with Low Cost under Fine-Grained Load Fluctuations

Dilina Dehigama, Shyam Jesalpura, Zeyu Xu, Marton Nemeth +3 more

The paper introduces Spandana, an architecture that decouples SLO enforcement from cost optimization in cloud-based online services, achieving high utilization, strict SLO adherence, and cost savings.

View →
cs.DSTheoreticalRecentJun 19, 2026

Online Stacking with a Few Load/Unload Points

Martin Olsen

A simple online algorithm is presented for the stacking problem to avoid shifts with a sufficient condition involving stacking area dimension, load/unload points, and maximum items.

View →
cs.LGcs.AIcs.DCEmpiricalRecentJul 16, 2026

An Auto-Scaling Approach for Serverless Environments Based on a Multi-Expert Consensus Mechanism

Mobina Kashaniyan, Mehrdad Ashtiani, Amirhossein Ghassemi

This paper proposes a dependency-aware autoscaling framework for serverless computing, integrating graph-based bottleneck identification, short-term workload forecasting, multi-model consensus, and co…

View →
cs.DCcs.DSEmpiricalRecentJun 27, 2026

Concurrent Splay-Based Tree

Vitaly Aksenov, Rene van Bevern, Artem Shilkin

This paper proposes a splay-like rotation design for concurrent binary search trees to preserve the main benefit of splaying on skewed workloads while reducing contention near the root.

View →
cs.NIEmpiricalRecentJul 24, 2026

Fewer Paths, Better Performance: Understanding the ZCube Topology through Braess's Paradox

Li Chen

The ZCube topology, which eliminates path multiplicity and reduces switching hardware, delivers better performance for large model training and inference than traditional multipath datacenter networks…

View →
cs.DCEmpiricalRecentJul 2, 2026

Elasticity in Parallel Sparse Triangular Solve

Raphael S. Steiner, Christos K. Matzoros, Pál András Papp, Toni Böhnlein +1 more

This paper introduces Stale Synchronous Parallel mode of execution for parallel sparse triangular linear system solve and presents a scheduler that achieves geometric-mean speed-ups of 7-30% over Grow…

View →
cs.DCcs.AIcs.OSEmpiricalRecentJul 2, 2026

Fine-Grained Computation Offload for Off-the-Shelf Servers in Tens of Lines

Bojie Li

The paper proposes a method to improve the performance of fine-grained offloads on servers by overlapping the offload with other requests using server-side routing.

View →
cs.DCEmpiricalRecentJun 29, 2026

StreamGuard: Low-Overhead Resilience for Real-time HPC Data Streams

Hai Duc Nguyen, Bogdan Nicolae, Tekin Bicer, Amal Gueroudji +3 more

This paper presents two techniques, dynamic checkpointing and progress-aware load redistribution, to maintain forward progress and balanced execution in real-time scientific workflows using the produc…

View →
cs.OScs.ARcs.NIEmpiricalRecentJul 17, 2026

Rethinking Polling Efficiency in Service Core Network Stacks

Matheus Stolet, Simon Peter, Antoine Kaufmann

This paper argues that idle cores on contemporary multicore processors can return compute capacity and proposes a budget-centric view of service core systems.

View →
cs.PFcs.DCcs.OSEmpiricalRecentJun 22, 2026

LMS-AR: LMS Prediction-based Adaptive Regulator for Memory Bandwidth in Multicore Systems

Sudarshan Srinivasan, Deepak Gangadharan, Dip Goswami

This paper proposes LMS-AR, a memory bandwidth regulation mechanism for multi-core systems using a Linux kernel module with adaptive filtering for prediction and regulation.

View →
cs.DCEmpiricalRecentJun 19, 2026

rush: Scalable Asynchronous Distributed Computing via Shared State in R

Marc Becker, Bernd Bischl

An R package named rush is introduced, which provides a shared-state coordination layer for asynchronously parallelized iterative algorithms using a Redis database.

View →
cs.DCEmpiricalRecentJun 30, 2026

Performance Analysis in Parallel Programming Education: A Comparative Usability Study

Anna-Lena Roth, David James, Jonas Posner, Michael Kuhn

The paper introduces EduMPI, a learning support tool for simplifying cluster usage and performance analysis of MPI parallel programs for students.

View →
cs.ETcs.DCEmpiricalRecentJul 21, 2026

Examining QRMI as a Unified Interface for Quantum-HPC Integration

Thomas Badts, Tim Boyle, Claudio Carvalho, Antonio Córcoles +24 more

The paper presents the Quantum Resource Management Interface (QRMI) as a standardized, vendor-agnostic middleware layer for integrating quantum resources into high-performance computing environments,…

View →
cs.DSTheoreticalRecentJul 17, 2026

Revisiting Real-Time Interval and Throughput Maximization

Allan Borodin, Changdao He, Nadim Mottu

The paper extends results for interval scheduling to the more general throughput problem in the real-time model with constant competitive ratios for specific weight functions and advance notice.

View →
cs.DCEmpiricalRecentJun 29, 2026

Energy-Aware Scheduling for Serverless LLM Serving on Shared GPUs

Tianyu Wang, Gourav Rattihalli, Aditya Dhakal, Longfei Shangguan +1 more

This paper presents Festina, a profiling-guided, power-aware control plane for minimizing energy consumption in serverless large language model (LLM) serving.

View →