ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “Understanding of load balancing concepts”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.DScs.DMTheoreticalRecentJul 28, 2026

Stochastic Load Balancing with Machine Reservations

David Alemán Espinosa, Naveen Garg, Sharat Ibrahimpur, Neil Olver +1 more

A new stochastic load balancing model is introduced that allows for a tradeoff between non-adaptive policies and performance, with results showing a 2-reservation approximation to the omniscient optim…

View →
cs.DSTheoreticalRecentJun 19, 2026

Online Stacking with a Few Load/Unload Points

Martin Olsen

A simple online algorithm is presented for the stacking problem to avoid shifts with a sufficient condition involving stacking area dimension, load/unload points, and maximum items.

View →
cs.NIEmpiricalRecentJul 24, 2026

Fewer Paths, Better Performance: Understanding the ZCube Topology through Braess's Paradox

Li Chen

The ZCube topology, which eliminates path multiplicity and reduces switching hardware, delivers better performance for large model training and inference than traditional multipath datacenter networks…

View →
cs.NIEmpiricalRecentJun 26, 2026

Host-Driven Flowlet Balancing with Segment Routing over IPv6

Ryo Nakamura, Hiroki Kano, Tomoko Okuzawa

This paper proposes a host-driven method for flowlet balancing using Segment Routing over IPv6 (SRv6), reducing tail latency by 15% and 33% compared to random flowlet balancing and ECMP, respectively.

View →
cs.DCEmpiricalRecentJun 29, 2026

Spandana: Reconciling Strict SLOs with Low Cost under Fine-Grained Load Fluctuations

Dilina Dehigama, Shyam Jesalpura, Zeyu Xu, Marton Nemeth +3 more

The paper introduces Spandana, an architecture that decouples SLO enforcement from cost optimization in cloud-based online services, achieving high utilization, strict SLO adherence, and cost savings.

View →
cs.DCcs.DSEmpiricalRecentJun 27, 2026

Concurrent Splay-Based Tree

Vitaly Aksenov, Rene van Bevern, Artem Shilkin

This paper proposes a splay-like rotation design for concurrent binary search trees to preserve the main benefit of splaying on skewed workloads while reducing contention near the root.

View →
cs.OScs.ARcs.NIEmpiricalRecentJul 17, 2026

Rethinking Polling Efficiency in Service Core Network Stacks

Matheus Stolet, Simon Peter, Antoine Kaufmann

This paper argues that idle cores on contemporary multicore processors can return compute capacity and proposes a budget-centric view of service core systems.

View →
cs.DCcs.AIcs.OSEmpiricalRecentJul 2, 2026

Fine-Grained Computation Offload for Off-the-Shelf Servers in Tens of Lines

Bojie Li

The paper proposes a method to improve the performance of fine-grained offloads on servers by overlapping the offload with other requests using server-side routing.

View →
cs.NIcs.DCcs.LGEmpiricalRecentJul 28, 2026

Incast-Free MoE Rate-Based Scheduling

Evyatar Cohen, Jose Yallouz, Alexander Shpiner, Mark Silberstein +2 more

This paper proposes a proactive fair scheduling framework to prevent fabric oversubscription and eliminate incast in Mixture of Experts (MoE) architectures, demonstrating consistent link utilization a…

View →
cs.LGcs.AIcs.DCEmpiricalRecentJul 16, 2026

An Auto-Scaling Approach for Serverless Environments Based on a Multi-Expert Consensus Mechanism

Mobina Kashaniyan, Mehrdad Ashtiani, Amirhossein Ghassemi

This paper proposes a dependency-aware autoscaling framework for serverless computing, integrating graph-based bottleneck identification, short-term workload forecasting, multi-model consensus, and co…

View →
cs.DCcs.OScs.PFEmpiricalRecentJul 18, 2026

Hardware-Transparent I/O Governance in Disaggregated Heterogeneous Storage

Rajarshi Chowdhury, Akshay Shah, Sue K. Lee

The I/O Resource Manager (IORM) is presented as a multi-stage distributed scheduler to maintain consistent performance and enforce global I/O limits in shared-nothing disaggregated storage clusters.

View →
cs.DCEmpiricalRecentJun 30, 2026

Performance Analysis in Parallel Programming Education: A Comparative Usability Study

Anna-Lena Roth, David James, Jonas Posner, Michael Kuhn

The paper introduces EduMPI, a learning support tool for simplifying cluster usage and performance analysis of MPI parallel programs for students.

View →
cs.PFcs.AREmpiricalRecentJul 16, 2026

Campaign Diagrams: Visualizing the March Through the Phases of a Workload

Toluwanimi O. Odemuyiwa, John D. Owens, Michael Pellauer, Joel S. Emer

This paper introduces campaign diagrams, a visualization technique for analyzing resource utilization and identifying bottlenecks in modern workloads.

View →
cs.DCEmpiricalRecentJul 2, 2026

Elasticity in Parallel Sparse Triangular Solve

Raphael S. Steiner, Christos K. Matzoros, Pál András Papp, Toni Böhnlein +1 more

This paper introduces Stale Synchronous Parallel mode of execution for parallel sparse triangular linear system solve and presents a scheduler that achieves geometric-mean speed-ups of 7-30% over Grow…

View →
cs.DSTheoreticalRecentJul 17, 2026

Revisiting Real-Time Interval and Throughput Maximization

Allan Borodin, Changdao He, Nadim Mottu

The paper extends results for interval scheduling to the more general throughput problem in the real-time model with constant competitive ratios for specific weight functions and advance notice.

View →
cs.CCcs.PFcs.PLTheoreticalRecentJun 29, 2026

The Fourth-Root Complexity of Data Movement

Chen Ding

This paper analyzes data-access cost in a memory hierarchy and shows it scales with the fourth root of data size, predicting scalability.

View →
cs.DSTheoreticalRecentJun 22, 2026

Is competitive online paging an artifact?

Enoch Peserico, Michele Scquizzato

The Sleator-Tarjan model used in competitive analysis of paging is incorrect, and this error undermines the performance predictions for online paging algorithms.

View →