Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Moinuddin Qureshi

Moinuddin Qureshi

1 indexed paper

Recent (6 mo)
1
With code
0
Influential cites
0
Benchmarked
0

Publications per year

1
26

Top categories

Distributed×1

Frequent co-authors

Hritvik Taneja1×
Anish Saxena1×
Abhishek Revinipati1×
Jae Hyung Ju1×
Neal C. Crago1×

Research Timeline

2026
SiFAR: Synchronization-Free All-Reduce for Low-Latency LLM Inference

This paper proposes Synchronization-Free All-Reduce (SiFAR) to reduce All-Reduce latency and improve end-to-end throughput in low-latency inference systems.

Highlighted terms show continued research focus across papers

Papers

cs.DCEmpiricalRecentJul 9, 2026

SiFAR: Synchronization-Free All-Reduce for Low-Latency LLM Inference

Hritvik Taneja, Anish Saxena, Abhishek Revinipati, Jae Hyung Ju +2 more

This paper proposes Synchronization-Free All-Reduce (SiFAR) to reduce All-Reduce latency and improve end-to-end throughput in low-latency inference systems.

View →