Anish Saxena
2 indexed papers
Recent (6 mo)
2With code
0Influential cites
0Benchmarked
0Publications per year
226
Top categories
Distributed×1Architecture×1
Frequent co-authors
Research Timeline
2026
TileLens: Efficiently Using Large-Granularity Memory Systems with Transparent Two-Dimensional Memory Layout
This paper proposes TileLens, a system to mitigate read amplification in Large-Granularity Memory Systems (LGMS) for Large Language Model (LLM) inference by adopting a tile-major layout.
SiFAR: Synchronization-Free All-Reduce for Low-Latency LLM Inference
This paper proposes Synchronization-Free All-Reduce (SiFAR) to reduce All-Reduce latency and improve end-to-end throughput in low-latency inference systems.
Highlighted terms show continued research focus across papers