Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:
ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Home/Authors/Hyunwoo Oh

Hyunwoo Oh

2 indexed papers

Recent (6 mo)
2
With code
0
Influential cites
0
Benchmarked
0

Publications per year

2
26

Top categories

Architecture×2ML×2OS×2

Frequent co-authors

Suyeon Jang2×
Hanning Chen2×
Sanggeon Yun2×
Ryozo Masukawa2×
Mohsen Imani2×
KyungIn Nam1×

Research Timeline

2026
ExaGEMM: Exploration Framework for CPU-Driven ML Inference via Associative In-Register Computing for Low-Bit GEMM

The paper presents ExaGEMM, a framework for designing and exploring CPU-native low-bit GEMM via register-resident LUT execution.

PolyQ: Codesigning End-to-End Quantization Framework for Scalable Edge CPU LLM Inference

PolyQ is a compiler/quantization co-design for activation-aware channel-wise bit allocation on CPUs, providing stable quality scaling and energy efficiency for fine-grained fractional-bit CPU deployment.

Highlighted terms show continued research focus across papers

Papers

cs.ARcs.LGcs.OSEmpiricalRecentJul 16, 2026

ExaGEMM: Exploration Framework for CPU-Driven ML Inference via Associative In-Register Computing for Low-Bit GEMM

Hyunwoo Oh, Suyeon Jang, Hanning Chen, Sanggeon Yun +2 more

The paper presents ExaGEMM, a framework for designing and exploring CPU-native low-bit GEMM via register-resident LUT execution.

View →
cs.LGcs.ARcs.OSEmpirical
Recent
Jul 16, 2026

PolyQ: Codesigning End-to-End Quantization Framework for Scalable Edge CPU LLM Inference

Hyunwoo Oh, Suyeon Jang, Hanning Chen, KyungIn Nam +3 more

PolyQ is a compiler/quantization co-design for activation-aware channel-wise bit allocation on CPUs, providing stable quality scaling and energy efficiency for fine-grained fractional-bit CPU deployme…

View →