Scholar
Gunho Park
Google Scholar ID: 3l_lwHkAAAAJ
NAVER Cloud
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
373
H-index
6
i10-index
5
Publications
15
Co-authors
0
Publications
7 items
SlimWise: Decoupling Expert Pruning Across Prefill and Decode for Efficient MoE Serving
2026
Cited
0
Affine-Scaled Attention: Towards Flexible and Stable Transformer Attention
2026
Cited
0
CodeGEMM: A Codebook-Centric Approach to Efficient GEMM in Quantized LLMs
2025
Cited
0
AnyBCQ: Hardware Efficient Flexible Binary-Coded Quantization for Multi-Precision LLMs
2025
Cited
0
Faster Inference of LLMs using FP8 on the Intel Gaudi
2025
Cited
0
FIGLUT: An Energy-Efficient Accelerator Design for FP-INT GEMM Using Look-Up Tables
2025
Cited
0
An Investigation of FP8 Across Accelerators for LLM Inference
2025
Cited
0