publications
publications by categories in reversed chronological order. generated by jekyll-scholar.
2026
- Under Review
HERALD: High Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval2026 - NeurIPS
MAGE: All-[MASK] Block Already Knows Where to Look in Block Diffusion LLMNeurIPS 2026& ICML 2026 AdaptFM Workshop (Best Paper Award)
2025
- ICCAD
STAR: Improving Lifetime and Performance of High-Capacity Modern SSDs Using State-Aware RandomizerICCAD 2025U.S. Patent Application with SK hynix (No. 19/569899; filed March 17, 2026) - CAL
AiDE: Attention-FFN Disaggregated Execution for Cost-Effective LLM Decoding on CXL-PNMIEEE CAL, 2025