I am an M.S. student at Efficient Computing Lab and Machine Learning Lab at POSTECH.
I am currently working on large model optimization, advised by Eunhyeok Park. My recent research interests include developing more effective methods for Parameter-Efficient Fine-Tuning (PEFT), as well as improving the efficiency of large model inference. More broadly, I am interested in understanding and exploiting low-dimensional structures in large language models to make both training and inference more efficient. I am particularly interested in methods that reduce computational and memory costs while preserving model performance.

Taehyeon Kim, Eunhyeok Park
The Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026 Accepted
TaRA is a training-aware LoRA initialization that matches combined low-rank adapter gradients to full-rank gradients, improving optimization and consistently outperforming prior SVD-based initializations with minimal overhead.
Taehyeon Kim, Eunhyeok Park
The Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026 Accepted
TaRA is a training-aware LoRA initialization that matches combined low-rank adapter gradients to full-rank gradients, improving optimization and consistently outperforming prior SVD-based initializations with minimal overhead.

Jinseop Yeom*, Taehyeon Kim*, Eunhyeok Park (* equal contribution)
The Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026 Accepted
MOB-KV makes adaptive low-rank KV cache compression practical by replacing costly online SVD with offline key bases, combining low-rank key compression and value quantization for strong memory-accuracy-latency trade-offs.
Jinseop Yeom*, Taehyeon Kim*, Eunhyeok Park (* equal contribution)
The Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026 Accepted
MOB-KV makes adaptive low-rank KV cache compression practical by replacing costly online SVD with offline key bases, combining low-rank key compression and value quantization for strong memory-accuracy-latency trade-offs.