Lee, Dongsoo
- LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
2022/06/20 by Gunho Park, Baeseong Park, Park, Gunho +18 · 1 voice · 16 citations
Engineering · Computer Science · #cs.DC #cs.CL
- DFX: A Low-latency Multi-FPGA Appliance for Accelerating Transformer-based Text Generation
2022/09/22 by Seongmin Hong, Hong, Seongmin, Seungjae Moon +11 · 10 citations
Computer Science · #Topic Modeling #Advanced Neural Network Applications #Parallel Computing and Optimization Techniques
- No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization
2024/02/28 by Yang, June Yong, Kim, Byeongwook, Bae, Jeongin +5 · 11 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization
2023/05/23 by Kim, Jeonghoon, Lee, Jung Hyun, Kim, Sungdong +4 · 7 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
2023/06/01 by Jung Hyun Lee, Jeonghoon Kim, Lee, Jung Hyun +5 · 5 citations
Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- BiQGEMM: Matrix Multiplication with Lookup Table For Binary-Coding-based Quantized DNNs
2020/05/20 by Yongkweon Jeon, Jeon, Yongkweon, Baeseong Park +9 · 4 citations
Computer Science · #Parallel Computing and Optimization Techniques #Topic Modeling #Advanced Neural Network Applications
- To FP8 and Back Again: Quantifying Reduced Precision Effects on LLM Training Stability
2024/05/29 by Joonhyung Lee, Lee, Joonhyung, Jeongin Bae +7 · 4 citations
Health Professions · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Quality and Safety in Healthcare
- HyperCLOVA X Technical Report
2024/04/02 by Kang Min Yoo, Yoo, Kang Min, Jaegeun Han +480 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Topic Modeling
- Structured Compression by Weight Encryption for Unstructured Pruning and Quantization
2019/05/24 by Kwon, Se Jung, Lee, Dongsoo, Kim, Byeongwook +3 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- FIGLUT: An Energy-Efficient Accelerator Design for FP-INT GEMM Using Look-Up Tables
2025/03/10 by Park, Gunho, Kwon, Hyeokjun, Kim, Jiwoo +4 · 4 citations
#FOS: Computer and information sciences #Hardware Architecture (cs.AR)
- AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models
2022/10/08 by Se Jung Kwon, Kwon, Se Jung, Jeonghoon Kim +17 · 1 citation
Computer Science · #Advanced Neural Network Applications #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
- DropBP: Accelerating Fine-Tuning of Large Language Models by Dropping Backward Propagation
2024/02/27 by Sunghyeon Woo, Woo, Sunghyeon, Baeseong Park +11 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
- LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices
2024/07/16 by Jung Hyun Lee, Lee, Jung Hyun, Jeonghoon Kim +11 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
- Debunking the CUDA Myth Towards GPU-based AI Systems
2024/12/31 by Lee, Yunjae, Lim, Juntaek, Bang, Jehyeon +10 · 1 citation
#Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Hardware Architecture (cs.AR) #Parallel #and Cluster Computing (cs.DC)
- An Inquiry into Datacenter TCO for LLM Inference with FP8
2025/02/03 by Jiwoo Kim, Joonhyung Lee, Kim, Jiwoo +11 · 1 citation
Physics and Astronomy · Engineering · #Magnetic confinement fusion research #Particle accelerators and beam dynamics
- CodeGEMM: A Codebook-Centric Approach to Efficient GEMM in Quantized LLMs
2025/12/19 by Park, Gunho, Bae, Jeongin, Kim, Byeongwook +5 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)