vix.ing · top · new · best · stats · spec

Lee, Dongsoo

  1. LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
    2022/06/20 by Gunho Park, Baeseong Park, Park, Gunho +18 · 1 voice · 16 citations
    Engineering · Computer Science · #cs.DC #cs.CL
  2. DFX: A Low-latency Multi-FPGA Appliance for Accelerating Transformer-based Text Generation
    2022/09/22 by Seongmin Hong, Hong, Seongmin, Seungjae Moon +11 · 10 citations
    Computer Science · #Topic Modeling #Advanced Neural Network Applications #Parallel Computing and Optimization Techniques
  3. No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization
    2024/02/28 by Yang, June Yong, Kim, Byeongwook, Bae, Jeongin +5 · 11 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  4. Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization
    2023/05/23 by Kim, Jeonghoon, Lee, Jung Hyun, Kim, Sungdong +4 · 7 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  5. FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
    2023/06/01 by Jung Hyun Lee, Jeonghoon Kim, Lee, Jung Hyun +5 · 5 citations
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  6. BiQGEMM: Matrix Multiplication with Lookup Table For Binary-Coding-based Quantized DNNs
    2020/05/20 by Yongkweon Jeon, Jeon, Yongkweon, Baeseong Park +9 · 4 citations
    Computer Science · #Parallel Computing and Optimization Techniques #Topic Modeling #Advanced Neural Network Applications
  7. To FP8 and Back Again: Quantifying Reduced Precision Effects on LLM Training Stability
    2024/05/29 by Joonhyung Lee, Lee, Joonhyung, Jeongin Bae +7 · 4 citations
    Health Professions · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Quality and Safety in Healthcare
  8. HyperCLOVA X Technical Report
    2024/04/02 by Kang Min Yoo, Yoo, Kang Min, Jaegeun Han +480 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Topic Modeling
  9. Structured Compression by Weight Encryption for Unstructured Pruning and Quantization
    2019/05/24 by Kwon, Se Jung, Lee, Dongsoo, Kim, Byeongwook +3 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  10. FIGLUT: An Energy-Efficient Accelerator Design for FP-INT GEMM Using Look-Up Tables
    2025/03/10 by Park, Gunho, Kwon, Hyeokjun, Kim, Jiwoo +4 · 4 citations
    #FOS: Computer and information sciences #Hardware Architecture (cs.AR)
  11. AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models
    2022/10/08 by Se Jung Kwon, Kwon, Se Jung, Jeonghoon Kim +17 · 1 citation
    Computer Science · #Advanced Neural Network Applications #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
  12. DropBP: Accelerating Fine-Tuning of Large Language Models by Dropping Backward Propagation
    2024/02/27 by Sunghyeon Woo, Woo, Sunghyeon, Baeseong Park +11 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
  13. LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices
    2024/07/16 by Jung Hyun Lee, Lee, Jung Hyun, Jeonghoon Kim +11 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Topic Modeling
  14. Debunking the CUDA Myth Towards GPU-based AI Systems
    2024/12/31 by Lee, Yunjae, Lim, Juntaek, Bang, Jehyeon +10 · 1 citation
    #Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Hardware Architecture (cs.AR) #Parallel #and Cluster Computing (cs.DC)
  15. An Inquiry into Datacenter TCO for LLM Inference with FP8
    2025/02/03 by Jiwoo Kim, Joonhyung Lee, Kim, Jiwoo +11 · 1 citation
    Physics and Astronomy · Engineering · #Magnetic confinement fusion research #Particle accelerators and beam dynamics
  16. CodeGEMM: A Codebook-Centric Approach to Efficient GEMM in Quantized LLMs
    2025/12/19 by Park, Gunho, Bae, Jeongin, Kim, Byeongwook +5 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)