vix.ing · top · new · best · stats · spec

Quanlu Zhang

  1. You Only Cache Once: Decoder-Decoder Architectures for Language Models
    2024/05/08 by Yutao Sun, Sun, Yutao, Li Dong +15 · 2 voices · 39 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #cs.CL
  2. Efficient Large Language Models: A Survey
    2023/12/06 by Zhongwei Wan, Xin Wang, Wan, Zhongwei +23 · 1 voice · 50 citations
    Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Materials Science #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL
  3. Deeper Insights into Weight Sharing in Neural Architecture Search
    2020/01/06 by Yuge Zhang, Zejun Lin, Zhang, Yuge +13 · 10 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  4. LadaBERT: Lightweight Adaptation of BERT through Hybrid Model Compression
    2020/04/08 by Yihuan Mao, Mao, Yihuan, Yujing Wang +15 · 4 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  5. PIT: Optimization of Dynamic Sparse Deep Learning Models via Permutation Invariant Transformation
    2023/01/26 by Ningxin Zheng, Huiqiang Jiang, Zheng, Ningxin +18 · 5 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Advanced Neural Network Applications #Algorithms and Data Compression #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
  6. RLinf: Flexible and Efficient Large-scale Reinforcement Learning via Macro-to-Micro Flow Transformation
    2025/09/19 by Chao Yu, Yu, Chao, Yuanqing Wang +52 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel #Reinforcement Learning in Robotics #and Cluster Computing (cs.DC)
  7. SuperScaler: Supporting Flexible DNN Parallelization via a Unified Abstraction
    2023/01/21 by Zhiqi Lin, Lin, Zhiqi, Youshan Miao +23 · 1 citation
    Computer Science · Engineering · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Parallel #Parallel Computing and Optimization Techniques #and Cluster Computing (cs.DC)