vix.ing · top · new · best · stats · spec

Tao, Keda

  1. RadioDiff: An Effective Generative Diffusion Model for Sampling-Free Dynamic Radio Map Construction
    2024/08/16 by Xiucheng Wang, Keda Tao, Wang, Xiucheng +11 · 32 citations
    Computer Science · Engineering · #FOS: Computer and information sciences #FOS: Electrical engineering #Indoor and Outdoor Localization Technologies #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and Audio Processing #Systems and Control (eess.SY) #electronic engineering #information engineering
  2. DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models
    2024/11/22 by Can Qin, Tao, Keda, Haoxuan You +6 · 26 citations
    Computer Science · #Video Analysis and Summarization #Advanced Data Compression Techniques #Multimodal Machine Learning Applications
  3. HoliTom: Holistic Token Merging for Fast Video Large Language Models
    2025/05/27 by Shao, Kele, Can Qin, Tao, Keda +6 · 20 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications
  4. A Survey of Token Compression for Efficient Multimodal Large Language Models
    2025/07/27 by Shao, Kele, Kejia Zhang, Tao, Keda +13 · 14 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications #Speech Recognition and Synthesis
  5. Plug-and-Play 1.x-Bit KV Cache Quantization for Video Large Language Models
    2025/03/20 by Keda Tao, Tao, Keda, Haoxuan You +6 · 6 citations
    Computer Science · #Advanced Data Compression Techniques #Advanced Image and Video Retrieval Techniques #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  6. PhotoArtAgent: Intelligent Photo Retouching with Language Model-Based Artist Agents
    2025/05/29 by Chen, Haoyu, Tao, Keda, Wang, Yizao +3 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs
    2025/01/31 by Zhang, Kejia, Tao, Keda, Tang, Jiasheng +1 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  8. Overcoming False Illusions in Real-World Face Restoration with Multi-Modal Guided Diffusion Model
    2024/10/05 by Keda Tao, Tao, Keda, Jinjin Gu +7 · 2 citations
    Computer Science · Engineering · #Advanced Image Processing Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Medical Imaging and Analysis
  9. Is Oracle Pruning the True Oracle?
    2024/11/28 by Sicheng Feng, Feng, Sicheng, Keda Tao +3 · 3 citations
    Computer Science · #Blockchain Technology Applications and Security #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  10. OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models
    2025/11/18 by Tao, Keda, Shao, Kele, Yu, Bohan +3 · 3 citations
    Computer Science · #Speech and Audio Processing #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing