vix.ing · top · new · best · stats · spec

Xiaoyu Yang

  1. Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
    2023/09/15 by Wei Kang, Kang, Wei, Xiaoyu Yang +12 · 33 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Zipformer: A faster and better encoder for automatic speech recognition
    2023/10/17 by Zengwei Yao, Yao, Zengwei, Liyong Guo +14 · 24 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
  3. Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards
    2024/05/22 by Xiaoyu Yang, Yang, Xiaoyu, Jie Lü +3 · 8 citations
    Computer Science · #Advanced Text Analysis Techniques #Data Management and Algorithms #Web Data Mining and Analysis
  4. SemEval-2020 Task 5: Counterfactual Recognition
    2020/08/02 by Xiaoyu Yang, Stephen Obadinma, Yang, Xiaoyu +9 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
  5. Segmentation and Vascular Vectorization for Coronary Artery by Geometry-based Cascaded Neural Network
    2023/05/07 by Xiaoyu Yang, Lijian Xu, Yang, Xiaoyu +9 · 3 citations
    Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #Coronary Interventions and Diagnostics #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Medical Image Segmentation Techniques #Radiomics and Machine Learning in Medical Imaging #electronic engineering #information engineering
  6. Knowledge Distillation for Neural Transducers from Large Self-Supervised Pre-trained Models
    2021/10/07 by Xiaoyu Yang, Yang, Xiaoyu, Qiujia Li +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  7. k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
    2024/11/26 by Yifan Yang, Yang, Yifan, Jianheng Zhuo +20 · 4 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  8. Predicting Multi-Codebook Vector Quantization Indexes for Knowledge Distillation
    2022/10/31 by Liyong Guo, Guo, Liyong, Xiaoyu Yang +20 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  9. Enhancing Visual Grounding and Generalization: A Multi-Task Cycle Training Approach for Vision-Language Models
    2023/11/21 by Xiaoyu Yang, Yang, Xiaoyu, Lijian Xu +6 · 2 citations
    Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Digital Imaging for Blood Diseases
  10. Bioprocess-inspired fabrication of materials with new structures and functions
    2019/05/21 by Jingjing Xie, Hang Ping, Tiening Tan +5 · 1 citation
    Materials Science · Biochemistry, Genetics and Molecular Biology · #Diatoms and Algae Research #Calcium Carbonate Crystallization and Inhibition #Protist diversity and phylogeny
  11. LibriheavyMix: A 20,000-Hour Dataset for Single-Channel Reverberant Multi-Talker Speech Separation, ASR and Speaker Diarization
    2024/09/01 by Zengrui Jin, Jin, Zengrui, Yifan Yang +22 · 2 citations
    Computer Science · Health Professions · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Infant Health and Development #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  12. Fast and parallel decoding for transducer
    2022/10/31 by Wei Kang, Liyong Guo, Kang, Wei +14 · 1 citation
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #DNA and Biological Computing #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Network Packet Processing and Optimization #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  13. Blank-regularized CTC for Frame Skipping in Neural Transducer
    2023/05/19 by Yifan Yang, Yang, Yifan, Xiaoyu Yang +14 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Neural Networks and Applications #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Engineering synthetic phosphorylation signaling networks in human cells
    2025/01/02 by Xiaoyu Yang, Jason W. Rocks, Kaiyi Jiang +9 · 1 voice · 2 citations
    Biochemistry, Genetics and Molecular Biology · #Receptor Mechanisms and Signaling #Single-cell and spatial transcriptomics #Gene Regulatory Network Analysis
  15. Walking the Tightrope: Disentangling Beneficial and Detrimental Drifts in Non-Stationary Custom-Tuning
    2025/05/19 by Xiaoyu Yang, Yang, Xiaoyu, Jie Lu +3 · 2 citations
    Computer Science · #Advanced Graph Neural Networks #Computer Vision and Pattern Recognition (cs.CV) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  16. MT2KD: Towards A General-Purpose Encoder for Speech, Speaker, and Audio Events
    2024/09/25 by Xiaoyu Yang, Yang, Xiaoyu, Qiujia Li +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  17. SALMONN-2: Advancing General-Purpose Hearing Abilities with Self-Supervised Representations
    2026/07/19 by Xiaoyu Yang, Xuenan Xu, Wenyi Yu +10
    #eess.AS
  18. MV-Bench: Benchmarking Multimodal Large Language Models for Coordinated Multi-View Interface Construction
    2026/07/22 by Yue Zhao, Hongxu Liu, Feiyu Wang +5
    #cs.CV #cs.HC
  19. Final assessment of radioactive impurities in the JUNO detector
    2026/07/20 by Thomas Adam, Fengpeng An, Costas Andreopoulos +569
    #physics.ins-det #hep-ex