vix.ing · top · new · best · stats · spec

Zou, Yuexian

  1. Diffsound: Discrete Diffusion Model for Text-to-sound Generation
    2022/07/20 by Dongchao Yang, Jianwei Yu, Yang, Dongchao +11 · 40 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  2. HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
    2023/05/04 by Dongchao Yang, Songxiang Liu, Yang, Dongchao +9 · 42 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation
    2021/06/13 by Fenglin Liu, Liu, Fenglin, Xian Wu +7 · 21 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  4. VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model
    2025/01/21 by Xianwei Zhuang, Yuxin Xie, Zhuang, Xianwei +11 · 24 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  5. On the Worst Prompt Performance of Large Language Models
    2024/06/08 by Bowen Cao, Cao, Bowen, Deng Cai +7 · 13 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
  6. BrushEdit: All-In-One Image Inpainting and Editing
    2024/12/13 by Li, Yaowei, Bian, Yuxuan, Ju, Xuan +5 · 18 citations
    #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. Competence-based Multimodal Curriculum Learning for Medical Report Generation
    2022/06/24 by Fenglin Liu, Ge Shen, Liu, Fenglin +4 · 7 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  8. Image Conductor: Precision Control for Interactive Video Synthesis
    2024/06/21 by Yaowei Li, Li, Yaowei, Xintao Wang +13 · 13 citations
    Computer Science · #Advanced Vision and Imaging #Artificial Intelligence (cs.AI) #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM) #Video Coding and Compression Technologies
  9. On Pursuit of Designing Multi-modal Transformer for Video Grounding
    2021/09/13 by Meng Cao, Long Chen, Cao, Meng +7 · 6 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Video Analysis and Summarization
  10. CoLA: Weakly-Supervised Temporal Action Localization with Snippet Contrastive Learning
    2021/03/30 by Can Zhang, Meng Cao, Zhang, Can +7 · 6 citations
    Computer Science · #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Anomaly Detection Techniques and Applications
  11. End-to-end Spoken Conversational Question Answering: Task, Dataset and Model
    2022/04/29 by Chenyu You, Nuo Chen, You, Chenyu +9 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  12. Contrastive Attention for Automatic Chest X-ray Report Generation
    2021/06/13 by Fenglin Liu, Liu, Fenglin, Changchang Yin +11 · 5 citations
    Computer Science · Medicine · #COVID-19 diagnosis using AI #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Topic Modeling
  13. CAR: Controllable Autoregressive Modeling for Visual Generation
    2024/10/07 by Ziyu Yao, Yao, Ziyu, Jialin Li +15 · 11 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Retrieval and Classification Techniques
  14. Target Confusion in End-to-end Speaker Extraction: Analysis and Approaches
    2022/04/04 by Zhao, Zifeng, Yang, Dongchao, Gu, Rongzhi +2 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  15. Rethinking Skip Connection with Layer Normalization in Transformers and ResNets
    2021/05/15 by Fenglin Liu, Xuancheng Ren, Liu, Fenglin +7 · 4 citations
    Computer Science · #Advanced Neural Network Applications #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG)
  16. Unify, Align and Refine: Multi-Level Semantic Alignment for Radiology Report Generation
    2023/03/28 by Yaowei Li, Bang Yang, Li, Yaowei +9 · 5 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Biomedical Text Mining and Ontologies #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
  17. Unsupervised Pre-training for Temporal Action Localization Tasks
    2022/03/25 by Zhang, Can, Yang, Tianyu, Weng, Junwu +3 · 4 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  18. G2L: Semantically Aligned and Uniform Video Grounding via Geodesic and Game Theory
    2023/07/26 by Hongxiang Li, Meng Cao, Li, Hongxiang +9 · 5 citations
    Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
  19. Enhancing End-to-End Multi-channel Speech Separation via Spatial Feature Learning
    2020/03/09 by Rongzhi Gu, Gu, Rongzhi, Shixiong Zhang +13 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  20. Improving Target Sound Extraction with Timestamp Information
    2022/04/02 by Helin Wang, Wang, Helin, Dongchao Yang +7 · 4 citations
    Computer Science · Engineering · #Acoustic Wave Phenomena Research #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  21. LocVTP: Video-Text Pre-training for Temporal Localization
    2022/07/21 by Meng Cao, Tianyu Yang, Cao, Meng +9 · 4 citations
    Computer Science · Biochemistry, Genetics and Molecular Biology · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Cancer-related molecular mechanisms research
  22. Towards Joint Intent Detection and Slot Filling via Higher-order Attention
    2021/09/18 by Dongsheng Chen, Zhiqi Huang, Chen, Dongsheng +7 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Speech and dialogue systems
  23. LAE: Language-Aware Encoder for Monolingual and Multilingual ASR
    2022/06/05 by Jinchuan Tian, Jianwei Yu, Tian, Jinchuan +9 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  24. Comparison of Spearman's rho and Kendall's tau in Normal and Contaminated Normal Models
    2010/11/09 by Weichao Xu, Xu, Weichao, Yunhe Hou +5 · 2 citations
    Mathematics · #Advanced Statistical Methods and Models #FOS: Computer and information sciences #Information Theory (cs.IT) #Statistical Distribution Estimation and Applications #Statistical Methods and Bayesian Inference
  25. Embracing Language Inclusivity and Diversity in CLIP through Continual Language Learning
    2024/01/30 by Yang, Bang, Dai, Yong, Cheng, Xuxin +3 · 4 citations
    #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Information Retrieval (cs.IR)
  26. WorldGPT: A Sora-Inspired Video AI Agent as Rich World Models from Text and Image Inputs
    2024/03/10 by Deshun Yang, Yang, Deshun, Luhui Hu +13 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Retrieval and Classification Techniques #Multimodal Machine Learning Applications
  27. Aligning Source Visual and Target Language Domains for Unpaired Video Captioning
    2022/11/22 by Fenglin Liu, Xian Wu, Liu, Fenglin +9 · 3 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cancer-related molecular mechanisms research #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  28. SpecAugment++: A Hidden Space Data Augmentation Method for Acoustic Scene Classification
    2021/03/31 by Helin Wang, Wang, Helin, Yuexian Zou +3 · 2 citations
    Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  29. O2NA: An Object-Oriented Non-Autoregressive Approach for Controllable Video Captioning
    2021/08/05 by Fenglin Liu, Xuancheng Ren, Liu, Fenglin +9 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Video Analysis and Summarization
  30. VARGPT-v1.1: Improve Visual Autoregressive Large Unified Model via Iterative Instruction Tuning and Reinforcement Learning
    2025/04/03 by Xianwei Zhuang, Zhuang, Xianwei, Yuxin Xie +13 · 12 citations
    Computer Science · Biochemistry, Genetics and Molecular Biology · #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications #Cell Image Analysis Techniques
  31. Towards Data Distillation for End-to-end Spoken Conversational Question Answering
    2020/10/18 by Chenyu You, Nuo Chen, You, Chenyu +7 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Signal Processing (eess.SP) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  32. All You Need is a Second Look: Towards Arbitrary-Shaped Text Detection
    2021/06/24 by Meng Cao, Can Zhang, Cao, Meng +5 · 2 citations
    Computer Science · Engineering · #Handwritten Text Recognition Techniques #Image Processing and 3D Reconstruction #Vehicle License Plate Recognition
  33. Knowledge Distillation for Improved Accuracy in Spoken Question Answering
    2020/10/21 by You, Chenyu, Chen, Nuo, Zou, Yuexian · 2 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  34. Retrieval is Accurate Generation
    2024/02/27 by Bowen Cao, Cao, Bowen, Deng Cai +11 · 3 citations
    Computer Science · #AI-based Problem Solving and Planning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning and Algorithms #Speech and dialogue systems
  35. VisionGPT: Vision-Language Understanding Agent Using Generalized Multimodal Framework
    2024/03/14 by Christopher B. Kelly, Luhui Hu, Kelly, Chris +17 · 3 citations
    Computer Science · Social Sciences · #Speech and dialogue systems #Religious Tourism and Spaces #AI in Service Interactions
  36. Temporal-Spatial Neural Filter: Direction Informed End-to-End Multi-channel Target Speech Separation
    2020/01/02 by Rongzhi Gu, Yuexian Zou, Gu, Rongzhi +1 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  37. PoseRAC: Pose Saliency Transformer for Repetitive Action Counting
    2023/03/15 by Ziyu Yao, Yao, Ziyu, Xuxin Cheng +3 · 2 citations
    Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Stroke Rehabilitation and Recovery #Virtual Reality Applications and Impacts
  38. Contextualized Attention-based Knowledge Transfer for Spoken Conversational Question Answering
    2020/10/21 by You, Chenyu, Chen, Nuo, Zou, Yuexian · 2 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  39. Self-supervised Dialogue Learning for Spoken Conversational Question Answering
    2021/06/04 by Chen, Nuo, You, Chenyu, Zou, Yuexian · 2 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #electronic engineering #information engineering
  40. VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification
    2025/01/11 by Xianwei Zhuang, Zhuang, Xianwei, Zhihong Zhu +7 · 6 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Generative Adversarial Networks and Image Synthesis #Digital Media Forensic Detection
  41. PIN: A Novel Parallel Interactive Network for Spoken Language Understanding
    2020/09/28 by Peilin Zhou, Zhou, Peilin, Zhiqi Huang +5 · 2 citations
    Computer Science · #Topic Modeling #Speech and dialogue systems #Multimodal Machine Learning Applications
  42. Unsupervised Multi-Target Domain Adaptation for Acoustic Scene Classification
    2021/05/21 by Dongchao Yang, Helin Wang, Yang, Dongchao +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  43. Do we really have to filter out random noise in pre-training data for language models?
    2025/02/10 by Jinghan Ru, Ru, Jinghan, Yuxin Xie +9 · 5 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  44. A Graph-based Interactive Reasoning for Human-Object Interaction Detection
    2020/07/14 by Dongming Yang, Yang, Dongming, Yuexian Zou +1 · 1 citation
    Computer Science · #Advanced Graph Neural Networks #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications
  45. Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding
    2025/01/19 by Z. J. Chen, Chen, Zhanpeng, Mingxiao Li +9 · 4 citations
    Computer Science · #Multimodal Machine Learning Applications #Speech and dialogue systems
  46. A Global-local Attention Framework for Weakly Labelled Audio Tagging
    2021/02/03 by Helin Wang, Yuexian Zou, Wang, Helin +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #Video Analysis and Summarization #electronic engineering #information engineering
  47. Joint Multiple Intent Detection and Slot Filling via Self-distillation
    2021/08/18 by Lisong Chen, Chen, Lisong, Peilin Zhou +3 · 1 citation
    Computer Science · #Topic Modeling #Text and Document Classification Technologies #Multimodal Machine Learning Applications
  48. Text Anchor Based Metric Learning for Small-footprint Keyword Spotting
    2021/08/12 by Wang, Li, Gu, Rongzhi, Chen, Nuo +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Sound (cs.SD)
  49. Adaptive Bi-directional Attention: Exploring Multi-Granularity Representations for Machine Reading Comprehension
    2020/12/20 by Nuo Chen, Fenglin Liu, Chen, Nuo +7 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  50. NoreSpeech: Knowledge Distillation based Conditional Diffusion Model for Noise-robust Expressive TTS
    2022/11/04 by Dongchao Yang, Songxiang Liu, Yang, Dongchao +9 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  51. ATRI: Mitigating Multilingual Audio Text Retrieval Inconsistencies by Reducing Data Distribution Errors
    2025/02/20 by Yuguo Yin, Yin, Yuguo, Yuxin Xie +13 · 5 citations
    Arts and Humanities · Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  52. DiMBERT: Learning Vision-Language Grounded Representations with Disentangled Multimodal-Attention
    2022/10/28 by Fenglin Liu, Liu, Fenglin, Xian Wu +11 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  53. Towards Unified All-Neural Beamforming for Time and Frequency Domain Speech Separation
    2022/12/16 by Rongzhi Gu, Gu, Rongzhi, Shixiong Zhang +5 · 1 citation
    Computer Science · Engineering · #Acoustic Wave Phenomena Research #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  54. Improving Weakly Supervised Sound Event Detection with Causal Intervention
    2023/03/10 by Yifei Xin, Dongchao Yang, Xin, Yifei +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  55. Improving Text-Audio Retrieval by Text-aware Attention Pooling and Prior Matrix Revised Loss
    2023/03/10 by Yifei Xin, Dongchao Yang, Xin, Yifei +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  56. ML-LMCL: Mutual Learning and Large-Margin Contrastive Learning for Improving ASR Robustness in Spoken Language Understanding
    2023/11/19 by Xuxin Cheng, Bowen Cao, Cheng, Xuxin +9 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
  57. Correspondence Matters for Video Referring Expression Comprehension
    2022/07/21 by Cao, Meng, Jiang, Ji, Chen, Long +1 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  58. BlobCtrl: Taming Controllable Blob for Element-level Image Editing
    2025/03/17 by Yaowei Li, Lingen Li, Li, Yaowei +15 · 2 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Digital Image Processing Techniques #FOS: Computer and information sciences #Medical Image Segmentation Techniques #Modular Robots and Swarm Intelligence #Multimedia (cs.MM)
  59. DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
    2024/09/16 by Yifei Xin, Xin, Yifei, Xuxin Cheng +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  60. Audio-text Retrieval with Transformer-based Hierarchical Alignment and Disentangled Cross-modal Representation
    2024/09/14 by Yifei Xin, Xin, Yifei, Zhihong Zhu +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering