Bin Ma
- M2MeT: The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Challenge
2021/10/14 by Fan Yu, Shiliang Zhang, Yu, Fan +21 · 35 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Topic Modeling
- The similarity metric
2001/11/20 by Ming Li, Li, Ming, Xin Chen +7 · 11 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Algorithms and Data Compression #Combinatorics (math.CO) #Computability, Logic, AI Algorithms #Computational Complexity (cs.CC) #Computational Engineering #Computer Vision and Pattern Recognition (cs.CV) #Data Analysis #E.4 #FOS: Biological sciences #FOS: Computer and information sciences #FOS: Mathematics #FOS: Physical sciences #Finance #Fractal and DNA sequence analysis #Genomics (q-bio.GN) #J.3 #Metric Geometry (math.MG) #Statistical Mechanics (cond-mat.stat-mech) #Statistics Theory (math.ST) #Statistics and Probability (physics.data-an) #and Science (cs.CE)
- PEAKS: powerful software for peptide de novo sequencing by tandem mass spectrometry
2003/09/19 by Bin Ma, Kaizhong Zhang, Christopher Hendrie +4 · 6 citations
Biochemistry, Genetics and Molecular Biology · Chemistry · #Advanced Proteomics Techniques and Applications #Genomics and Phylogenetic Studies #Mass Spectrometry Techniques and Applications
- FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
2022/06/15 by Shengkui Zhao, Zhao, Shengkui, Bin Ma +5 · 18 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Novor: Real-Time Peptide de Novo Sequencing Software
2015/06/29 by Bin Ma · 10 citations
Biochemistry, Genetics and Molecular Biology · Chemistry · #Advanced Proteomics Techniques and Applications #Machine Learning in Bioinformatics #vaccines and immunoinformatics approaches
- MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation
2023/12/19 by Shengkui Zhao, Zhao, Shengkui, Yukun Ma +17 · 21 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- MossFormer: Pushing the Performance Limit of Monaural Speech Separation using Gated Single-Head Transformer with Convolution-Augmented Joint Self-Attentions
2023/02/23 by Shengkui Zhao, Bin Ma, Zhao, Shengkui +1 · 12 citations
Computer Science · Engineering · #Speech and Audio Processing #Speech Recognition and Synthesis #Ultrasonics and Acoustic Wave Propagation
- Robust Identity Perceptual Watermark Against Deepfake Face Swapping
2023/11/02 by Tianyi Wang, Wang, Tianyi, Mengxiao Huang +7 · 7 citations
Computer Science · #Digital Media Forensic Detection #Advanced Steganography and Watermarking Techniques #Face recognition and analysis
- Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
2022/02/08 by Fan Yu, Shiliang Zhang, Yu, Fan +29 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Biochar as construction materials for achieving carbon neutrality
2022/10/11 by Yuying Zhang, Mingjing He, Lei Wang +6 · 1 voice · 3 citations
Engineering · Environmental Science · Materials Science · #Concrete and Cement Materials Research #Magnesium Oxide Properties and Applications #Smart Materials for Construction
- Towards Audio Codec-based Speech Separation
2024/06/18 by Jia Qi Yip, Shengkui Zhao, Yip, Jia Qi +7 · 5 citations
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- On The Closest String and Substring Problems
2000/02/17 by Ming Li, Bin Ma, Li, Ming +3 · 1 citation
Biochemistry, Genetics and Molecular Biology · Computer Science · #Algorithms and Data Compression #Computational Complexity (cs.CC) #Computational Engineering #DNA and Biological Computing #F.2 #FOS: Computer and information sciences #Finance #J.3 #Machine Learning and Algorithms #and Science (cs.CE) #cs.CC #cs.CE
- Radial Oxygen Loss Triggers Diel Fluctuation of Cadmium Dissolution in the Rhizosphere of Rice
2024/08/07 by Binbin Wu, Jingyi Wang, Hengyi Dai +7 · 6 citations
Environmental Science · Agricultural and Biological Sciences · #Heavy metals in environment #Plant Stress Responses and Tolerance #Aluminum toxicity and tolerance in plants and animals
- Adaptive Knowledge Distillation between Text and Speech Pre-trained Models
2023/03/07 by Jinjie Ni, Yukun Ma, Ni, Jinjie +17 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Natural Language Processing Techniques
- Learning Disentangled Representations for Counterfactual Regression via Mutual Information Minimization
2022/06/02 by Ming‐Yuan Cheng, Xinru Liao, Cheng, Mingyuan +9 · 2 citations
Computer Science · Mathematics · #Advanced Causal Inference Techniques #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data
- I2CR: Improving Noise Robustness on Keyword Spotting Using Inter-Intra Contrastive Regularization
2022/09/14 by Dianwen Ng, Jia Qi Yip, Ng, Dianwen +15 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Multi-Modal Hybrid Deep Neural Network for Speech Enhancement
2016/06/15 by Zhenhua Wu, Wu, Zhenzhou, Sunil Sivadas +7 · 1 citation
Computer Science · Health Professions · #FOS: Computer and information sciences #Infant Health and Development #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing
- Speech Separation using Neural Audio Codecs with Embedding Loss
2024/11/27 by Jia Qi Yip, Yip, Jia Qi, Chin Yuen Kwok +5 · 4 citations
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Immunofluorescent characterization of innervation and nerve-immune cell neighborhood in mouse thymus
2019/06/22 by Huda A. M. Al-Shalan, Dailun Hu, Philip K. Nicholls +2 · 1 citation
- FD-Bench: A Full-Duplex Benchmarking Pipeline Designed for Full Duplex Spoken Dialogue Systems
2025/07/25 by Y Peng, Peng, Yizhou, Yi Chao +11 · 5 citations
Computer Science · Decision Sciences · #Speech and dialogue systems #Personal Information Management and User Behavior #Topic Modeling
- Towards Natural and Controllable Cross-Lingual Voice Conversion Based on Neural TTS Model and Phonetic Posteriorgram
2021/02/03 by Shengkui Zhao, Zhao, Shengkui, Hao Wang +5 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Natural Language Processing Techniques
- ClearerVoice-Studio: Bridging Advanced Speech Processing Research and Practical Deployment
2025/06/24 by Shengkui Zhao, Zhao, Shengkui, Bin Ma +2 · 4 citations
Computer Science · #Speech and dialogue systems
- Nitrous oxide sources, mechanisms and mitigation
2025/08/11 by Guibing Zhu, Hao Shi, Lei Zhong +27 · 1 voice · 2 citations
Chemical Engineering · Environmental Science · #Odor and Emission Control Technologies #Wastewater Treatment and Nitrogen Removal #Water Treatment and Disinfection
- HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution
2025/01/17 by Shengkui Zhao, Kun Zhou, Zhao, Shengkui +8 · 3 citations
Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Ultrasonics and Acoustic Wave Propagation #electronic engineering #information engineering
- D2Former: A Fully Complex Dual-Path Dual-Decoder Conformer Network using Joint Complex Masking and Complex Spectral Mapping for Monaural Speech Enhancement
2023/02/23 by Shengkui Zhao, Bin Ma, Zhao, Shengkui +1 · 1 citation
Computer Science · Engineering · #Speech and Audio Processing #Speech Recognition and Synthesis #Advanced Adaptive Filtering Techniques
- Dynamic Prompting of Frozen Text-to-Image Diffusion Models for Panoptic Narrative Grounding
2024/09/12 by Hongyu Li, Tianrui Hui, Li, Hongyu +13 · 2 citations
Arts and Humanities · Computer Science · #Artificial Intelligence in Games #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Narrative Theory and Analysis #Topic Modeling
- Multi-band Frequency Reconstruction for Neural Psychoacoustic Coding
2025/05/12 by Dianwen Ng, Ng, Dianwen, Kun Zhou +9 · 3 citations
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- SPGM: Prioritizing Local Features for enhanced speech separation performance
2023/09/22 by Jia Qi Yip, Yip, Jia Qi, Shengkui Zhao +19 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
PatentAgent: Intelligent Agent for Automated Pharmaceutical Patent Analysis
2024/10/25 by Xin Wang, Wang, Xin, Yifan Zhang +13 · 2 citations
Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Innovative Microfluidic and Catalytic Techniques Innovation #Machine Learning (cs.LG)
- Are Soft Prompts Good Zero-shot Learners for Speech Recognition?
2023/09/18 by Dianwen Ng, Ng, Dianwen, Chong Zhang +17 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Fun-ASR Technical Report
2025/09/15 by Keyu An, Yanni Chen, An, Keyu +62 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Independent language modeling architecture for end-to-end ASR
2019/11/25 by Van Tung Pham, Haihua Xu, Pham, Van Tung +13 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Music and Audio Processing
- Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation
2025/12/29 by Shaocong Xu, Songlin Wei, Xu, Shaocong +27 · 1 citation
Computer Science · #Generative Adversarial Networks and Image Synthesis #Advanced Vision and Imaging #Image Enhancement Techniques
- Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering
2026/07/22 by Junyu Dai, Xinyue Fan, Weiqin Li +14 · 1 citation
Computer Science · Engineering · #cs.AI #cs.SD #eess.AS
- Early Near-Infrared Excess and Rapid Disk-Corona Evolution in the Tidal Disruption Event 2024aepd
2026/07/16 by Yongxin Wu, Yanan Wang, Thomas M. Reynolds +38
#astro-ph.HE
- Bayesian Repetition Penalty: A Principled Adjacent-Conditional Framework for Reversing Attention Collapse in Autoregressive Language Models
2026/07/17 by Wenjie Fan, Bin Ma, Dong Li
#cs.AI
- TF-MossFormer: Integrating Convolution Gated Local-Global Attentions for Enhanced Time-Frequency Domain Monaural Speech Separation
2026/07/23 by Shengkui Zhao, Zexu Pan, Haoxu Wang +3
#cs.SD