vix.ing · top · new · best · stats · spec

Wang, Joanna

  1. Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model
    2025/06/10 by Ailin Huang, Huang, Ailin, Bingxin Li +143 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering