vix.ing · top · new · best · stats · spec

Junbo Zhang

  1. Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within an Open Agentic Learning Ecosystem
    2025/12/31 by Weixun Wang, XiaoXiao Xu, Wanhe An +86 · 22 voices · 2 citations
    #cs.AI #cs.CL
  2. Deep Spatio-Temporal Residual Networks for Citywide Crowd Flows Prediction
    2016/10/01 by Junbo Zhang, Zhang, Junbo, Yu Zheng +3 · 31 citations
    Computer Science · Engineering · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #Evacuation and Crowd Dynamics #FOS: Computer and information sciences #Machine Learning (cs.LG) #Traffic Prediction and Management Techniques
  3. Spatio-Temporal Graph Neural Networks for Predictive Learning in Urban Computing: A Survey
    2023/03/25 by Guangyin Jin, Jin, Guangyin, Yuxuan Liang +10 · 25 citations
    Engineering · Social Sciences · #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Smart Cities and Technologies #Traffic Prediction and Management Techniques
  4. Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning?
    2022/12/16 by Runpei Dong, Dong, Runpei, Zekun Qi +13 · 23 citations
    Computer Science · Earth and Planetary Sciences · #3D Surveying and Cultural Heritage #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences
  5. speechocean762: An Open-Source Non-native English Speech Corpus For Pronunciation Assessment
    2021/04/03 by Junbo Zhang, Zhiwen Zhang, Zhang, Junbo +15 · 15 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  6. AirFormer: Predicting Nationwide Air Quality in China with Transformers
    2022/11/29 by Yuxuan Liang, Yutong Xia, Liang, Yuxuan +13 · 17 citations
    Computer Science · Environmental Science · #Air Quality Monitoring and Forecasting #Air Quality and Health Impacts #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Signal Processing (eess.SP) #Solar Radiation and Photovoltaics #electronic engineering #information engineering
  7. Scaling up masked audio encoder learning for general audio classification
    2024/06/11 by Heinrich Dinkel, Dinkel, Heinrich, Zhiyong Yan +9 · 25 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  8. CED: Consistent ensemble distillation for audio tagging
    2023/08/23 by Heinrich Dinkel, Dinkel, Heinrich, Yongqing Wang +7 · 12 citations
    Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  9. Unified Vision-Language-Action Model
    2025/06/24 by Yuqi Wang, Wang, Yuqi, Xinghang Li +13 · 32 citations
    Arts and Humanities · Engineering · Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Geographic Information Systems Studies #Media, Religion, Digital Communication #Robotics (cs.RO) #Robotics and Automated Systems
  10. AV-SepFormer: Cross-Attention SepFormer for Audio-Visual Target Speaker Extraction
    2023/06/25 by Jiuxin Lin, Lin, Jiuxin, Xinyu Cai +17 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Predicting Citywide Crowd Flows Using Deep Spatio-Temporal Residual Networks
    2017/01/10 by Junbo Zhang, Zhang, Junbo, Yu Zheng +9 · 3 citations
    Computer Science · Engineering · Social Sciences · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Traffic Prediction and Management Techniques
  12. Attention-based End-to-End Models for Small-Footprint Keyword Spotting
    2018/03/29 by Changhao Shan, Shan, Changhao, Junbo Zhang +5 · 4 citations
    Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Natural Language Processing Techniques
  13. AutoSTL: Automated Spatio-Temporal Multi-Task Learning
    2023/04/16 by Zijian Zhang, Zhang, Zijian, Xiangyu Zhao +9 · 2 citations
    Engineering · Social Sciences · #Artificial Intelligence (cs.AI) #Automated Road and Building Extraction #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Traffic Prediction and Management Techniques
  14. HiSTGNN: Hierarchical Spatio-temporal Graph Neural Networks for Weather Forecasting
    2022/01/22 by Minbo Ma, Ma, Minbo, Peng Xie +11 · 1 citation
    Engineering · Environmental Science · #Artificial Intelligence (cs.AI) #Energy Load and Power Forecasting #FOS: Computer and information sciences #Hydrological Forecasting Using AI #Machine Learning (cs.LG)
  15. Sequence-to-sequence Models for Small-Footprint Keyword Spotting
    2018/11/01 by Haitong Zhang, Junbo Zhang, Zhang, Haitong +3 · 1 citation
    Computer Science · #Advanced Text Analysis Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  16. A spectral data release for 104 Type II Supernovae from the Tsinghua Supernova Group
    2024/01/11 by Han Lin, Lin, Han, Xiaofeng Wang +68 · 1 citation
    Physics and Astronomy · #Gamma-ray bursts and supernovae
  17. X-ARES: A Comprehensive Framework for Assessing Audio Encoder Performance
    2025/05/22 by Junbo Zhang, Zhang, Junbo, Heinrich Dinkel +9 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  18. Efficient Speech Enhancement via Embeddings from Pre-trained Generative Audioencoders
    2025/06/13 by Xingwei Sun, Sun, Xingwei, Heinrich Dinkel +8 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering