Junbo Zhang
- Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within an Open Agentic Learning Ecosystem
2025/12/31 by Weixun Wang, XiaoXiao Xu, Wanhe An +86 · 22 voices · 2 citations
#cs.AI #cs.CL
- Deep Spatio-Temporal Residual Networks for Citywide Crowd Flows Prediction
2016/10/01 by Junbo Zhang, Zhang, Junbo, Yu Zheng +3 · 31 citations
Computer Science · Engineering · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #Evacuation and Crowd Dynamics #FOS: Computer and information sciences #Machine Learning (cs.LG) #Traffic Prediction and Management Techniques
- Spatio-Temporal Graph Neural Networks for Predictive Learning in Urban Computing: A Survey
2023/03/25 by Guangyin Jin, Jin, Guangyin, Yuxuan Liang +10 · 25 citations
Engineering · Social Sciences · #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Smart Cities and Technologies #Traffic Prediction and Management Techniques
- Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning?
2022/12/16 by Runpei Dong, Dong, Runpei, Zekun Qi +13 · 23 citations
Computer Science · Earth and Planetary Sciences · #3D Surveying and Cultural Heritage #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences
- speechocean762: An Open-Source Non-native English Speech Corpus For Pronunciation Assessment
2021/04/03 by Junbo Zhang, Zhiwen Zhang, Zhang, Junbo +15 · 15 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- AirFormer: Predicting Nationwide Air Quality in China with Transformers
2022/11/29 by Yuxuan Liang, Yutong Xia, Liang, Yuxuan +13 · 17 citations
Computer Science · Environmental Science · #Air Quality Monitoring and Forecasting #Air Quality and Health Impacts #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Signal Processing (eess.SP) #Solar Radiation and Photovoltaics #electronic engineering #information engineering
- Scaling up masked audio encoder learning for general audio classification
2024/06/11 by Heinrich Dinkel, Dinkel, Heinrich, Zhiyong Yan +9 · 25 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- CED: Consistent ensemble distillation for audio tagging
2023/08/23 by Heinrich Dinkel, Dinkel, Heinrich, Yongqing Wang +7 · 12 citations
Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Unified Vision-Language-Action Model
2025/06/24 by Yuqi Wang, Wang, Yuqi, Xinghang Li +13 · 32 citations
Arts and Humanities · Engineering · Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Geographic Information Systems Studies #Media, Religion, Digital Communication #Robotics (cs.RO) #Robotics and Automated Systems
- AV-SepFormer: Cross-Attention SepFormer for Audio-Visual Target Speaker Extraction
2023/06/25 by Jiuxin Lin, Lin, Jiuxin, Xinyu Cai +17 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Predicting Citywide Crowd Flows Using Deep Spatio-Temporal Residual Networks
2017/01/10 by Junbo Zhang, Zhang, Junbo, Yu Zheng +9 · 3 citations
Computer Science · Engineering · Social Sciences · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Traffic Prediction and Management Techniques
- Attention-based End-to-End Models for Small-Footprint Keyword Spotting
2018/03/29 by Changhao Shan, Shan, Changhao, Junbo Zhang +5 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Natural Language Processing Techniques
- AutoSTL: Automated Spatio-Temporal Multi-Task Learning
2023/04/16 by Zijian Zhang, Zhang, Zijian, Xiangyu Zhao +9 · 2 citations
Engineering · Social Sciences · #Artificial Intelligence (cs.AI) #Automated Road and Building Extraction #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Traffic Prediction and Management Techniques
- HiSTGNN: Hierarchical Spatio-temporal Graph Neural Networks for Weather Forecasting
2022/01/22 by Minbo Ma, Ma, Minbo, Peng Xie +11 · 1 citation
Engineering · Environmental Science · #Artificial Intelligence (cs.AI) #Energy Load and Power Forecasting #FOS: Computer and information sciences #Hydrological Forecasting Using AI #Machine Learning (cs.LG)
- Sequence-to-sequence Models for Small-Footprint Keyword Spotting
2018/11/01 by Haitong Zhang, Junbo Zhang, Zhang, Haitong +3 · 1 citation
Computer Science · #Advanced Text Analysis Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- A spectral data release for 104 Type II Supernovae from the Tsinghua Supernova Group
2024/01/11 by Han Lin, Lin, Han, Xiaofeng Wang +68 · 1 citation
Physics and Astronomy · #Gamma-ray bursts and supernovae
- X-ARES: A Comprehensive Framework for Assessing Audio Encoder Performance
2025/05/22 by Junbo Zhang, Zhang, Junbo, Heinrich Dinkel +9 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Efficient Speech Enhancement via Embeddings from Pre-trained Generative Audioencoders
2025/06/13 by Xingwei Sun, Sun, Xingwei, Heinrich Dinkel +8 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering