Han Bao
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1695 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- Will Large-scale Generative Models Corrupt Future Datasets?
2022/11/15 by Ryuichiro Hataya, Han Bao, Hataya, Ryuichiro +3 · 1 voice · 9 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning and Data Classification #cs.CV
- FlatQuant: Flatness Matters for LLM Quantization
2024/10/12 by Yuxuan Sun, Ruikang Liu, Sun, Yuxuan +23 · 21 citations
Engineering · Physics and Astronomy · Medicine · #Advancements in Photolithography Techniques #Magnetic confinement fusion research #Medical Imaging Techniques and Applications
- Imitation Learning from Imperfect Demonstration
2019/01/27 by Yueh-Hua Wu, Wu, Yueh-Hua, Nontawat Charoenphakdee +7 · 8 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Embarrassingly Simple Text Watermarks
2023/10/13 by Ryoma Sato, Sato, Ryoma, Yuki Takezawa +7 · 1 voice · 6 citations
Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Adversarial Robustness in Machine Learning
- TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
2025/01/28 by Makoto Shing, Kou Misaki, Shing, Makoto +7 · 2 voices · 6 citations
Computer Science · #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling #cs.AI #cs.CL #cs.LG
- On the Surrogate Gap between Contrastive and Supervised Losses
2021/10/06 by Han Bao, Yoshihiro Nagano, Bao, Han +3 · 3 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Self-attention Networks Localize When QK-eigenspectrum Concentrates
2024/02/03 by Han Bao, Ryuichiro Hataya, Bao, Han +3 · 4 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Neural Networks and Reservoir Computing
- Classification from Pairwise Similarities/Dissimilarities and Unlabeled Data via Empirical Risk Minimization
2019/04/26 by Takuya Shimada, Shimada, Takuya, Han Bao +5 · 2 citations
Computer Science · #Anomaly Detection Techniques and Applications #Machine Learning and Data Classification #Machine Learning and Algorithms
- Parameter-free Clipped Gradient Descent Meets Polyak
2024/05/23 by Yuki Takezawa, Takezawa, Yuki, Han Bao +7 · 4 citations
Computer Science · Decision Sciences · Medicine · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Medical Imaging Techniques and Applications #Optimization and Control (math.OC) #Stochastic Gradient Optimization Techniques
- Necessary and Sufficient Watermark for Large Language Models
2023/10/02 by Yuki Takezawa, Takezawa, Yuki, Ryoma Sato +7 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Calibrated Surrogate Maximization of Linear-fractional Utility in Binary Classification
2019/05/29 by Han Bao, Bao, Han, Masashi Sugiyama +1 · 1 citation
Computer Science · Mathematics · #Machine Learning and Algorithms #Imbalanced Data Classification Techniques #Statistical Methods and Inference
- AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
2024/10/28 by Han Bao, Bao, Han, Yue Huang +13 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications
- Calibrated Surrogate Losses for Adversarially Robust Classification
2020/05/28 by Han Bao, Clayton Scott, Bao, Han +3 · 1 citation
Computer Science · Mathematics · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Statistical Methods and Inference
- Dynamical Effects of Neuron Activation Gradient on Hopfield Neural Network: Numerical Analyses and Hardware Experiments
2019/04/01 by Bocheng Bao, Chengjie Chen, Han Bao +3 · 1 citation
Computer Science · Physics and Astronomy · #Neural Networks Stability and Synchronization #stochastic dynamics and bifurcation #Chaos control and synchronization
- Zipfian Whitening
2024/11/01 by Sho Yokoi, Yokoi, Sho, Han Bao +5 · 2 citations
Agricultural and Biological Sciences · Biochemistry, Genetics and Molecular Biology · Chemistry · #Computation and Language (cs.CL) #Dye analysis and toxicity #FOS: Computer and information sciences #Garlic and Onion Studies #Machine Learning (cs.LG) #Machine Learning (stat.ML) #melanin and skin pigmentation
- Beyond Exponential Graph: Communication-Efficient Topologies for Decentralized Learning via Finite-time Convergence
2023/05/19 by Yuki Takezawa, Ryoma Sato, Takezawa, Yuki +7 · 1 citation
Computer Science · #Cooperative Communication and Network Coding #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Parallel #Privacy-Preserving Technologies in Data #Stochastic Gradient Optimization Techniques #and Cluster Computing (cs.DC)
- Referee-Meta-Learning for Fast Adaptation of Locational Fairness
2024/02/20 by Weiye Chen, Chen, Weiye, Yiqun Xie +11 · 1 citation
Social Sciences · Economics, Econometrics and Finance · Decision Sciences · #Ethics and Social Impacts of AI #Insurance and Financial Risk Management #Impact of AI and Big Data on Business and Society
- Generating multi-scroll chaotic attractor in a three-dimensional memristive neuron model
2024/04/01 by Ruoyu Ding, Han Bao, Ning Wang +2 · 1 citation
- MemoHarness: Agent Harnesses That Learn from Experience
2026/07/14 by Yue Huang, Wenjie Wang, Han Bao +7 · 2 voices
#cs.AI #cs.CL