Huazheng Wang
- Solving Verbal Comprehension Questions in IQ Test by Knowledge-Powered Word Embedding
2015/05/29 by Huazheng Wang, Wang, Huazheng, Fei Tian +8 · 2 voices · 1 citation
Computer Science · #Intelligent Tutoring Systems and Adaptive Learning #Topic Modeling #cs.CL #cs.IR #cs.LG
- AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
2024/03/02 by Yifan Zeng, Zeng, Yifan, Yiran Wu +7 · 27 citations
Computer Science · Engineering · #Advanced Malware Detection Techniques #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Smart Grid Security and Resilience
- A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence
2025/07/28 by Huan-ang Gao, Gao, Huan-ang, Jiayi Geng +51 · 2 voices · 55 citations
#cs.AI
- Embodied LLM Agents Learn to Cooperate in Organized Teams
2024/03/19 by Xudong Guo, Kaixuan Huang, Guo, Xudong +15 · 14 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation #Multiagent Systems (cs.MA)
- LLM-RankFusion: Mitigating Intrinsic Inconsistency in LLM-based Ranking
2024/05/31 by Yifan Zeng, Ojas Tendolkar, Zeng, Yifan +9 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Imbalanced Data Classification Techniques #Information Retrieval (cs.IR)
- Incentivized Exploration for Multi-Armed Bandits under Reward Drift
2019/11/12 by Zhiyuan Liu, Huazheng Wang, Liu, Zhiyuan +7 · 2 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Adversarial Attacks on Online Learning to Rank with Stochastic Click Models
2023/05/30 by Zichen Wang, Wang, Zichen, Rishab Balasubramanian +9 · 2 citations
Computer Science · Decision Sciences · #Spam and Phishing Detection #Advanced Bandit Algorithms Research #Machine Learning and Algorithms
- A Common Pitfall of Margin-based Language Model Alignment: Gradient Entanglement
2024/10/17 by Hui Yuan, Yifan Zeng, Yuan, Hui +9 · 1 voice · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.AI #cs.CL #cs.LG
- SimpleDoc: Multi-Modal Document Understanding with Dual-Cue Page Retrieval and Iterative Refinement
2025/06/16 by Chelsi Jain, Yan Wu, Jain, Chelsi +12 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Image Retrieval and Classification Techniques #Video Analysis and Summarization
- Tree Search-Based Evolutionary Bandits for Protein Sequence Optimization
2024/01/08 by Jiahao Qiu, Qiu, Jiahao, Hui Yuan +9 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Advanced Multi-Objective Optimization Algorithms #Biomolecules (q-bio.BM) #Evolutionary Algorithms and Applications #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Stealthy Adversarial Attacks on Stochastic Multi-Armed Bandits
2024/02/21 by Zhiwei Wang, Wang, Zhiwei, Huazheng Wang +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Divide, Optimize, Merge: Fine-Grained LLM Agent Optimization at Scale
2025/05/06 by Jiale Liu, Yifan Zeng, Liu, Jiale +12 · 2 citations
Computer Science · #Advanced Neural Network Applications #Machine Learning and Data Classification #Reinforcement Learning in Robotics
- From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement
2026/07/26 by Qinsi Wang, Jing Shi, Huazheng Wang +8
#cs.AI