vix.ing · top · new · best · stats · spec

Mu, Ronghui

  1. A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation
    2023/05/19 by Xiaowei Huang, Wenjie Ruan, Huang, Xiaowei +31 · 20 citations
    Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Software Engineering Research #Artificial Intelligence in Healthcare and Education
  2. Safeguarding Large Language Models: A Survey
    2024/06/03 by Yi Dong, Ronghui Mu, Dong, Yi +21 · 27 citations
    Computer Science · #Privacy-Preserving Technologies in Data
  3. Building Guardrails for Large Language Models
    2024/02/02 by Yi Dong, Dong, Yi, Ronghui Mu +15 · 18 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling
  4. Randomized Adversarial Training via Taylor Expansion
    2023/03/19 by Gaojie Jin, Xinping Yi, Jin, Gaojie +7 · 5 citations
    Computer Science · #Advanced Neural Network Applications #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  5. Towards Fairness-Aware Adversarial Learning
    2024/02/27 by Yanghao Zhang, Tianle Zhang, Zhang, Yanghao +7 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  6. Certified Policy Smoothing for Cooperative Multi-Agent Reinforcement Learning
    2022/12/22 by Ronghui Mu, Wenjie Ruan, Mu, Ronghui +7 · 2 citations
    Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Safety Systems Engineering in Autonomy
  7. Enhancing Robust Fairness via Confusional Spectral Regularization
    2025/01/22 by Gaojie Jin, Sihao Wu, Jin, Gaojie +7 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG)
  8. Reward Certification for Policy Smoothed Reinforcement Learning
    2023/12/11 by Mu, Ronghui, Marcolino, Leandro Soriano, Zhang, Tianle +3 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  9. Safety of Embodied Navigation: A Survey
    2025/08/07 by Zixia Wang, Wang, Zixia, Jia Hu +3 · 2 citations
    Engineering · Psychology · #Artificial Intelligence (cs.AI) #Evacuation and Crowd Dynamics #FOS: Computer and information sciences #Maritime Navigation and Safety #Robotics (cs.RO) #Social Robot Interaction and HRI