Ronghui Mu
- A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation
2023/05/19 by Xiaowei Huang, Huang, Xiaowei, Wenjie Ruan +31 · 14 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Software Engineering Research #Artificial Intelligence in Healthcare and Education
- Building Guardrails for Large Language Models
2024/02/02 by Yi Dong, Ronghui Mu, Dong, Yi +15 · 16 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling
- Safeguarding Large Language Models: A Survey
2024/06/03 by Yi Dong, Dong, Yi, Ronghui Mu +21 · 19 citations
Computer Science · #Privacy-Preserving Technologies in Data
- Towards Fairness-Aware Adversarial Learning
2024/02/27 by Yanghao Zhang, Tianle Zhang, Zhang, Yanghao +7 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Enhancing Robust Fairness via Confusional Spectral Regularization
2025/01/22 by Gaojie Jin, Jin, Gaojie, Sihao Wu +7 · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Safety of Embodied Navigation: A Survey
2025/08/07 by Zixia Wang, Wang, Zixia, Jia Hu +3 · 2 citations
Engineering · Psychology · #Artificial Intelligence (cs.AI) #Evacuation and Crowd Dynamics #FOS: Computer and information sciences #Maritime Navigation and Safety #Robotics (cs.RO) #Social Robot Interaction and HRI