vix.ing · top · new · best · stats · spec

Gaojie Jin

  1. Building Guardrails for Large Language Models
    2024/02/02 by Yi Dong, Ronghui Mu, Dong, Yi +15 · 17 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling
  2. A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation
    2023/05/19 by Xiaowei Huang, Huang, Xiaowei, Wenjie Ruan +31 · 14 citations
    Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Software Engineering Research #Artificial Intelligence in Healthcare and Education
  3. Safeguarding Large Language Models: A Survey
    2024/06/03 by Yi Dong, Ronghui Mu, Dong, Yi +21 · 19 citations
    Computer Science · #Privacy-Preserving Technologies in Data
  4. SAFARI: Versatile and Efficient Evaluations for Robustness of Interpretability
    2022/08/19 by Wei Huang, Xingyu Zhao, Huang, Wei +5 · 4 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare
  5. Enhancing Robust Fairness via Confusional Spectral Regularization
    2025/01/22 by Gaojie Jin, Jin, Gaojie, Sihao Wu +7 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG)