vix.ing · top · new · best · stats · spec

Che, Zora

  1. Can Watermarking Large Language Models Prevent Copyrighted Text Generation and Hide Training Data?
    2024/07/24 by Michael-Andrei Panaitescu-Liess, Zora Che, Panaitescu-Liess, Michael-Andrei +15 · 2 voices · 4 citations
    #cs.LG
  2. Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities
    2025/02/03 by Zora Che, Stephen Casper, Che, Zora +27 · 10 citations
    Computer Science · #Security and Verification in Computing #Advanced Malware Detection Techniques #Network Security and Intrusion Detection
  3. SAIL: Self-Improving Efficient Online Alignment of Large Language Models
    2024/06/21 by Mucong Ding, Souradip Chakraborty, Ding, Mucong +13 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  4. EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
    2024/10/06 by Aakriti Agrawal, Mucong Ding, Agrawal, Aakriti +11 · 4 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
  5. Transferring Fairness under Distribution Shifts via Fair Consistency Regularization
    2022/06/26 by Bang An, Zora Che, An, Bang +5 · 1 citation
    Computer Science · Social Sciences · #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data #Retirement, Disability, and Employment
  6. AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
    2025/04/29 by Zikui Cai, Cai, Zikui, Shayan Shabihi +13 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Security and Verification in Computing
  7. EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
    2025/05/28 by Agrawal, Aakriti, Ding, Mucong, Che, Zora +6 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)