vix.ing · top · new · best · stats · spec

D. O. Glukhov

  1. LLM Censorship: A Machine Learning Challenge or a Computer Security Problem?
    2023/07/20 by David Glukhov, D. O. Glukhov, Ilia Shumailov +8 · 1 voice · 6 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling #cs.AI #cs.CL #cs.CR #cs.LG
  2. Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
    2024/07/02 by D. O. Glukhov, Glukhov, David, Ziwen Han +7 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning