D. O. Glukhov
- LLM Censorship: A Machine Learning Challenge or a Computer Security Problem?
2023/07/20 by David Glukhov, D. O. Glukhov, Ilia Shumailov +8 · 1 voice · 6 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling #cs.AI #cs.CL #cs.CR #cs.LG
- Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
2024/07/02 by D. O. Glukhov, Glukhov, David, Ziwen Han +7 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning