vix.ing · top · new · best · stats · spec

Harsh Chaudhari

  1. The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
    2025/10/10 by Milad Nasr, Nasr, Milad, Nicholas Carlini +27 · 4 voices · 17 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Network Security and Intrusion Detection #Security and Verification in Computing #cs.CR #cs.LG
  2. Measuring memorization in language models via probabilistic extraction
    2024/10/25 by Jamie Hayes, Marika Swanberg, Hayes, Jamie +11 · 13 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
  3. Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation
    2025/02/01 by Ali Naseh, Naseh, Ali, Yuefeng Peng +9 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #Domain Adaptation and Few-Shot Learning #Expert finding and Q&A systems #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Topic Modeling
  4. SNAP: Efficient Extraction of Private Properties with Poisoning
    2022/08/25 by Harsh Chaudhari, John Abascal, Chaudhari, Harsh +9 · 2 citations
    Computer Science · Medicine · #Privacy-Preserving Technologies in Data #Adversarial Robustness in Machine Learning #Artificial Intelligence in Healthcare and Education
  5. UTrace: Poisoning Forensics for Private Collaborative Learning
    2024/09/23 by Evan Rose, Hidde Lycklama, Rose, Evan +8 · 1 citation
    Psychology · Social Sciences · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Innovative Teaching and Learning Methods #Machine Learning (cs.LG) #Problem and Project Based Learning #Wikis in Education and Collaboration
  6. Cascading Adversarial Bias from Injection to Distillation in Language Models
    2025/05/30 by Harsh Chaudhari, Chaudhari, Harsh, Jamie Hayes +9 · 2 voices · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling #cs.CR #cs.LG