vix.ing · top · new · best · stats · spec

Das, Nilanjana

  1. Human-Interpretable Adversarial Prompt Attack on Large Language Models with Situational Context
    2024/07/19 by Nilanjana Das, Das, Nilanjana, Edward Raff +3 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Topic Modeling
  2. Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
    2024/12/20 by Das, Nilanjana, Raff, Edward, Chadha, Aman +1 · 3 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences