vix.ing · top · new · best · stats · spec

Melissa Kazemi Rad

  1. Refining Input Guardrails: Enhancing LLM-as-a-Judge Efficiency Through Chain-of-Thought Fine-Tuning and Alignment
    2025/01/22 by Melissa Kazemi Rad, Huy Nghiem, Rad, Melissa Kazemi +7 · 9 citations
    Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  2. Building Safe GenAI Applications: An End-to-End Overview of Red Teaming for Large Language Models
    2025/03/03 by Alberto Purpura, Sahil Wadhwa, Purpura, Alberto +11 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Hate Speech and Cyberbullying Detection #Advanced Malware Detection Techniques
  3. DynaGuard: A Dynamic Guardian Model With User-Defined Policies
    2025/09/02 by Monte Hoover, Vatsal Baherwani, Hoover, Monte +17 · 2 citations
    Social Sciences · #Access Control and Trust #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)