Melissa Kazemi Rad
- Refining Input Guardrails: Enhancing LLM-as-a-Judge Efficiency Through Chain-of-Thought Fine-Tuning and Alignment
2025/01/22 by Melissa Kazemi Rad, Huy Nghiem, Rad, Melissa Kazemi +7 · 9 citations
Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Building Safe GenAI Applications: An End-to-End Overview of Red Teaming for Large Language Models
2025/03/03 by Alberto Purpura, Sahil Wadhwa, Purpura, Alberto +11 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Hate Speech and Cyberbullying Detection #Advanced Malware Detection Techniques
- DynaGuard: A Dynamic Guardian Model With User-Defined Policies
2025/09/02 by Monte Hoover, Vatsal Baherwani, Hoover, Monte +17 · 2 citations
Social Sciences · #Access Control and Trust #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)