Basharat, Arslan
- Defending Against Social Engineering Attacks in the Age of LLMs
2024/06/18 by Lin Ai, Tharindu Kumarage, Ai, Lin +27 · 7 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Law, AI, and Intellectual Property
- Language Models are Alignable Decision-Makers: Dataset and Application to the Medical Triage Domain
2024/06/10 by Brian Hu, Hu, Brian, Bill Ray +11 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Healthcare #Topic Modeling
- Personalized Attacks of Social Engineering in Multi-turn Conversations: LLM Agents for Simulation and Detection
2025/03/18 by Tharindu Kumarage, Kumarage, Tharindu, Cameron Johnson +17 · 1 voice · 4 citations
Computer Science · Psychology · #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #Deception detection and forensic psychology #FOS: Computer and information sciences #Information and Cyber Security #cs.CL #cs.CR
- Aligning Machiavellian Agents: Behavior Steering via Test-Time Policy Shaping
2025/11/14 by Dena F. Mujtaba, Mujtaba, Dena, Brian Hu +5 · 2 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Control (management) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Function (biology) #Key (lock) #Maximization #Reinforcement learning #Scalability
- Steerable Pluralism: Pluralistic Alignment via Few-Shot Comparative Regression
2025/08/11 by Jadie Adams, Brian Hu, Adams, Jadie +12 · 2 citations
Social Sciences · Computer Science · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #Mobile Crowdsensing and Crowdsourcing