Simon Lermen
- Large-scale online deanonymization with LLMs
LLMs enable scalable, high-precision deanonymization of pseudonymous online accounts using unstructured text.
2026/02/18 by Simon Lermen, Daniel Paleka, Joshua Swanson +3 · 85 voices · 1 citation
#cs.CR #cs.AI #cs.LG
- LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
2023/10/31 by Simon Lermen, Charlie Rogers-Smith, Lermen, Simon +4 · 2 voices · 31 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence in Healthcare and Education #cs.AI #cs.LG
- Evaluating Large Language Models' Capability to Launch Fully Automated Spear Phishing Campaigns: Validated on Human Subjects
2024/11/30 by Fred Heiding, Heiding, Fred, Simon Lermen +7 · 5 voices · 4 citations
Computer Science · Social Sciences · #Spam and Phishing Detection #Misinformation and Its Impacts #Topic Modeling
- BadLlama: cheaply removing safety fine-tuning from Llama 2-Chat 13B
2023/10/31 by P.R. Gade, Gade, Pranav, Simon Lermen +5 · 3 citations
Computer Science · #Hate Speech and Cyberbullying Detection #Adversarial Robustness in Machine Learning
- Deceptive Automated Interpretability: Language Models Coordinating to Fool Oversight Systems
2025/04/10 by Simon Lermen, Mateusz Dziemian, Lermen, Simon +3 · 1 voice · 2 citations
#cs.AI #cs.CL