Mitchell Gordon
- A Roadmap to Pluralistic Alignment
2024/02/07 by Taylor Sorensen, Jared Moore, Sorensen, Taylor +21 · 1 voice · 35 citations
Social Sciences · Computer Science · Medicine · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #Artificial Intelligence in Healthcare and Education
- Localizing Paragraph Memorization in Language Models
2024/03/28 by Niklas Stoehr, Mitchell Gordon, Stoehr, Niklas +5 · 7 citations
Computer Science · #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
2024/08/28 by Jillian Fisher, Fisher, Jillian, Skyler Hallinan +9 · 3 citations
Computer Science · #Authorship Attribution and Profiling #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
- MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
2025/10/18 by Yu Ying Chiu, Michael S. Lee, Chiu, Yu Ying +30 · 2 citations
Computer Science · Neuroscience · Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Psychology of Moral and Emotional Judgment
- A Roadmap to Impactful Pluralistic Alignment Research
2026/07/24 by Elinor Poole-Dayan, Jillian Fisher, Atoosa Kasirzadeh +3
#cs.AI