Wes Gurnee
- External Invariants: A Cryptographic Trust Architecture for Institutional AI Inference
2024/06/17 by Andy Arditi, Arditi, Andy, Oscar Obeso +11 · 20 voices · 190 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG
- Language Models Represent Space and Time
2023/10/03 by Wes Gurnee, Max Tegmark, Gurnee, Wes +1 · 12 voices · 46 citations
Computer Science · Social Sciences · #Language and cultural evolution #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG
- Finding Neurons in a Haystack: Case Studies with Sparse Probing
2023/05/02 by Wes Gurnee, Gurnee, Wes, Neel Nanda +10 · 2 voices · 43 citations
Computer Science · #Explainable Artificial Intelligence (XAI) #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.LG
- Not All Language Model Features Are One-Dimensionally Linear
2024/05/23 by Joshua Engels, Engels, Joshua, Eric J. Michaud +7 · 36 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
- Confidence Regulation Neurons in Language Models
2024/06/24 by Alessandro Stolfo, Stolfo, Alessandro, Ben Wu +11 · 18 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- The Remarkable Robustness of LLMs: Stages of Inference?
2024/06/27 by Vedang Lad, Lad, Vedang, Lee, Jin Hwa +4 · 18 citations
Social Sciences · #Artificial Intelligence in Law
- Learning Sparse Nonlinear Dynamics via Mixed-Integer Optimization
2022/06/01 by Dimitris Bertsimas, Wes Gurnee, Bertsimas, Dimitris +1 · 5 citations
Engineering · Physics and Astronomy · #Fault Detection and Control Systems #Model Reduction and Neural Networks #Control Systems and Identification
- When Models Manipulate Manifolds: The Geometry of a Counting Task
2026/01/08 by Wes Gurnee, Emmanuel Ameisen, Isaac Kauvar +4 · 4 voices · 2 citations
#cs.LG
- Multilevel Interpretability Of Artificial Neural Networks: Leveraging Framework And Methods From Neuroscience
2024/08/22 by Zhonghao He, He, Zhonghao, Jascha Achterberg +31 · 1 voice · 3 citations
Computer Science · #Explainable Artificial Intelligence (XAI)
- Fairmandering: A column generation heuristic for fairness-optimized political districting
2021/03/21 by Wes Gurnee, David B. Shmoys, Gurnee, Wes +1 · 2 citations
Economics, Econometrics and Finance · Social Sciences · #Game Theory and Voting Systems #Electoral Systems and Political Participation #Local Government Finance and Decentralization
- Verbalizable Representations Form a Global Workspace in Language Models
2026/07/16 by Wes Gurnee, Nicholas Sofroniew, Adam Pearce +13 · 2 voices · 8 citations
#cs.CL #cs.AI #cs.LG