vix.ing · top · new · best · stats · spec

Been Kim

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1374 citations
    #cs.CL #cs.AI
  2. Acquisition of Chess Knowledge in AlphaZero
    2021/11/17 by Thomas McGrath, Andrei Kapishnikov, Nenad Tomašev +5 · 5 voices · 4 citations
    #cs.AI #stat.ML
  3. Interpretability Beyond Feature Attribution: Quantitative Testing with Concept Activation Vectors (TCAV)
    2017/11/30 by Been Kim, Martin Wattenberg, Justin Gilmer +4 · 3 voices · 71 citations
    #stat.ML
  4. Towards A Rigorous Science of Interpretable Machine Learning
    2017/02/28 by Finale Doshi‐Velez, Been Kim, Doshi-Velez, Finale +1 · 144 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  5. Getting aligned on representational alignment
    2023/10/18 by Ilia Sucholutsky, Lukas Muttenthaler, Sucholutsky, Ilia +66 · 10 voices · 34 citations
    Biochemistry, Genetics and Molecular Biology · Materials Science · Neuroscience · #Bioinformatics and Genomic Networks #Machine Learning in Materials Science #Neural dynamics and brain function #cs.AI #cs.LG #cs.NE #q-bio.NC
  6. Sanity Checks for Saliency Maps
    2018/10/08 by Julius Adebayo, Justin Gilmer, Adebayo, Julius +9 · 130 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cell Image Analysis Techniques #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Visual Attention and Saliency Detection
  7. Concept Bottleneck Models
    2020/07/09 by Pang Wei Koh, Koh, Pang Wei, Thao Nguyen +11 · 136 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Machine Learning in Bioinformatics #Metabolomics and Mass Spectrometry Studies #Machine Learning in Healthcare
  8. SmoothGrad: removing noise by adding noise
    2017/06/12 by Daniel Smilkov, Nikhil Thorat, Smilkov, Daniel +7 · 113 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Anomaly Detection Techniques and Applications #Cell Image Analysis Techniques #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  9. Video models are zero-shot learners and reasoners
    2025/09/24 by Thaddäus Wiedemer, Yuxuan Li, Wiedemer, Thaddäus +15 · 7 voices · 56 citations
    #cs.LG #cs.AI #cs.CV #cs.RO
  10. Bridging the Human-AI Knowledge Gap: Concept Discovery and Transfer in AlphaZero
    2023/10/25 by Lisa Schut, Nenad Tomašev, Schut, Lisa +10 · 6 voices · 4 citations
    Computer Science · Economics, Econometrics and Finance · #Data Visualization and Analytics #Sports Analytics and Performance #Time Series Analysis and Forecasting #cs.AI #cs.HC #cs.LG #stat.ML
  11. Learning how to explain neural networks: PatternNet and PatternAttribution
    2017/05/16 by Pieter Jan Kindermans, Kristof T. Schütt, Kindermans, Pieter-Jan +11 · 28 citations
    Computer Science · #Neural Networks and Applications
  12. Does Localization Inform Editing? Surprising Differences in Causality-Based Localization vs. Knowledge Editing in Language Models
    2023/01/10 by Peter Hase, Mohit Bansal, Hase, Peter +5 · 1 voice · 18 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Explainable Artificial Intelligence (XAI)
  13. On Completeness-aware Concept-Based Explanations in Deep Neural Networks
    2019/10/17 by Chih‐Kuan Yeh, Been Kim, Yeh, Chih-Kuan +9 · 12 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  14. To Trust Or Not To Trust A Classifier
    2018/05/30 by Heinrich Jiang, Been Kim, Jiang, Heinrich +5 · 11 citations
    Computer Science · Mathematics · #Topological and Geometric Data Analysis #Statistical Methods and Inference
  15. Model evaluation for extreme risks
    2023/05/24 by Toby Shevlane, Sebastian Farquhar, Shevlane, Toby +39 · 16 citations
    Computer Science · #Software Engineering Research #Software Reliability and Analysis Research #Information and Cyber Security
  16. QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
    2025/03/28 by Belinda Z. Li, Li, Belinda Z., Been Kim +3 · 4 voices · 13 citations
    #cs.AI #cs.CL #cs.LG
  17. Human-Centered Tools for Coping with Imperfect Algorithms during Medical Decision-Making
    2019/02/08 by Carrie J. Cai, Cai, Carrie J., Emily Reif +19 · 7 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · Medicine · #AI in cancer detection #Artificial Intelligence in Healthcare and Education #Biomedical Text Mining and Ontologies #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning in Healthcare
  18. Towards Realistic Individual Recourse and Actionable Explanations in Black-Box Decision Making Systems
    2019/07/22 by Shalmali Joshi, Oluwasanmi Koyejo, Joshi, Shalmali +7 · 11 citations
    Decision Sciences · Computer Science · #Scientific Computing and Data Management #Explainable Artificial Intelligence (XAI) #Bayesian Modeling and Causal Inference
  19. We Can't Understand AI Using our Existing Vocabulary
    2025/02/11 by John Hewitt, Robert Geirhos, Hewitt, John +3 · 3 voices · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Topic Modeling
  20. Explaining Classifiers with Causal Concept Effect (CaCE)
    2019/07/16 by Yash Goyal, Amir Feder, Goyal, Yash +5 · 6 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  21. Benchmarking Attribution Methods with Relative Feature Importance
    2019/07/23 by Mengjiao Yang, Yang, Mengjiao, Been Kim +1 · 6 citations
    Computer Science · #Explainable Artificial Intelligence (XAI) #Adversarial Robustness in Machine Learning #Machine Learning and Data Classification
  22. Don't trust your eyes: on the (un)reliability of feature visualizations
    2023/06/07 by Robert Geirhos, Roland S. Zimmermann, R. Zimmermann +8 · 1 voice · 4 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cell Image Analysis Techniques #Explainable Artificial Intelligence (XAI) #Neural Networks and Applications
  23. Interpreting Black Box Predictions using Fisher Kernels
    2018/10/23 by Rajiv Khanna, Been Kim, Khanna, Rajiv +5 · 4 citations
    Computer Science · #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Gaussian Processes and Bayesian Inference #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  24. Because we have LLMs, we Can and Should Pursue Agentic Interpretability
    2025/06/13 by Been Kim, Kim, Been, John Hewitt +7 · 1 voice · 2 citations
    Computer Science · Social Sciences · #Multi-Agent Systems and Negotiation #Natural Language Processing Techniques #European and International Law Studies
  25. Human-Centered Concept Explanations for Neural Networks
    2022/02/25 by Chih‐Kuan Yeh, Yeh, Chih-Kuan, Been Kim +3 · 3 citations
    Computer Science · Materials Science · #Explainable Artificial Intelligence (XAI) #Machine Learning in Materials Science #Adversarial Robustness in Machine Learning
  26. State2Explanation: Concept-Based Explanations to Benefit Agent Learning and User Understanding
    2023/09/21 by Devleena Das, Sonia Chernova, Das, Devleena +3 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Bayesian Modeling and Causal Inference #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  27. Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
    2024/12/09 by Meera Hahn, Hahn, Meera, Wenjun Zeng +11 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Retrieval and Classification Techniques #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  28. How new data permeates LLM knowledge and how to dilute it
    2025/04/13 by Chen Sun, Sun, Chen, Renat Aksitov +12 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Library Science and Information Systems #Natural Language Processing Techniques #Semantic Web and Ontologies
  29. Visual prompt engineering for video models
    2026/07/28 by Robert Geirhos, Yuxuan Li, Thaddäus Wiedemer +7 · 1 voice
    #cs.CV #cs.AI