Madsen, Andreas
- Are self-explanations from Large Language Models faithful?
2024/01/15 by Andreas Nygaard Madsen, Madsen, Andreas, Sarath Chandar +3 · 18 citations
Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Topic Modeling
- Interpretability Needs a New Paradigm
2024/05/08 by Andreas Nygaard Madsen, Andreas Madsen, Himabindu Lakkaraju +6 · 1 voice · 4 citations
Health Professions · #Interpreting and Communication in Healthcare
- Neural Arithmetic Units
2020/01/14 by Madsen, Andreas, Johansen, Alexander Rosenberg · 2 citations
#FOS: Computer and information sciences #Neural and Evolutionary Computing (cs.NE)
- Evaluating the Faithfulness of Importance Measures in NLP by Recursively Masking Allegedly Important Tokens and Retraining
2021/10/15 by Madsen, Andreas, Meade, Nicholas, Adlakha, Vaibhav +1 · 2 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Faithfulness Measurable Masked Language Models
2023/10/11 by Andreas Nygaard Madsen, Siva Reddy, Madsen, Andreas +3 · 2 citations
Computer Science · #Topic Modeling #Explainable Artificial Intelligence (XAI) #Natural Language Processing Techniques
- Measuring Arithmetic Extrapolation Performance
2019/10/04 by Madsen, Andreas, Johansen, Alexander Rosenberg · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)