vix.ing · top · new · best · stats · spec

Moninder Singh

  1. AI Fairness 360: An Extensible Toolkit for Detecting, Understanding, and\n Mitigating Unwanted Algorithmic Bias
    2018/10/03 by Rachel Bellamy, Kuntal Dey, Bellamy, Rachel K. E. +35 · 50 citations
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences
  2. One Explanation Does Not Fit All: A Toolkit and Taxonomy of AI Explainability Techniques
    2019/09/06 by Vijay Arya, Rachel Bellamy, Arya, Vijay +36 · 15 citations
    Computer Science · Decision Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (stat.ML) #Scientific Computing and Data Management
  3. Alignment Studio: Aligning Large Language Models to Particular Contextual Regulations
    2024/03/08 by Swapnaja Achintalwar, Achintalwar, Swapnaja, Ioana Baldini +35 · 5 citations
    Computer Science · #Topic Modeling
  4. Ranking Large Language Models without Ground Truth
    2024/02/21 by Amit Dhurandhar, Dhurandhar, Amit, Rahul Nair +7 · 1 voice · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #cs.AI #cs.CL #cs.LG
  5. SocialStigmaQA: A Benchmark to Uncover Stigma Amplification in Generative Language Models
    2023/12/12 by Manish Nagireddy, Lamogha Chiazor, Nagireddy, Manish +5 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Topic Modeling
  6. Your fairness may vary: Pretrained language model fairness in toxic text classification
    2021/08/03 by Ioana Baldini, Baldini, Ioana, Dennis Wei +7 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG)
  7. Reasoning about concepts with LLMs: Inconsistencies abound
    2024/05/30 by Rosario Uceda‐Sosa, Uceda-Sosa, Rosario, Karthikeyan Natesan Ramamurthy +5 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Semantic Web and Ontologies
  8. On the Safety of Interpretable Machine Learning: A Maximum Deviation Approach
    2022/11/02 by Dennis Wei, Wei, Dennis, Rahul Nair +9 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)