vix.ing · top · new · best · stats · spec

Nouha Dziri

  1. Faith and Fate: Limits of Transformers on Compositionality
    2023/05/29 by Nouha Dziri, Dziri, Nouha, Ximing Lu +29 · 13 voices · 72 citations
    Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Natural Language Processing Techniques #Topic Modeling
  2. Self-Refine: Iterative Refinement with Self-Feedback
    2023/03/30 by Aman Madaan, Madaan, Aman, Niket Tandon +29 · 1 voice · 572 citations
    Computer Science · #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling #cs.AI #cs.CL #cs.LG
  3. 2 OLMo 2 Furious
    2024/12/31 by Team OLMo, Pete Walsh, P N Walsh +85 · 9 voices · 87 citations
    Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Natural Language Processing Techniques
  4. Tulu 3: Pushing Frontiers in Open Language Model Post-Training
    2024/11/22 by Nathan Lambert, Jacob Morrison, Lambert, Nathan +44 · 9 voices · 240 citations
    Computer Science · #Natural Language Processing Techniques
  5. The Generative AI Paradox: "What It Can Create, It May Not Understand"
    2023/10/31 by Peter West, Ximing Lu, West, Peter +27 · 5 voices · 13 citations
    Computer Science · Social Sciences · Medicine · #cs.AI #cs.CL #cs.CV #cs.LG
  6. Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
    2025/10/27 by Liwei Jiang, Yuanjun Chai, Jiang, Liwei +17 · 15 voices · 13 citations
    #cs.CL
  7. Fine-Grained Human Feedback Gives Better Rewards for Language Model Training
    2023/06/02 by Zeqiu Wu, Wu, Zeqiu, Yushi Hu +15 · 1 voice · 33 citations
    Computer Science · #Natural Language Processing Techniques #Text Readability and Simplification #Topic Modeling #cs.CL
  8. RewardBench: Evaluating Reward Models for Language Modeling
    2024/03/20 by Nathan Lambert, Lambert, Nathan, Valentina Pyatkin +21 · 68 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  9. AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
    2024/10/05 by Ximing Lu, Melanie Sclar, Lu, Ximing +19 · 8 voices · 13 citations
    #cs.CL
  10. A Roadmap to Pluralistic Alignment
    2024/02/07 by Taylor Sorensen, Jared Moore, Sorensen, Taylor +21 · 1 voice · 30 citations
    Social Sciences · Computer Science · Medicine · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #Artificial Intelligence in Healthcare and Education
  11. WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
    2024/06/26 by Liwei Jiang, Jiang, Liwei, Kavel Rao +19 · 50 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences
  12. The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning
    2023/12/04 by Bill Yuchen Lin, Lin, Bill Yuchen, Abhilasha Ravichander +13 · 33 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Text Readability and Simplification
  13. FaithDial: A Faithful Benchmark for Information-Seeking Dialogue
    2022/04/22 by Nouha Dziri, Dziri, Nouha, Ehsan Kamalloo +11 · 15 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
  14. On the Origin of Hallucinations in Conversational Models: Is it the Datasets or the Models?
    2022/04/17 by Nouha Dziri, Dziri, Nouha, Sivan Milton +7 · 13 citations
    Social Sciences · Psychology · Computer Science · #Misinformation and Its Impacts #Mental Health via Writing #Machine Learning in Healthcare
  15. The Art of Saying No: Contextual Noncompliance in Language Models
    2024/07/02 by Faeze Brahman, Brahman, Faeze, Sachin Kumar +26 · 2 voices · 15 citations
    Computer Science · #Hate Speech and Cyberbullying Detection #Multi-Agent Systems and Negotiation #cs.AI #cs.CL #cs.HC
  16. Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
    2023/10/12 by Linlu Qiu, Qiu, Linlu, Liwei Jiang +19 · 9 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
  17. Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
    2024/07/10 by Kaitlyn Zhou, Zhou, Kaitlyn, Jena D. Hwang +9 · 2 voices · 4 citations
    Decision Sciences · Computer Science · #Complex Systems and Decision Making #Software Engineering Techniques and Practices #Cognitive Science and Mapping
  18. Evaluating Attribution in Dialogue Systems: The BEGIN Benchmark
    2021/04/30 by Nouha Dziri, Dziri, Nouha, Hannah Rashkin +5 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  19. OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety
    2025/07/08 by Sanidhya Vijayvargiya, Vijayvargiya, Sanidhya, Akshay Soni +11 · 12 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Security and Verification in Computing #Explainable Artificial Intelligence (XAI)
  20. The Singapore Consensus on Global AI Safety Research Priorities
    2025/06/25 by Yoshua Bengio, Tegan Maharaj, Bengio, Yoshua +171 · 2 voices · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #cs.AI #cs.CY
  21. SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
    2024/10/22 by Jing‐Jing Li, Li, Jing-Jing, Valentina Pyatkin +17 · 6 citations
    Decision Sciences · Engineering · Computer Science · #Risk and Safety Analysis #Safety Systems Engineering in Autonomy #Adversarial Robustness in Machine Learning
  22. What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
    2023/10/24 by Kavel Rao, Liwei Jiang, Rao, Kavel +13 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Topic Modeling #Explainable Artificial Intelligence (XAI)
  23. Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
    2025/04/16 by Youfa Sun, Sun, Yiyou, Bai, Haoyue +9 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  24. To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
    2024/10/15 by Sean McGregor, Allyson Ettinger, McGregor, Sean +24 · 2 voices · 1 citation
    Computer Science · #cs.CY #cs.LG #cs.SE
  25. RL Grokking Recipe: How Does RL Unlock and Transfer New Algorithms in LLMs?
    2025/09/25 by Yiyou Sun, Sun, Yiyou, Yuhan Cao +11 · 6 citations
    Computer Science · #Distributed and Parallel Computing Systems