vix.ing · top · new · best · stats · spec

Lambert, Nathan

  1. 2 OLMo 2 Furious
    2024/12/31 by Team OLMo, P N Walsh, Pete Walsh +85 · 9 voices · 86 citations
    Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Natural Language Processing Techniques
  2. Tulu 3: Pushing Frontiers in Open Language Model Post-Training
    2024/11/22 by Nathan Lambert, Jacob Morrison, Lambert, Nathan +44 · 9 voices · 233 citations
    Computer Science · #Natural Language Processing Techniques
  3. Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
    2024/09/25 by Matt Deitke, Christopher Clark, Deitke, Matt +100 · 4 voices · 150 citations
    Computer Science · #Semantic Web and Ontologies #Speech and dialogue systems
  4. Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
    2024/01/31 by Luca Soldaini, Soldaini, Luca, Rodney Kinney +69 · 80 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
  5. OLMo: Accelerating the Science of Language Models
    2024/02/01 by Dirk Groeneveld, Iz Beltagy, Groeneveld, Dirk +83 · 76 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  6. WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
    2024/06/26 by Han, Seungju, Rao, Kavel, Ettinger, Allyson +5 · 88 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  7. RewardBench: Evaluating Reward Models for Language Modeling
    2024/03/20 by Nathan Lambert, Lambert, Nathan, Valentina Pyatkin +21 · 65 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  8. OLMoE: Open Mixture-of-Experts Language Models
    2024/09/03 by Niklas Muennighoff, Muennighoff, Niklas, Luca Soldaini +45 · 1 voice · 60 citations
    Computer Science · #Expert finding and Q&A systems #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG
  9. Zephyr: Direct Distillation of LM Alignment
    2023/10/25 by Lewis Tunstall, Tunstall, Lewis, Edward Beeching +25 · 45 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Recommender Systems and Techniques #Speech and dialogue systems #Topic Modeling
  10. A Survey on Data Selection for Language Models
    2024/02/26 by Alon Albalak, Yanai Elazar, Albalak, Alon +25 · 40 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
  11. Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback
    2024/04/16 by Vincent Conitzer, Rachel Freedman, Conitzer, Vincent +21 · 1 voice · 23 citations
    #cs.LG #cs.AI #cs.CL #cs.CY #cs.GT
  12. Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
    2023/11/17 by Hamish Ivison, Ivison, Hamish, Yizhong Wang +19 · 2 voices · 17 citations
    #cs.CL
  13. Spurious Rewards: Rethinking Training Signals in RLVR
    2025/06/12 by Rulin Shao, Shao, Rulin, Shuyue Stella Li +25 · 59 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
  14. Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
    2024/06/13 by Hamish Ivison, Yizhong Wang, Ivison, Hamish +15 · 15 citations
    Computer Science · #Bayesian Modeling and Causal Inference #Computation and Language (cs.CL) #FOS: Computer and information sciences
  15. RewardBench 2: Advancing Reward Model Evaluation
    2025/06/02 by Saumya Malik, Malik, Saumya, Valentina Pyatkin +11 · 31 citations
    Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Intelligent Tutoring Systems and Adaptive Learning
  16. Investigating Compounding Prediction Errors in Learned Dynamics Models
    2022/03/17 by Nathan Lambert, Lambert, Nathan, Kristofer S. J. Pister +3 · 5 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  17. M-RewardBench: Evaluating Reward Models in Multilingual Settings
    2024/10/20 by Srishti Gureja, Lester James V. Miranda, Gureja, Srishti +16 · 11 citations
    Health Professions · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Interpreting and Communication in Healthcare #Machine Learning (cs.LG)
  18. Generalizing Verifiable Instruction Following
    2025/07/03 by Valentina Pyatkin, Pyatkin, Valentina, Saumya Malik +13 · 24 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  19. On the Importance of Hyperparameter Optimization for Model-based Reinforcement Learning
    2021/02/26 by Zhang, Baohe, Rajan, Raghu, Pineda, Luis +5 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Systems and Control (eess.SY) #electronic engineering #information engineering
  20. Towards a Framework for Openness in Foundation Models: Proceedings from the Columbia Convening on Openness in Artificial Intelligence
    2024/05/17 by Basdevant, Adrien, François, Camille, Storchan, Victor +13 · 4 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Software Engineering (cs.SE)
  21. A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
    2023/10/10 by Wei, Ran, Lambert, Nathan, McDonald, Anthony +2 · 2 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  22. Learning Generalizable Locomotion Skills with Hierarchical Reinforcement Learning
    2019/09/26 by Li, Tianyu, Lambert, Nathan, Calandra, Roberto +2 · 1 citation
    #FOS: Computer and information sciences #Robotics (cs.RO)
  23. Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
    2025/11/03 by Sharan Maiya, Maiya, Sharan, Henning Bartsch +5 · 2 voices · 2 citations
    Computer Science · Psychology · #cs.CL #cs.AI #cs.LG
  24. Reward Reports for Reinforcement Learning
    2022/04/22 by Gilbert, Thomas Krendl, Lambert, Nathan, Dean, Sarah +2 · 1 citation
    #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. Self-Directed Synthetic Dialogues and Revisions Technical Report
    2024/07/25 by Nathan Lambert, Lambert, Nathan, Hailey Schoelkopf +9 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Speech and dialogue systems
  26. The Challenges of Exploration for Offline Reinforcement Learning
    2022/01/27 by Lambert, Nathan, Wulfmeier, Markus, Whitney, William +5 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  27. Confidence-Building Measures for Artificial Intelligence: Workshop Proceedings
    2023/08/01 by Sarah Shoker, Andrew W. Reddie, Shoker, Sarah +43 · 1 citation
    Decision Sciences · Computer Science · #Scientific Computing and Data Management #Research Data Management Practices
  28. The History and Risks of Reinforcement Learning and Human Feedback
    2023/10/20 by Lambert, Nathan, Gilbert, Thomas Krendl, Zick, Tom · 1 citation
    #Computers and Society (cs.CY) #FOS: Computer and information sciences
  29. The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
    2023/10/31 by Nathan Lambert, Roberto Calandra, Lambert, Nathan +1 · 1 citation
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics