Lambert, Nathan
- 2 OLMo 2 Furious
2024/12/31 by Team OLMo, P N Walsh, Pete Walsh +85 · 9 voices · 86 citations
Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Natural Language Processing Techniques
- Tulu 3: Pushing Frontiers in Open Language Model Post-Training
2024/11/22 by Nathan Lambert, Jacob Morrison, Lambert, Nathan +44 · 9 voices · 233 citations
Computer Science · #Natural Language Processing Techniques
- Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
2024/09/25 by Matt Deitke, Christopher Clark, Deitke, Matt +100 · 4 voices · 150 citations
Computer Science · #Semantic Web and Ontologies #Speech and dialogue systems
- Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
2024/01/31 by Luca Soldaini, Soldaini, Luca, Rodney Kinney +69 · 80 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
- OLMo: Accelerating the Science of Language Models
2024/02/01 by Dirk Groeneveld, Iz Beltagy, Groeneveld, Dirk +83 · 76 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques
- WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
2024/06/26 by Han, Seungju, Rao, Kavel, Ettinger, Allyson +5 · 88 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- RewardBench: Evaluating Reward Models for Language Modeling
2024/03/20 by Nathan Lambert, Lambert, Nathan, Valentina Pyatkin +21 · 65 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques
- OLMoE: Open Mixture-of-Experts Language Models
2024/09/03 by Niklas Muennighoff, Muennighoff, Niklas, Luca Soldaini +45 · 1 voice · 60 citations
Computer Science · #Expert finding and Q&A systems #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG
- Zephyr: Direct Distillation of LM Alignment
2023/10/25 by Lewis Tunstall, Tunstall, Lewis, Edward Beeching +25 · 45 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Recommender Systems and Techniques #Speech and dialogue systems #Topic Modeling
- A Survey on Data Selection for Language Models
2024/02/26 by Alon Albalak, Yanai Elazar, Albalak, Alon +25 · 40 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback
2024/04/16 by Vincent Conitzer, Rachel Freedman, Conitzer, Vincent +21 · 1 voice · 23 citations
#cs.LG #cs.AI #cs.CL #cs.CY #cs.GT
- Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
2023/11/17 by Hamish Ivison, Ivison, Hamish, Yizhong Wang +19 · 2 voices · 17 citations
#cs.CL
- Spurious Rewards: Rethinking Training Signals in RLVR
2025/06/12 by Rulin Shao, Shao, Rulin, Shuyue Stella Li +25 · 59 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
- Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
2024/06/13 by Hamish Ivison, Yizhong Wang, Ivison, Hamish +15 · 15 citations
Computer Science · #Bayesian Modeling and Causal Inference #Computation and Language (cs.CL) #FOS: Computer and information sciences
- RewardBench 2: Advancing Reward Model Evaluation
2025/06/02 by Saumya Malik, Malik, Saumya, Valentina Pyatkin +11 · 31 citations
Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Intelligent Tutoring Systems and Adaptive Learning
- Investigating Compounding Prediction Errors in Learned Dynamics Models
2022/03/17 by Nathan Lambert, Lambert, Nathan, Kristofer S. J. Pister +3 · 5 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- M-RewardBench: Evaluating Reward Models in Multilingual Settings
2024/10/20 by Srishti Gureja, Lester James V. Miranda, Gureja, Srishti +16 · 11 citations
Health Professions · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Interpreting and Communication in Healthcare #Machine Learning (cs.LG)
- Generalizing Verifiable Instruction Following
2025/07/03 by Valentina Pyatkin, Pyatkin, Valentina, Saumya Malik +13 · 24 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- On the Importance of Hyperparameter Optimization for Model-based Reinforcement Learning
2021/02/26 by Zhang, Baohe, Rajan, Raghu, Pineda, Luis +5 · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Systems and Control (eess.SY) #electronic engineering #information engineering
- Towards a Framework for Openness in Foundation Models: Proceedings from the Columbia Convening on Openness in Artificial Intelligence
2024/05/17 by Basdevant, Adrien, François, Camille, Storchan, Victor +13 · 4 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Software Engineering (cs.SE)
- A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
2023/10/10 by Wei, Ran, Lambert, Nathan, McDonald, Anthony +2 · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Learning Generalizable Locomotion Skills with Hierarchical Reinforcement Learning
2019/09/26 by Li, Tianyu, Lambert, Nathan, Calandra, Roberto +2 · 1 citation
#FOS: Computer and information sciences #Robotics (cs.RO)
- Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
2025/11/03 by Sharan Maiya, Maiya, Sharan, Henning Bartsch +5 · 2 voices · 2 citations
Computer Science · Psychology · #cs.CL #cs.AI #cs.LG
- Reward Reports for Reinforcement Learning
2022/04/22 by Gilbert, Thomas Krendl, Lambert, Nathan, Dean, Sarah +2 · 1 citation
#Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Self-Directed Synthetic Dialogues and Revisions Technical Report
2024/07/25 by Nathan Lambert, Lambert, Nathan, Hailey Schoelkopf +9 · 2 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Speech and dialogue systems
- The Challenges of Exploration for Offline Reinforcement Learning
2022/01/27 by Lambert, Nathan, Wulfmeier, Markus, Whitney, William +5 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Confidence-Building Measures for Artificial Intelligence: Workshop Proceedings
2023/08/01 by Sarah Shoker, Andrew W. Reddie, Shoker, Sarah +43 · 1 citation
Decision Sciences · Computer Science · #Scientific Computing and Data Management #Research Data Management Practices
- The History and Risks of Reinforcement Learning and Human Feedback
2023/10/20 by Lambert, Nathan, Gilbert, Thomas Krendl, Zick, Tom · 1 citation
#Computers and Society (cs.CY) #FOS: Computer and information sciences
- The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
2023/10/31 by Nathan Lambert, Roberto Calandra, Lambert, Nathan +1 · 1 citation
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics