Nathan Lambert
- 2 OLMo 2 Furious
2024/12/31 by Team OLMo, OLMo, Team, P N Walsh +85 · 9 voices · 86 citations
Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Natural Language Processing Techniques
- Tulu 3: Pushing Frontiers in Open Language Model Post-Training
2024/11/22 by Nathan Lambert, Lambert, Nathan, Jacob Morrison +44 · 9 voices · 233 citations
Computer Science · #Natural Language Processing Techniques
- Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
2024/09/25 by Matt Deitke, Deitke, Matt, Christopher Clark +100 · 4 voices · 150 citations
Computer Science · #Semantic Web and Ontologies #Speech and dialogue systems
- Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
2024/01/31 by Luca Soldaini, Rodney Kinney, Soldaini, Luca +69 · 80 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
- OLMo: Accelerating the Science of Language Models
2024/02/01 by Dirk Groeneveld, Groeneveld, Dirk, Iz Beltagy +83 · 76 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques
- RewardBench: Evaluating Reward Models for Language Modeling
2024/03/20 by Nathan Lambert, Lambert, Nathan, Valentina Pyatkin +21 · 65 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques
- OLMoE: Open Mixture-of-Experts Language Models
2024/09/03 by Niklas Muennighoff, Muennighoff, Niklas, Luca Soldaini +45 · 1 voice · 60 citations
Computer Science · #Expert finding and Q&A systems #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG
- Zephyr: Direct Distillation of LM Alignment
2023/10/25 by Lewis Tunstall, Edward Beeching, Tunstall, Lewis +25 · 46 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Recommender Systems and Techniques #Speech and dialogue systems #Topic Modeling
- A Survey on Data Selection for Language Models
2024/02/26 by Alon Albalak, Yanai Elazar, Albalak, Alon +25 · 40 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback
2024/04/16 by Vincent Conitzer, Rachel Freedman, Conitzer, Vincent +21 · 1 voice · 23 citations
#cs.LG #cs.AI #cs.CL #cs.CY #cs.GT
- Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
2023/11/17 by Hamish Ivison, Ivison, Hamish, Yizhong Wang +19 · 2 voices · 17 citations
#cs.CL
- Reinforcement Learning from Human Feedback
2025/04/16 by Nathan Lambert · 9 voices · 18 citations
#cs.LG
- Spurious Rewards: Rethinking Training Signals in RLVR
2025/06/12 by Rulin Shao, Shao, Rulin, Shuyue Stella Li +25 · 59 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
- Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
2024/06/13 by Hamish Ivison, Ivison, Hamish, Yizhong Wang +15 · 15 citations
Computer Science · #Bayesian Modeling and Causal Inference #Computation and Language (cs.CL) #FOS: Computer and information sciences
- RewardBench 2: Advancing Reward Model Evaluation
2025/06/02 by Saumya Malik, Valentina Pyatkin, Malik, Saumya +11 · 31 citations
Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Intelligent Tutoring Systems and Adaptive Learning
- Low Level Control of a Quadrotor with Deep Model-Based Reinforcement\n Learning
2019/01/11 by Nathan Lambert, Lambert, Nathan O., Daniel Drew +9 · 4 citations
Computer Science · Physics and Astronomy · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Neural Networks and Reservoir Computing #Reinforcement Learning in Robotics #Robotics (cs.RO)
- Investigating Compounding Prediction Errors in Learned Dynamics Models
2022/03/17 by Nathan Lambert, Lambert, Nathan, Kristofer S. J. Pister +3 · 5 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- M-RewardBench: Evaluating Reward Models in Multilingual Settings
2024/10/20 by Srishti Gureja, Gureja, Srishti, Lester James V. Miranda +16 · 11 citations
Health Professions · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Interpreting and Communication in Healthcare #Machine Learning (cs.LG)
- Generalizing Verifiable Instruction Following
2025/07/03 by Valentina Pyatkin, Saumya Malik, Pyatkin, Valentina +13 · 24 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- Learning Accurate Long-term Dynamics for Model-based Reinforcement\n Learning
2020/12/16 by Nathan Lambert, Lambert, Nathan O., Albert Wilcox +7 · 2 citations
Computer Science · Decision Sciences · Engineering · #Reinforcement Learning in Robotics #Simulation Techniques and Applications #Advanced Control Systems Optimization
- Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
2025/11/03 by Sharan Maiya, Henning Bartsch, Maiya, Sharan +5 · 2 voices · 2 citations
Computer Science · Psychology · #cs.CL #cs.AI #cs.LG
- Self-Directed Synthetic Dialogues and Revisions Technical Report
2024/07/25 by Nathan Lambert, Lambert, Nathan, Hailey Schoelkopf +9 · 2 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Speech and dialogue systems
- Confidence-Building Measures for Artificial Intelligence: Workshop Proceedings
2023/08/01 by Sarah Shoker, Andrew W. Reddie, Shoker, Sarah +43 · 1 citation
Decision Sciences · Computer Science · #Scientific Computing and Data Management #Research Data Management Practices
- The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
2023/10/31 by Nathan Lambert, Roberto Calandra, Lambert, Nathan +1 · 1 citation
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Meta-Reinforcement Learning with Self-Reflection for Agentic Search
2026/03/11 by Teng Xiao, Yige Yuan, Hamish Ivison +6 · 1 voice · 1 citation
#cs.LG #cs.CL