Gerald Shen
- Llama-Nemotron: Efficient Reasoning Models
2025/05/02 by Akhiad Bercovich, Bercovich, Akhiad, Itay Levy +224 · 1 voice · 29 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Topic Modeling
- HelpSteer2: Open-source dataset for training top-performing reward models
2024/06/12 by Zhilin Wang, Yi Dong, Wang, Zhilin +15 · 37 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare
- NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
2024/05/02 by Gerald Shen, Shen, Gerald, Zhilin Wang +23 · 12 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computational Physics and Python Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Time Series Analysis and Forecasting
- NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model
2025/08/20 by NVIDIA, Aarti Basant, : +305 · 23 citations
Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Big Data and Digital Economy #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel Computing and Optimization Techniques
- Scaling Up RL: Unlocking Diverse Reasoning in LLMs via Prolonged Training
2025/07/16 by Mingjie Liu, Liu, Mingjie, Shizhe Diao +40 · 12 citations
Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
2025/12/23 by NVIDIA, Aaron Blakeman, : +623 · 1 voice · 2 citations
#cs.CL #cs.AI #cs.LG
- Reward-aware Preference Optimization: A Unified Mathematical Framework for Model Alignment
2025/01/31 by Shengyang Sun, Sun, Shengyang, Yian Zhang +27 · 1 voice · 2 citations
Computer Science · #Advanced Database Systems and Queries #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Semantic Web and Ontologies #cs.CL #cs.LG
- Elucidating Optimal Reward-Diversity Tradeoffs in Text-to-Image Diffusion Models
2024/09/09 by Rohit Jena, Jena, Rohit, Ali Taghibakhshi +9 · 1 citation
Social Sciences · #Media Influence and Politics