Shrimai Prabhumoye
- Self-Refine: Iterative Refinement with Self-Feedback
2023/03/30 by Aman Madaan, Niket Tandon, Madaan, Aman +29 · 1 voice · 560 citations
Computer Science · #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling #cs.AI #cs.CL #cs.LG
- Llama-Nemotron: Efficient Reasoning Models
2025/05/02 by Akhiad Bercovich, Itay Levy, Bercovich, Akhiad +224 · 1 voice · 29 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Topic Modeling
- Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model
2022/01/28 by Shaden Smith, Smith, Shaden, Mostofa Patwary +36 · 23 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- Nemotron-4 15B Technical Report
2024/02/26 by Jupinder Parmar, Parmar, Jupinder, Shrimai Prabhumoye +51 · 1 voice · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.AI #cs.CL #cs.LG
- Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining
2024/12/18 by Feng, Steven, Shrimai Prabhumoye, Kezhi Kong +10 · 9 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Machine Learning and Data Classification #Neural Networks and Applications
- Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
2025/05/26 by Jaehun Jung, Seungju Han, Jung, Jaehun +17 · 10 citations
Engineering · #Reservoir Engineering and Simulation Methods
- MIND: Math Informed syNthetic Dialogues for Pretraining LLMs
2024/10/15 by Syeda Nahida Akter, Shrimai Prabhumoye, Akter, Syeda Nahida +13 · 4 citations
Computer Science · #Natural Language Processing Techniques #Mathematics, Computing, and Information Processing #Semantic Web and Ontologies
- Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
2025/04/06 by Ximing Lu, Seungju Han, Lu, Ximing +19 · 5 citations
Computer Science · #Intelligent Tutoring Systems and Adaptive Learning #Explainable Artificial Intelligence (XAI) #Advanced Graph Neural Networks
- AgentKit: Structured LLM Reasoning with Dynamic Graphs
2024/04/17 by Yue Wu, Wu, Yue, Yewen Fan +15 · 2 citations
Computer Science · Decision Sciences · #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Peer-to-Peer Network Technologies
- Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
2025/12/23 by NVIDIA, Aaron Blakeman, : +623 · 1 voice · 2 citations
#cs.CL #cs.AI #cs.LG
- SPRING: Studying the Paper and Reasoning to Play Games
2023/05/24 by Yue Wu, Shrimai Prabhumoye, Wu, Yue +13 · 1 citation
Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Educational Games and Gamification #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- Front-Loading Reasoning: The Synergy between Pretraining and Post-Training Data
2025/09/26 by Syeda Nahida Akter, Akter, Syeda Nahida, Shrimai Prabhumoye +11 · 6 citations
Computer Science · #Intelligent Tutoring Systems and Adaptive Learning
- I love your chain mail! Making knights smile in a fantasy game world:\n Open-domain goal-oriented dialogue agents
2020/02/07 by Shrimai Prabhumoye, Margaret Li, Prabhumoye, Shrimai +11 · 2 citations
Computer Science · #Speech and dialogue systems #Topic Modeling #AI in Service Interactions
- Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
2025/08/20 by Rabeeh Karimi Mahabadi, Sanjeev Satheesh, Mahabadi, Rabeeh Karimi +9 · 1 voice · 5 citations
#cs.CL #cs.AI #cs.LG
- RLP: Reinforcement as a Pretraining Objective
2025/09/26 by Ali Hatamizadeh, Syeda Nahida Akter, Hatamizadeh, Ali +13 · 4 citations
Psychology · #Behavioral and Psychological Studies
- Robostral Navigate
2026/07/22 by Arjun Majumdar, Abhijeet Somani, Aditi Kabra +212
#cs.RO #cs.AI