Zhilin Wang
- Llama-Nemotron: Efficient Reasoning Models
2025/05/02 by Akhiad Bercovich, Itay Levy, Bercovich, Akhiad +224 · 1 voice · 29 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Topic Modeling
- HelpSteer2: Open-source dataset for training top-performing reward models
2024/06/12 by Zhilin Wang, Yi Dong, Wang, Zhilin +15 · 37 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare
- SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF
2023/10/09 by Yi Dong, Dong, Yi, Zhilin Wang +7 · 12 citations
Computer Science · Social Sciences · #Topic Modeling #Natural Language Processing Techniques #Computational and Text Analysis Methods
- NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
2024/05/02 by Gerald Shen, Zhilin Wang, Shen, Gerald +23 · 12 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computational Physics and Python Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Time Series Analysis and Forecasting
- Diverging Preferences: When do Annotators Disagree and do Models Know?
2024/10/18 by Michael JQ Zhang, Zhang, Michael JQ, Zhilin Wang +15 · 9 citations
Computer Science · #Speech and dialogue systems
- Humanoid Agents: Platform for Simulating Human-like Generative Agents
2023/10/09 by Zhilin Wang, Yu Ying Chiu, Wang, Zhilin +3 · 4 citations
Psychology · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Social Robot Interaction and HRI
- Adversarial Training of Reward Models
2025/04/08 by Alexander Bukharin, Bukharin, Alexander, Haifeng Qian +15 · 4 citations
Computer Science · #Adversarial Robustness in Machine Learning #Topic Modeling #Hate Speech and Cyberbullying Detection
- Reward-aware Preference Optimization: A Unified Mathematical Framework for Model Alignment
2025/01/31 by Shengyang Sun, Yian Zhang, Sun, Shengyang +27 · 1 voice · 2 citations
Computer Science · #Advanced Database Systems and Queries #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Semantic Web and Ontologies #cs.CL #cs.LG
- Effects of Lactobacillus reuteri LR1 on the growth performance, intestinal morphology, and intestinal barrier function in weaned pigs
2018/04/12 by Hongbo Yi, Li Wang, Yunxia Xiong +5 · 25 citations
Agricultural and Biological Sciences · Biochemistry, Genetics and Molecular Biology · #Animal Nutrition and Physiology #Probiotics and Fermented Foods #Gut microbiota and health