vix.ing · top · new · best · stats · spec

Feng, Yihao

  1. LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning
    2023/06/05 by Bo Liu, Liu, Bo, Yifeng Zhu +11 · 291 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications
  2. UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild
    2023/05/18 by Can Qin, Shu Zhang, Qin, Can +23 · 34 citations
    Computer Science · #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning
  3. APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
    2024/06/26 by Zuxin Liu, Liu, Zuxin, Thai Hoang +31 · 43 citations
    Computer Science · Decision Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Scientific Computing and Data Management #Software Engineering (cs.SE) #Time Series Analysis and Forecasting
  4. HIVE: Harnessing Human Feedback for Instructional Visual Editing
    2023/03/16 by Shu Zhang, Xinyi Yang, Zhang, Shu +21 · 25 citations
    Computer Science · #Multimodal Machine Learning Applications
  5. xLAM: A Family of Large Action Models to Empower AI Agent Systems
    2024/09/05 by Jianguo Zhang, Zhang, Jianguo, Tian Lan +41 · 32 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation
  6. FAMO: Fast Adaptive Multitask Optimization
    2023/06/06 by Bo Liu, Yihao Feng, Liu, Bo +5 · 16 citations
    Computer Science · #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and ELM
  7. Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward
    2024/04/01 by Ruohong Zhang, Liangke Gui, Zhang, Ruohong +19 · 21 citations
    Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Computational and Text Analysis Methods #Computer Vision and Pattern Recognition (cs.CV) #Educational and Technological Research #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  8. Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
    2024/11/06 by Haolin Chen, Chen, Haolin, Yihao Feng +19 · 2 voices · 12 citations
    Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #I.2.7 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG #stat.ML
  9. Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization
    2023/08/04 by Weiran Yao, Shelby Heinecke, Yao, Weiran +27 · 10 citations
    Computer Science · #Topic Modeling #Multimodal Machine Learning Applications #Natural Language Processing Techniques
  10. FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability
    2024/02/28 by Congying Xia, Chen Xing, Xia, Congying +13 · 11 citations
    Computer Science · #Library Science and Information Systems
  11. Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
    2024/08/13 by Kexun Zhang, Weiran Yao, Zhang, Kexun +29 · 12 citations
    Computer Science · Social Sciences · #Open Source Software Innovations #Multi-Agent Systems and Negotiation #Ethics and Social Impacts of AI
  12. Learning to Draw Samples with Amortized Stein Variational Gradient Descent
    2017/07/20 by Yihao Feng, Feng, Yihao, Dilin Wang +3 · 7 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Gaussian Processes and Bayesian Inference #Adversarial Robustness in Machine Learning
  13. Longhorn: State Space Models are Amortized Online Learners
    2024/07/19 by Bo Liu, Rui Wang, Liu, Bo +9 · 10 citations
    Computer Science · #Intelligent Tutoring Systems and Adaptive Learning
  14. BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents
    2023/08/11 by Zhiwei Liu, Weiran Yao, Liu, Zhiwei +27 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation #Natural Language Processing Techniques #Topic Modeling
  15. AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning
    2024/02/23 by Jianguo Zhang, Zhang, Jianguo, Tian Lan +31 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation
  16. Incremental Few-shot Text Classification with Multi-round New Classes: Formulation, Dataset and System
    2021/04/24 by Congying Xia, Wenpeng Yin, Xia, Congying +5 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Text and Document Classification Technologies #Topic Modeling
  17. Unsupervised Out-of-Domain Detection via Pre-trained Transformers
    2021/06/02 by Xu, Keyang, Ren, Tongzheng, Zhang, Shikun +2 · 3 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  18. Two Methods For Wild Variational Inference
    2016/11/30 by Qiang Liu, Yihao Feng, Liu, Qiang +1 · 5 citations
    Computer Science · Physics and Astronomy · Mathematics · #Gaussian Processes and Bayesian Inference #Model Reduction and Neural Networks #Markov Chains and Monte Carlo Methods
  19. A Kernel Loss for Solving the Bellman Equation
    2019/05/25 by Yihao Feng, Feng, Yihao, Lihong Li +3 · 4 citations
    Computer Science · Physics and Astronomy · #Adaptive Dynamic Programming Control #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Reinforcement Learning in Robotics
  20. Metric Residual Networks for Sample Efficient Goal-Conditioned Reinforcement Learning
    2022/08/17 by Bo Liu, Yihao Feng, Liu, Bo +5 · 3 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  21. Accountable Off-Policy Evaluation With Kernel Bellman Statistics
    2020/08/15 by Yihao Feng, Tongzheng Ren, Feng, Yihao +5 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms
  22. Apple Intelligence Foundation Language Models: Tech Report 2025
    2025/07/17 by Ethan Li, Li, Ethan, Anders Larsen +487 · 14 citations
    Computer Science · #Advanced Malware Detection Techniques #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Big Data and Digital Economy #FOS: Computer and information sciences #Machine Learning (cs.LG)
  23. Preference-grounded Token-level Guidance for Language Model Fine-tuning
    2023/06/01 by Shentao Yang, Yang, Shentao, Shujian Zhang +9 · 2 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
  24. REX: Rapid Exploration and eXploitation for AI Agents
    2023/07/18 by Murthy, Rithesh, Heinecke, Shelby, Niebles, Juan Carlos +12 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. A Behavior Regularized Implicit Policy for Offline Reinforcement Learning
    2022/02/19 by Yang, Shentao, Wang, Zhendong, Zheng, Huangjie +2 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  26. Regularizing a Model-based Policy Stationary Distribution to Stabilize Offline Reinforcement Learning
    2022/06/14 by Yang, Shentao, Feng, Yihao, Zhang, Shujian +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  27. A Unified Framework for Alternating Offline Model Training and Policy Learning
    2022/10/12 by Yang, Shentao, Zhang, Shujian, Feng, Yihao +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  28. Text2Data: Low-Resource Data Generation with Textual Control
    2024/02/08 by Shiyu Wang, Wang, Shiyu, Yihao Feng +15 · 1 citation
    Computer Science · #Semantic Web and Ontologies #Natural Language Processing Techniques