vix.ing · top · new · best · stats · spec

Hong, Yuzhong

  1. GVPO: Group Variance Policy Optimization for Large Language Model Post-Training
    2025/04/28 by Kaicheng Zhang, Y. Hong, Zhang, Kaichen +9 · 8 citations
    Computer Science · Medicine · #Topic Modeling #Domain Adaptation and Few-Shot Learning #Artificial Intelligence in Healthcare and Education
  2. Energy-Based Preference Model Offers Better Offline Alignment than the Bradley-Terry Preference Model
    2024/12/18 by Yuzhong Hong, Hanshan Zhang, Hong, Yuzhong +7 · 2 citations
    Economics, Econometrics and Finance · #Economic and Environmental Valuation
  3. Preference-Oriented Supervised Fine-Tuning: Favoring Target Model Over Aligned Large Language Models
    2024/12/17 by Yuchen Fan, Yuzhong Hong, Fan, Yuchen +9 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis