vix.ing · top · new · best · stats · spec

Idan Shenfeld

  1. Self-Distillation Enables Continual Learning
    2026/01/27 by Idan Shenfeld, Mehul Damani, Jonas Hübotter +1 · 12 voices · 17 citations
    #cs.LG
  2. RL's Razor: Why Online Reinforcement Learning Forgets Less
    2025/09/04 by Idan Shenfeld, Shenfeld, Idan, Jyothish Pari +3 · 1 voice · 40 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.LG
  3. From Imitation to Refinement -- Residual RL for Precise Assembly
    2024/07/23 by Lars Ankile, Anthony Simeonov, Ankile, Lars +7 · 23 citations
    Engineering · #Manufacturing Process and Optimization #Advanced Surface Polishing Techniques #Additive Manufacturing and 3D Printing Technologies
  4. JUICER: Data-Efficient Imitation Learning for Robotic Assembly
    2024/04/04 by Lars Ankile, Ankile, Lars, Anthony Simeonov +5 · 15 citations
    Engineering · #Additive Manufacturing and 3D Printing Technologies #FOS: Computer and information sciences #Machine Learning (cs.LG) #Manufacturing Process and Optimization #Robot Manipulation and Learning #Robotics (cs.RO)
  5. Reinforcement Learning via Self-Distillation
    2026/01/28 by Jonas Hübotter, Frederike Lübeck, Lejs Behric +8 · 5 voices · 12 citations
    #cs.LG #cs.AI
  6. Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
    2024/10/07 by Mehul Damani, Damani, Mehul, Idan Shenfeld +7 · 14 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Neural Networks and Reservoir Computing
  7. Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
    2025/07/22 by Mehul Damani, Isha Puri, Damani, Mehul +11 · 1 voice · 25 citations
    Computer Science · #Intelligent Tutoring Systems and Adaptive Learning
  8. Value Augmented Sampling for Language Model Alignment and Personalization
    2024/05/10 by Seungwook Han, Idan Shenfeld, Han, Seungwook +7 · 7 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
  9. The Future of Open Human Feedback
    2024/08/15 by Shachar Don-Yehiya, Ben Burtenshaw, Don-Yehiya, Shachar +37 · 4 citations
    Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human-Automation Interaction and Safety #Human-Computer Interaction (cs.HC)
  10. KL-Regularized RLHF with Multiple Reference Models: Exact Solutions and Sample Complexity
    2025/02/03 by Gholamali Aminian, Amir R. Asadi, Aminian, Gholamali +5 · 1 citation
    Engineering · #Engineering Applied Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  11. Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis
    2025/07/08 by Gholamali Aminian, Idan Shenfeld, Aminian, Gholamali +7 · 2 citations
    Computer Science · Decision Sciences · Mathematics · #Explainable Artificial Intelligence (XAI) #Advanced Bandit Algorithms Research #Advanced Causal Inference Techniques