vix.ing · top · new · best · stats · spec

Florian Strub

  1. Mastering the game of Stratego with model-free multiagent reinforcement learning
    2022/06/30 by Julien Pérolat, Julien Perolat, Bart De Vylder +38 · 6 voices · 21 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Artificial Intelligence in Games #Advanced Bandit Algorithms Research
  2. Bootstrap your own latent: A new approach to self-supervised Learning
    2020/06/13 by Jean-Bastien Grill, Grill, Jean-Bastien, Florian Strub +25 · 313 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  3. Deep Reinforcement Learning and the Deadly Triad
    2018/12/06 by Hado van Hasselt, Yotam Doron, van Hasselt, Hado +9 · 27 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Age of Information Optimization #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  4. BYOL works even without batch statistics
    2020/10/20 by Pierre H. Richemond, Richemond, Pierre H., Jean-Bastien Grill +19 · 1 voice · 3 citations
    Computer Science · Mathematics · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CV #cs.LG #stat.ML
  5. Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier
    2024/12/05 by John Dang, Shivalika Singh, Dang, John +90 · 2 voices · 23 citations
    Arts and Humanities · Computer Science · #Second Language Learning and Teaching #cs.CL
  6. HoME: a Household Multimodal Environment
    2017/11/29 by Simon Brodeur, Ethan Perez, Brodeur, Simon +15 · 11 citations
    Computer Science · Psychology · #Speech and dialogue systems #Social Robot Interaction and HRI #AI in Service Interactions
  7. Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion
    2024/06/27 by Yannis Flet-Berliac, Flet-Berliac, Yannis, Nathan Grinsztajn +18 · 5 citations
    Decision Sciences · #Educational Assessment and Improvement
  8. Countering Language Drift with Seeded Iterated Learning
    2020/03/28 by Yuchen Lu, Soumye Singhal, Lu, Yuchen +7 · 1 citation
    Computer Science · #Topic Modeling #Speech and dialogue systems #Natural Language Processing Techniques
  9. Hybrid Collaborative Filtering with Autoencoders
    2016/03/02 by Florian Strub, Jérémie Mary, Strub, Florian +3 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Pose and Action Recognition #Information Retrieval (cs.IR) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Recommender Systems and Techniques
  10. Supervised Seeded Iterated Learning for Interactive Language Learning
    2020/10/06 by Yuchen Lu, Lu, Yuchen, Soumye Singhal +7 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Speech and dialogue systems #Topic Modeling
  11. Countering Reward Over-optimization in LLM with Demonstration-Guided Reinforcement Learning
    2024/04/30 by Mathieu Rita, Rita, Mathieu, Florian Strub +9 · 2 citations
    Computer Science · Engineering · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Scheduling and Optimization Algorithms
  12. Developing, Evaluating and Scaling Learning Agents in Multi-Agent Environments
    2022/09/22 by Ian Gemp, Gemp, Ian, T Anthony +51 · 1 citation
    Computer Science · #Reinforcement Learning in Robotics
  13. Emergent Communication: Generalization and Overfitting in Lewis Games
    2022/09/30 by Mathieu Rita, Rita, Mathieu, Tallec, Corentin +9 · 1 citation
    Biochemistry, Genetics and Molecular Biology · Physics and Astronomy · Social Sciences · #Computation and Language (cs.CL) #DNA and Biological Computing #FOS: Computer and information sciences #Information Theory (cs.IT) #Language and cultural evolution #Multiagent Systems (cs.MA) #Origins and Evolution of Life
  14. The Edge of Orthogonality: A Simple View of What Makes BYOL Tick
    2023/02/09 by Pierre H. Richemond, Richemond, Pierre H., Allison Tam +9 · 1 citation
    Computer Science · #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  15. Learning Nash Equilibrium for General-Sum Markov Games from Batch Data
    2016/06/28 by Julien Pérolat, Florian Strub, Pérolat, Julien +5 · 1 citation
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Game Theory and Applications #Advanced Bandit Algorithms Research
  16. ShiQ: Bringing back Bellman to LLMs
    2025/05/16 by Pierre Clavier, Clavier, Pierre, Nathan Grinsztajn +19 · 2 citations
    Medicine · Computer Science · #Artificial Intelligence in Healthcare and Education #Topic Modeling #Natural Language Processing Techniques