vix.ing · top · new · best · stats · spec

Piot, Bilal

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1398 citations
    #cs.CL #cs.AI
  2. Noisy Networks for Exploration
    2017/06/30 by Meire Fortunato, Fortunato, Meire, Mohammad Gheshlaghi Azar +22 · 2 voices · 73 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Advanced Bandit Algorithms Research
  3. A General Theoretical Paradigm to Understand Learning from Human Preferences
    2023/10/18 by Mohammad Gheshlaghi Azar, Azar, Mohammad Gheshlaghi, Mark Rowland +11 · 3 voices · 215 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research
  4. Deep Q-learning from Demonstrations
    2017/04/12 by Todd Hester, Matej Vecerik, Hester, Todd +27 · 2 voices · 50 citations
    Computer Science · Economics, Econometrics and Finance · #Reinforcement Learning in Robotics #Software Engineering Research #Sports Analytics and Performance #cs.AI #cs.LG
  5. Bootstrap your own latent: A new approach to self-supervised Learning
    2020/06/13 by Jean-Bastien Grill, Grill, Jean-Bastien, Florian Strub +25 · 543 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  6. Direct Language Model Alignment from Online AI Feedback
    2024/02/07 by Shangmin Guo, Guo, Shangmin, Biao Zhang +23 · 1 voice · 63 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling
  7. Gemma 2: Improving Open Language Models at a Practical Size
    2024/07/31 by Gemma Team, Morgane Rivière, Shreya Pathak +292 · 581 citations
    Computer Science · #Natural Language Processing Techniques
  8. Gemma 3 Technical Report
    2025/03/25 by Aishwarya Kamath, Gemma Team, Kamath, Aishwarya +418 · 3 voices · 625 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
  9. Rainbow: Combining Improvements in Deep Reinforcement Learning
    2017/10/06 by Matteo Hessel, Hessel, Matteo, Joseph Modayil +17 · 179 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  10. Agent57: Outperforming the Atari Human Benchmark
    2020/03/30 by Adrià Puigdomènech Badia, Bilal Piot, Badia, Adrià Puigdomènech +11 · 1 voice · 21 citations
    Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.LG #stat.ML
  11. Observational Learning by Reinforcement Learning
    2017/06/20 by Diana Borsa, Bilal Piot, Borsa, Diana +5 · 1 voice · 5 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Evolutionary Algorithms and Applications
  12. Shaking the foundations: delusions in sequence models for interaction and control
    2021/10/20 by Pedro A. Ortega, Markus Kunesch, Ortega, Pedro A. +37 · 2 voices · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Topic Modeling #cs.AI #cs.LG
  13. Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
    2017/07/27 by Todd Hester, Vecerik, Mel, Hester, Todd +16 · 39 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Robot Manipulation and Learning
  14. Nash Learning from Human Feedback
    2023/12/01 by Rémi Munos, Munos, Rémi, Michal Valko +31 · 1 voice · 47 citations
    #stat.ML #cs.AI #cs.GT #cs.LG #cs.MA
  15. Never Give Up: Learning Directed Exploration Strategies
    2020/02/14 by Adrià Puigdomènech Badia, Pablo Sprechmann, Badia, Adrià Puigdomènech +18 · 27 citations
    Computer Science · #Artificial Intelligence in Games #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  16. Generalized Preference Optimization: A Unified Approach to Offline Alignment
    2024/02/08 by Yunhao Tang, Zhaohan Daniel Guo, Tang, Yunhao +17 · 37 citations
    Computer Science · Decision Sciences · #Artificial Intelligence (cs.AI) #Constraint Satisfaction and Optimization #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Criteria Decision Making
  17. BYOL works even without batch statistics
    2020/10/20 by Pierre H. Richemond, Richemond, Pierre H., Jean-Bastien Grill +19 · 1 voice · 6 citations
    Computer Science · Mathematics · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications #cs.CV #cs.LG #stat.ML
  18. Multi-turn Reinforcement Learning from Preference Human Feedback
    2024/05/23 by Lior Shani, Shani, Lior, Aviv Rosenberg +23 · 1 voice · 25 citations
    Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #cs.LG
  19. BYOL-Explore: Exploration by Bootstrapped Prediction
    2022/06/16 by Zhaohan Daniel Guo, Guo, Zhaohan Daniel, Shantanu Thakoor +25 · 14 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics
  20. Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning
    2020/04/30 by Daniel Guo, Guo, Daniel, Bernardo Ávila Pires +11 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  21. RRM: Robust Reward Model Training Mitigates Reward Hacking
    2024/09/20 by Tianqi Liu, Wei Xiong, Liu, Tianqi +33 · 27 citations
    Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Automation Interaction and Safety
  22. Neural Predictive Belief Representations
    2018/11/15 by Zhaohan Daniel Guo, Guo, Zhaohan Daniel, Mohammad Gheshlaghi Azar +7 · 13 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Domain Adaptation and Few-Shot Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  23. Hindsight Credit Assignment
    2019/12/05 by Anna Harutyunyan, Harutyunyan, Anna, Will Dabney +19 · 10 citations
    Computer Science · Business, Management and Accounting · #Reinforcement Learning in Robotics #Financial Distress and Bankruptcy Prediction
  24. Acme: A Research Framework for Distributed Reinforcement Learning
    2020/06/01 by Matt Hoffman, Hoffman, Matthew W., Bobak Shahriari +40 · 12 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics
  25. Human Alignment of Large Language Models through Online Preference Optimisation
    2024/03/13 by Daniele Calandriello, Calandriello, Daniele, Daniel Guo +23 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  26. Building Math Agents with Multi-Turn Iterative Preference Learning
    2024/09/04 by Wei Xiong, Xiong, Wei, Chengshuai Shi +23 · 18 citations
    Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Rough Sets and Fuzzy Logic
  27. Offline Regularised Reinforcement Learning for Large Language Models Alignment
    2024/05/29 by Pierre Harvey Richemond, Yunhao Tang, Richemond, Pierre Harvey +33 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  28. Understanding Self-Predictive Learning for Reinforcement Learning
    2022/12/06 by Tang, Yunhao, Guo, Zhaohan Daniel, Richemond, Pierre Harvey +13 · 8 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  29. Observe and Look Further: Achieving Consistent Performance on Atari
    2018/05/29 by Tobias Pohlen, Bilal Piot, Pohlen, Tobias +23 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  30. Geometric Entropic Exploration
    2021/01/06 by Zhaohan Daniel Guo, Mohammad Gheshlaghi Azar, Guo, Zhaohan Daniel +17 · 4 citations
    Computer Science · #Advanced Multi-Objective Optimization Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms
  31. End-to-end optimization of goal-driven and visually grounded dialogue systems Harm de Vries
    2017/03/15 by Florian Strub, Harm de Vries, Strub, Florian +9 · 5 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Speech and dialogue systems #Topic Modeling
  32. Playing the Game of Universal Adversarial Perturbations
    2018/09/20 by Julien Pérolat, Mateusz Malinowski, Perolat, Julien +5 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  33. Learning from negative feedback, or positive feedback or both
    2024/10/05 by Abbas Abdolmaleki, Abdolmaleki, Abbas, Bilal Piot +21 · 6 citations
    Computer Science · #Semantic Web and Ontologies
  34. The Edge of Orthogonality: A Simple View of What Makes BYOL Tick
    2023/02/09 by Pierre H. Richemond, Allison Tam, Richemond, Pierre H. +9 · 3 citations
    Computer Science · #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  35. Learning Nash Equilibrium for General-Sum Markov Games from Batch Data
    2016/06/28 by Julien Pérolat, Pérolat, Julien, Florian Strub +5 · 2 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Game Theory and Applications #Advanced Bandit Algorithms Research
  36. Unlocking the Power of Representations in Long-term Novelty-based Exploration
    2023/05/02 by Alaa Saade, Steven Kapturowski, Saade, Alaa +15 · 1 citation
    Computer Science · #Anomaly Detection Techniques and Applications #Artificial Intelligence in Games #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Time Series Analysis and Forecasting