Piot, Bilal
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1398 citations
#cs.CL #cs.AI
- Noisy Networks for Exploration
2017/06/30 by Meire Fortunato, Fortunato, Meire, Mohammad Gheshlaghi Azar +22 · 2 voices · 73 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Advanced Bandit Algorithms Research
- A General Theoretical Paradigm to Understand Learning from Human Preferences
2023/10/18 by Mohammad Gheshlaghi Azar, Azar, Mohammad Gheshlaghi, Mark Rowland +11 · 3 voices · 215 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research
- Deep Q-learning from Demonstrations
2017/04/12 by Todd Hester, Matej Vecerik, Hester, Todd +27 · 2 voices · 50 citations
Computer Science · Economics, Econometrics and Finance · #Reinforcement Learning in Robotics #Software Engineering Research #Sports Analytics and Performance #cs.AI #cs.LG
- Bootstrap your own latent: A new approach to self-supervised Learning
2020/06/13 by Jean-Bastien Grill, Grill, Jean-Bastien, Florian Strub +25 · 543 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
- Direct Language Model Alignment from Online AI Feedback
2024/02/07 by Shangmin Guo, Guo, Shangmin, Biao Zhang +23 · 1 voice · 63 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling
- Gemma 2: Improving Open Language Models at a Practical Size
2024/07/31 by Gemma Team, Morgane Rivière, Shreya Pathak +292 · 581 citations
Computer Science · #Natural Language Processing Techniques
- Gemma 3 Technical Report
2025/03/25 by Aishwarya Kamath, Gemma Team, Kamath, Aishwarya +418 · 3 voices · 625 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
- Rainbow: Combining Improvements in Deep Reinforcement Learning
2017/10/06 by Matteo Hessel, Hessel, Matteo, Joseph Modayil +17 · 179 citations
Computer Science · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Agent57: Outperforming the Atari Human Benchmark
2020/03/30 by Adrià Puigdomènech Badia, Bilal Piot, Badia, Adrià Puigdomènech +11 · 1 voice · 21 citations
Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.LG #stat.ML
- Observational Learning by Reinforcement Learning
2017/06/20 by Diana Borsa, Bilal Piot, Borsa, Diana +5 · 1 voice · 5 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Evolutionary Algorithms and Applications
- Shaking the foundations: delusions in sequence models for interaction and control
2021/10/20 by Pedro A. Ortega, Markus Kunesch, Ortega, Pedro A. +37 · 2 voices · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Topic Modeling #cs.AI #cs.LG
- Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
2017/07/27 by Todd Hester, Vecerik, Mel, Hester, Todd +16 · 39 citations
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Robot Manipulation and Learning
- Nash Learning from Human Feedback
2023/12/01 by Rémi Munos, Munos, Rémi, Michal Valko +31 · 1 voice · 47 citations
#stat.ML #cs.AI #cs.GT #cs.LG #cs.MA
- Never Give Up: Learning Directed Exploration Strategies
2020/02/14 by Adrià Puigdomènech Badia, Pablo Sprechmann, Badia, Adrià Puigdomènech +18 · 27 citations
Computer Science · #Artificial Intelligence in Games #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Generalized Preference Optimization: A Unified Approach to Offline Alignment
2024/02/08 by Yunhao Tang, Zhaohan Daniel Guo, Tang, Yunhao +17 · 37 citations
Computer Science · Decision Sciences · #Artificial Intelligence (cs.AI) #Constraint Satisfaction and Optimization #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Criteria Decision Making
- BYOL works even without batch statistics
2020/10/20 by Pierre H. Richemond, Richemond, Pierre H., Jean-Bastien Grill +19 · 1 voice · 6 citations
Computer Science · Mathematics · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications #cs.CV #cs.LG #stat.ML
- Multi-turn Reinforcement Learning from Preference Human Feedback
2024/05/23 by Lior Shani, Shani, Lior, Aviv Rosenberg +23 · 1 voice · 25 citations
Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #cs.LG
- BYOL-Explore: Exploration by Bootstrapped Prediction
2022/06/16 by Zhaohan Daniel Guo, Guo, Zhaohan Daniel, Shantanu Thakoor +25 · 14 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics
- Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning
2020/04/30 by Daniel Guo, Guo, Daniel, Bernardo Ávila Pires +11 · 15 citations
Computer Science · #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- RRM: Robust Reward Model Training Mitigates Reward Hacking
2024/09/20 by Tianqi Liu, Wei Xiong, Liu, Tianqi +33 · 27 citations
Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Automation Interaction and Safety
- Neural Predictive Belief Representations
2018/11/15 by Zhaohan Daniel Guo, Guo, Zhaohan Daniel, Mohammad Gheshlaghi Azar +7 · 13 citations
Computer Science · #Adversarial Robustness in Machine Learning #Domain Adaptation and Few-Shot Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Hindsight Credit Assignment
2019/12/05 by Anna Harutyunyan, Harutyunyan, Anna, Will Dabney +19 · 10 citations
Computer Science · Business, Management and Accounting · #Reinforcement Learning in Robotics #Financial Distress and Bankruptcy Prediction
- Acme: A Research Framework for Distributed Reinforcement Learning
2020/06/01 by Matt Hoffman, Hoffman, Matthew W., Bobak Shahriari +40 · 12 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics
- Human Alignment of Large Language Models through Online Preference Optimisation
2024/03/13 by Daniele Calandriello, Calandriello, Daniele, Daniel Guo +23 · 15 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Building Math Agents with Multi-Turn Iterative Preference Learning
2024/09/04 by Wei Xiong, Xiong, Wei, Chengshuai Shi +23 · 18 citations
Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Rough Sets and Fuzzy Logic
- Offline Regularised Reinforcement Learning for Large Language Models Alignment
2024/05/29 by Pierre Harvey Richemond, Yunhao Tang, Richemond, Pierre Harvey +33 · 15 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Understanding Self-Predictive Learning for Reinforcement Learning
2022/12/06 by Tang, Yunhao, Guo, Zhaohan Daniel, Richemond, Pierre Harvey +13 · 8 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Observe and Look Further: Achieving Consistent Performance on Atari
2018/05/29 by Tobias Pohlen, Bilal Piot, Pohlen, Tobias +23 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Geometric Entropic Exploration
2021/01/06 by Zhaohan Daniel Guo, Mohammad Gheshlaghi Azar, Guo, Zhaohan Daniel +17 · 4 citations
Computer Science · #Advanced Multi-Objective Optimization Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms
- End-to-end optimization of goal-driven and visually grounded dialogue systems Harm de Vries
2017/03/15 by Florian Strub, Harm de Vries, Strub, Florian +9 · 5 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Speech and dialogue systems #Topic Modeling
- Playing the Game of Universal Adversarial Perturbations
2018/09/20 by Julien Pérolat, Mateusz Malinowski, Perolat, Julien +5 · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Learning from negative feedback, or positive feedback or both
2024/10/05 by Abbas Abdolmaleki, Abdolmaleki, Abbas, Bilal Piot +21 · 6 citations
Computer Science · #Semantic Web and Ontologies
- The Edge of Orthogonality: A Simple View of What Makes BYOL Tick
2023/02/09 by Pierre H. Richemond, Allison Tam, Richemond, Pierre H. +9 · 3 citations
Computer Science · #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- Learning Nash Equilibrium for General-Sum Markov Games from Batch Data
2016/06/28 by Julien Pérolat, Pérolat, Julien, Florian Strub +5 · 2 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Game Theory and Applications #Advanced Bandit Algorithms Research
- Unlocking the Power of Representations in Long-term Novelty-based Exploration
2023/05/02 by Alaa Saade, Steven Kapturowski, Saade, Alaa +15 · 1 citation
Computer Science · #Anomaly Detection Techniques and Applications #Artificial Intelligence in Games #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Time Series Analysis and Forecasting