vix.ing · top · new · best · stats · spec

Pablo Samuel Castro

  1. Deep Reinforcement Learning at the Edge of the Statistical Precipice
    2021/08/30 by Rishabh Agarwal, Max Schwarzer, Agarwal, Rishabh +7 · 65 citations
    Computer Science · #Reinforcement Learning in Robotics #Evolutionary Algorithms and Applications #Advanced Multi-Objective Optimization Algorithms
  2. Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks
    2023/06/24 by Maxime Chevalier-Boisvert, Bolun Dai, Chevalier-Boisvert, Maxime +15 · 42 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Context-Aware Activity Recognition Systems #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  3. Stop Regressing: Training Value Functions via Classification for Scalable Deep RL
    2024/03/06 by Jesse Farebrother, Jordi Orbay, Farebrother, Jesse +21 · 3 voices · 33 citations
    Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.AI #cs.LG #stat.ML
  4. Bigger, Better, Faster: Human-level Atari with human-level efficiency
    2023/05/30 by Max Schwarzer, Johan Obando-Ceron, Schwarzer, Max +10 · 1 voice · 27 citations
    Computer Science · #Anomaly Detection Techniques and Applications #Human Pose and Action Recognition #Reinforcement Learning in Robotics #cs.AI #cs.LG
  5. The Dormant Neuron Phenomenon in Deep Reinforcement Learning
    2023/02/24 by Ghada Sokar, Sokar, Ghada, Rishabh Agarwal +5 · 21 citations
    Neuroscience · Computer Science · #Neural dynamics and brain function #Reinforcement Learning in Robotics #Neural Networks and Applications
  6. Scalable methods for computing state similarity in deterministic Markov\n Decision Processes
    2019/11/21 by Pablo Samuel Castro, Castro, Pablo Samuel · 9 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Bayesian Modeling and Causal Inference #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms
  7. The State of Sparse Training in Deep Reinforcement Learning
    2022/06/17 by Laura Graesser, Graesser, Laura, Utku Evci +5 · 11 citations
    Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Machine Learning and ELM
  8. Mixtures of Experts Unlock Parameter Scaling for Deep RL
    2024/02/13 by Johan Obando-Ceron, Ghada Sokar, Obando-Ceron, Johan +15 · 13 citations
    Computer Science · Physics and Astronomy · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and ELM #Model Reduction and Neural Networks
  9. Methods for computing state similarity in Markov Decision Processes
    2012/06/27 by Norman Ferns, Ferns, Norman, Pablo Samuel Castro +5 · 3 citations
    Engineering · Computer Science · #Fault Detection and Control Systems #AI-based Problem Solving and Planning #Reinforcement Learning in Robotics
  10. The Difficulty of Passive Learning in Deep Reinforcement Learning
    2021/10/26 by Georg Ostrovski, Pablo Samuel Castro, Ostrovski, Georg +3 · 1 voice · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.AI #cs.LG
  11. MICo: Improved representations via sampling-based state similarity for Markov decision processes
    2021/06/03 by Pablo Samuel Castro, Tyler Kastner, Castro, Pablo Samuel +5 · 5 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Bayesian Modeling and Causal Inference #Data Stream Mining Techniques #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  12. In value-based deep reinforcement learning, a pruned network is a good network
    2024/02/19 by Johan Obando-Ceron, Obando-Ceron, Johan, Aaron Courville +3 · 1 voice · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #cs.AI #cs.LG
  13. A density estimation perspective on learning from pairwise human preferences
    2023/11/23 by Vincent Dumoulin, Dumoulin, Vincent, Daniel D. Johnson +7 · 6 citations
    Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
  14. Revisiting Rainbow: Promoting more Insightful and Inclusive Deep\n Reinforcement Learning Research
    2020/11/20 by Johan Obando-Ceron, Obando-Ceron, Johan S., Pablo Samuel Castro +1 · 4 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  15. A Comparative Analysis of Expected and Distributional Reinforcement\n Learning
    2019/01/30 by Clare Lyle, Pablo Samuel Castro, Lyle, Clare +3 · 3 citations
    Computer Science · #Advanced Multi-Objective Optimization Algorithms #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  16. Meta-World+: An Improved, Standardized, RL Benchmark
    2025/05/16 by Reginald McLean, McLean, Reginald, Evangelos Chatzaroulas +21 · 10 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
  17. Small batch deep reinforcement learning
    2023/10/05 by Johan Obando-Ceron, Obando-Ceron, Johan, Marc G. Bellemare +3 · 5 citations
    Engineering · Neuroscience · #Artificial Intelligence (cs.AI) #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #Machine Learning (cs.LG) #Muscle activation and electromyography studies
  18. On the consistency of hyper-parameter selection in value-based deep reinforcement learning
    2024/06/25 by Johan Obando-Ceron, Obando-Ceron, Johan, João G. M. Araújo +5 · 1 voice · 5 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Scheduling and Optimization Algorithms #cs.AI #cs.LG
  19. A general class of surrogate functions for stable and efficient reinforcement learning
    2021/08/12 by Sharan Vaswani, Olivier Bachem, Vaswani, Sharan +15 · 2 citations
    Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  20. Multi-Task Reinforcement Learning Enables Parameter Scaling
    2025/03/07 by Reginald McLean, McLean, Reginald, Evangelos Chatzaroulas +9 · 1 voice · 4 citations
    Engineering · #Digital Transformation in Industry #Scheduling and Optimization Algorithms #cs.AI #cs.LG
  21. Mixture of Experts in a Mixture of RL settings
    2024/06/26 by Timon Willi, Willi, Timon, Johan Obando-Ceron +7 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Bayesian Methods and Mixture Models #Distributed Sensor Networks and Detection Algorithms #Expert finding and Q&A systems #FOS: Computer and information sciences #Machine Learning (cs.LG)
  22. Discovering Symbolic Cognitive Models from Human and Animal Behavior
    2025/02/06 by Pablo Samuel Castro, Nenad Tomašev, Ankit Anand +14 · 1 voice · 3 citations
    Computer Science · Social Sciences · #Evolutionary Algorithms and Applications #Reinforcement Learning in Robotics #Language and cultural evolution
  23. Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
    2025/05/29 by Jiashun Liu, Zihao Wu, Liu, Jiashun +9 · 5 citations
    Neuroscience · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural dynamics and brain function #Neuroscience and Neural Engineering
  24. Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
    2025/06/18 by Roger Creus Castanyer, Castanyer, Roger Creus, Johan Obando-Ceron +11 · 1 voice · 6 citations
    #cs.LG
  25. Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
    2025/05/31 by Hongyao Tang, Johan Obando-Ceron, Tang, Hongyao +7 · 4 citations
    Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Muscle activation and electromyography studies
  26. Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
    2024/10/02 by Ghada Sokar, Johan Obando-Ceron, Sokar, Ghada +7 · 2 citations
    Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Machine Learning (cs.LG)
  27. CALE: Continuous Arcade Learning Environment
    2024/10/31 by Jesse Farebrother, Farebrother, Jesse, Pablo Samuel Castro +1 · 1 voice · 2 citations
    #cs.LG #cs.AI
  28. Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
    2025/05/23 by Ghada Sokar, Pablo Samuel Castro, Sokar, Ghada +1 · 3 voices · 1 citation
    #cs.LG #cs.AI
  29. Performing Structured Improvisations with pre-trained Deep Learning Models
    2019/04/30 by Pablo Samuel Castro, Castro, Pablo Samuel · 2 citations
    Computer Science · Neuroscience · #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception
  30. The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
    2025/06/03 by Walter Mayor, Johan Obando-Ceron, Mayor, Walter +5 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Software-Defined Networks and 5G #Stochastic Gradient Optimization Techniques
  31. Discovering Differences in Strategic Behavior Between Humans and LLMs
    2026/02/10 by Caroline Wang, Daniel Kasenberg, Kim Stachenfeld +1 · 3 voices
    #cs.AI #cs.CL #cs.CY #cs.HC
  32. Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
    2025/10/02 by Johan Obando-Ceron, Liu, Jiashun, Obando-Ceron, Johan +16 · 2 citations
    Decision Sciences · #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG)
  33. Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents
    2025/10/15 by Johan Obando-Ceron, Walter Mayor, Obando-Ceron, Johan +9 · 1 voice · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Reinforcement Learning in Robotics #Robotics (cs.RO) #cs.AI #cs.LG #cs.RO