Precup, Doina
- Training Language Models to Self-Correct via Reinforcement Learning
2024/09/19 by Aviral Kumar, Vincent Zhuang, Kumar, Aviral +36 · 2 voices · 76 citations
Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
- Deep Reinforcement Learning that Matters
2017/09/19 by Peter Henderson, Riashat Islam, Henderson, Peter +10 · 4 voices · 86 citations
Computer Science · #Evolutionary Algorithms and Applications #Reinforcement Learning in Robotics #cs.LG #stat.ML
- Off-Policy Deep Reinforcement Learning without Exploration
2018/12/07 by Fujimoto, Scott, Meger, David, Precup, Doina · 96 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Flow Network based Generative Models for Non-Iterative Diverse Candidate Generation
2021/06/08 by Emmanuel Bengio, Moksh Jain, Bengio, Emmanuel +7 · 49 citations
Materials Science · Computer Science · #Machine Learning in Materials Science #Machine Learning and Algorithms #Topic Modeling
- The Option-Critic Architecture
2016/09/16 by Bacon, Pierre-Luc, Harb, Jean, Precup, Doina · 31 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- Conditional Computation in Neural Networks for faster models
2015/11/19 by Emmanuel Bengio, Pierre‐Luc Bacon, Bengio, Emmanuel +5 · 32 citations
Computer Science · #Adversarial Robustness in Machine Learning #Machine Learning and Algorithms #Advanced Neural Network Applications
- Towards Continual Reinforcement Learning: A Review and Perspectives
2020/12/25 by Khimya Khetarpal, Khetarpal, Khimya, Matthew Riemer +5 · 23 citations
Biochemistry, Genetics and Molecular Biology · #Viral Infectious Diseases and Gene Expression in Insects
- Algorithms for multi-armed bandit problems
2014/02/25 by Volodymyr Kuleshov, Kuleshov, Volodymyr, Doina Precup +1 · 19 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Optimization and Search Problems
- Nash Learning from Human Feedback
2023/12/01 by Rémi Munos, Michal Valko, Munos, Rémi +31 · 1 voice · 25 citations
#stat.ML #cs.AI #cs.GT #cs.LG #cs.MA
- Metrics for Finite Markov Decision Processes
2012/07/11 by Norm Ferns, Ferns, Norman, Prakash Panangaden +3 · 13 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Gradient Starvation: A Learning Proclivity in Neural Networks
2020/11/18 by Mohammad Pezeshki, Sékou-Oumar Kaba, Pezeshki, Mohammad +9 · 20 citations
Computer Science · Physics and Astronomy · #Adversarial Robustness in Machine Learning #Model Reduction and Neural Networks #Domain Adaptation and Few-Shot Learning
- What can I do here? A Theory of Affordances in Reinforcement Learning
2020/06/26 by Khimya Khetarpal, Zafarali Ahmed, Khetarpal, Khimya +7 · 2 voices · 1 citation
Computer Science · Decision Sciences · Mathematics · #Artificial Intelligence (cs.AI) #Complex Systems and Decision Making #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #cs.AI #cs.LG #stat.ML
- A Definition of Continual Reinforcement Learning
2023/07/20 by David Abel, Abel, David, André Barreto +9 · 14 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Reproducibility of Benchmarked Deep Reinforcement Learning Tasks for Continuous Control
2017/08/10 by Islam, Riashat, Henderson, Peter, Gomrokchi, Maziar +1 · 7 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- When Do Graph Neural Networks Help with Node Classification? Investigating the Impact of Homophily Principle on Node Distinguishability
2023/04/25 by Luan, Sitao, Hua, Chenqing, Xu, Minkai +6 · 12 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Social and Information Networks (cs.SI)
- Mixtures of Experts Unlock Parameter Scaling for Deep RL
2024/02/13 by Obando-Ceron, Johan, Sokar, Ghada, Willi, Timon +6 · 13 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- On the Expressivity of Markov Reward
2021/11/01 by Abel, David, Dabney, Will, Harutyunyan, Anna +4 · 8 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Revisiting Heterophily For Graph Neural Networks
2022/10/14 by Sitao Luan, Luan, Sitao, Chenqing Hua +13 · 9 citations
Computer Science · Materials Science · #Advanced Graph Neural Networks #Bayesian Modeling and Causal Inference #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Social and Information Networks (cs.SI)
- Invariant Causal Prediction for Block MDPs
2020/03/12 by Amy Zhang, Zhang, Amy, Clare Lyle +13 · 7 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- AndroidEnv: A Reinforcement Learning Platform for Android
2021/05/27 by Toyama, Daniel, Hamel, Philippe, Gergely, Anita +6 · 7 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation
2022/04/19 by Lee, Jongmin, Paduraru, Cosmin, Mankowitz, Daniel J. +4 · 8 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- A Survey of Exploration Methods in Reinforcement Learning
2021/09/01 by Susan Amin, Maziar Gomrokchi, Amin, Susan +7 · 7 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- For SALE: State-Action Representation Learning for Deep Reinforcement Learning
2023/06/04 by Scott Fujimoto, Fujimoto, Scott, Wei-Di Chang +9 · 9 citations
Computer Science · Engineering · Neuroscience · #Advanced Memory and Neural Computing #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural dynamics and brain function #Reinforcement Learning in Robotics
- Independently Controllable Features
2017/03/22 by Emmanuel Bengio, Valentin Thomas, Bengio, Emmanuel +7 · 9 citations
Computer Science · #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms #Neural Networks and Applications
- Exploring Uncertainty Measures in Deep Networks for Multiple Sclerosis\n Lesion Detection and Segmentation
2018/08/03 by T. R. Gopalakrishnan Nair, Doina Precup, Nair, Tanya +5 · 5 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · Engineering · #AI in cancer detection #Cell Image Analysis Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing Techniques and Applications
- Learning with Pseudo-Ensembles
2014/12/16 by Bachman, Philip, Alsharif, Ouais, Precup, Doina · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
- Hindsight Credit Assignment
2019/12/05 by Anna Harutyunyan, Harutyunyan, Anna, Will Dabney +19 · 7 citations
Computer Science · Business, Management and Accounting · #Reinforcement Learning in Robotics #Financial Distress and Bankruptcy Prediction
- When Waiting is not an Option : Learning Options with a Deliberation\n Cost
2017/09/13 by Jean Harb, Pierre‐Luc Bacon, Harb, Jean +5 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics
- Methods for computing state similarity in Markov Decision Processes
2012/06/27 by Norman Ferns, Ferns, Norman, Pablo Samuel Castro +5 · 3 citations
Engineering · Computer Science · #Fault Detection and Control Systems #AI-based Problem Solving and Planning #Reinforcement Learning in Robotics
- Self-supervised Learning of Distance Functions for Goal-Conditioned Reinforcement Learning
2019/07/05 by Srinivas Venkattaramanujam, Eric Crawford, Venkattaramanujam, Srinivas +5 · 4 citations
Computer Science · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Why Should I Trust You, Bellman? The Bellman Error is a Poor Replacement for Value Error
2022/01/28 by Fujimoto, Scott, Meger, David, Precup, Doina +2 · 5 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Is Heterophily A Real Nightmare For Graph Neural Networks To Do Node Classification?
2021/09/12 by Sitao Luan, Chenqing Hua, Luan, Sitao +13 · 5 citations
Computer Science · Physics and Astronomy · #Advanced Graph Neural Networks #Topic Modeling #Complex Network Analysis Techniques
- MaestroMotif: Skill Design from Artificial Intelligence Feedback
2024/12/11 by Martin Klissarov, Mikael Henaff, Klissarov, Martin +21 · 1 voice · 8 citations
Business, Management and Accounting · #AI and HR Technologies #cs.AI #cs.CL #cs.LG
- Break the Ceiling: Stronger Multi-scale Deep Graph Convolutional Networks
2019/06/05 by Sitao Luan, Luan, Sitao, Mingde Zhao +5 · 4 citations
Computer Science · Materials Science · Physics and Astronomy · #Advanced Graph Neural Networks #Artificial Intelligence (cs.AI) #Complex Network Analysis Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning in Materials Science
- Interference and Generalization in Temporal Difference Learning
2020/03/13 by Emmanuel Bengio, Bengio, Emmanuel, Joëlle Pineau +3 · 4 citations
Computer Science · Physics and Astronomy · #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Neural Networks and Applications
- Agency Is Frame-Dependent
2025/02/06 by David Abel, Abel, David, André Barreto +30 · 4 voices · 1 citation
Neuroscience · Psychology · #Action Observation and Synchronization #Embodied and Extended Cognition #Free Will and Agency #cs.AI
- Randomized Exploration for Reinforcement Learning with General Value Function Approximation
2021/06/15 by Ishfaq, Haque, Cui, Qiwen, Nguyen, Viet +5 · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Code as Reward: Empowering Reinforcement Learning with VLMs
2024/02/07 by David Venuto, Venuto, David, Sami Nur Islam +9 · 7 citations
Computer Science · Engineering · #Elevator Systems and Control #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering Research
- Combined Reinforcement Learning via Abstract Representations
2018/09/12 by François-Lavet, Vincent, Bengio, Yoshua, Precup, Doina +1 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- The Heterophilic Graph Learning Handbook: Benchmarks, Models, Theoretical Analysis, Applications and Challenges
2024/07/12 by Sitao Luan, Chenqing Hua, Luan, Sitao +25 · 1 voice · 6 citations
Computer Science · #Advanced Graph Neural Networks #FOS: Computer and information sciences #Machine Learning (cs.LG) #Social and Information Networks (cs.SI) #cs.LG #cs.SI
- MUDiff: Unified Diffusion for Complete Molecule Generation
2023/04/28 by Chenqing Hua, Hua, Chenqing, Sitao Luan +11 · 5 citations
Materials Science · Computer Science · Biochemistry, Genetics and Molecular Biology · #Machine Learning in Materials Science #Computational Drug Discovery Methods #Protein Structure and Dynamics
- Continuous MDP Homomorphisms and Homomorphic Policy Gradient
2022/09/15 by Rezaei-Shoshtari, Sahand, Zhao, Rosie, Panangaden, Prakash +2 · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- The Stable Entropy Hypothesis and Entropy-Aware Decoding: An Analysis and Algorithm for Robust Natural Language Generation
2023/02/14 by Arora, Kushal, O'Donnell, Timothy J., Precup, Doina +2 · 4 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- The Option Keyboard: Combining Skills in Reinforcement Learning
2021/06/24 by Barreto, André, Borsa, Diana, Hou, Shaobo +8 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Learning Safe Policies with Expert Guidance
2018/05/21 by Jessie Huang, Huang, Jessie, Fa Wu +5 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Policy Gradient Methods in the Presence of Symmetries and State Abstractions
2023/05/09 by Panangaden, Prakash, Rezaei-Shoshtari, Sahand, Zhao, Rosie +2 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- An Equivalence between Loss Functions and Non-Uniform Sampling in Experience Replay
2020/07/12 by Fujimoto, Scott, Meger, David, Precup, Doina · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Discrete Probabilistic Inference as Control in Multi-path Environments
2024/02/15 by Tristan Deleu, Deleu, Tristan, Padideh Nouri +7 · 5 citations
Computer Science · #Advanced Database Systems and Queries #FOS: Computer and information sciences #Machine Learning (cs.LG)
- CryCeleb: A Speaker Verification Dataset Based on Infant Cry Sounds
2023/05/01 by Budaghyan, David, Onu, Charles C., Gorin, Arsenii +2 · 3 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Language Agents Mirror Human Causal Reasoning Biases. How Can We Help Them Think Like Scientists?
2025/05/14 by Anthony GX-Chen, Dongyan Lin, GX-Chen, Anthony +11 · 3 voices · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
- Flexible Option Learning
2021/12/06 by Klissarov, Martin, Precup, Doina · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Prediction and Control in Continual Reinforcement Learning
2023/12/18 by Nishanth Anand, Doina Precup, Anand, Nishanth +1 · 3 citations
Computer Science · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
2023/05/29 by Ishfaq, Haque, Lan, Qingfeng, Xu, Pan +4 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Bayesian Q-learning With Imperfect Expert Demonstrations
2022/10/01 by Che, Fengdi, Zhu, Xiru, Precup, Doina +2 · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Towards Safe Mechanical Ventilation Treatment Using Deep Offline Reinforcement Learning
2022/10/05 by Flemming Kondrup, Thomas Jiralerspong, Kondrup, Flemming +13 · 2 citations
Medicine · Computer Science · #Respiratory Support and Mechanisms #Machine Learning in Healthcare #Cardiac Arrest and Resuscitation
- Plasticity as the Mirror of Empowerment
2025/05/15 by David Abel, Abel, David, Michael Bowling +30 · 3 voices · 3 citations
Neuroscience · Psychology · #cs.AI #cs.LG
- On the Challenges of using Reinforcement Learning in Precision Drug Dosing: Delay and Prolongedness of Action Effects
2023/01/02 by Sumana Basu, Basu, Sumana, Marc‐André Legault +5 · 2 citations
Biochemistry, Genetics and Molecular Biology · Medicine · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Pharmaceutical studies and practices #Receptor Mechanisms and Signaling #Schizophrenia research and treatment
- Practical Kernel-Based Reinforcement Learning
2014/07/21 by André M. S. Barreto, Doina Precup, Barreto, André M. S. +3 · 1 citation
Computer Science · #49L20 (Secondary) #68T05 (Primary) #90C40 #93E20 #93E35 #Advanced Multi-Objective Optimization Algorithms #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #G.3 #I.2.6 #I.2.8 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- A Canonical Form for Weighted Automata and Applications to Approximate Minimization
2015/01/27 by Balle, Borja, Panangaden, Prakash, Precup, Doina · 1 citation
#FOS: Computer and information sciences #Formal Languages and Automata Theory (cs.FL)
- Parseval Regularization for Continual Reinforcement Learning
2024/12/10 by Wesley Chung, Lynn Cherif, Chung, Wesley +5 · 4 citations
Computer Science · #Machine Learning and ELM
- Soft Condorcet Optimization for Ranking of General Agents
2024/10/31 by Marc Lanctot, Lanctot, Marc, Kate Larson +17 · 1 voice · 4 citations
Computer Science · Decision Sciences · #Advanced Algebra and Logic #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Criteria Decision Making #Multiagent Systems (cs.MA) #cs.LG #cs.MA
- Learnings Options End-to-End for Continuous Action Tasks
2017/11/30 by Martin Klissarov, Pierre‐Luc Bacon, Klissarov, Martin +5 · 1 citation
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Auction Theory and Applications
- Disentangling the independently controllable factors of variation by interacting with the world
2018/02/26 by Thomas, Valentin, Bengio, Emmanuel, Fedus, William +6 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
2025/01/29 by Ishfaq, Haque, Wang, Guangyuan, Islam, Sami Nur +1 · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Connecting Weighted Automata and Recurrent Neural Networks through Spectral Learning
2018/07/04 by Rabusseau, Guillaume, Li, Tianyu, Precup, Doina · 1 citation
#FOS: Computer and information sciences #Formal Languages and Automata Theory (cs.FL) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- The Termination Critic
2019/02/26 by Harutyunyan, Anna, Dabney, Will, Borsa, Diana +3 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Neural Transfer Learning for Cry-based Diagnosis of Perinatal Asphyxia
2019/06/24 by Onu, Charles C., Lebensold, Jonathan, Hamilton, William L. +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- An Empirical Study of Batch Normalization and Group Normalization in Conditional Computation
2019/07/31 by Michalski, Vincent, Voleti, Vikram, Kahou, Samira Ebrahimi +4 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Revisit Policy Optimization in Matrix Form
2019/09/19 by Luan, Sitao, Chang, Xiao-Wen, Precup, Doina · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Options of Interest: Temporal Abstraction with Interest Functions
2020/01/01 by Khimya Khetarpal, Khetarpal, Khimya, Martin Klissarov +7 · 1 citation
Computer Science · Engineering · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Control Systems Optimization #Auction Theory and Applications
- Marginalized State Distribution Entropy Regularization in Policy Optimization
2019/12/11 by Riashat Islam, Islam, Riashat, Zafarali Ahmed +3 · 1 citation
Computer Science · Decision Sciences · Engineering · #Advanced Control Systems Optimization #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Simulation Techniques and Applications
- Option-Critic in Cooperative Multi-agent Systems
2019/11/28 by Chakravorty, Jhelum, Ward, Nadeem, Roy, Julien +4 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Multiagent Systems (cs.MA) #Optimization and Control (math.OC) #Systems and Control (eess.SY) #electronic engineering #information engineering
- Policy Evaluation Networks
2020/02/26 by Jean Harb, Tom Schaul, Harb, Jean +5 · 1 citation
Computer Science · Engineering · #Reinforcement Learning in Robotics #Infrastructure Resilience and Vulnerability Analysis #Machine Learning and Algorithms
- Learning to Prove from Synthetic Theorems
2020/06/19 by Eser Aygün, Zafarali Ahmed, Aygün, Eser +13 · 1 citation
Computer Science · #FOS: Computer and information sciences #I.2.3 #Intelligent Tutoring Systems and Adaptive Learning #Logic in Computer Science (cs.LO) #Machine Learning (cs.LG) #Machine Learning and Algorithms #Teaching and Learning Programming
- Learning Successor Features the Simple Way
2024/10/29 by Raymond Chua, Arna Ghosh, Chua, Raymond +7 · 1 voice · 2 citations
Computer Science · #AI in Service Interactions #cs.LG
- Connecting Weighted Automata, Tensor Networks and Recurrent Neural Networks through Spectral Learning
2020/10/19 by Li, Tianyu, Precup, Doina, Rabusseau, Guillaume · 1 citation
#FOS: Computer and information sciences #Formal Languages and Automata Theory (cs.FL) #Machine Learning (cs.LG)
- Locally Persistent Exploration in Continuous Control Tasks with Sparse Rewards
2020/12/26 by Amin, Susan, Gomrokchi, Maziar, Aboutalebi, Hossein +2 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Membership Inference Attacks Against Temporally Correlated Data in Deep Reinforcement Learning
2021/09/08 by Maziar Gomrokchi, Gomrokchi, Maziar, Susan Amin +7 · 1 citation
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Electrostatic Discharge in Electronics
- Correcting Momentum in Temporal Difference Learning
2021/06/07 by Bengio, Emmanuel, Pineau, Joelle, Precup, Doina · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Proving Theorems using Incremental Learning and Hindsight Experience Replay
2021/12/20 by Eser Aygün, Aygün, Eser, Laurent Orseau +13 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #I.2.3 #Logic in Computer Science (cs.LO) #Logic, Reasoning, and Knowledge #Logic, programming, and type systems #Topic Modeling
- Single-Shot Pruning for Offline Reinforcement Learning
2021/12/31 by Samin Yeasar Arnob, Arnob, Samin Yeasar, Riyasat Ohib +5 · 1 citation
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Machine Learning and ELM #Reinforcement Learning in Robotics
- EnzymeFlow: Generating Reaction-specific Enzyme Catalytic Pockets through Flow Matching and Co-Evolutionary Dynamics
2024/10/01 by Hua, Chenqing, Liu, Yong, Zhang, Dinghuai +6 · 2 citations
#Artificial Intelligence (cs.AI) #Computational Engineering #FOS: Biological sciences #FOS: Computer and information sciences #Finance #Machine Learning (cs.LG) #Quantitative Methods (q-bio.QM) #and Science (cs.CE)
- Discovering Temporal Structure: An Overview of Hierarchical Reinforcement Learning
2025/06/16 by Martin Klissarov, Akhil Bagaria, Klissarov, Martin +9 · 1 voice · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Music and Audio Processing #cs.AI
- Learning how to Interact with a Complex Interface using Hierarchical Reinforcement Learning
2022/04/21 by Comanici, Gheorghe, Glaese, Amelia, Gergely, Anita +5 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Complete the Missing Half: Augmenting Aggregation Filtering with Diversification for Graph Convolutional Neural Networks
2022/12/21 by Luan, Sitao, Zhao, Mingde, Hua, Chenqing +2 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- More Efficient Randomized Exploration for Reinforcement Learning via Approximate Sampling
2024/06/18 by Haque Ishfaq, Ishfaq, Haque, Yixin Tan +13 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Distributed Sensor Networks and Detection Algorithms #Energy Efficient Wireless Sensor Networks #FOS: Computer and information sciences #Machine Learning (cs.LG)
- An Empirical Study of the Effectiveness of Using a Replay Buffer on Mode Discovery in GFlowNets
2023/07/15 by Vemgal, Nikhil, Lau, Elaine, Precup, Doina · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- A Consciousness-Inspired Planning Agent for Model-Based Reinforcement Learning
2021/06/03 by Mingde Zhao, Zhen Liu, Zhao, Mingde +9 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics #Topic Modeling
- SCAR: Shapley Credit Assignment for More Efficient RLHF
2025/05/26 by Meng Cao, Cao, Meng, Shuyuan Zhang +5 · 4 citations
Computer Science · #Distributed and Parallel Computing Systems
- SVRG for Policy Evaluation with Fewer Gradient Evaluations
2019/06/09 by Zilun Peng, Peng, Zilun, Touati, Ahmed +4 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Stochastic Gradient Optimization Techniques
- Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
2023/11/03 by Carr, Jonathan Colaço, Panangaden, Prakash, Precup, Doina · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Effective Protein-Protein Interaction Exploration with PPIretrieval
2024/02/06 by Hua, Chenqing, Coley, Connor, Wolf, Guy +2 · 1 citation
#Artificial Intelligence (cs.AI) #Biomolecules (q-bio.BM) #Computational Engineering #FOS: Biological sciences #FOS: Computer and information sciences #Finance #Machine Learning (cs.LG) #and Science (cs.CE)
- Offline Multitask Representation Learning for Reinforcement Learning
2024/03/18 by Haque Ishfaq, Thanh Nguyen-Tang, Ishfaq, Haque +11 · 1 citation
Computer Science · #Reinforcement Learning in Robotics
- Finding Increasingly Large Extremal Graphs with AlphaZero and Tabu Search
2023/11/06 by Mehrabian, Abbas, Anand, Ankit, Kim, Hyunjik +16 · 1 citation
#Artificial Intelligence (cs.AI) #Discrete Mathematics (cs.DM) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- On the Convergence of Bounded Agents
2023/07/20 by David Abel, Abel, David, André Sales Barreto +9 · 1 citation
Computer Science · Decision Sciences · #Artificial Intelligence (cs.AI) #Computability, Logic, AI Algorithms #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Game Theory and Applications #Machine Learning (cs.LG)
- Uncovering a Universal Abstract Algorithm for Modular Addition in Neural Networks
2025/05/23 by Gavin McCracken, McCracken, Gavin, Gabriela Moisescu-Pareja +7 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #Neural Networks and Applications
- QGFN: Controllable Greediness with Action Values
2024/02/07 by Lau, Elaine, Lu, Stephen Zhewen, Pan, Ling +2 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- On the Privacy of Selection Mechanisms with Gaussian Noise
2024/02/09 by Lebensold, Jonathan, Precup, Doina, Balle, Borja · 1 citation
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Reward Propagation Using Graph Convolutional Networks
2020/10/06 by Klissarov, Martin, Precup, Doina · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Generative Active Learning for the Search of Small-molecule Protein Binders
2024/05/02 by Korablyov, Maksym, Liu, Cheng-Hao, Jain, Moksh +31 · 1 citation
#Artificial Intelligence (cs.AI) #Biomolecules (q-bio.BM) #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Robust Reward Modeling via Causal Rubrics
2025/06/19 by Srivastava, Pragya, Singh, Harman, Madhavan, Rahul +9 · 3 citations
Computer Science · #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Topic Modeling
- Functional Acceleration for Policy Mirror Descent
2024/07/23 by Chelu, Veronica, Precup, Doina · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments
2025/05/31 by Luo, Ziyan, Ni, Tianwei, Bacon, Pierre-Luc +2 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Leveraging Lexical Resources for Learning Entity Embeddings in Multi-Relational Data
2016/05/18 by Teng Long, Ryan Lowe, Long, Teng +5 · 1 citation
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Advanced Graph Neural Networks
- Reaction-conditioned De Novo Enzyme Design with GENzyme
2024/11/10 by Hua, Chenqing, Lu, Jiarui, Liu, Yong +7 · 1 citation
#Artificial Intelligence (cs.AI) #Biomolecules (q-bio.BM) #FOS: Biological sciences #FOS: Computer and information sciences