vix.ing · top · new · best · stats · spec

Martha White

  1. Empirical Design in Reinforcement Learning
    2023/04/03 by Andrew Patterson, Andrew D. Patterson, Patterson, Andrew +6 · 3 voices · 8 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Software Engineering Research #Viral Infectious Diseases and Gene Expression in Insects #cs.AI #cs.LG
  2. Maxmin Q-learning: Controlling the Estimation Bias of Q-learning
    2020/02/16 by Qingfeng Lan, Lan, Qingfeng, Yangchen Pan +5 · 15 citations
    Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Adversarial Robustness in Machine Learning
  3. Meta-Learning Representations for Continual Learning
    2019/05/29 by Khurram Javed, Martha White, Javed, Khurram +1 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications
  4. Nonparametric semi-supervised learning of class proportions
    2016/01/08 by Shantanu Jain, Martha White, Jain, Shantanu +5 · 4 citations
    Computer Science · #Machine Learning and Data Classification #Machine Learning and Algorithms #Bayesian Methods and Mixture Models
  5. Investigating the Properties of Neural Network Representations in Reinforcement Learning
    2022/03/30 by Han Wang, Erfan Miahi, Wang, Han +13 · 4 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  6. Is Ockham’s razor losing its edge? New perspectives on the principle of model parsimony
    2025/01/27 by Marina Dubova, Suyog Chandramouli, Gerd Gigerenzer +12 · 2 voices · 5 citations
    Computer Science · Decision Sciences · #Data Visualization and Analytics #Scientific Computing and Data Management #Explainable Artificial Intelligence (XAI)
  7. Investigating the Histogram Loss in Regression
    2024/02/20 by Ehsan Imani, Kai Luedemann, Imani, Ehsan +7 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification
  8. Recovering True Classifier Performance in Positive-Unlabeled Learning
    2017/02/02 by Shantanu Jain, Martha White, Jain, Shantanu +3 · 2 citations
    Computer Science · #Machine Learning and Data Classification #Imbalanced Data Classification Techniques #Machine Learning and Algorithms
  9. Position: Benchmarking is Limited in Reinforcement Learning Research
    2024/06/23 by Scott M. Jordan, Jordan, Scott M., Adam White +7 · 3 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Methodology (stat.ME) #Open Source Software Innovations
  10. Unifying task specification in reinforcement learning
    2016/09/07 by Martha White, White, Martha · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Advanced Multi-Objective Optimization Algorithms #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics
  11. Incremental Truncated LSTD
    2015/11/26 by Clement Gehring, Gehring, Clement, Yangchen Pan +3 · 1 citation
    Computer Science · Engineering · Physics and Astronomy · #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Reinforcement Learning in Robotics
  12. Accelerated Gradient Temporal Difference Learning
    2016/11/28 by Yangchen Pan, Pan, Yangchen, Adam White +3 · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Control Systems and Identification #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Metaheuristic Optimization Algorithms Research #Optimal Power Flow Distribution
  13. Learning Sparse Representations in Reinforcement Learning with Sparse Coding
    2017/07/26 by Lei Le, Raksha Kumaraswamy, Le, Lei +3 · 1 citation
    Computer Science · Engineering · Neuroscience · #Reinforcement Learning in Robotics #Advanced Memory and Neural Computing #Neural dynamics and brain function
  14. High-confidence error estimates for learned value functions
    2018/08/28 by Touqir Sajed, Sajed, Touqir, Wesley Chung +3 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  15. PC-Gym: Benchmark Environments For Process Control Problems
    2024/10/29 by Maximilian Bloor, Bloor, Maximilian, José Torraca +15 · 4 citations
    Engineering · #Advanced Manufacturing and Logistics Optimization #FOS: Electrical engineering #Manufacturing Process and Optimization #Scheduling and Optimization Algorithms #Systems and Control (eess.SY) #electronic engineering #information engineering
  16. An Off-policy Policy Gradient Theorem Using Emphatic Weightings
    2018/11/22 by Ehsan Imani, Eric Graves, Imani, Ehsan +3 · 1 citation
    Computer Science · #Adaptive Dynamic Programming Control #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  17. Optimizing for the Future in Non-Stationary MDPs
    2020/05/17 by Yash Chandak, Georgios Theocharous, Chandak, Yash +9 · 1 citation
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Reinforcement Learning in Robotics #Advanced Multi-Objective Optimization Algorithms
  18. Selective Dyna-style Planning Under Limited Model Capacity
    2020/07/05 by Zaheer Abbas, Abbas, Zaheer, Samuel Sokota +5 · 1 citation
    Computer Science · #Artificial Intelligence in Games #AI-based Problem Solving and Planning #Robotic Path Planning Algorithms
  19. Understanding and Mitigating the Limitations of Prioritized Experience Replay
    2020/07/19 by Yangchen Pan, Jincheng Mei, Pan, Yangchen +11 · 1 citation
    Computer Science · Neuroscience · #Age of Information Optimization #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Functional Brain Connectivity Studies #Machine Learning (cs.LG) #Neural dynamics and brain function
  20. The In-Sample Softmax for Offline Reinforcement Learning
    2023/02/28 by Chenjun Xiao, Xiao, Chenjun, Han Wang +7 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  21. Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence
    2023/01/27 by Lingwei Zhu, Zhu, Lingwei, Zheng Chen +4 · 1 citation
    Decision Sciences · #Decision-Making and Behavioral Economics
  22. Averaging n-step Returns Reduces Variance in Reinforcement Learning
    2024/02/06 by Brett Daley, Daley, Brett, Martha White +3 · 1 citation
    Decision Sciences · Social Sciences · #Experimental Behavioral Economics Studies #FOS: Computer and information sciences #Innovation Diffusion and Forecasting #Machine Learning (cs.LG)
  23. Deep Policy Gradient Methods Without Batch Updates, Target Networks, or Replay Buffers
    2024/11/22 by Gautham Vasan, Vasan, Gautham, Mohamed Elsayed +13 · 2 citations
    Computer Science · #Advanced Data Storage Technologies #Age of Information Optimization #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Robotics (cs.RO) #Stochastic Gradient Optimization Techniques #Systems and Control (eess.SY) #electronic engineering #information engineering
  24. Gradient Temporal-Difference Learning with Regularized Corrections
    2020/07/01 by Sina Ghiassian, Ghiassian, Sina, Andrew Patterson +8 · 1 citation
    Computer Science · Physics and Astronomy · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Neural Networks and Applications
  25. Rethinking the Foundations for Continual Reinforcement Learning
    2025/04/10 by Esraa Elelimy, Elelimy, Esraa, David Szepesvari +5 · 2 voices · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.AI #cs.LG
  26. Investigating the Interplay of Prioritized Replay and Generalization
    2024/07/12 by Parham Mohammad Panahi, Andrew D. Patterson, Panahi, Parham Mohammad +5 · 1 citation
    Psychology · #Communication in Education and Healthcare
  27. Endpoint Replay: Compressing the Recency Buffer in Deep Reinforcement Learning
    2026/07/27 by Parham Mohammad Panahi, Armin Ashrafi, Haoyu Du +3
    #cs.LG