Piché, Alexandre
- Causal Discovery with Language Models as Imperfect Experts
2023/07/05 by Long, Stephanie, Piché, Alexandre, Zantedeschi, Valentina +2 · 11 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Self-Evolving Curriculum for LLM Reasoning
2025/05/20 by Xiaoyin Chen, Chen, Xiaoyin, James J. Lu +15 · 30 citations
Computer Science · #Digital Rights Management and Security #Open Education and E-Learning #Information Systems Education and Curriculum Development
- Can large language models build causal graphs?
2023/03/07 by Stephanie Long, Tibor Schuster, Long, Stephanie +3 · 8 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Epigenetics and DNA Methylation #FOS: Computer and information sciences #Machine Learning in Healthcare #Topic Modeling
- Bridging the Gap Between Target Networks and Functional Regularization
2021/06/04 by Piché, Alexandre, Thomas, Valentin, Pardinas, Rafael +4 · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Reward Estimation for Variance Reduction in Deep Reinforcement Learning
2018/05/09 by Joshua Romoff, Romoff, Joshua, Peter Henderson +7 · 1 citation
Computer Science · #Advanced Multi-Objective Optimization Algorithms #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Mastering the Unsupervised Reinforcement Learning Benchmark from Pixels
2022/09/24 by Sai Rajeswar, Pietro Mazzaglia, Rajeswar, Sai +11 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- BigCharts-R1: Enhanced Chart Reasoning with Visual Reinforcement Finetuning
2025/08/13 by Masry, Ahmed, Puri, Abhay, Hashemi, Masoud +13 · 5 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- How to Train Your LLM Web Agent: A Statistical Diagnosis
2025/07/05 by Dheeraj Vattikonda, Vattikonda, Dheeraj, S. Ravichandran +29 · 5 citations
Computer Science · #Digital Rights Management and Security
- PipelineRL: Faster On-policy Reinforcement Learning for Long Sequence Generation
2025/09/23 by Alexandre Piché, Ehsan Kamalloo, Piché, Alexandre +7 · 3 citations
Computer Science · #Evolutionary Algorithms and Applications #Speech Recognition and Synthesis #Reinforcement Learning in Robotics