Szymon Sidor
- Competitive Programming with Large Reasoning Models
2025/02/03 by OpenAI, :, Ahmed El-Kishky +50 · 13 voices · 64 citations
Decision Sciences · Economics, Econometrics and Finance · #Economic theories and models #Multi-Criteria Decision Making #cs.AI #cs.CL #cs.LG
- GPT-4 Technical Report
2023/03/15 by Josh Achiam, OpenAI, Steven Adler +556 · 6 voices · 4330 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
- Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
2022/03/07 by Greg Yang, Yang, Greg, J. Edward Hu +19 · 5 voices · 57 citations
Computer Science · Physics and Astronomy · #Advanced Neural Network Applications #Computational Physics and Python Applications #Disordered Systems and Neural Networks (cond-mat.dis-nn) #FOS: Computer and information sciences #FOS: Physical sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Parallel Computing and Optimization Techniques #cond-mat.dis-nn #cs.LG #cs.NE
- Evolution Strategies as a Scalable Alternative to Reinforcement Learning
2017/03/10 by Tim Salimans, Jonathan Ho, Salimans, Tim +7 · 1 voice · 201 citations
Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Metaheuristic Optimization Algorithms Research #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics #cs.AI #cs.LG #cs.NE #stat.ML
- Dota 2 with Large Scale Deep Reinforcement Learning
2019/12/13 by OpenAI, Christopher Berner, : +50 · 1 voice · 181 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG #stat.ML
- Learning dexterous in-hand manipulation
2019/11/18 by OpenAI: Marcin Andrychowicz, OpenAI Marcin Andrychowicz, Bowen Baker +16 · 173 citations
Engineering · Computer Science · #Robot Manipulation and Learning #Reinforcement Learning in Robotics #Teleoperation and Haptic Systems
- Parameter Space Noise for Exploration
2017/06/06 by Matthias Plappert, Rein Houthooft, Plappert, Matthias +15 · 79 citations
Computer Science · Economics, Econometrics and Finance · #Reinforcement Learning in Robotics #Evolutionary Algorithms and Applications #Sports Analytics and Performance
- Emergent Complexity via Multi-Agent Competition
2017/10/10 by Trapit Bansal, Bansal, Trapit, Jakub Pachocki +7 · 32 citations
Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Reinforcement Learning in Robotics
- Schema Networks: Zero-shot Transfer with a Generative Causal Model of Intuitive Physics
2017/06/14 by Ken Kansky, Tom Silver, Kansky, Ken +17 · 21 citations
Computer Science · #Reinforcement Learning in Robotics #Explainable Artificial Intelligence (XAI) #Adversarial Robustness in Machine Learning