vix.ing · top · new · best · stats · spec

Szymon Sidor

  1. Competitive Programming with Large Reasoning Models
    2025/02/03 by OpenAI, :, Ahmed El-Kishky +50 · 13 voices · 64 citations
    Decision Sciences · Economics, Econometrics and Finance · #Economic theories and models #Multi-Criteria Decision Making #cs.AI #cs.CL #cs.LG
  2. GPT-4 Technical Report
    2023/03/15 by Josh Achiam, OpenAI, Steven Adler +556 · 6 voices · 4330 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
  3. Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
    2022/03/07 by Greg Yang, Yang, Greg, J. Edward Hu +19 · 5 voices · 57 citations
    Computer Science · Physics and Astronomy · #Advanced Neural Network Applications #Computational Physics and Python Applications #Disordered Systems and Neural Networks (cond-mat.dis-nn) #FOS: Computer and information sciences #FOS: Physical sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Parallel Computing and Optimization Techniques #cond-mat.dis-nn #cs.LG #cs.NE
  4. Evolution Strategies as a Scalable Alternative to Reinforcement Learning
    2017/03/10 by Tim Salimans, Jonathan Ho, Salimans, Tim +7 · 1 voice · 201 citations
    Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Metaheuristic Optimization Algorithms Research #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics #cs.AI #cs.LG #cs.NE #stat.ML
  5. Dota 2 with Large Scale Deep Reinforcement Learning
    2019/12/13 by OpenAI, Christopher Berner, : +50 · 1 voice · 181 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG #stat.ML
  6. Learning dexterous in-hand manipulation
    2019/11/18 by OpenAI: Marcin Andrychowicz, OpenAI Marcin Andrychowicz, Bowen Baker +16 · 173 citations
    Engineering · Computer Science · #Robot Manipulation and Learning #Reinforcement Learning in Robotics #Teleoperation and Haptic Systems
  7. Parameter Space Noise for Exploration
    2017/06/06 by Matthias Plappert, Rein Houthooft, Plappert, Matthias +15 · 79 citations
    Computer Science · Economics, Econometrics and Finance · #Reinforcement Learning in Robotics #Evolutionary Algorithms and Applications #Sports Analytics and Performance
  8. Emergent Complexity via Multi-Agent Competition
    2017/10/10 by Trapit Bansal, Bansal, Trapit, Jakub Pachocki +7 · 32 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Reinforcement Learning in Robotics
  9. Schema Networks: Zero-shot Transfer with a Generative Causal Model of Intuitive Physics
    2017/06/14 by Ken Kansky, Tom Silver, Kansky, Ken +17 · 21 citations
    Computer Science · #Reinforcement Learning in Robotics #Explainable Artificial Intelligence (XAI) #Adversarial Robustness in Machine Learning