vix.ing · top · new · best · stats · spec

Jakob Foerster

  1. The Hanabi challenge: A new frontier for AI research
    2019/02/01 by Nolan Bard, Jakob Foerster, Jakob N. Foerster +14 · 4 voices · 43 citations
    Computer Science · Decision Sciences · #Artificial Intelligence in Games #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research
  2. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
    2024/08/12 by Chris Lu, Lu, Chris, Cong Lu +10 · 19 voices · 165 citations
    Decision Sciences · #Scientific Computing and Data Management
  3. Perfectly Secure Steganography Using Minimum Entropy Coupling
    2022/10/24 by Christian Schroeder de Witt, de Witt, Christian Schroeder, Samuel Sokota +7 · 3 voices · 7 citations
    #cs.CR #cs.AI #cs.MM
  4. The StarCraft Multi-Agent Challenge
    2019/02/11 by Mikayel Samvelyan, Samvelyan, Mikayel, Tabish Rashid +17 · 1 voice · 58 citations
    Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multiagent Systems (cs.MA) #cs.LG #cs.MA #stat.ML
  5. Learning to Communicate to Solve Riddles with Deep Distributed Recurrent\n Q-Networks
    2016/02/08 by Jakob N. Foerster, Jakob Foerster, Yannis M. Assael +7 · 3 voices · 5 citations
    Computer Science · Engineering · #Distributed Control Multi-Agent Systems #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics #cs.AI #cs.LG
  6. Counterfactual Multi-Agent Policy Gradients
    2017/05/24 by Jakob Foerster, Foerster, Jakob, Gregory Farquhar +7 · 1 voice · 72 citations
    Computer Science · Engineering · #Fuel Cells and Related Materials #Reinforcement Learning in Robotics #cs.AI #cs.MA
  7. Learning to Communicate with Deep Multi-Agent Reinforcement Learning
    2016/05/21 by Jakob Foerster, Foerster, Jakob N., Yannis Assael +5 · 75 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Reinforcement Learning in Robotics #Machine Learning and Algorithms
  8. QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
    2018/03/30 by Tabish Rashid, Mikayel Samvelyan, Rashid, Tabish +9 · 45 citations
    Computer Science · Decision Sciences · #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Multiagent Systems (cs.MA) #Open Source Software Innovations #Reinforcement Learning in Robotics
  9. Hyperagents
    2026/03/19 by Jenny Zhang, Bingchen Zhao, Wannan Yang +5 · 11 voices · 2 citations
    Computer Science · #cs.AI
  10. TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
    2024/10/04 by Jonathan Cook, Tim Rocktäschel, Cook, Jonathan +7 · 1 voice · 8 citations
    Computer Science · #Mathematics, Computing, and Information Processing #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.HC #cs.LG
  11. Stabilising Experience Replay for Deep Multi-Agent Reinforcement Learning
    2017/02/28 by Jakob Foerster, Foerster, Jakob, Nantas Nardelli +11 · 23 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Smart Grid Energy Management #Smart Grid Security and Resilience
  12. BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
    2024/11/20 by Davide Paglieri, Paglieri, Davide, Bartłomiej Cupiał +25 · 4 voices · 33 citations
    Computer Science · #Multi-Agent Systems and Negotiation #Semantic Web and Ontologies #cs.AI
  13. Multi-Agent Risks from Advanced AI
    2025/02/19 by Lewis Hammond, Hammond, Lewis, Alan Chan +89 · 3 voices · 36 citations
    Social Sciences · #Ethics and Social Impacts of AI #cs.AI #cs.CY #cs.ET #cs.LG #cs.MA
  14. Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
    2024/02/26 by Mikayel Samvelyan, Samvelyan, Mikayel, Sharath Chandra Raparthy +21 · 33 citations
    Computer Science · #Advanced Malware Detection Techniques #Artificial Intelligence in Games #Information and Cyber Security
  15. Simplifying Deep Temporal Difference Learning
    2024/07/05 by Matteo Gallici, Mattie Fellows, Gallici, Matteo +15 · 2 voices · 24 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #cs.LG
  16. "Other-Play" for Zero-Shot Coordination
    2020/03/06 by Hengyuan Hu, Hu, Hengyuan, Adam Lerer +5 · 16 citations
    Computer Science · Economics, Econometrics and Finance · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Sports Analytics and Performance
  17. SMACv2: An Improved Benchmark for Cooperative Multi-Agent Reinforcement Learning
    2022/12/14 by Benjamin J. Ellis, Ellis, Benjamin, Skander Moalla +12 · 19 citations
    Computer Science · #Reinforcement Learning in Robotics #Machine Learning and Data Classification #Adaptive Dynamic Programming Control
  18. JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
    2023/11/16 by Alexander R. Rutherford, Rutherford, Alexander, Benjamin J. Ellis +38 · 23 citations
    Computer Science · Social Sciences · #Reinforcement Learning in Robotics #Artificial Intelligence in Games #Digital Games and Media
  19. Towards end-to-end automation of AI research
    2026/03/25 by Chris Lu, Cong Lu, R. T. Lange +5 · 2 voices · 20 citations
    Biochemistry, Genetics and Molecular Biology · Decision Sciences · Materials Science · #Cell Image Analysis Techniques #Machine Learning in Materials Science #Scientific Computing and Data Management
  20. Evolution Strategies at the Hyperscale
    2025/11/20 by Bidipta Sarkar, Mattie Fellows, Sarkar, Bidipta +38 · 4 voices · 1 citation
    Computer Science · Mathematics · #Metaheuristic Optimization Algorithms Research #Stochastic Gradient Optimization Techniques #Tensor decomposition and applications #cs.AI #cs.LG
  21. Structured State Space Models for In-Context Reinforcement Learning
    2023/03/07 by Chris Xiaoxuan Lu, Yannick Schroecker, Lu, Chris +11 · 13 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
  22. Bayesian Action Decoder for Deep Multi-Agent Reinforcement Learning
    2018/11/04 by Jakob Foerster, Foerster, Jakob N., Francis Song +13 · 8 citations
    Computer Science · Engineering · #Anomaly Detection Techniques and Applications #Artificial Immune Systems Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Reinforcement Learning in Robotics
  23. Replay-Guided Adversarial Environment Design
    2021/10/06 by Minqi Jiang, Jiang, Minqi, Michael D. Dennis +9 · 9 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
  24. Autodata: An agentic data scientist to create high quality synthetic data
    2026/06/24 by Ilia Kulikov, Chenxi Whitehouse, Tianhao Wu +12 · 4 voices
    #cs.AI #cs.CL #cs.LG
  25. Generative AI for End-to-End Limit Order Book Modelling: A Token-Level Autoregressive Generative Model of Message Flow Using a Deep State Space Network
    2023/08/23 by Peer Nagy, Sascha Frey, Nagy, Peer +11 · 9 citations
    Economics, Econometrics and Finance · Decision Sciences · #Financial Markets and Investment Strategies #Stock Market Forecasting Methods #Complex Systems and Time Series Analysis
  26. Nocturne: a scalable driving benchmark for bringing multi-agent learning one step closer to the real world
    2022/06/20 by Eugene Vinitsky, Nathan Lichtlé, Vinitsky, Eugene +7 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  27. Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
    2024/09/12 by Alisia Lupidi, Carlos Gemmell, Lupidi, Alisia +13 · 9 citations
    Computer Science · #Anomaly Detection Techniques and Applications #Time Series Analysis and Forecasting #Video Analysis and Summarization
  28. Equivariant Networks for Zero-Shot Coordination
    2022/10/21 by Darius Muglich, Muglich, Darius, Christian Schroeder de Witt +7 · 5 citations
    Computer Science · #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Topic Modeling
  29. Off-Belief Learning
    2021/03/06 by Hengyuan Hu, Hu, Hengyuan, Adam Lerer +11 · 4 citations
    Computer Science · #Reinforcement Learning in Robotics #Machine Learning and Algorithms #Topic Modeling
  30. A New Formalism, Method and Open Issues for Zero-Shot Coordination
    2021/06/11 by Johannes Treutlein, Michael D. Dennis, Treutlein, Johannes +5 · 4 citations
    Computer Science · Social Sciences · #Reinforcement Learning in Robotics #Ethics and Social Impacts of AI
  31. Illusory Attacks: Information-Theoretic Detectability Matters in Adversarial Attacks
    2022/07/20 by Tim Franzmeyer, Franzmeyer, Tim, João F. Henriques +10 · 4 citations
    Computer Science · #Adversarial Robustness in Machine Learning
  32. Policy-Guided Diffusion
    2024/04/09 by Matthew Thomas Jackson, Jackson, Matthew Thomas, Michael Matthews +9 · 6 citations
    Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Generative Adversarial Networks and Image Synthesis
  33. Near to Mid-term Risks and Opportunities of Open-Source Generative AI
    2024/04/25 by Francisco Eiras, Eiras, Francisco, Aleksandar Petrov +45 · 2 voices · 4 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.LG
  34. Discovering Preference Optimization Algorithms with and for Large Language Models
    2024/06/12 by Chris Lu, Lu, Chris, Samuel Holt +11 · 6 citations
    Computer Science · #Data Management and Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG)
  35. The Complexity Dynamics of Grokking
    2024/12/13 by Branton DeMoss, DeMoss, Branton, Silvia Sapora +7 · 1 voice · 6 citations
    Computer Science · Engineering · #Metal Forming Simulation Techniques #Vibration and Dynamic Analysis #cs.LG
  36. PARDEN, Can You Repeat That? Defending against Jailbreaks via Repetition
    2024/05/13 by Ziyang Zhang, Zhang, Ziyang, Qizhen Zhang +3 · 5 citations
    Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Criminal Justice and Corrections Analysis #Criminal Law and Evidence #FOS: Computer and information sciences #I.2.7 #Jury Decision Making Processes
  37. Multi-Agent Common Knowledge Reinforcement Learning
    2018/10/27 by Christian A. Schroeder de Witt, de Witt, Christian A. Schroeder, Jakob Foerster +9 · 3 citations
    Computer Science · Social Sciences · #Reinforcement Learning in Robotics #Experimental Behavioral Economics Studies #Evolutionary Game Theory and Cooperation
  38. Measuring what Matters: Construct Validity in Large Language Model Benchmarks
    2025/11/03 by Andrew M. Bean, Bean, Andrew M., Angelika Romanou +78 · 10 citations
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #Computational and Text Analysis Methods #FOS: Computer and information sciences #Topic Modeling
  39. No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
    2024/08/27 by Alexander R. Rutherford, Michael Beukman, Rutherford, Alexander +9 · 5 citations
    Social Sciences · Decision Sciences · Mathematics · #Educational Assessment and Pedagogy #Educational Assessment and Improvement #Statistics Education and Methodologies
  40. Scaling Opponent Shaping to High Dimensional Games
    2023/12/19 by Akbir Khan, Khan, Akbir, Timon Willi +13 · 3 citations
    Computer Science · Economics, Econometrics and Finance · Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Mental Health Research Topics #Sports Analytics and Performance #Time Series Analysis and Forecasting
  41. COLA: Consistent Learning with Opponent-Learning Awareness
    2022/03/08 by Timon Willi, Willi, Timon, Alistair Letcher +5 · 2 citations
    Decision Sciences · Social Sciences · #Artificial Intelligence (cs.AI) #Computer Science and Game Theory (cs.GT) #Evolutionary Game Theory and Cooperation #Experimental Behavioral Economics Studies #FOS: Computer and information sciences #Game Theory and Applications #Machine Learning (cs.LG)
  42. Human-AI Coordination via Human-Regularized Search and Learning
    2022/10/11 by Hengyuan Hu, David Wu, Hu, Hengyuan +7 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Reinforcement Learning in Robotics
  43. Mixture of Experts in a Mixture of RL settings
    2024/06/26 by Timon Willi, Willi, Timon, Johan Obando-Ceron +7 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Bayesian Methods and Mixture Models #Distributed Sensor Networks and Detection Algorithms #Expert finding and Q&A systems #FOS: Computer and information sciences #Machine Learning (cs.LG)
  44. On the interaction between supervision and self-play in emergent\n communication
    2020/02/03 by Ryan Lowe, Abhinav Gupta, Lowe, Ryan +7 · 4 citations
    Social Sciences · Computer Science · #Language and cultural evolution #Reinforcement Learning in Robotics #Topic Modeling
  45. Exploring Zero-Shot Emergent Communication in Embodied Multi-Agent Populations
    2020/10/29 by Kalesha Bullard, Franziska Meier, Bullard, Kalesha +7 · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Modular Robots and Swarm Intelligence #Multiagent Systems (cs.MA) #Reinforcement Learning in Robotics #Robot Manipulation and Learning
  46. Rethinking Out-of-Distribution Detection for Reinforcement Learning: Advancing Methods for Evaluation and Detection
    2024/04/10 by Linas Nasvytis, Nasvytis, Linas, Kai Sandbrink +7 · 2 citations
    Computer Science · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG)
  47. Adversarial Cheap Talk
    2022/11/20 by Chris Xiaoxuan Lu, Timon Willi, Lu, Chris +5 · 1 citation
    Computer Science · Biochemistry, Genetics and Molecular Biology · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Bacillus and Francisella bacterial research
  48. Similarity-based cooperative equilibrium
    2022/11/26 by Caspar Oesterheld, Oesterheld, Caspar, Johannes Treutlein +7 · 1 citation
    Decision Sciences · Social Sciences · #91A10 (Primary) 91A05 91A26 91A35 (Secondary) #Artificial Intelligence (cs.AI) #Computer Science and Game Theory (cs.GT) #Evolutionary Game Theory and Cooperation #Experimental Behavioral Economics Studies #FOS: Computer and information sciences #Game Theory and Applications #I.2.11 #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
  49. LOB-Bench: Benchmarking Generative AI for Finance -- an Application to Limit Order Book Data
    2025/02/13 by Peer Nagy, S. Frey, Nagy, Peer +12 · 3 citations
    Decision Sciences · #Computational Engineering #Computational Finance (q-fin.CP) #FOS: Computer and information sciences #FOS: Economics and business #Finance #Machine Learning (cs.LG) #Stock Market Forecasting Methods #Trading and Market Microstructure (q-fin.TR) #and Science (cs.CE)
  50. Risks and Opportunities of Open-Source Generative AI
    2024/05/14 by Francisco Eiras, Eiras, Francisco, Aleksander Petrov +47 · 1 citation
    Computer Science · Decision Sciences · Social Sciences · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Scientific Computing and Data Management
  51. An Optimisation Framework for Unsupervised Environment Design
    2025/05/27 by Nathan Monette, Monette, Nathan, Alistair Letcher +11 · 3 citations
    Engineering · #BIM and Construction Integration
  52. A Clean Slate for Offline Reinforcement Learning
    2025/04/15 by Matthew Thomas Jackson, Jackson, Matthew Thomas, Uljad Berdica +8 · 1 voice · 2 citations
    Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Stochastic Gradient Optimization Techniques #cs.AI #cs.LG #cs.RO
  53. EvIL: Evolution Strategies for Generalisable Imitation Learning
    2024/06/15 by Silvia Sapora, Sapora, Silvia, Gokul Swamy +7 · 1 citation
    Computer Science · Engineering · #Human Pose and Action Recognition #Reinforcement Learning in Robotics #Robot Manipulation and Learning
  54. HelloFresh: LLM Evaluations on Streams of Real-World Human Editorial Actions across X Community Notes and Wikipedia edits
    2024/06/05 by Tim Franzmeyer, Aleksandar Shtedritski, Franzmeyer, Tim +9 · 1 citation
    Social Sciences · Computer Science · #Wikis in Education and Collaboration #Topic Modeling #Advanced Text Analysis Techniques
  55. Beyond the Boundaries of Proximal Policy Optimization
    2024/11/01 by Charlie B. Tan, Edan Toledo, Tan, Charlie B. +7 · 1 citation
    Economics, Econometrics and Finance · #Artificial Intelligence (cs.AI) #Economic Policies and Impacts #FOS: Computer and information sciences #Machine Learning (cs.LG)
  56. BAM! Just Like That: Simple and Efficient Parameter Upcycling for Mixture of Experts
    2024/08/15 by Qizhen Zhang, Zhang, Qizhen, Nikolas Gritsch +19 · 1 citation
    Computer Science · Business, Management and Accounting · #Time Series Analysis and Forecasting #Customer churn and segmentation #Data Stream Mining Techniques
  57. AI & Human Co-Improvement for Safer Co-Superintelligence
    2025/12/05 by Jason Weston, Weston, Jason, Jakob Foerster +1 · 3 citations
    Social Sciences · Medicine · Physics and Astronomy · #Ethics and Social Impacts of AI #Artificial Intelligence in Healthcare and Education #Space Science and Extraterrestrial Life
  58. Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
    2024/12/22 by Benjamin J. Ellis, Matthew Jackson, Ellis, Benjamin +11 · 1 citation
    Computer Science · #Multi-Agent Systems and Negotiation #Logic, Reasoning, and Knowledge