vix.ing · top · new · best · stats · spec

Bobak Shahriari

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1358 citations
    #cs.CL #cs.AI
  2. Taking the Human Out of the Loop: A Review of Bayesian Optimization
    2015/12/10 by Bobak Shahriari, Kevin Swersky, Ziyu Wang +2 · 310 citations
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Advanced Multi-Objective Optimization Algorithms #Gaussian Processes and Bayesian Inference
  3. Gemma 3 Technical Report
    2025/03/25 by Gemma Team, Aishwarya Kamath, Johan Ferret +418 · 3 voices · 466 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
  4. Gemma 2: Improving Open Language Models at a Practical Size
    2024/07/31 by Gemma Team, Morgane Rivière, Shreya Pathak +292 · 307 citations
    Computer Science · #Natural Language Processing Techniques
  5. Gemma: Open Models Based on Gemini Research and Technology
    2024/03/13 by Thomas Mésnard, Gemma Team, Cassidy Hardin +203 · 148 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
  6. Gemma 4 Technical Report
    2026/07/02 by Gemma Team, Sherif El Abd, Vaibhav Aggarwal +320 · 7 voices · 8 citations
    #cs.CL #cs.AI
  7. Critic Regularized Regression
    2020/06/26 by Ziyu Wang, Wang, Ziyu, Alexander Novikov +19 · 12 citations
    Computer Science · #Adaptive Dynamic Programming Control #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  8. On Multi-objective Policy Optimization as a Tool for Reinforcement Learning: Case Studies in Offline RL and Finetuning
    2021/06/15 by Abbas Abdolmaleki, Abdolmaleki, Abbas, Sandy H. Huang +25 · 2 citations
    Computer Science · #Advanced Multi-Objective Optimization Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  9. Making Efficient Use of Demonstrations to Solve Hard Exploration Problems
    2019/09/03 by Tom Le Paine, Paine, Tom Le, Çağlar Gülçehre +25 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Optimization and Search Problems #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms
  10. Learning from negative feedback, or positive feedback or both
    2024/10/05 by Abbas Abdolmaleki, Bilal Piot, Abdolmaleki, Abbas +21 · 5 citations
    Computer Science · #Semantic Web and Ontologies