vix.ing · top · new · best · stats · spec

Sertan Girgin

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1393 citations
    #cs.CL #cs.AI
  2. Gemma 3 Technical Report
    2025/03/25 by Gemma Team, Aishwarya Kamath, Kamath, Aishwarya +418 · 3 voices · 532 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
  3. Gemma 2: Improving Open Language Models at a Practical Size
    2024/07/31 by Gemma Team, Morgane Rivière, Riviere, Morgane +292 · 406 citations
    Computer Science · #Natural Language Processing Techniques
  4. Gemma: Open Models Based on Gemini Research and Technology
    2024/03/13 by Gemma Team, Thomas Mésnard, Mesnard, Thomas +203 · 192 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
  5. RecurrentGemma: Moving Past Transformers for Efficient Open Language Models
    2024/04/11 by Aleksandar Botev, Soham De, Botev, Aleksandar +127 · 1 voice · 11 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
  6. Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation
    2021/06/24 by C. Daniel Freeman, Freeman, C. Daniel, Erik Frey +9 · 63 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Robotic Locomotion and Control #Modeling and Simulation Systems
  7. Gemma 4 Technical Report
    2026/07/02 by Gemma Team, Sherif El Abd, Vaibhav Aggarwal +320 · 7 voices · 8 citations
    #cs.CL #cs.AI
  8. Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
    2023/02/07 by Eugene Kharitonov, Kharitonov, Eugene, Damien Vincent +15 · 34 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  9. Nash Learning from Human Feedback
    2023/12/01 by Rémi Munos, Michal Valko, Munos, Rémi +31 · 1 voice · 28 citations
    #stat.ML #cs.AI #cs.GT #cs.LG #cs.MA
  10. What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
    2020/06/10 by Marcin Andrychowicz, Andrychowicz, Marcin, Anton Raichuk +21 · 17 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics #Software Engineering Research
  11. WARP: On the Benefits of Weight Averaged Rewarded Policies
    2024/06/24 by Alexandre Ramé, Ramé, Alexandre, Johan Ferret +17 · 1 voice · 8 citations
    Computer Science · Health Professions · Medicine · #Health Promotion and Cardiovascular Prevention #Obesity and Health Practices #cs.AI #cs.LG
  12. Learning in Mean Field Games: A Survey
    2022/05/25 by Mathieu Laurière, Sarah Perrin, Laurière, Mathieu +13 · 1 voice · 10 citations
    Economics, Econometrics and Finance · #Sports Analytics and Performance #cs.AI #cs.GT #cs.LG #math.OC
  13. RLDS: an Ecosystem to Generate, Share and Use Datasets in Reinforcement Learning
    2021/11/04 by Sabela Ramos, Ramos, Sabela, Sertan Girgin +21 · 11 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Scientific Computing and Data Management #Data Stream Mining Techniques
  14. MusicRL: Aligning Music Generation to Human Preferences
    2024/02/06 by Geoffrey Cideron, Sertan Girgin, Cideron, Geoffrey +25 · 12 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  15. Scalable Deep Reinforcement Learning Algorithms for Mean Field Games
    2022/03/22 by Mathieu Laurière, Sarah Perrin, Laurière, Mathieu +19 · 8 citations
    Decision Sciences · Economics, Econometrics and Finance · #FOS: Computer and information sciences #FOS: Mathematics #Game Theory and Applications #Innovation Diffusion and Forecasting #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Sports Analytics and Performance
  16. BOND: Aligning LLMs with Best-of-N Distillation
    2024/07/19 by Pier Giuseppe Sessa, Robert Dadashi, Sessa, Pier Giuseppe +37 · 12 citations
    Engineering · #Advanced Control Systems Optimization
  17. What Matters for Adversarial Imitation Learning?
    2021/06/01 by Manu Orsini, Orsini, Manu, Anton Raichuk +17 · 5 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics #Robot Manipulation and Learning
  18. Diversity-Rewarded CFG Distillation
    2024/10/08 by Geoffrey Cideron, Andrea Agostinelli, Cideron, Geoffrey +13 · 3 citations
    Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Process Optimization and Integration
  19. Get Back Here: Robust Imitation by Return-to-Distribution Planning
    2023/05/02 by Geoffrey Cideron, Baruch Tabanpour, Cideron, Geoffrey +15 · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO) #Systems and Control (eess.SY) #electronic engineering #information engineering