vix.ing · top · new · best · stats · spec

Girgin, Sertan

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1393 citations
    #cs.CL #cs.AI
  2. Gemma 3 Technical Report
    2025/03/25 by Gemma Team, Aishwarya Kamath, Kamath, Aishwarya +418 · 3 voices · 504 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
  3. Gemma 2: Improving Open Language Models at a Practical Size
    2024/07/31 by Morgane Rivière, Gemma Team, Shreya Pathak +292 · 349 citations
    Computer Science · #Natural Language Processing Techniques
  4. Gemma: Open Models Based on Gemini Research and Technology
    2024/03/13 by Thomas Mésnard, Gemma Team, Cassidy Hardin +203 · 172 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
  5. RecurrentGemma: Moving Past Transformers for Efficient Open Language Models
    2024/04/11 by Aleksandar Botev, Botev, Aleksandar, Soham De +127 · 1 voice · 11 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
  6. Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation
    2021/06/24 by C. Daniel Freeman, Erik Frey, Freeman, C. Daniel +9 · 59 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Robotic Locomotion and Control #Modeling and Simulation Systems
  7. Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
    2023/02/07 by Eugene Kharitonov, Damien Vincent, Kharitonov, Eugene +15 · 30 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  8. Nash Learning from Human Feedback
    2023/12/01 by Rémi Munos, Munos, Rémi, Michal Valko +31 · 1 voice · 26 citations
    #stat.ML #cs.AI #cs.GT #cs.LG #cs.MA
  9. What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
    2020/06/10 by Marcin Andrychowicz, Anton Raichuk, Andrychowicz, Marcin +21 · 10 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics #Software Engineering Research
  10. Learning in Mean Field Games: A Survey
    2022/05/25 by Mathieu Laurière, Laurière, Mathieu, Sarah Perrin +13 · 1 voice · 9 citations
    Economics, Econometrics and Finance · #Sports Analytics and Performance #cs.AI #cs.GT #cs.LG #math.OC
  11. WARP: On the Benefits of Weight Averaged Rewarded Policies
    2024/06/24 by Alexandre Ramé, Ramé, Alexandre, Johan Ferret +17 · 1 voice · 6 citations
    Computer Science · Health Professions · Medicine · #Health Promotion and Cardiovascular Prevention #Obesity and Health Practices #cs.AI #cs.LG
  12. RLDS: an Ecosystem to Generate, Share and Use Datasets in Reinforcement Learning
    2021/11/04 by Sabela Ramos, Ramos, Sabela, Sertan Girgin +21 · 9 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Scientific Computing and Data Management #Data Stream Mining Techniques
  13. MusicRL: Aligning Music Generation to Human Preferences
    2024/02/06 by Cideron, Geoffrey, Girgin, Sertan, Verzetti, Mauro +11 · 10 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  14. Scalable Deep Reinforcement Learning Algorithms for Mean Field Games
    2022/03/22 by Mathieu Laurière, Laurière, Mathieu, Sarah Perrin +19 · 7 citations
    Decision Sciences · Economics, Econometrics and Finance · #FOS: Computer and information sciences #FOS: Mathematics #Game Theory and Applications #Innovation Diffusion and Forecasting #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Sports Analytics and Performance
  15. What Matters for Adversarial Imitation Learning?
    2021/06/01 by Manu Orsini, Orsini, Manu, Anton Raichuk +17 · 5 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics #Robot Manipulation and Learning
  16. BOND: Aligning LLMs with Best-of-N Distillation
    2024/07/19 by Pier Giuseppe Sessa, Robert Dadashi, Sessa, Pier Giuseppe +37 · 10 citations
    Engineering · #Advanced Control Systems Optimization
  17. Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback
    2023/05/31 by Roit, Paul, Ferret, Johan, Shani, Lior +16 · 6 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  18. Acme: A Research Framework for Distributed Reinforcement Learning
    2020/06/01 by Hoffman, Matthew W., Shahriari, Bobak, Aslanides, John +36 · 4 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  19. Continuous Control with Action Quantization from Demonstrations
    2021/10/19 by Dadashi, Robert, Hussenot, Léonard, Vincent, Damien +4 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  20. Diversity-Rewarded CFG Distillation
    2024/10/08 by Geoffrey Cideron, Andrea Agostinelli, Cideron, Geoffrey +13 · 3 citations
    Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Process Optimization and Integration
  21. Hyperparameter Selection for Imitation Learning
    2021/05/25 by Hussenot, Leonard, Andrychowicz, Marcin, Vincent, Damien +11 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  22. Solving N-player dynamic routing games with congestion: a mean field approach
    2021/10/22 by Cabannes, Theophile, Lauriere, Mathieu, Perolat, Julien +7 · 1 citation
    #Dynamical Systems (math.DS) #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Multiagent Systems (cs.MA) #Networking and Internet Architecture (cs.NI) #Optimization and Control (math.OC) #Systems and Control (eess.SY) #electronic engineering #information engineering
  23. Decoding a Neural Retriever's Latent Space for Query Suggestion
    2022/10/21 by Adolphs, Leonard, Huebscher, Michelle Chen, Buck, Christian +4 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  24. Get Back Here: Robust Imitation by Return-to-Distribution Planning
    2023/05/02 by Geoffrey Cideron, Cideron, Geoffrey, Baruch Tabanpour +15 · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO) #Systems and Control (eess.SY) #electronic engineering #information engineering