vix.ing · top · new · best · stats · spec

Gat, Itai

  1. The Llama 3 Herd of Models
    2024/07/31 by Grattafiori, Aaron, Dubey, Abhimanyu, Jauhri, Abhinav +556 · 2853 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  2. Flow Matching Guide and Code
    2024/12/09 by Yaron Lipman, Lipman, Yaron, Marton Havasi +18 · 10 voices · 82 citations
    Decision Sciences · Computer Science · #Simulation Techniques and Applications #Software System Performance and Reliability
  3. Code Llama: Open Foundation Models for Code
    2023/08/24 by Baptiste Rozière, Jonas Gehring, Rozière, Baptiste +48 · 316 citations
    Computer Science · #Model-Driven Software Engineering Techniques #Advanced Database Systems and Queries #Semantic Web and Ontologies
  4. Simple and Controllable Music Generation
    2023/06/08 by Jade Copet, Copet, Jade, Felix Kreuk +13 · 2 voices · 104 citations
    Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #cs.AI #cs.LG #cs.SD #eess.AS
  5. Discrete Flow Matching
    2024/07/22 by Gat, Itai, Remez, Tal, Shaul, Neta +5 · 72 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. EXPRESSO: A Benchmark and Analysis of Discrete Expressive Speech Resynthesis
    2023/08/10 by Tu Anh Nguyen, Nguyen, Tu Anh, Wei-Ning Hsu +23 · 24 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  7. Textually Pretrained Speech Language Models
    2023/05/22 by Michael Hassid, Tal Remez, Hassid, Michael +21 · 22 citations
    Computer Science · #Topic Modeling #Speech Recognition and Synthesis #Natural Language Processing Techniques
  8. Spirit LM: Interleaved Spoken and Written Language Model
    2024/02/08 by Nguyen, Tu Anh, Muller, Benjamin, Yu, Bokai +13 · 26 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  9. D-Flow: Differentiating through Flows for Controlled Generation
    2024/02/21 by Ben-Hamu, Heli, Puny, Omri, Gat, Itai +3 · 25 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  10. Diverse and Aligned Audio-to-Video Generation via Text-to-Video Model Adaptation
    2023/09/28 by Guy Yariv, Yariv, Guy, Itai Gat +9 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Music and Audio Processing #Video Analysis and Summarization
  11. Flow Matching with General Discrete Paths: A Kinetic-Optimal Perspective
    2024/12/04 by Neta Shaul, Shaul, Neta, Itai Gat +15 · 1 voice · 18 citations
    Engineering · #Fluid Dynamics and Turbulent Flows #Lattice Boltzmann Simulation Studies
  12. Masked Audio Generation using a Single Non-Autoregressive Transformer
    2024/01/09 by Alon Ziv, Ziv, Alon, Itai Gat +15 · 14 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  13. Accelerated Sampling from Masked Diffusion Models via Entropy Bounded Unmasking
    2025/05/30 by Heli Ben-Hamu, Ben-Hamu, Heli, Itai Gat +7 · 38 citations
    Mathematics · Physics and Astronomy · #Statistical Methods and Inference #Model Reduction and Neural Networks #Numerical methods in inverse problems
  14. Perceptual Score: What Data Modalities Does Your Model Perceive?
    2021/10/27 by Gat, Itai, Schwartz, Idan, Schwing, Alexander · 7 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimedia (cs.MM)
  15. Speech Emotion Recognition using Self-Supervised Features
    2022/02/07 by Morais, Edmilson, Hoory, Ron, Zhu, Weizhong +3 · 7 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  16. Generator Matching: Generative modeling with arbitrary Markov processes
    2024/10/27 by Peter Holderrieth, Holderrieth, Peter, Marton Havasi +15 · 16 citations
    Engineering · #Electric Power System Optimization
  17. Set Block Decoding is a Language Model Inference Accelerator
    2025/09/04 by Gat, Itai, Ben-Hamu, Heli, Havasi, Marton +6 · 1 voice · 7 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  18. Edit Flows: Flow Matching with Edit Operations
    2025/06/10 by Havasi, Marton, Karrer, Brian, Gat, Itai +1 · 18 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  19. Augmentation Invariant Discrete Representation for Generative Spoken Language Modeling
    2022/09/30 by Gat, Itai, Kreuk, Felix, Nguyen, Tu Anh +5 · 4 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #electronic engineering #information engineering
  20. Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
    2024/10/11 by Buu Phan, Brandon Amos, Phan, Buu +9 · 1 voice · 5 citations
    #cs.CL #cs.LG
  21. AudioToken: Adaptation of Text-Conditioned Diffusion Models for Audio-to-Image Generation
    2023/05/22 by Guy Yariv, Yariv, Guy, Itai Gat +7 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  22. Layer Collaboration in the Forward-Forward Algorithm
    2023/05/21 by Guy Lorberbom, Itai Gat, Lorberbom, Guy +7 · 3 citations
    Computer Science · #Advanced Neural Network Applications #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE)
  23. Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music Generation
    2024/06/16 by Or Tal, Alon Ziv, Tal, Or +7 · 4 citations
    Computer Science · Engineering · #Music and Audio Processing #Music Technology and Sound Studies #Human Motion and Animation
  24. Speaker Normalization for Self-supervised Speech Emotion Recognition
    2022/02/02 by Gat, Itai, Aronowitz, Hagai, Zhu, Weizhong +2 · 2 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. Transition Matching: Scalable and Flexible Generative Modeling
    2025/06/30 by Neta Shaul, Uriel Singer, Shaul, Neta +5 · 1 voice · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.AI #cs.LG
  26. Removing Bias in Multi-modal Classifiers: Regularization by Maximizing Functional Entropies
    2020/10/21 by Gat, Itai, Schwartz, Idan, Schwing, Alexander +1 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)