vix.ing · top · new · best · stats · spec

Choshen, Leshem

  1. TIES-Merging: Resolving Interference When Merging Models
    2023/06/02 by Prateek Yadav, Yadav, Prateek, Derek Tam +7 · 130 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Topic Modeling #Multimodal Machine Learning Applications
  2. Knowledge is a Region in Weight Space for Fine-tuned Language Models
    2023/02/09 by Almog Gueta, Gueta, Almog, Elad Venezian +9 · 3 voices · 3 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Explainable Artificial Intelligence (XAI)
  3. tinyBenchmarks: evaluating LLMs with fewer examples
    2024/02/22 by Felipe Maia Polo, Lucas Weber, Polo, Felipe Maia +9 · 45 citations
    Computer Science · #Natural Language Processing Techniques #Mathematics, Computing, and Information Processing
  4. Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
    2024/12/04 by Singh, Shivalika, Romanou, Angelika, Fourrier, Clémentine +21 · 45 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  5. Model merging with SVD to tie the Knots
    2024/10/25 by George Stoica, Pratik Ramesh, Stoica, George +7 · 1 voice · 23 citations
    Computer Science · #Image Processing and 3D Reconstruction
  6. Efficient Benchmarking of Language Models
    2023/08/22 by Yotam Perlitz, Perlitz, Yotam, Elron Bandel +16 · 2 voices · 9 citations
    Computer Science · #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling #cs.AI #cs.CL #cs.CV #cs.LG
  7. Asymmetry in Low-Rank Adapters of Foundation Models
    2024/02/26 by Jiacheng Zhu, Kristjan Greenewald, Zhu, Jiacheng +15 · 15 citations
    Engineering · Environmental Science · #FOS: Computer and information sciences #Geotechnical and Geomechanical Engineering #Landslides and related hazards #Machine Learning (cs.LG)
  8. Fusing finetuned models for better pretraining
    2022/04/06 by Leshem Choshen, Choshen, Leshem, Elad Venezian +5 · 9 citations
    Computer Science · #Machine Learning and Algorithms #Machine Learning and Data Classification #Domain Adaptation and Few-Shot Learning
  9. Elements of World Knowledge (EWoK): A Cognition-Inspired Framework for Evaluating Basic World Knowledge in Language Models
    2024/05/15 by Anna A. Ivanova, Aalok Sathe, Ivanova, Anna A. +37 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
  10. Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
    2023/01/27 by Warstadt, Alex, Choshen, Leshem, Mueller, Aaron +3 · 8 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  11. Efficient multi-prompt evaluation of LLMs
    2024/05/27 by Felipe Maia Polo, Polo, Felipe Maia, Ronald Xu +15 · 12 citations
    Computer Science · #Algorithms and Data Compression #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Network Packet Processing and Optimization #VLSI and Analog Circuit Testing
  12. NumeroLogic: Number Encoding for Enhanced LLMs' Numerical Reasoning
    2024/03/30 by Schwartz, Eli, Choshen, Leshem, Shtok, Joseph +3 · 10 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  13. BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
    2025/02/15 by Lucas Charpentier, Charpentier, Lucas, Leshem Choshen +25 · 1 voice · 13 citations
    #cs.CL
  14. DisentQA: Disentangling Parametric and Contextual Knowledge with Counterfactual Question Answering
    2022/11/10 by Ella Neeman, Roee Aharoni, Neeman, Ella +9 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  15. Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
    2024/12/06 by Hu, Michael Y., Mueller, Aaron, Ross, Candace +7 · 13 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  16. On the Weaknesses of Reinforcement Learning for Neural Machine Translation
    2019/07/03 by Choshen, Leshem, Fox, Lior, Aizenbud, Zohar +1 · 4 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  17. Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
    2024/04/29 by Andreas Waldis, Yotam Perlitz, Waldis, Andreas +7 · 9 citations
    Computer Science · #Natural Language Processing Techniques
  18. Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
    2025/07/22 by Mehul Damani, Isha Puri, Damani, Mehul +11 · 1 voice · 24 citations
    Computer Science · #Intelligent Tutoring Systems and Adaptive Learning
  19. A Survey on Model MoErging: Recycling and Routing Among Specialized Experts for Collaborative Learning
    2024/08/13 by Prateek Yadav, Yadav, Prateek, Colin Raffel +14 · 9 citations
    Psychology · #Innovative Teaching and Learning Methods
  20. Are You Convinced? Choosing the More Convincing Evidence with a Siamese Network
    2019/07/21 by Gleize, Martin, Shnarch, Eyal, Choshen, Leshem +4 · 3 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  21. Label-Efficient Model Selection for Text Generation
    2024/02/12 by Shir Ashury-Tahan, Ariel Gera, Ashury-Tahan, Shir +9 · 1 voice · 4 citations
    #cs.CL #cs.LG
  22. Q2: Evaluating Factual Consistency in Knowledge-Grounded Dialogues\n via Question Generation and Question Answering
    2021/04/16 by Or Honovich, Leshem Choshen, Honovich, Or +9 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  23. Naturally Occurring Feedback is Common, Extractable and Useful
    2024/07/15 by Don-Yehiya, Shachar, Choshen, Leshem, Abend, Omri · 6 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  24. The Grammar-Learning Trajectories of Neural Language Models
    2021/09/13 by Choshen, Leshem, Hacohen, Guy, Weinshall, Daphna +1 · 3 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
    2024/12/09 by Felipe Maia Polo, Polo, Felipe Maia, Seamus Somerstep +7 · 1 voice · 6 citations
    Computer Science · #Online Learning and Analytics
  26. [Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
    2024/04/09 by Choshen, Leshem, Cotterell, Ryan, Hu, Michael Y. +7 · 5 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  27. Automatic Metric Validation for Grammatical Error Correction
    2018/04/30 by Choshen, Leshem, Abend, Omri · 2 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  28. Do These LLM Benchmarks Agree? Fixing Benchmark Evaluation with BenchBench
    2024/07/18 by Yotam Perlitz, Perlitz, Yotam, Ariel Gera +13 · 5 citations
    Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences
  29. Unitxt: Flexible, Shareable and Reusable Data Preparation and Evaluation for Generative AI
    2024/01/25 by Elron Bandel, Bandel, Elron, Yotam Perlitz +21 · 3 citations
    Computer Science · #Advanced Text Analysis Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
  30. Lossless and Near-Lossless Compression for Foundation Models
    2024/04/05 by Hershcovitch, Moshik, Choshen, Leshem, Wood, Andrew +4 · 3 citations
    #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
  31. DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation
    2025/03/03 by Habba, Eliya, Arviv, Ofir, Itzhak, Itay +5 · 5 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  32. Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
    2024/08/20 by Maxim Ifergan, Leshem Choshen, Ifergan, Maxim +7 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Semantic Web and Ontologies
  33. Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
    2024/06/17 by Rickard Brüel‐Gabrielsson, Brüel-Gabrielsson, Rickard, Jiacheng Zhu +11 · 3 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Distributed #FOS: Computer and information sciences #IoT Networks and Protocols #IoT and Edge/Fog Computing #Machine Learning (cs.LG) #Parallel #Underwater Vehicles and Communication Systems #and Cluster Computing (cs.DC)
  34. A Hitchhiker's Guide to Scaling Law Estimation
    2024/10/15 by Leshem Choshen, Choshen, Leshem, Jacob Andreas +2 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Natural Language Processing Techniques #Topic Modeling
  35. Data Contamination Report from the 2024 CONDA Shared Task
    2024/07/31 by Oscar Sainz, Sainz, Oscar, Iker García-Ferrero +53 · 3 citations
    Decision Sciences · Computer Science · #Data Quality and Management #Advanced Data Storage Technologies
  36. The Future of Open Human Feedback
    2024/08/15 by Shachar Don-Yehiya, Ben Burtenshaw, Don-Yehiya, Shachar +37 · 4 citations
    Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human-Automation Interaction and Safety #Human-Computer Interaction (cs.HC)
  37. ZipNN: Lossless Compression for AI Models
    2024/11/07 by Hershcovitch, Moshik, Wood, Andrew, Choshen, Leshem +7 · 2 voices
    #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
  38. DORA The Explorer: Directed Outreaching Reinforcement Action-Selection
    2018/04/11 by Choshen, Leshem, Fox, Lior, Loewenstein, Yonatan · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  39. Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
    2023/11/13 by Zaman, Kerem, Choshen, Leshem, Srivastava, Shashank · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  40. Pretraining Language Models for Diachronic Linguistic Change Discovery
    2025/04/07 by Elisabeth Fittschen, Fittschen, Elisabeth, Sabrina Li +7 · 3 voices · 2 citations
    #cs.CL
  41. ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
    2023/11/22 by Prateek Yadav, Yadav, Prateek, Leshem Choshen +5 · 2 citations
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
  42. Learning to combine Grammatical Error Corrections
    2019/06/10 by Kantor, Yoav, Katz, Yoav, Choshen, Leshem +5 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  43. Unsupervised Expressive Rules Provide Explainability and Assist Human Experts Grasping New Domains
    2020/10/19 by Shnarch, Eyal, Choshen, Leshem, Moshkowich, Guy +2 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG)
  44. Classifying Syntactic Errors in Learner Language
    2020/10/21 by Choshen, Leshem, Nikolaev, Dmitry, Berzak, Yevgeni +1 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  45. ComSum: Commit Messages Summarization and Meaning Preservation
    2021/08/23 by Choshen, Leshem, Amit, Idan · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE)
  46. Cluster & Tune: Boost Cold Start Performance in Text Classification
    2022/03/20 by Shnarch, Eyal, Gera, Ariel, Halfon, Alon +4 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  47. Genie: Achieving Human Parity in Content-Grounded Datasets Generation
    2024/01/25 by Asaf Yehudai, Yehudai, Asaf, Boaz Carmeli +13 · 1 voice · 1 citation
    #cs.CL #cs.AI #cs.LG
  48. The ShareLM Collection and Plugin: Contributing Human-Model Chats for the Benefit of the Community
    2024/08/15 by Shachar Don-Yehiya, Leshem Choshen, Don-Yehiya, Shachar +3 · 1 citation
    Decision Sciences · #Personal Information Management and User Behavior
  49. NeurIPS 2023 LLM Efficiency Fine-tuning Competition
    2025/03/13 by Saroufim, Mark, Perlitz, Yotam, Choshen, Leshem +11 · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  50. The Mighty ToRR: A Benchmark for Table Reasoning and Robustness
    2025/02/26 by Shir Ashury-Tahan, Ashury-Tahan, Shir, Yifan Mai +20 · 1 voice · 2 citations
    Computer Science · Decision Sciences · #Computation and Language (cs.CL) #Data Quality and Management #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning and Data Classification #cs.CL
  51. How Safe is Your Safety Metric? Automatic Concatenation Tests for Metric Reliability
    2024/08/22 by Ora Nova Fandina, Leshem Choshen, Fandina, Ora Nova +9 · 1 citation
    Computer Science · Engineering · Decision Sciences · #Software Reliability and Analysis Research #Fault Detection and Control Systems #Risk and Safety Analysis