Leshem Choshen
- TIES-Merging: Resolving Interference When Merging Models
2023/06/02 by Prateek Yadav, Derek Tam, Yadav, Prateek +7 · 129 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #Topic Modeling #Multimodal Machine Learning Applications
- Knowledge is a Region in Weight Space for Fine-tuned Language Models
2023/02/09 by Almog Gueta, Gueta, Almog, Elad Venezian +9 · 3 voices · 3 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Explainable Artificial Intelligence (XAI)
- tinyBenchmarks: evaluating LLMs with fewer examples
2024/02/22 by Felipe Maia Polo, Polo, Felipe Maia, Lucas Weber +9 · 44 citations
Computer Science · #Natural Language Processing Techniques #Mathematics, Computing, and Information Processing
- Model merging with SVD to tie the Knots
2024/10/25 by George Stoica, Stoica, George, Pratik Ramesh +7 · 1 voice · 23 citations
Computer Science · #Image Processing and 3D Reconstruction
- Efficient Benchmarking of Language Models
2023/08/22 by Yotam Perlitz, Perlitz, Yotam, Elron Bandel +16 · 2 voices · 9 citations
Computer Science · #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling #cs.AI #cs.CL #cs.CV #cs.LG
- Asymmetry in Low-Rank Adapters of Foundation Models
2024/02/26 by Jiacheng Zhu, Kristjan Greenewald, Zhu, Jiacheng +15 · 15 citations
Engineering · Environmental Science · #FOS: Computer and information sciences #Geotechnical and Geomechanical Engineering #Landslides and related hazards #Machine Learning (cs.LG)
- Fusing finetuned models for better pretraining
2022/04/06 by Leshem Choshen, Choshen, Leshem, Elad Venezian +5 · 9 citations
Computer Science · #Machine Learning and Algorithms #Machine Learning and Data Classification #Domain Adaptation and Few-Shot Learning
- Elements of World Knowledge (EWoK): A Cognition-Inspired Framework for Evaluating Basic World Knowledge in Language Models
2024/05/15 by Anna A. Ivanova, Ivanova, Anna A., Aalok Sathe +37 · 14 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
- Efficient multi-prompt evaluation of LLMs
2024/05/27 by Felipe Maia Polo, Ronald Xu, Polo, Felipe Maia +15 · 12 citations
Computer Science · #Algorithms and Data Compression #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Network Packet Processing and Optimization #VLSI and Analog Circuit Testing
- BabyLM Turns 3: Call for papers for the 2025 BabyLM workshop
2025/02/15 by Lucas Charpentier, Leshem Choshen, Charpentier, Lucas +25 · 1 voice · 13 citations
#cs.CL
- Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
2024/04/29 by Andreas Waldis, Yotam Perlitz, Waldis, Andreas +7 · 9 citations
Computer Science · #Natural Language Processing Techniques
- Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
2025/07/22 by Mehul Damani, Damani, Mehul, Isha Puri +11 · 1 voice · 24 citations
Computer Science · #Intelligent Tutoring Systems and Adaptive Learning
- DisentQA: Disentangling Parametric and Contextual Knowledge with Counterfactual Question Answering
2022/11/10 by Ella Neeman, Roee Aharoni, Neeman, Ella +9 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- A Survey on Model MoErging: Recycling and Routing Among Specialized Experts for Collaborative Learning
2024/08/13 by Prateek Yadav, Colin Raffel, Yadav, Prateek +14 · 9 citations
Psychology · #Innovative Teaching and Learning Methods
- Label-Efficient Model Selection for Text Generation
2024/02/12 by Shir Ashury-Tahan, Ashury-Tahan, Shir, Ariel Gera +9 · 1 voice · 4 citations
#cs.CL #cs.LG
- Q2: Evaluating Factual Consistency in Knowledge-Grounded Dialogues\n via Question Generation and Question Answering
2021/04/16 by Or Honovich, Honovich, Or, Leshem Choshen +9 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Do These LLM Benchmarks Agree? Fixing Benchmark Evaluation with BenchBench
2024/07/18 by Yotam Perlitz, Ariel Gera, Perlitz, Yotam +13 · 5 citations
Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
2024/12/09 by Felipe Maia Polo, Polo, Felipe Maia, Seamus Somerstep +7 · 1 voice · 5 citations
Computer Science · #Online Learning and Analytics
- Unitxt: Flexible, Shareable and Reusable Data Preparation and Evaluation for Generative AI
2024/01/25 by Elron Bandel, Yotam Perlitz, Bandel, Elron +21 · 3 citations
Computer Science · #Advanced Text Analysis Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
- Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
2024/08/20 by Maxim Ifergan, Leshem Choshen, Ifergan, Maxim +7 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Semantic Web and Ontologies
- Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
2024/06/17 by Rickard Brüel‐Gabrielsson, Jiacheng Zhu, Brüel-Gabrielsson, Rickard +11 · 3 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Distributed #FOS: Computer and information sciences #IoT Networks and Protocols #IoT and Edge/Fog Computing #Machine Learning (cs.LG) #Parallel #Underwater Vehicles and Communication Systems #and Cluster Computing (cs.DC)
- A Hitchhiker's Guide to Scaling Law Estimation
2024/10/15 by Leshem Choshen, Choshen, Leshem, Zhang, Yang +2 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Natural Language Processing Techniques #Topic Modeling
- Data Contamination Report from the 2024 CONDA Shared Task
2024/07/31 by Oscar Sainz, Sainz, Oscar, Iker García-Ferrero +53 · 3 citations
Decision Sciences · Computer Science · #Data Quality and Management #Advanced Data Storage Technologies
- The Future of Open Human Feedback
2024/08/15 by Shachar Don-Yehiya, Don-Yehiya, Shachar, Ben Burtenshaw +37 · 4 citations
Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human-Automation Interaction and Safety #Human-Computer Interaction (cs.HC)
- Pretraining Language Models for Diachronic Linguistic Change Discovery
2025/04/07 by Elisabeth Fittschen, Fittschen, Elisabeth, Sabrina Li +7 · 3 voices · 2 citations
#cs.CL
- ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
2023/11/22 by Prateek Yadav, Yadav, Prateek, Leshem Choshen +5 · 2 citations
Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- Automated Discovery Has No Universally Superior Harness
2026/07/20 by Akshat Gupta, Jermaine Lei, Alexander Lu +2 · 2 voices
#cs.CL #cs.AI
- Do LLMs Benefit From Their Own Words?
2026/02/27 by Jenny Y. Huang, Leshem Choshen, Ramon Astudillo +2 · 1 voice · 3 citations
#cs.CL #cs.AI
- The ShareLM Collection and Plugin: Contributing Human-Model Chats for the Benefit of the Community
2024/08/15 by Shachar Don-Yehiya, Leshem Choshen, Don-Yehiya, Shachar +3 · 1 citation
Decision Sciences · #Personal Information Management and User Behavior
- CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data
2026/01/25 by Pedro Ortiz Suarez, Laurie Burchell, Catherine Arnett +94 · 2 voices · 1 citation
#cs.CL
- The Mighty ToRR: A Benchmark for Table Reasoning and Robustness
2025/02/26 by Shir Ashury-Tahan, Ashury-Tahan, Shir, Yifan Mai +20 · 1 voice · 2 citations
Computer Science · Decision Sciences · #Computation and Language (cs.CL) #Data Quality and Management #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning and Data Classification #cs.CL
- How Safe is Your Safety Metric? Automatic Concatenation Tests for Metric Reliability
2024/08/22 by Ora Nova Fandina, Fandina, Ora Nova, Leshem Choshen +9 · 1 citation
Computer Science · Engineering · Decision Sciences · #Software Reliability and Analysis Research #Fault Detection and Control Systems #Risk and Safety Analysis