vix.ing · top · new · best · stats · spec

Rotem Dror

  1. State of What Art? A Call for Multi-Prompt LLM Evaluation
    2023/12/31 by Moran Mizrahi, Mizrahi, Moran, Guy Kaplan +9 · 48 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling
  2. The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
    2025/01/19 by Nitay Calderon, Roi Reichart, Calderon, Nitay +3 · 1 voice · 19 citations
    #cs.CL #cs.AI #cs.HC
  3. A Statistical Analysis of Summarization Evaluation Metrics using Resampling Methods
    2021/03/31 by Daniel Deutsch, Deutsch, Daniel, Rotem Dror +3 · 6 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Advanced Text Analysis Techniques
  4. On the Limitations of Reference-Free Evaluations of Generated Text
    2022/10/22 by Daniel Deutsch, Rotem Dror, Deutsch, Daniel +3 · 5 citations
    Computer Science · Decision Sciences · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Scientific Computing and Data Management #Software Engineering Research #Topic Modeling
  5. The Eval4NLP 2023 Shared Task on Prompting Large Language Models as Explainable Metrics
    2023/10/30 by Christoph Leiter, Leiter, Christoph, Juri Opitz +9 · 1 citation
    Computer Science · Social Sciences · #Topic Modeling #Natural Language Processing Techniques #Computational and Text Analysis Methods