vix.ing · top · new · best · stats · spec

Ding, Jinru

  1. MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models
    2024/06/24 by Mianxin Liu, Jinru Ding, Liu, Mianxin +35 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Healthcare
  2. LLM-Mini-CEX: Automatic Evaluation of Large Language Model for Diagnostic Conversation
    2023/08/15 by Shi, Xiaoming, Xu, Jie, Ding, Jinru +9 · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  3. MedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine
    2023/05/12 by Jie Xu, Lu Lu, Xu, Jie +23 · 1 citation
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Healthcare #Social Media in Health Education