Jinru Ding
- MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models
2024/06/24 by Mianxin Liu, Jinru Ding, Liu, Mianxin +35 · 9 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Healthcare
- MedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine
2023/05/12 by Jie Xu, Xu, Jie, Lu Lu +23 · 1 citation
Computer Science · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Healthcare #Social Media in Health Education