Craig Swift
- GAIA: a benchmark for General AI Assistants
2023/11/21 by Grégoire Mialon, Mialon, Grégoire, Clémentine Fourrier +9 · 6 voices · 346 citations
Computer Science · Medicine · #Artificial Intelligence in Healthcare and Education #Explainable Artificial Intelligence (XAI) #Topic Modeling #cs.AI #cs.CL
- Testing Language Model Agents Safely in the Wild
2023/11/17 by Silen Naihin, Naihin, Silen, David Atkinson +13 · 16 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation #Topic Modeling