vix.ing · top · new · best · stats · spec

Abdelhak at SemEval-2024 Task 9 : Decoding Brainteasers, The Efficacy of Dedicated Models Versus ChatGPT

2024/02/24 by Abdelhak Kelious, Kelious, Abdelhak, Mounir Okirim +1
Computer Science · Medicine · Neuroscience · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #Machine Learning in Healthcare

paper · pdf · doi:10.48550/arxiv.2403.00809

openalex publication_date 2024/02/24 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

This study introduces a dedicated model aimed at solving the BRAINTEASER task 9 , a novel challenge designed to assess models lateral thinking capabilities through sentence and word puzzles. Our model demonstrates remarkable efficacy, securing Rank 1 in sentence puzzle solving during the test phase with an overall score of 0.98. Additionally, we explore the comparative performance of ChatGPT, specifically analyzing how variations in temperature settings affect its ability to engage in lateral thinking and problem-solving. Our findings indicate a notable performance disparity between the dedicated model and ChatGPT, underscoring the potential of specialized approaches in enhancing creative reasoning in AI.

Related