Skip to main content
. 2024 Feb 29;10:e1893. doi: 10.7717/peerj-cs.1893

Table 5. The accuracy of GPT-4 on the GAT Arabic and English questions compared to ChatGPT.

AR and EN refer to Arabic and English respectively. The best performance is indicated in bold.

Question type EN (GPT-4) AR (GPT-4) EN (ChatGPT) AR (ChatGPT)
Reading comprehension 86.81% 74.29% 80.22% 55.71%
Analogy 73.39% 57.02% 54.03% 37.19%
Contextual error 83.52% 63.37% 68.13% 38.61%
Sentence completion 87.33% 75.47% 82.67% 35.85%
Micro-average 82.68% 67.74% 71.49% 42.74%
Macro-average 82.76% 67.54% 71.26% 41.84%