|国家预印本平台
首页|TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs

TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs

TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs

来源:Arxiv_logoArxiv
英文摘要

We introduce TurBLiMP, the first Turkish benchmark of linguistic minimal pairs, designed to evaluate the linguistic abilities of monolingual and multilingual language models (LMs). Covering 16 linguistic phenomena with 1000 minimal pairs each, TurBLiMP fills an important gap in linguistic evaluation resources for Turkish. In designing the benchmark, we give extra attention to two properties of Turkish that remain understudied in current syntactic evaluations of LMs, namely word order flexibility and subordination through morphological processes. Our experiments on a wide range of LMs and a newly collected set of human acceptability judgments reveal that even cutting-edge Large LMs still struggle with grammatical phenomena that are not challenging for humans, and may also exhibit different sensitivities to word order and morphological complexity compared to humans.

Ezgi Ba?ar、Francesca Padovani、Jaap Jumelet、Arianna Bisazza

语言学阿尔泰语系(突厥-蒙古-通古斯语系)

Ezgi Ba?ar,Francesca Padovani,Jaap Jumelet,Arianna Bisazza.TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs[EB/OL].(2025-06-16)[2025-07-16].https://arxiv.org/abs/2506.13487.点此复制

评论