MultiBLiMP v2 - Language Overview

Subject-Verb agreement prediction through decision trees. Click a language row (or the Results button) to open its create_pairs results below. The small chevron after N Keep (or the Diagnostics button) is separate: it opens just the bucket breakdown, inline, right where those items dropped out of the pipeline.
Overview

Languages shown in green have categorical agreement throughout — all samples share the label Yes.

The following languages were omitted for having too few or uninformative agreement labels (label distribution shown): Abaza (69), Afrikaans (1805), Akuntsu (139), Albanian (193), Alemannic (884), Armenian (687), Aromanian (35), Assamese (59), Azerbaijani (83), Bambara (1763), Basque (2298), Bavarian (816), Bengali (17), Bokota (237), Bororo (13328), Brahui (65), Bulgarian (6490), Buryat (520), Cantonese (598), Cappadocian (231), Catalan (17593), Cebuano (127), Central Kurdish (45), Chinese (15147), Chintang (958), Chukchi (211), Classical Armenian (3844), Classical Chinese (39438), Coptic (1140), Croatian (4157), Czech (10880), Danish (5776), Dutch (22261), English (19135), Erzya (1462), Esperanto (185), Estonian (24807), Faroese (386), Finnish (18975), French (14302), Galician (1935), Georgian (3719), German (149531), Gheg (516), Gorontalo (1), Gothic (3528), Guarani (1), Gujarati (107), Gwichin (27), Haitian Creole (6403), Highland Puebla Nahuatl (657), Hittite (77), Hungarian (1955), Ika (170), Indonesian (7251), Irish (5407), Japanese (6473), Javanese (784), Kaapor (72), Karelian (177), Karo (256), Kazakh (673), Khoekhoe (2205), Kiche (533), Komi Permyak (108), Komi Zyrian (601), Korean (23796), Kyrgyz (2045), Latgalian (16), Latvian (14174), Ligurian (231), Livvi (106), Low Saxon (1549), Luxembourgish (20), Macedonian (55), Madi (15), Makurap (10), Malayalam (126), Maltese (1626), Manx (2384), Mbya Guarani (74), Middle French (5989), Moksha (306), Munduruku (62), Naga (355), Naija (9634), Neapolitan (8), Nenets (33), Nepali (31), Nheengatu (1605), North Sami (2440), Northern Kurdish (589), Northwest Gbaya (367), Norwegian Bokmaal (17120), Norwegian Nynorsk (unk: 16335; No: 1), Occitan (921), Old English (19), Old French (15398), Old Georgian (177), Old Irish (10), Old Occitan (1653), Old Turkish (2), Ottoman Turkish (1177), Paumari (23), Persian (24045), Pesh (83), Phrygian (115), Pomak (646), Romanian (9577), Ruuli (272), Scottish Gaelic (5768), Serbian (2126), Shanghainese (670), Sicilian (548), Sinhala (82), Skolt Sami (212), Slovenian (4016), South Levantine Arabic (23), Southern Kurdish (110), Swedish (15079), Tagalog (107), Tatar (83), Teko (165), Telugu (922), Thai (5729), Tswana (20), Tupinamba (72), Turkish (30741), Umbrian (27), Upper Sorbian (unk: 392; No: 1), Uyghur (2720), Uzbek (971), Veps (82), Vietnamese (3326), Warlpiri (56), Welsh (1702), Western Armenian (2141), Western Sierra Puebla Nahuatl (1061), Wolof (4238), Xavante (158), Xibe (561), Yakut (167), Yiddish (2187), Yoruba (721), Yupik (163), Zaar (336).

Keep threshold: entropy < 0.12
Coverage 0%100%
Item loss 0%100%
Language
Distribution
Base Entropy
Reduced Entropy
Δ Entropy
DT Acc%
N RAW
N KEEP
N PAIRS