MultiBLiMP v2 - Language Overview

Subject-Verb agreement prediction through decision trees. Click a language row (or the Results button) to open its create_pairs results below. The small chevron after N Keep (or the Diagnostics button) is separate: it opens just the bucket breakdown, inline, right where those items dropped out of the pipeline.
Overview

Languages shown in green have categorical agreement throughout — all samples share the label Yes.

The following languages were omitted for having too few or uninformative agreement labels (label distribution shown): Afrikaans (374), Albanian (25), Armenian (1027), Assamese (8), Assyrian (1), Bambara (29), Bengali (9), Breton (165), Dutch (2949), Estonian (2631), German (18923), Gheg (203), Hausa (6), Hittite (1), Kazakh (31), Kiche (8), Komi Zyrian (1), Low Saxon (254), Malayalam (4), Middle French (963), Naija (60), Neapolitan (2), Northern Kurdish (4), Odia (4), Old English (1), Sinhala (29), Skolt Sami (19), Tamil (185), Uyghur (39), Veps (3), Western Armenian (614), Xibe (127).

Keep threshold: entropy < 0.12
Coverage 0%100%
Item loss 0%100%
Language
Distribution
Base Entropy
Reduced Entropy
Δ Entropy
DT Acc%
N RAW
N KEEP
N PAIRS