Explore model rankings on the BananaMind text benchmark
Chance-normalized evaluation of compact language models.