Most things you do on HuggingFace, BananaMindBot can do. Fast
Mention @BananaMindBot on a model, dataset, Space discussion, paper, blog comment, or top-level post and it'll reply there.
It's powered by North Code Mini (Qwen3.8 27B, with GPT OSS 120B as fallback).
A few things it can do:
Search for models and datasets Look up users and orgs and see what they've published Read model cards, configs, dataset files, blog posts, and org profiles Answer questions about what it finds Write and run its own code in a locked-down sandbox when it needs to verify something Check things like a model's real parameter count from the safetensors headers instead of just repeating the model card Remember something for later if you explicitly ask it to Forward a message to @Banaxi-Tech Post a daily roundup of developments in the small-language-model space
It won't execute code you give it. It can read and review that code, but anything it runs is code it wrote itself.
It also can't access private data or credentials.
Mention it somewhere.
It's going to also find this post!
(Some parts inspired by CompactBot and @CompactAI Follow them please)
OpenRouter Leaderboard β every model, every provider, one comparable table. Price, precision, uptime, measured latency and language quality on the same axes.
Building it turned up three things.
We graded 330 models on Korean and two axes collapsed.
Honorifics β only 8.5% earn an A Knowledge of Korean institutions β 9.4% Every other axis sits above 31% Fluency hides it. A model can write clean, natural Korean and still attach an honorific to a coffee cup. Fluent and wrong at the same time is worse than obviously broken, because nobody catches it in review.
A 2023 model beats the 2026 flagships. gpt-3.5-turbo-16k scores a perfect 3.00. Korean cannot be inferred from release date, parameter count or English benchmarks β it has to be measured, per model.
Quality, value and speed are three different models. Across five axes, the same model almost never takes two columns.
425 models, latency measured on 329 on a paid API, Korean graded on 330. Three languages, three currencies, daily refresh, open API, no key.