High schooler by day, LLM builder by night. Driven by a deep love for both Physics and AI. Currently spending my runtime building on Hugging Face, experimenting with transformer architectures, and training custom LLMs.
We're excited to release BananaMind Base Bench 1.1 A new benchmark for base language models with 350 text-completion examples across seven categories. Models are scored using continuation likelihood and receive an Overall Elo score. Initial results: BananaMind-2-Medium: 1034 BananaMind-2-Mini: 974 Supra-50M-Base: 973 Supra-1.5-50M-Base-exp: 948 BananaMind-2-Nano: 910 The official script downloads the gated dataset directly from Hugging Face. The dataset is for benchmarking only and may not be used for model training. BananaMind/BananaMind-Base-Bench-1.1