IrokoBench Collection a human-translated benchmark dataset for 16 African languages covering three tasks: NLI, MMLU and MGSM • 6 items • Updated May 31, 2024 • 21
Pebble Collection LLM360's medium and smaller sized model series, ideal for research and efficient deployment • 5 items • Updated 1 day ago • 2
MegaMath Collection MegaMath, the largest open math pre-training dataset curated from diverse, math-focused sources, with over 300B tokens. • 4 items • Updated 1 day ago • 5
K2 Horizon Collection K2 Horizon models, datasets, and supporting resources • 13 items • Updated 1 day ago • 140
view article Article Meet North Micro Vision: A 2.4B Native-Resolution Vision-Language Model CohereLabs • Aug 12 • 42
Foundation-Sec-8B Collection Foundation-Sec-8B models and quantizations. • 8 items • Updated Jan 28 • 15
Inkling Collection Inkling is a versatile, customizable model that reasons over text, images, audio, with variable and efficient thinking effort. • 4 items • Updated Jul 27 • 72
AfroBench Collection Large Scale Benchmark of Large Language Models on African Languages • 21 items • Updated Jul 1 • 6
AfrIFact Collection a multilingual information retrieval, evidence retrieval and fact checking benchmark covering healthcare, culturally grounded content • 4 items • Updated May 29 • 1
AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages Paper • 2601.06395 • Published Jan 10 • 5
view article Article OlmoEarth v1.1: A more efficient family of Earth observation models allenai • May 19 • 27
TIPSv2 Collection TIPSv2 foundational vision-language models. Webpage: https://gdm-tipsv2.github.io/ • 9 items • Updated Jul 21 • 46