HF1BitLLM/Llama3-8B-1.58-100B-tokens Text Generation • 3B • Updated Sep 19, 2024 • 1.52k • 225
LLMs achieve adult human performance on higher-order theory of mind tasks Paper • 2405.18870 • Published May 29, 2024 • 17
lightblue/suzume-llama-3-8B-japanese Text Generation • 8B • Updated Jun 2, 2024 • 21 • • 25
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention Paper • 2404.07143 • Published Apr 10, 2024 • 111
CohereLabs/c4ai-command-r-plus-4bit Text Generation • 105B • Updated Apr 16, 2025 • 800 • 261