Zetta ΞΆ: An Efficient Closed-Loop Embodied Harness for Self-Evolving Physical Intelligence Paper β’ 2608.16590 β’ Published Aug 17 β’ 152
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling Paper β’ 2511.20785 β’ Published Nov 25, 2025 β’ 188
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook π 3.31k The secrets to building world-class LLMs
view article Article SmolVLA: Efficient Vision-Language-Action Model trained on Lerobot Community Data +7 danaaubakirova, andito, merve, ariG23498, fracapuano, loubnabnl, pcuenq, mshukor, cadene β’ Jun 3, 2025 β’ 376
Running on CPU Upgrade 14.1k Open LLM Leaderboard π 14.1k Track, rank and evaluate open LLMs and chatbots
Qwen/Qwen2.5-VL-7B-Instruct Image-Text-to-Text β’ 8B β’ Updated Apr 6, 2025 β’ 6.71M β’ β’ 1.72k
Qwen2.5-VL Collection Vision-language model series based on Qwen2.5 β’ 10 items β’ Updated Mar 2 β’ 571
meta-llama/Meta-Llama-3-8B-Instruct Text Generation β’ 8B β’ Updated Jun 18, 2025 β’ 1.18M β’ β’ 5.1k
view article Article seemore: Implement a Vision Language Model from Scratch AviSoori1x β’ Jun 23, 2024 β’ 112