GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published 2 days ago • 30
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Paper • 2608.05000 • Published 3 days ago • 51
Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts Paper • 2608.00574 • Published 7 days ago • 7
LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-GGUF Image-Text-to-Text • 35B • Updated about 4 hours ago • 333k • 425
nota-ai/Solar-Open2-250B-Nota-NVFP4 Text Generation • 145B • Updated about 17 hours ago • 69.8k • 174