OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 11 days ago • 149
Running on Zero Agents 836 MiniMax H3 Turbo LoRA 🎬 836 Video generation with a synchronized soundtrack
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 8 days ago • 130
Running 144 HF Viewer · Model Architecture Explorer 🟩 144 Interactive architecture graph for any HF model
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 19 days ago • 701
Running on Zero MCP Featured 2.76k Qwen-Image-Edit-2511-LoRAs-Fast 🎃 2.76k Demo of the Collection of Qwen Image Edit LoRAs
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published Jul 16 • 145
Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget Paper • 2607.13125 • Published Jul 18 • 139