MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Paper • 2607.11562 • Published 17 days ago • 76
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 14 days ago • 141
greghavens/kimi-k3-coding-and-debugging-traces Viewer • Updated about 13 hours ago • 3.91k • 5.11k • 51
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF Text Generation • 1B • Updated 16 days ago • 304k • 308
Running on Zero Agents Featured 47 Follow the Mean (FLUX.2) 🪷 47 Training-free reference-guided generation with FLUX.2-klein
Running on Zero Agents Featured 1.16k OmniVoice 🌍 1.16k High-quality voice cloning TTS for 600+ languages
Paused 361 Gemma-4-E4B-Uncensored-HauhauCS-Aggressive-Q5_K_P 🔥 361 This Space uses ''HauhauCS/Gemma-4-E4B-Uncensored-Hauh...''.
Running on Zero Agents Featured 310 LTX 2.3 Studio 🎬 310 Generate videos from text, images, audio, or style
ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents Paper • 2604.11784 • Published Apr 13 • 143
Runtime error Agents Featured 72 NAG Wan2-1-fast 🏢 72 Demo of Normalized Attention Guidance for 4 steps Wan2.1
Running on Zero Agents Featured 1.1k InfiniteYou-FLUX 📸 1.1k Flexible Photo Recrafting While Preserving Your Identity