MobileWan: Closing the Quality Gap for Mobile Video Diffusion Paper • 2607.06173 • Published 20 days ago • 3
ComfyUI Abliterated Text Encoders Collection Abliterated text encoders for ComfyUI image and video generation. Includes GGUF and Safetensors. • 7 items • Updated 11 days ago • 4
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization Paper • 2601.01554 • Published Jan 4 • 65
MOSS Transcribe Collection A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription. • 4 items • Updated 16 days ago • 13
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Paper • 2607.14187 • Published 12 days ago • 31
view article Article NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval nvidia • 10 days ago • 57
Laguna S 2.1 Collection Our most capable model to date, designed for long-horizon work. • 12 items • Updated 4 days ago • 27
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Paper • 2607.09581 • Published 17 days ago • 6
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation Paper • 2606.02441 • Published Jun 1 • 2
ZUNA Collection Brain-Computer Interface models for reconstruction, interpolation, and downstream tasks • 2 items • Updated 13 days ago • 4
Qwen 3.6 - Reg/Uncensored 9b, 12b, 21b, 27b, 40B Collection Fine tuned Qwen 3.6 models, including source and GGUF from 9B and up. 9B,12B, 21B and 40B are custom built by me. Tuning via Unsloth on local hardware • 12 items • Updated 3 days ago • 15
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding Paper • 2503.06796 • Published Mar 9, 2025 • 2
GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors Paper • 2606.05160 • Published Jun 3 • 9
UniVR: Thinking in Visual Space for Unified Visual Reasoning Paper • 2607.12800 • Published 13 days ago • 32
Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes Paper • 2607.13188 • Published 13 days ago • 34
MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Paper • 2607.14189 • Published 12 days ago • 34