ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models Paper • 2609.18487 • Published 18 days ago • 48
Marionette: Predicting World States, Rendering Geometry, Painting Appearance Paper • 2608.14530 • Published Aug 14 • 35
Semantic Browsing: Controllable Diversity for Image Generation Paper • 2606.23679 • Published Jun 22 • 20
SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers Paper • 2605.22668 • Published May 21 • 43
LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation Paper • 2605.18739 • Published May 18 • 117
Prox-E: Fine-Grained 3D Shape Editing via Primitive-Based Abstractions Paper • 2604.23774 • Published Apr 29 • 18
Video Analysis and Generation via a Semantic Progress Function Paper • 2604.22554 • Published Apr 24 • 59
ParetoSlider: Diffusion Models Post-Training for Continuous Reward Control Paper • 2604.20816 • Published Apr 22 • 15
Think in Strokes, Not Pixels: Process-Driven Image Generation via Interleaved Reasoning Paper • 2604.04746 • Published Apr 8 • 72
LTX-2.3 Collection LTX-2.3 base models, quantized models and accompanying LoRAs and IC-LoRAs • 10 items • Updated 24 days ago • 72
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers Paper • 2603.28762 • Published Mar 30 • 22
AVControl: Efficient Framework for Training Audio-Visual Controls Paper • 2603.24793 • Published Mar 25 • 30
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework Paper • 2512.03041 • Published Dec 2, 2025 • 65
SemanticMoments: Training-Free Motion Similarity via Third Moment Features Paper • 2602.09146 • Published Feb 9 • 22
view article Article Reachy Mini - The Open-Source Robot for Today's and Tomorrow's AI Builders thomwolf, matthieu-lapeyre • Jul 9, 2025 • 816
view article Article Introducing Daggr: Chain apps programmatically, inspect visually +3 merve, ysharma, abidlabs, hysts, pcuenq • Jan 29 • 107