WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting Paper • 2605.11696 • Published May 12 • 3
WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting Paper • 2605.11696 • Published May 12 • 3
General Multimodal Protein Design Enables DNA-Encoding of Chemistry Paper • 2604.05181 • Published Apr 6 • 27
XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision Paper • 2601.21688 • Published Jan 29
Revealing Subtle Phenotypes in Small Microscopy Datasets Using Latent Diffusion Models Paper • 2502.09665 • Published Feb 12, 2025
CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video Paper • 2603.04291 • Published Mar 4 • 16
CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video Paper • 2603.04291 • Published Mar 4 • 16
SemanticMoments: Training-Free Motion Similarity via Third Moment Features Paper • 2602.09146 • Published Feb 9 • 22
BFTBrain: Adaptive BFT Consensus with Reinforcement Learning Paper • 2408.06432 • Published Aug 12, 2024
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation Paper • 2502.11897 • Published Feb 17, 2025
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer Paper • 2601.16515 • Published Jan 23 • 15
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer Paper • 2601.16515 • Published Jan 23 • 15
CaricatureGS: Exaggerating 3D Gaussian Splatting Faces With Gaussian Curvature Paper • 2601.03319 • Published Jan 6 • 54
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training Paper • 2512.06864 • Published Dec 7, 2025 • 13
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training Paper • 2512.06864 • Published Dec 7, 2025 • 13
Efficiently Reconstructing Dynamic Scenes One D4RT at a Time Paper • 2512.08924 • Published Dec 9, 2025 • 25
Unified Speech-Text Pre-training for Speech Translation and Recognition Paper • 2204.05409 • Published Apr 11, 2022
data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language Paper • 2202.03555 • Published Feb 7, 2022