Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation Paper • 2609.20744 • Published 17 days ago • 53
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 13 days ago • 72
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation Paper • 2609.08084 • Published 26 days ago • 72
Explicit Layer Modeling for Video Object Insertion and Layer Decomposition Paper • 2607.25802 • Published Jul 28 • 8
Native and Compact Structured Latents for 3D Generation Paper • 2512.14692 • Published Dec 16, 2025 • 7
AVControl: Efficient Framework for Training Audio-Visual Controls Paper • 2603.24793 • Published Mar 25 • 30
HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos Paper • 2605.17543 • Published May 19 • 13
Ovis1.6 Collection With 29B parameters, Ovis1.6-Gemma2-27B achieves exceptional performance in the OpenCompass benchmark, ranking among the top-tier open-source MLLMs. • 5 items • Updated Nov 26, 2024 • 13
State-Of-The-Art Korean-RAG LM Collection Markr AI's RAG LLM (based on Ko-Mixtral) • 5 items • Updated Feb 11, 2024 • 2