S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation Paper • 2607.15686 • Published 13 days ago • 16
OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models Paper • 2601.21639 • Published Jan 29 • 52
view article Article Qwen-Image-i2L: Training Strategies for Image-to-LoRA Generation kelseye • Dec 16, 2025 • 59