TakalaWang/Discussion-Phi-4-multimodal-instruct-audio-dimp-reasoning Text Generation • 6B • Updated May 15, 2025 • 87 • 3
All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts Paper • 2609.24058 • Published 5 days ago • 42
Complex KDA: Understanding and Enhancing the Expressivity of Kimi Delta Attention Paper • 2609.24797 • Published 5 days ago • 11
nezahatkorkmaz/Turkish-medical-visual-question-answering-LLaVa-dataset Viewer • Updated Mar 19, 2025 • 316 • 147 • 10
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 9 days ago • 70
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 12 days ago • 50