Model weights of paper "Let ViT Speak: Generative Language-Image Pre-training"
Yan Fang
YanFang
AI & ML interests
Computer Vision, Incremental Learning, semi-supervised learning
Recent Activity
upvoted a paper 22 days ago
Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States liked a model about 1 month ago
Qwen/Qwen3.8-Flash-NextOrganizations
None yet