vid analysis
updated
prithivMLmods/Qwen3-VL-8B-Abliterated-Caption-it
Image-Text-to-Text
• 9B • Updated • 176
• 37
mradermacher/Qwen3-VL-8B-NSFW-Caption-V4.5-GGUF
8B • Updated • 13.2k
• 87
prithivMLmods/Qwen3-VisionCaption-2B
Image-Text-to-Text
• 2B • Updated • 41
• 6
msrcam/Qwen3-VL-2B-Instruct-heretic
ghost-actual/Qwen3.5-4B-Claude-Opus-4.6-Distilled-heretic
Text Generation
• 5B • Updated • 55
• 4
Image-Text-to-Text
• 3B • Updated • 38
• 5
Image-Text-to-Text
• 5B • Updated • 7.82M
• • 1.01k
bobber/routangseng-qwen35-0.8b-abliterated-onnx
Image-Text-to-Text
• Updated • 21
bobber/routangseng-0.8b-hottake-onnx
Image-Text-to-Text
• Updated • 17
Caplin43/multimodal-vision-language-mini
Image-to-Text
• Updated • 18
Ytgetahun/visual-narrator-llm
Image-to-Text
• Updated
Ytgetahun/visual-narrator-vlm
0.2B • Updated • 9
allenai/MolmoPoint-Vid-4B
Video-Text-to-Text
• 5B • Updated • 606
• 13
TencentARC/ARC-Qwen-Video-7B-Narrator
Video-Text-to-Text
• 9B • Updated • 69
• 12
Video-Text-to-Text
• Updated • 41
• 32
Image-Text-to-Text
• 5B • Updated • 52.4k
• 54
DAMO-NLP-SG/VideoLLaMA3-2B
Video-Text-to-Text
• 2B • Updated • 3.9k
• 21
u94fmn391j/SAVANT-scene-description-lora
Image-to-Text
• Updated • 8
VINAY-UMRETHE/SigMamba-V1-Large
Video Classification
• 0.9B • Updated • 56
• 6
qoranet/QORA-Vision-Video
Video Classification
• Updated • 22
sumit7488/TimesFormer_Baseline
Video Classification
• 0.1B • Updated • 13
StreamFormer/streamformer-timesformer
Video Classification
• 0.1B • Updated • 39
• 5
facebook/vjepa2-vitg-fpc64-256
Video Classification
• 1B • Updated • 87k
• 57
Image-Text-to-Text
• 3B • Updated • 4.39k
• 203
Video-Text-to-Text
• Updated • 2
BidirLM/BidirLM-Omni-2.5B-Embedding
Sentence Similarity
• 2B • Updated • 1.51k
• 50
Image-Text-to-Text
• Updated • 104
• 14
Video-Text-to-Text
• 0.9B • Updated • 33
• 1
Video-Text-to-Text
• 2B • Updated • 2.95k
• 599
LongVie 2: Multimodal Controllable Ultra-Long Video World Model
Paper
• 2512.13604
• Published • 76
Image-Text-to-Text
• 4B • Updated • 110k
• 3.09k
prithivMLmods/Gemma4-BLIP3o-Captioner-5B
Image-Text-to-Text
• 5B • Updated • 1.19k
• 3
lewiswatson/Frame2KG-LFM-2.5-450m-JSON
Image-Text-to-Text
• 0.4B • Updated • 221
69.7M • Updated • 16.7k
• 2
prithivMLmods/jpt-4b-GGUF
Image-Text-to-Text
• 4B • Updated • 806
• 2