steven hobbs
shobbs
·
AI & ML interests
Vision
HRL
Edge
Recent Activity
updated a collection 14 days ago
storytime updated a collection 21 days ago
bio updated a collection about 2 months ago
visionOrganizations
papers
embed RAG
small and fast
bio
video llm llava
-
NVILA: Efficient Frontier Visual Language Models
Paper • 2412.04468 • Published • 61 -
unsloth/GLM-4.1V-9B-Thinking-GGUF
Image-Text-to-Text • 9B • Updated • 7.54k • 45 -
zai-org/GLM-4.5V
Image-Text-to-Text • 108B • Updated • 41.1k • • 722 -
dots-studio/dots.vlm1.inst
Image-Text-to-Text • 672B • Updated • 200 • 84
arm
Mobile use aka smart phone actions dataset
storytime
think and learn
-
deepseek-ai/DeepSeek-R1-0528
Text Generation • 685B • Updated • 158k • • 2.46k -
unsloth/ERNIE-4.5-300B-A47B-PT-GGUF
Text Generation • 299B • Updated • 875 • 9 -
Qwen/Qwen3-235B-A22B-Instruct-2507-FP8
Text Generation • 235B • Updated • 420k • 150 -
cerebras/GLM-4.6-REAP-218B-A32B-FP8
Text Generation • 218B • Updated • 74 • 44
NSFW
vision
-
google/paligemma2-28b-pt-896
Image-Text-to-Text • 28B • Updated • 55 • 52 -
lmstudio-community/olmOCR-7B-0225-preview-GGUF
Image-Text-to-Text • 8B • Updated • 378 • 13 -
vidore/colqwen2.5-v0.2
Visual Document Retrieval • Updated • 196k • 103 -
vidore/colpali-v1.3
Visual Document Retrieval • Updated • 35.4k • 100
image art
video
Mobile use aka smart phone actions dataset
papers
storytime
embed RAG
think and learn
-
deepseek-ai/DeepSeek-R1-0528
Text Generation • 685B • Updated • 158k • • 2.46k -
unsloth/ERNIE-4.5-300B-A47B-PT-GGUF
Text Generation • 299B • Updated • 875 • 9 -
Qwen/Qwen3-235B-A22B-Instruct-2507-FP8
Text Generation • 235B • Updated • 420k • 150 -
cerebras/GLM-4.6-REAP-218B-A32B-FP8
Text Generation • 218B • Updated • 74 • 44
small and fast
NSFW
bio
vision
-
google/paligemma2-28b-pt-896
Image-Text-to-Text • 28B • Updated • 55 • 52 -
lmstudio-community/olmOCR-7B-0225-preview-GGUF
Image-Text-to-Text • 8B • Updated • 378 • 13 -
vidore/colqwen2.5-v0.2
Visual Document Retrieval • Updated • 196k • 103 -
vidore/colpali-v1.3
Visual Document Retrieval • Updated • 35.4k • 100
video llm llava
-
NVILA: Efficient Frontier Visual Language Models
Paper • 2412.04468 • Published • 61 -
unsloth/GLM-4.1V-9B-Thinking-GGUF
Image-Text-to-Text • 9B • Updated • 7.54k • 45 -
zai-org/GLM-4.5V
Image-Text-to-Text • 108B • Updated • 41.1k • • 722 -
dots-studio/dots.vlm1.inst
Image-Text-to-Text • 672B • Updated • 200 • 84
image art
arm