steven hobbs
shobbs
·
AI & ML interests
Vision
HRL
Edge
Recent Activity
updated a collection 1 day ago
vision updated a collection 1 day ago
small and fast updated a collection 3 days ago
image artOrganizations
papers
embed RAG
small and fast
bio
video llm llava
-
NVILA: Efficient Frontier Visual Language Models
Paper • 2412.04468 • Published • 62 -
unsloth/GLM-4.1V-9B-Thinking-GGUF
Image-Text-to-Text • 9B • Updated • 2.38k • 43 -
zai-org/GLM-4.5V
Image-Text-to-Text • 108B • Updated • 150k • • 719 -
dots-studio/dots.vlm1.inst
Image-Text-to-Text • 672B • Updated • 46 • 83
arm
Mobile use aka smart phone actions dataset
storytime
think and learn
-
deepseek-ai/DeepSeek-R1-0528
Text Generation • 685B • Updated • 246k • • 2.46k -
unsloth/ERNIE-4.5-300B-A47B-PT-GGUF
Text Generation • 299B • Updated • 2.25k • 9 -
Qwen/Qwen3-235B-A22B-Instruct-2507-FP8
Text Generation • 235B • Updated • 88.1k • 147 -
cerebras/GLM-4.6-REAP-218B-A32B-FP8
Text Generation • 218B • Updated • 29 • 44
NSFW
vision
-
google/paligemma2-28b-pt-896
Image-Text-to-Text • 28B • Updated • 500 • 52 -
lmstudio-community/olmOCR-7B-0225-preview-GGUF
Image-Text-to-Text • 8B • Updated • 422 • 13 -
vidore/colqwen2.5-v0.2
Visual Document Retrieval • Updated • 44.7k • 99 -
vidore/colpali-v1.3
Visual Document Retrieval • Updated • 18.3k • 99
image art
video
Mobile use aka smart phone actions dataset
papers
storytime
embed RAG
think and learn
-
deepseek-ai/DeepSeek-R1-0528
Text Generation • 685B • Updated • 246k • • 2.46k -
unsloth/ERNIE-4.5-300B-A47B-PT-GGUF
Text Generation • 299B • Updated • 2.25k • 9 -
Qwen/Qwen3-235B-A22B-Instruct-2507-FP8
Text Generation • 235B • Updated • 88.1k • 147 -
cerebras/GLM-4.6-REAP-218B-A32B-FP8
Text Generation • 218B • Updated • 29 • 44
small and fast
NSFW
bio
vision
-
google/paligemma2-28b-pt-896
Image-Text-to-Text • 28B • Updated • 500 • 52 -
lmstudio-community/olmOCR-7B-0225-preview-GGUF
Image-Text-to-Text • 8B • Updated • 422 • 13 -
vidore/colqwen2.5-v0.2
Visual Document Retrieval • Updated • 44.7k • 99 -
vidore/colpali-v1.3
Visual Document Retrieval • Updated • 18.3k • 99
video llm llava
-
NVILA: Efficient Frontier Visual Language Models
Paper • 2412.04468 • Published • 62 -
unsloth/GLM-4.1V-9B-Thinking-GGUF
Image-Text-to-Text • 9B • Updated • 2.38k • 43 -
zai-org/GLM-4.5V
Image-Text-to-Text • 108B • Updated • 150k • • 719 -
dots-studio/dots.vlm1.inst
Image-Text-to-Text • 672B • Updated • 46 • 83
image art
arm