Johann-Peter Hartmann PRO
johannhartmann
AI & ML interests
LLMs, Local LLMs, Transformers, Image Processing, Audio Processing, E-Commerce
Recent Activity
liked a model 2 days ago
openeurollm/prelude liked a model 21 days ago
Soofi-Project/Soofi-S-Isar-Preview upvoted a collection 21 days ago
Soofi S Beta ModelsOrganizations
Document & UI Intelligence
-
xlangai/Aguvis-7B-720P
8B • Updated • 31 • 9 -
Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
Paper • 2412.04454 • Published • 70 -
SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Paper • 2401.10935 • Published • 6 -
cckevinn/SeeClick
Text Generation • 10B • Updated • 373 • 18
Medical MultiModal
Multimodal models that have been trained on medical datasets.
Music
Computer Use Models
Document & UI Intelligence
-
xlangai/Aguvis-7B-720P
8B • Updated • 31 • 9 -
Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
Paper • 2412.04454 • Published • 70 -
SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Paper • 2401.10935 • Published • 6 -
cckevinn/SeeClick
Text Generation • 10B • Updated • 373 • 18
Multimodal Models
A collection of multimodal models for the gpu poor
Medical MultiModal
Multimodal models that have been trained on medical datasets.