HTR ByteDance/Sa2VA-4B Image-Text-to-Text • 4B • Updated Sep 8, 2025 • 825 • 99 Finnish-NLP/Ahma-2-4B-Instruct Text Generation • 4B • Updated Nov 25, 2025 • 6 • 4 black-forest-labs/FLUX.2-dev Image-to-Image • 32B • Updated Feb 17 • 1.36M • • 1.97k mistralai/Mistral-Large-Instruct-2407 123B • Updated Jul 28, 2025 • 5.03k • 865
Computer Vision Vision Grid Transformer for Document Layout Analysis Paper • 2308.14978 • Published Aug 29, 2023 • 4
HTR ByteDance/Sa2VA-4B Image-Text-to-Text • 4B • Updated Sep 8, 2025 • 825 • 99 Finnish-NLP/Ahma-2-4B-Instruct Text Generation • 4B • Updated Nov 25, 2025 • 6 • 4 black-forest-labs/FLUX.2-dev Image-to-Image • 32B • Updated Feb 17 • 1.36M • • 1.97k mistralai/Mistral-Large-Instruct-2407 123B • Updated Jul 28, 2025 • 5.03k • 865
Computer Vision Vision Grid Transformer for Document Layout Analysis Paper • 2308.14978 • Published Aug 29, 2023 • 4