alexey
kingograder
AI & ML interests
None yet
Recent Activity
updated a collection 4 days ago
ASR liked a model 7 days ago
yandex/AliceAI-Foundation-80B-A3B-Base updated a collection 7 days ago
Image GenerationOrganizations
None yet
3D
ASR
-
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 1.81M • • 1.13k -
Qwen/Qwen3-ASR-0.6B
Automatic Speech Recognition • 0.9B • Updated • 519k • • 365 -
nvidia/canary-1b-v2
Automatic Speech Recognition • 1.0B • Updated • 58.6k • 425 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 561k • • 1.16k
TTS
Image Generation
Text to image and Image to image collection. All* model awalible for commercial* purposes.
datasets
Video Generation
OCR
-
PaddlePaddle/PaddleOCR-VL-1.5
Image-Text-to-Text • 1.0B • Updated • 15.6k • 666 -
PaddlePaddle/PaddleOCR-VL
Image-Text-to-Text • 1.0B • Updated • 8.04k • 1.66k -
lightonai/LightOnOCR-2-1B
Image-Text-to-Text • 1B • Updated • 213k • • 834 -
lightonai/LightOnOCR-1B-1025
Image-to-Text • 1B • Updated • 35.1k • 256
LLM
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 10.6k • 231 -
tencent/Youtu-VL-4B-Instruct
Image-Text-to-Text • 5B • Updated • 1.71k • 162 -
OpenDataArena/MMFineReason-4B
Visual Question Answering • 5B • Updated • 83 • 15 -
OpenDataArena/MMFineReason-2B
Visual Question Answering • 2B • Updated • 224 • 8
interested
datasets
3D
Video Generation
ASR
-
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 1.81M • • 1.13k -
Qwen/Qwen3-ASR-0.6B
Automatic Speech Recognition • 0.9B • Updated • 519k • • 365 -
nvidia/canary-1b-v2
Automatic Speech Recognition • 1.0B • Updated • 58.6k • 425 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 561k • • 1.16k
OCR
-
PaddlePaddle/PaddleOCR-VL-1.5
Image-Text-to-Text • 1.0B • Updated • 15.6k • 666 -
PaddlePaddle/PaddleOCR-VL
Image-Text-to-Text • 1.0B • Updated • 8.04k • 1.66k -
lightonai/LightOnOCR-2-1B
Image-Text-to-Text • 1B • Updated • 213k • • 834 -
lightonai/LightOnOCR-1B-1025
Image-to-Text • 1B • Updated • 35.1k • 256
TTS
LLM
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 10.6k • 231 -
tencent/Youtu-VL-4B-Instruct
Image-Text-to-Text • 5B • Updated • 1.71k • 162 -
OpenDataArena/MMFineReason-4B
Visual Question Answering • 5B • Updated • 83 • 15 -
OpenDataArena/MMFineReason-2B
Visual Question Answering • 2B • Updated • 224 • 8
Image Generation
Text to image and Image to image collection. All* model awalible for commercial* purposes.