-
Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Model
Paper • 2407.07053 • Published • 46 -
LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Paper • 2407.12772 • Published • 35 -
VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Paper • 2407.11691 • Published • 17 -
MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models
Paper • 2408.02718 • Published • 60
Collections
Discover the best community collections!
Collections trending this week
-
PKU-Alignment/align-anything
Viewer • Updated • 69.4k • 5.51k • 49 -
PKU-Alignment/Align-Anything-Instruction-100K-zh
Viewer • Updated • 105k • 102 • 10 -
PKU-Alignment/Align-Anything-Instruction-100K
Viewer • Updated • 105k • 137 • 9 -
PKU-Alignment/Align-Anything-TI2T-Instruction-100K
Viewer • Updated • 103k • 326 • 1
-
Alibaba-NLP/gte-Qwen2-7B-instruct
Sentence Similarity • 8B • Updated • 109k • 483 -
Alibaba-NLP/gte-Qwen2-1.5B-instruct
Sentence Similarity • 2B • Updated • 502k • 239 -
Alibaba-NLP/gte-multilingual-base
Sentence Similarity • 0.3B • Updated • 1.51M • 379 -
Alibaba-NLP/gte-multilingual-reranker-base
Text Ranking • 0.3B • Updated • 269k • 192
-
🧩 DiffuseCraft Mod (SDXL/SD1.5 Models Text-to-Image)
🧩222Stunning images using stable diffusion.
-
Votepurchase Multiple Model (SD1.5/SDXL Text-to-Image)
🖼141Text-to-Image
-
FLUX LoRA the Explorer Mod
🏆160Generate AI images from text and optional sketches
-
🧩 DiffuseCraft
🧩241Stunning images using stable diffusion.
-
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
bigscience/bloom
Text Generation • 176B • Updated • 14.4k • 5.07k -
bigcode/starcoder
Text Generation • 16B • Updated • 16.7k • 2.99k -
nvidia/Llama3-ChatQA-1.5-8B
Text Generation • 8B • Updated • 11.9k • • 555
-
speakleash/Bielik-7B-v0.1
Text Generation • 7B • Updated • 2.95k • 75 -
speakleash/Bielik-7B-Instruct-v0.1
Text Generation • 7B • Updated • 3.02k • • 64 -
speakleash/Bielik-7B-Instruct-v0.1-GGUF
Text Generation • 7B • Updated • 3.3k • 15 -
speakleash/Bielik-7B-Instruct-v0.1-GPTQ
Text Generation • 7B • Updated • 1.46k • 5
-
Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Model
Paper • 2407.07053 • Published • 46 -
LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Paper • 2407.12772 • Published • 35 -
VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Paper • 2407.11691 • Published • 17 -
MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models
Paper • 2408.02718 • Published • 60
-
PKU-Alignment/align-anything
Viewer • Updated • 69.4k • 5.51k • 49 -
PKU-Alignment/Align-Anything-Instruction-100K-zh
Viewer • Updated • 105k • 102 • 10 -
PKU-Alignment/Align-Anything-Instruction-100K
Viewer • Updated • 105k • 137 • 9 -
PKU-Alignment/Align-Anything-TI2T-Instruction-100K
Viewer • Updated • 103k • 326 • 1
-
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
bigscience/bloom
Text Generation • 176B • Updated • 14.4k • 5.07k -
bigcode/starcoder
Text Generation • 16B • Updated • 16.7k • 2.99k -
nvidia/Llama3-ChatQA-1.5-8B
Text Generation • 8B • Updated • 11.9k • • 555
-
Alibaba-NLP/gte-Qwen2-7B-instruct
Sentence Similarity • 8B • Updated • 109k • 483 -
Alibaba-NLP/gte-Qwen2-1.5B-instruct
Sentence Similarity • 2B • Updated • 502k • 239 -
Alibaba-NLP/gte-multilingual-base
Sentence Similarity • 0.3B • Updated • 1.51M • 379 -
Alibaba-NLP/gte-multilingual-reranker-base
Text Ranking • 0.3B • Updated • 269k • 192
-
speakleash/Bielik-7B-v0.1
Text Generation • 7B • Updated • 2.95k • 75 -
speakleash/Bielik-7B-Instruct-v0.1
Text Generation • 7B • Updated • 3.02k • • 64 -
speakleash/Bielik-7B-Instruct-v0.1-GGUF
Text Generation • 7B • Updated • 3.3k • 15 -
speakleash/Bielik-7B-Instruct-v0.1-GPTQ
Text Generation • 7B • Updated • 1.46k • 5
-
🧩 DiffuseCraft Mod (SDXL/SD1.5 Models Text-to-Image)
🧩222Stunning images using stable diffusion.
-
Votepurchase Multiple Model (SD1.5/SDXL Text-to-Image)
🖼141Text-to-Image
-
FLUX LoRA the Explorer Mod
🏆160Generate AI images from text and optional sketches
-
🧩 DiffuseCraft
🧩241Stunning images using stable diffusion.