Ji-Xiang
Ji-Xiang
AI & ML interests
None yet
Recent Activity
updated a collection 1 day ago
Text Generation Inference liked a model 1 day ago
orcarouter/OrcaSAQ-2-27B liked a dataset 4 days ago
Contrastive-LM/deepswe-clm-train-embeddings-8kOrganizations
Docling
V-JEPA 2
-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 209k • 210 -
facebook/vjepa2-vith-fpc64-256
Video Classification • 0.7B • Updated • 2.31k • 24 -
facebook/vjepa2-vitg-fpc64-256
Video Classification • 1B • Updated • 99.1k • 57 -
facebook/vjepa2-vitg-fpc64-384
Video Classification • 1B • Updated • 3.99k • 46
GUI-Actor
1-bit Large Language Model (LLM)
GRPO datasets
Conversational Speech Model
OCR tools
Images Datasets
Critique Fine-Tuning (CFT) Datasets
Critique Fine-Tuning (CFT) Datasets
Test-time scaling Datasets
Please see below:
https://medium.com/@techsachin/s1-simple-test-time-scaling-approach-to-exceed-openais-o1-preview-performance-ec5a624c5d2f
Thinking/Reasoning Datasets
RLHF Datasets
Math Datasets
Multilingual-dataset
Retrieval-Augmented Generation (RAG) Dataset
Multilingual Large Language Models
-
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 800k • • 2.65k -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.53M • • 2.7k -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.19M • • 8.12k -
mistralai/Mistral-Small-24B-Base-2501
24B • Updated • 6.14k • 269
Recommended Datasets
Text-to-Video
Traditional-chinese-dataset
China models
China-dataset
unfiltered dataset
Edge Computing
Medical
GGUF Models
-
crusoeai/Llama-3-8B-Instruct-Gradient-1048k-GGUF
8B • Updated • 1.44k • 71 -
NousResearch/Hermes-2-Pro-Llama-3-8B-GGUF
8B • Updated • 2.28k • 165 -
ibm-granite/granite-34b-code-instruct-8k-GGUF
Text Generation • 34B • Updated • 161 • 8 -
ibm-granite/granite-20b-code-base-8k-GGUF
Text Generation • 20B • Updated • 112 • 5
Visual Question Answering
Multi Tasks
DPO datasets
SLM (small language models)
-
HuggingFaceTB/SmolLM-135M
Text Generation • 0.1B • Updated • 167k • 272 -
HuggingFaceTB/SmolLM-135M-Instruct
Text Generation • 0.1B • Updated • 31.2k • 145 -
HuggingFaceTB/SmolLM-360M-Instruct
Text Generation • 0.4B • Updated • 10.9k • 89 -
HuggingFaceTB/SmolLM-360M
Text Generation • 0.4B • Updated • 12.6k • 73
Vision-Language dataset
Dense Passage Retrieval (DPR) Datasets
background-removal
Try on
Agentic Coding
Image-Editing
-
HiDream-ai/HiDream-E1-1
Any-to-Any • 17B • Updated • 254 • 217 -
HiDream-ai/HiDream-E1-Full
Any-to-Any • 17B • Updated • 311 • 216 -
Qwen/Qwen-Image-Edit
Image-to-Image • 20B • Updated • 111k • • 2.53k - Running on CPU UpgradeAgents2.74k
Omni Image Editor
🖼2.74kImage edit, text to image, image upscale, remove watermark
Robotics
Reasoning models
Taiwanese Taigi Datasets
Image-Text-to-Text
Text Generation Inference
-
Qwen/QwQ-32B
Text Generation • 33B • Updated • 69.1k • • 2.97k -
deepseek-ai/DeepSeek-R1-Distill-Llama-70B
Text Generation • 71B • Updated • 72.5k • • 809 -
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B
Text Generation • 33B • Updated • 461k • • 1.63k -
microsoft/Phi-4-mini-flash-reasoning
Text Generation • 4B • Updated • 1.34k • 285
Video generation
General screen parsing tool
Reasoning datasets
RLVR Datasets
Reinforcement Learning from Verifiable Rewards (RLVR) Datasets
WebGPU
HTML to Markdown
Logical Reasoning Datasets
Object Detection
Image-to-Video
SFT Datasets
Coder LLM
- PausedAgentsFeatured1.73k
Qwen2.5 Coder Artifacts
🐢1.73kGenerate and preview app code from a text description
-
Qwen/Qwen2.5-72B-Instruct
Text Generation • 73B • Updated • 298k • • 1k -
Qwen/Qwen2.5-32B-Instruct
Text Generation • 33B • Updated • 1.94M • • 362 -
Qwen/Qwen2.5-14B-Instruct
Text Generation • 15B • Updated • 1.81M • • 372
Multimodal Language Models
Chinese models
-
MediaTek-Research/Breeze-7B-Instruct-v0_1
Text Generation • 7B • Updated • 601 • 91 -
MediaTek-Research/Breeze-7B-Base-v0_1
Text Generation • 7B • Updated • 92 • 23 -
MediaTek-Research/Breeze-7B-Instruct-v1_0
Text Generation • 7B • Updated • 1.14k • 67 -
YC-Chen/Breeze-7B-Instruct-v1_0-GGUF
Text Generation • 7B • Updated • 162 • 22
Uncensored models
common-dataset
Image Generator
- Running on ZeroAgentsFeatured1.15k
Playground V2.5
🌍1.15kGenerate highly aesthetic images
- Running on ZeroAgentsFeatured9.57k
FLUX.1 [dev]
🖥9.57kGenerate images from text prompts
- RunningAgentsFeatured919
Kolors Character With Flux
🤹919Kolors Character to keep character developed with Flux
-
franciszzj/Leffa
Image-to-Image • Updated • 344
Voice
Big Language Models
-
deepseek-ai/DeepSeek-V2-Chat
Text Generation • 236B • Updated • 19.3k • 463 -
deepseek-ai/DeepSeek-V2
Text Generation • 236B • Updated • 34.1k • 337 -
ibm-granite/granite-34b-code-base-8k-GGUF
Text Generation • 34B • Updated • 102 • 3 -
ibm-granite/granite-20b-code-base-8k-GGUF
Text Generation • 20B • Updated • 112 • 5
text-to-speech (TTS)
Chat
Vision
ORPO-DPO datasets
automatic speech recognition (ASR)
MoE
Audio-To-Text
Extreme Quantization
2-bit
Agentic Coding
Docling
Image-Editing
-
HiDream-ai/HiDream-E1-1
Any-to-Any • 17B • Updated • 254 • 217 -
HiDream-ai/HiDream-E1-Full
Any-to-Any • 17B • Updated • 311 • 216 -
Qwen/Qwen-Image-Edit
Image-to-Image • 20B • Updated • 111k • • 2.53k - Running on CPU UpgradeAgents2.74k
Omni Image Editor
🖼2.74kImage edit, text to image, image upscale, remove watermark
V-JEPA 2
-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 209k • 210 -
facebook/vjepa2-vith-fpc64-256
Video Classification • 0.7B • Updated • 2.31k • 24 -
facebook/vjepa2-vitg-fpc64-256
Video Classification • 1B • Updated • 99.1k • 57 -
facebook/vjepa2-vitg-fpc64-384
Video Classification • 1B • Updated • 3.99k • 46
Robotics
GUI-Actor
Reasoning models
1-bit Large Language Model (LLM)
Taiwanese Taigi Datasets
GRPO datasets
Image-Text-to-Text
Conversational Speech Model
Text Generation Inference
-
Qwen/QwQ-32B
Text Generation • 33B • Updated • 69.1k • • 2.97k -
deepseek-ai/DeepSeek-R1-Distill-Llama-70B
Text Generation • 71B • Updated • 72.5k • • 809 -
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B
Text Generation • 33B • Updated • 461k • • 1.63k -
microsoft/Phi-4-mini-flash-reasoning
Text Generation • 4B • Updated • 1.34k • 285
OCR tools
Video generation
Images Datasets
General screen parsing tool
Critique Fine-Tuning (CFT) Datasets
Critique Fine-Tuning (CFT) Datasets
Reasoning datasets
Test-time scaling Datasets
Please see below:
https://medium.com/@techsachin/s1-simple-test-time-scaling-approach-to-exceed-openais-o1-preview-performance-ec5a624c5d2f
RLVR Datasets
Reinforcement Learning from Verifiable Rewards (RLVR) Datasets
Thinking/Reasoning Datasets
WebGPU
RLHF Datasets
HTML to Markdown
Math Datasets
Logical Reasoning Datasets
Multilingual-dataset
Object Detection
Retrieval-Augmented Generation (RAG) Dataset
Image-to-Video
Multilingual Large Language Models
-
meta-llama/Llama-3.2-1B
Text Generation • 1B • Updated • 800k • • 2.65k -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.53M • • 2.7k -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.19M • • 8.12k -
mistralai/Mistral-Small-24B-Base-2501
24B • Updated • 6.14k • 269
SFT Datasets
Recommended Datasets
Coder LLM
- PausedAgentsFeatured1.73k
Qwen2.5 Coder Artifacts
🐢1.73kGenerate and preview app code from a text description
-
Qwen/Qwen2.5-72B-Instruct
Text Generation • 73B • Updated • 298k • • 1k -
Qwen/Qwen2.5-32B-Instruct
Text Generation • 33B • Updated • 1.94M • • 362 -
Qwen/Qwen2.5-14B-Instruct
Text Generation • 15B • Updated • 1.81M • • 372
Text-to-Video
Multimodal Language Models
Traditional-chinese-dataset
Chinese models
-
MediaTek-Research/Breeze-7B-Instruct-v0_1
Text Generation • 7B • Updated • 601 • 91 -
MediaTek-Research/Breeze-7B-Base-v0_1
Text Generation • 7B • Updated • 92 • 23 -
MediaTek-Research/Breeze-7B-Instruct-v1_0
Text Generation • 7B • Updated • 1.14k • 67 -
YC-Chen/Breeze-7B-Instruct-v1_0-GGUF
Text Generation • 7B • Updated • 162 • 22
China models
Uncensored models
China-dataset
common-dataset
unfiltered dataset
Image Generator
- Running on ZeroAgentsFeatured1.15k
Playground V2.5
🌍1.15kGenerate highly aesthetic images
- Running on ZeroAgentsFeatured9.57k
FLUX.1 [dev]
🖥9.57kGenerate images from text prompts
- RunningAgentsFeatured919
Kolors Character With Flux
🤹919Kolors Character to keep character developed with Flux
-
franciszzj/Leffa
Image-to-Image • Updated • 344
Edge Computing
Voice
Medical
Big Language Models
-
deepseek-ai/DeepSeek-V2-Chat
Text Generation • 236B • Updated • 19.3k • 463 -
deepseek-ai/DeepSeek-V2
Text Generation • 236B • Updated • 34.1k • 337 -
ibm-granite/granite-34b-code-base-8k-GGUF
Text Generation • 34B • Updated • 102 • 3 -
ibm-granite/granite-20b-code-base-8k-GGUF
Text Generation • 20B • Updated • 112 • 5
GGUF Models
-
crusoeai/Llama-3-8B-Instruct-Gradient-1048k-GGUF
8B • Updated • 1.44k • 71 -
NousResearch/Hermes-2-Pro-Llama-3-8B-GGUF
8B • Updated • 2.28k • 165 -
ibm-granite/granite-34b-code-instruct-8k-GGUF
Text Generation • 34B • Updated • 161 • 8 -
ibm-granite/granite-20b-code-base-8k-GGUF
Text Generation • 20B • Updated • 112 • 5
text-to-speech (TTS)
Visual Question Answering
Chat
Multi Tasks
Vision
DPO datasets
ORPO-DPO datasets
SLM (small language models)
-
HuggingFaceTB/SmolLM-135M
Text Generation • 0.1B • Updated • 167k • 272 -
HuggingFaceTB/SmolLM-135M-Instruct
Text Generation • 0.1B • Updated • 31.2k • 145 -
HuggingFaceTB/SmolLM-360M-Instruct
Text Generation • 0.4B • Updated • 10.9k • 89 -
HuggingFaceTB/SmolLM-360M
Text Generation • 0.4B • Updated • 12.6k • 73
automatic speech recognition (ASR)
Vision-Language dataset
MoE
Dense Passage Retrieval (DPR) Datasets
Audio-To-Text
background-removal
Extreme Quantization
Try on