Inference Providers
Active filters: web-llm
mlc-ai/Llama-2-7b-chat-hf-q4f16_1-MLC
Updated • 807
• 9
mlc-ai/Llama-2-7b-chat-hf-q4f32_1-MLC
Updated • 45
• 2
mlc-ai/Mistral-7B-Instruct-v0.2-q4f16_1-MLC
Updated • 992
• 4
mlc-ai/Llama-2-13b-chat-hf-q4f16_1-MLC
Updated • 113
• 1
mlc-ai/Llama-2-70b-chat-hf-q4f16_1-MLC
Updated • 5
• 2
mlc-ai/OpenHermes-2.5-Mistral-7B-q4f16_1-MLC
Updated • 87
• 1
mlc-ai/NeuralHermes-2.5-Mistral-7B-q4f16_1-MLC
Updated • 66
• 2
mlc-ai/WizardMath-7B-V1.0-q4f16_1-MLC
mlc-ai/WizardMath-13B-V1.0-q4f16_1-MLC
mlc-ai/WizardMath-70B-V1.0-q4f16_1-MLC
mlc-ai/RedPajama-INCITE-Chat-3B-v1-q4f16_1-MLC
Updated • 937
• 3
mlc-ai/RedPajama-INCITE-Chat-3B-v1-q4f32_1-MLC
Updated • 40
• 2
mlc-ai/WizardMath-7B-V1.1-q4f16_1-MLC
Updated • 810
• 2
mlc-ai/phi-1_5-q4f16_1-MLC
Updated • 200
Updated • 5
• 1
mlc-ai/gpt2-medium-q0f16-MLC
Updated • 7
• 1
mlc-ai/Mistral-7B-Instruct-v0.2-q3f16_1-MLC
Updated • 13
• 4
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q4f16_1-MLC
Updated • 280
• 1
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q4f32_1-MLC
Updated • 290
• 1
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q0f16-MLC
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q0f32-MLC
mlc-ai/phi-1_5-q4f32_1-MLC
Updated • 50
• 1
Updated • 10
• 2
mlc-ai/NeuralHermes-2.5-Mistral-7B-q3f16_1-MLC
mlc-ai/Qwen-7B-Chat-q4f16_1-MLC
Updated • 3
• 2