Inference Providers
Active filters: kv-cache
Dyluhn/Qwen3.8-Flash-Next-Uncensored-CED-Projector
stoyani/Qwen3.8-27B-FP8-KVcal
Image-Text-to-Text
• 28B • Updated • 55
• 1
fromthesky/PLDR-LLM-v51-104M
Text Generation
• 0.1B • Updated • 20
fromthesky/PLDR-LLM-v51-110M-1
Text Generation
• 0.1B • Updated • 24
fromthesky/PLDR-LLM-v51-110M-2
Text Generation
• 0.1B • Updated • 20
fromthesky/PLDR-LLM-v51-110M-3
Text Generation
• 0.1B • Updated • 174
fromthesky/PLDR-LLM-v51-110M-4
Text Generation
• 0.1B • Updated • 23
fromthesky/PLDR-LLM-v51-110M-5
Text Generation
• 0.1B • Updated • 14
fromthesky/PLDR-LLM-v51-DAG-110M
Text Generation
• 0.1B • Updated • 16
fromthesky/PLDR-LLM-v51G-106M-1
Text Generation
• 0.1B • Updated • 32
fromthesky/PLDR-LLM-v51G-106M-2
Text Generation
• 0.1B • Updated • 29
fromthesky/PLDR-LLM-v51G-106M-3
Text Generation
• 0.1B • Updated • 267
fromthesky/PLDR-LLM-v51G-106M-test
Text Generation
• 0.1B • Updated • 27
fromthesky/PLDR-LLM-v52-81M-FT-SC-1
Text Classification
• 81M • Updated • 7
fromthesky/PLDR-LLM-v52-81M-FT-QA-1
Question Answering
• 81M • Updated • 10
fromthesky/PLDR-LLM-v52-81M-FT-TC-1
Token Classification
• 81M • Updated • 12
fromthesky/PLDR-LLM-v52-110M-1
Text Generation
• 0.1B • Updated • 19
nm-testing/Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Llama-3.1-8B-Instruct-QKV-Cache-FP8-Per-Head
Updated
nm-testing/Llama-3.1-8B-Instruct-FP8-dynamic-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Llama-3.1-8B-Instruct-FP8-dynamic-QKV-Cache-FP8-Per-Head
Updated
nm-testing/Qwen3-32B-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Qwen3-32B-QKV-Cache-FP8-Per-Head
Updated
nm-testing/Qwen3-32B-FP8-dynamic-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Qwen3-32B-FP8-dynamic-QKV-Cache-FP8-Per-Head
Updated
nm-testing/Llama-3.3-70B-Instruct-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Llama-3.3-70B-Instruct-QKV-Cache-FP8-Per-Head
Updated
nm-testing/Llama-3.3-70B-Instruct-FP8-dynamic-QKV-Cache-FP8-Per-Tensor
Updated
nm-testing/Llama-3.3-70B-Instruct-FP8-dynamic-QKV-Cache-FP8-Per-Head
Updated