baseten/o200k-base-tiktoken
Updated
baseten/GLM-5.2-Vision-NVFP4
Image-Text-to-Text
• 381B • Updated • 2.03k
• 107
baseten/GLM-5.2-Vision-FP8
Image-Text-to-Text
• 754B • Updated • 92
• 3
baseten/glm-5-2-projector
baseten/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4-dequant-to-BF16
Text Generation
• 561B • Updated • 100
15.5M • Updated • 24
• 1
baseten/gemma-4-26b-a4b-it-sequence-classification
26B • Updated • 19
baseten/gemma-4-e2b-it-sequence-classification
Text Classification
• 5B • Updated • 5
baseten/distilled_8step_FLUX.2-dev
Text-to-Image
• 32B • Updated • 28
• 3
baseten/Wan2.2-T2V-A14B-LightX2V-V2.0-4step
Text-to-Video
• 14B • Updated • 7
baseten/Qwen3-1.7B-NVFP4-PTQ
baseten/Qwen3-4B-NVFP4-PTQ
2B • Updated • 69
• 1
baseten/Qwen-Image-2512-Pruned-50blocks
Text-to-Image
• 17B • Updated • 4
baseten/embedding-smol_llama-101M-GQA
76.6M • Updated • 10
baseten/qwen3-engine-30A3-repro
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/Llama-3.2-3B-Instruct-pythonic
Text Generation
• 3B • Updated • 29.9k
• baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_1_0_0rc6
Updated
8B • Updated • 5