baseten/kimi-k3-tokenizer
Updated • 45
• 2
baseten/o200k-base-tiktoken
Updated
baseten/GLM-5.2-Vision-NVFP4
Image-Text-to-Text
• 381B • Updated • 2.76k
• 128
baseten/GLM-5.2-Vision-FP8
Image-Text-to-Text
• 754B • Updated • 137
• 7
baseten/glm-5-2-projector
baseten/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4-dequant-to-BF16
Text Generation
• 561B • Updated • 172
15.5M • Updated • 23
• 1
baseten/gemma-4-26b-a4b-it-sequence-classification
26B • Updated • 19
baseten/gemma-4-e2b-it-sequence-classification
Text Classification
• 5B • Updated • 4
baseten/distilled_8step_FLUX.2-dev
Image-to-Image
• 32B • Updated • 27
• 3
baseten/Wan2.2-T2V-A14B-LightX2V-V2.0-4step
Text-to-Video
• 14B • Updated • 9
baseten/Qwen3-1.7B-NVFP4-PTQ
baseten/Qwen3-4B-NVFP4-PTQ
2B • Updated • 68
• 1
baseten/Qwen-Image-2512-Pruned-50blocks
Text-to-Image
• 17B • Updated • 4
baseten/embedding-smol_llama-101M-GQA
76.6M • Updated • 6
baseten/qwen3-engine-30A3-repro
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v2_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_turbo_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_251013_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/Llama-3.2-3B-Instruct-pythonic
Text Generation
• 3B • Updated • 31.2k
• baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_L4_0_21_0
Updated
baseten/whisper_trt_large_v3_250729_NVIDIA_H100_80GB_HBM3_MIG_3g_40gb_1_0_0rc6
Updated