Skip to main content

Supported models

The release ships its supported-models document beside the archives: supported_models.json, and supported_models.csv, the same document flat. It lists 109 checkpoints over 56 model families with 312 weight variants. This page lists the 80 of them whose weights carry a permissive license, the document's own class; the model card is where a checkpoint's license is read, and the runtime prints it at fetch, at load and in info. The whole document is a download of the platform with the release (Get ClikaRT).

A model is named by its Hugging Face repository id, which clikart-cli fetch <id> and AutoModel.from_pretrained(id) take as they are. Parameters and Context are read off the checkpoint. Weights lists the weight variants the release serves for the model, base first, with their format. Proven on names the backends the release's own run record covers for the model; an empty cell is a model the release lists without a recorded run.

Text generation and chat​

ModelFamilyParametersContextWeightsProven on
arianraje/qwen3-4b-gdn-hybrid-stage1-align-dtfixqwen3-nextabout 4.5 B40,960BF16 (safetensors)CPU, CUDA, Metal
DavidAU/Qwen3-MOE-4x0.6B-2.4B-Writing-Thunderqwenabout 1.5 B40,960BF16 (safetensors)CPU, CUDA, Vulkan, Metal
microsoft/phi-2phiabout 2.8 B2,048F16 (safetensors)CPU, CUDA, Metal
microsoft/Phi-3-mini-4k-instructphiabout 3.8 B4,096BF16 (safetensors)CPU, CUDA, Vulkan, Metal
microsoft/Phi-3.5-mini-instructphiabout 3.8 B131,072IQ1_M (gguf), IQ1_S (gguf), IQ2_XS (gguf), IQ3_XS (gguf), IQ4_XS (gguf), Q2_K (gguf), Q3_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf)CPU, CUDA, Vulkan
moonshotai/Kimi-Linear-48B-A3B-Instructkimi-linearabout 49.1 BBF16 (safetensors)CPU
openai/gpt-oss-20bgpt-ossabout 20.9 B131,072U8 (safetensors), MXFP4 (gguf)CPU, CUDA, Vulkan
OuteAI/Lite-Oute-1-300Mmistralabout 300 M4,096F32 (safetensors)CPU, CUDA, Metal
Qwen/Qwen2.5-0.5B-Instructqwenabout 494 M32,768BF16 (safetensors)CPU, CUDA, Vulkan, Metal
Qwen/Qwen3-0.6Bqwenabout 752 M40,960BF16 (gguf), F8_E4M3 (safetensors), IQ4_NL (gguf), IQ4_XS (gguf), Q2_K (gguf), Q2_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ1_M (gguf), UD-IQ1_S (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_XXS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf)CPU, CUDA, Vulkan
Qwen/Qwen3-0.6B-Baseqwenabout 596 M32,768Qwen3-Embedding-0.6B-f16 (gguf), Q8_0 (gguf)CPU, CUDA, Vulkan
SupraLabs/Supra2-100M-Baseqwenabout 101 M2,048F32 (safetensors)CPU, CUDA, Vulkan
zai-org/GLM-4-9B-0414glmabout 9.4 B32,768BF16 (safetensors), IQ2_M (gguf), IQ3_M (gguf), IQ3_XS (gguf), IQ3_XXS (gguf), IQ4_NL (gguf), IQ4_XS (gguf), Q2_K (gguf), Q2_K_L (gguf), Q3_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q3_K_XL (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_L (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_L (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q6_K_L (gguf), Q8_0 (gguf), THUDM_GLM-4-9B-0414-bf16 (gguf)CPU, CUDA, Vulkan, Metal
zai-org/GLM-4.5-Airglmabout 110.5 B131,072BF16 (safetensors)CPU

Vision-language chat​

ModelFamilyParametersContextWeightsProven on
baidu/Unlimited-OCRdeepseek_ocrabout 3.3 B32,768BF16 (safetensors)CPU, CUDA, Metal
deepseek-ai/DeepSeek-OCRdeepseek_ocrabout 3.3 B8,192BF16 (safetensors)CPU, CUDA, Metal
deepseek-ai/DeepSeek-OCR-2deepseek_ocrabout 3.4 B8,192BF16 (safetensors)CPU, CUDA, Metal
deepseek-community/DeepSeek-OCR-2deepseek_ocrabout 3.4 B8,192BF16 (safetensors)CPU, CUDA, Metal
florence-community/Florence-2-baseflorence2about 232 M1,024F16 (safetensors)CPU, CUDA, Vulkan, Metal
google/gemma-4-26B-A4Bgemmaabout 26.5 B262,144BF16 (safetensors)CPU, CUDA
google/gemma-4-26B-A4B-itgemmaabout 25.8 B262,144BF16 (safetensors)
google/gemma-4-26B-A4B-it-qat-q4_0-unquantizedgemmaabout 26.5 B262,144BF16 (safetensors)CUDA
google/gemma-4-31Bgemmaabout 32.7 B262,144BF16 (safetensors)CPU, CUDA
google/gemma-4-31B-itgemmaabout 31.3 B262,144BF16 (safetensors)CPU, CUDA, Vulkan
google/gemma-4-31B-it-qat-q4_0-unquantizedgemmaabout 32.7 B262,144BF16 (safetensors), I32 (safetensors), gemma-4-31B_q4_0-it (gguf)CPU, CUDA, Vulkan
meta-models/Muse-Glimmer-30Bmuse-glimmerabout 29.8 B131,072BF16 (safetensors)CPU, CUDA, Vulkan
microsoft/Florence-2-baseflorence2about 232 M1,024F16 (safetensors)CPU, CUDA, Vulkan, Metal
Qwen/Qwen2-VL-2B-Instructqwen-vlabout 2.2 B32,768BF16 (safetensors)CPU, CUDA, Vulkan, Metal
Qwen/Qwen3-VL-2B-Instructqwen-vlabout 2.1 B262,144BF16 (safetensors), F16 (gguf), Q4_K_M (gguf), Q8_0 (gguf)CPU, CUDA, Vulkan, Metal
Qwen/Qwen3-VL-30B-A3B-Instructqwen-vlabout 31.1 B262,144F8_E4M3 (safetensors)CPU, CUDA, Vulkan
Qwen/Qwen3.5-0.8Bqwen3-5about 873 M262,144BF16 (safetensors), IQ4_NL (gguf), IQ4_XS (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_XXS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf)CPU, CUDA, Vulkan, Metal
Qwen/Qwen3.5-35B-A3Bqwen3-5-moeabout 36 B262,144BF16 (gguf), I32 (safetensors), MXFP4_MOE (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_S (gguf), UD-IQ3_XXS (gguf), UD-IQ4_NL (gguf), UD-IQ4_XS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_L (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_S (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf)CPU, CUDA, Vulkan
thinkingmachines/Inkling-Smallinklingabout 266 BU8 (safetensors)CPU

Multimodal (any to any)​

ModelFamilyParametersContextWeightsProven on
google/gemma-4-12B-itgemmaabout 12 B262,144BF16 (safetensors)CPU, CUDA, Vulkan
google/gemma-4-12B-it-qat-q4_0-unquantizedgemmaabout 12 B262,144gemma-4-12b-it-qat-q4_0 (gguf)CPU, CUDA, Vulkan
google/gemma-4-E2B-itgemmaabout 5.1 B131,072BF16 (safetensors)CPU, CUDA, Vulkan, Metal
google/gemma-4-E2B-it-qat-q4_0-unquantizedgemmaabout 5.1 B131,072gemma-4-E2B_q4_0-it (gguf)CPU, CUDA, Vulkan
google/gemma-4-E4B-itgemmaabout 8 B131,072BF16 (safetensors)CPU, CUDA, Vulkan, Metal
google/gemma-4-E4B-it-qat-q4_0-unquantizedgemmaabout 7.9 B131,072gemma-4-E4B_q4_0-it (gguf)CPU, CUDA, Vulkan

Speech to text​

ModelFamilyParametersContextWeightsProven on
nvidia/canary-1b-v2canaryabout 979 MF32 (safetensors)
nvidia/parakeet-ctc-0.6bparakeetabout 609 MF32 (safetensors)
nvidia/parakeet-rnnt-0.6bparakeetabout 617 MF32 (safetensors)
nvidia/parakeet-tdt-0.6b-v3parakeetabout 627 MF32 (safetensors)
openai/whisper-large-v3-turbowhisperabout 809 Mwhisper-large-v3-turbo-q4_0 (gguf), whisper-large-v3-turbo-q4_1 (gguf), whisper-large-v3-turbo-q8_0 (gguf)CPU, CUDA
openai/whisper-tinywhisperabout 38 M448F32 (safetensors), F16 (gguf), Q4_K_M (gguf), Q5_K_M (gguf), Q6_K (gguf), Q8_0 (gguf)CPU, CUDA, Vulkan, Metal

Text to speech​

ModelFamilyParametersContextWeightsProven on
sesame/csm-1bcsmabout 1.6 B2,048F32 (safetensors)CPU, CUDA, Metal

Translation​

ModelFamilyParametersContextWeightsProven on
tencent/Hy-MT2-1.8Bhunyuanabout 2 B262,144BF16 (safetensors), Q4_K_M (gguf), Q6_K (gguf), Q8_0 (gguf)CPU, CUDA, Vulkan, Metal
tencent/Hy-MT2-30B-A3Bhy-v3about 30.1 B262,144F8_E4M3 (safetensors)CPU, CUDA, Vulkan

Text embedding​

ModelFamilyParametersContextWeightsProven on
TaylorAI/bge-micro-v2bertabout 17 M512F16 (safetensors)Metal

Feature extraction​

ModelFamilyParametersContextWeightsProven on
Qwen/Qwen3-Embedding-0.6Bqwen-embeddingabout 596 M32,768BF16 (safetensors)CPU, CUDA, Vulkan, Metal

Image embedding​

ModelFamilyParametersContextWeightsProven on
facebook/dinov2-smalldinov2about 22 MF32 (safetensors)CPU, CUDA, Vulkan, Metal

Reranking​

ModelFamilyParametersContextWeightsProven on
Qwen/Qwen3-Reranker-0.6Bqwen-rerankerabout 596 M40,960BF16 (safetensors)CPU, CUDA, Vulkan, Metal

Object detection​

ModelFamilyParametersContextWeightsProven on
facebook/detr-resnet-50detrabout 42 M1,024F32 (safetensors)CPU, CUDA, Vulkan, Metal
microsoft/table-transformer-detectiondetrabout 29 M1,024F32 (safetensors)CPU, CUDA, Vulkan, Metal
PekingU/rtdetr_r18vdrt-detrabout 20 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
PekingU/rtdetr_v2_r18vdrt-detrabout 20 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
Roboflow/rf-detr-nanorf-detrabout 30 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
ustc-community/dfine-nano-cocod-fineabout 4 MF32 (safetensors)CPU, CUDA, Vulkan, Metal

Open-vocabulary object detection​

ModelFamilyParametersContextWeightsProven on
google/owlv2-base-patch16owlvitabout 155 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
google/owlvit-base-patch32owlvitabout 153 M16F32 (safetensors)CPU, CUDA, Metal
IDEA-Research/grounding-dino-tinygrounding-dinoabout 172 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
openmmlab-community/mm_grounding_dino_tiny_o365v1_goldggrounding-dinoabout 173 M512F32 (safetensors)CPU, CUDA, Vulkan, Metal

Zero-shot image classification​

ModelFamilyParametersContextWeightsProven on
google/siglip-base-patch16-224siglipabout 203 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
google/siglip2-base-patch16-naflexsiglipabout 375 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
wkcn/TinyCLIP-ViT-8M-16-Text-3M-YFCC15Mclipabout 23 M77F32 (safetensors)CPU, CUDA, Vulkan, Metal

Image segmentation​

ModelFamilyParametersContextWeightsProven on
facebook/detr-resnet-50-panopticdetr-panoptic1,024torch (torch)CPU, CUDA, Vulkan, Metal
Roboflow/rf-detr-seg-nanorf-detr-segabout 34 MF32 (safetensors)CPU, CUDA, Vulkan, Metal

Depth estimation​

ModelFamilyParametersContextWeightsProven on
depth-anything/DA3-Smallda3about 34 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
depth-anything/Depth-Anything-V2-Small-hfdepth-anythingabout 25 MF32 (safetensors)CPU, CUDA, Vulkan, Metal
Intel/dpt-largedptabout 342 MF32 (safetensors)CPU, CUDA, Vulkan, Metal

Image to text​

ModelFamilyParametersContextWeightsProven on
zai-org/GLM-OCRglmabout 1.3 B131,072BF16 (safetensors)CPU, CUDA, Vulkan, Metal

Text to image​

ModelFamilyParametersContextWeightsProven on
Qwen/Qwen-Imageqwenimageabout 20.4 BBF16 (safetensors)CPU, CUDA, Vulkan
Tongyi-MAI/Z-Image-Turboz-imageabout 6.2 BF32 (safetensors)CPU, CUDA, Vulkan

Audio-language chat​

ModelFamilyParametersContextWeightsProven on
Qwen/Qwen2-Audio-7B-Instructqwen2-audioabout 8.4 BBF16 (safetensors)CPU, CUDA, Vulkan, Metal

Image to 3D​

ModelFamilyParametersContextWeightsProven on
depth-anything/DA3-BASEda3about 135 MF32 (safetensors)CPU, CUDA, Vulkan, Metal

Text to video​

ModelFamilyParametersContextWeightsProven on
Wan-AI/Wan2.1-T2V-1.3B-Diffuserswanabout 1.4 BF32 (safetensors)CPU, CUDA, Vulkan
Wan-AI/Wan2.2-TI2V-5B-Diffuserswanabout 5 BF32 (safetensors)CPU, CUDA, Vulkan

Other​

ModelFamilyParametersContextWeightsProven on
facebook/m2m100_418Mm2m-1001,024torch (torch)CPU, CUDA, Vulkan, Metal
fromziro/ZeroS-v0.1-150Mqwen3-5about 152 M2,048F32 (safetensors)CPU, CUDA, Metal
rpatel622/mamba2-130m-hf-Q8_0-GGUFmamba2about 168 Mmamba2-130m-hf-q8_0 (gguf), mamba2-130m-q8_0 (gguf)CPU, CUDA, Vulkan