Supported models
The release ships its supported-models document beside the archives: supported_models.json, and supported_models.csv, the same document flat. It lists 109 checkpoints over 56 model families with 312 weight variants. This page lists the 80 of them whose weights carry a permissive license, the document's own class; the model card is where a checkpoint's license is read, and the runtime prints it at fetch, at load and in info. The whole document is a download of the platform with the release (Get ClikaRT).
A model is named by its Hugging Face repository id, which clikart-cli fetch <id> and AutoModel.from_pretrained(id) take as they are. Parameters and Context are read off the checkpoint. Weights lists the weight variants the release serves for the model, base first, with their format. Proven on names the backends the release's own run record covers for the model; an empty cell is a model the release lists without a recorded run.
Text generation and chat
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| arianraje/qwen3-4b-gdn-hybrid-stage1-align-dtfix | qwen3-next | about 4.5 B | 40,960 | BF16 (safetensors) | CPU, CUDA, Metal |
| DavidAU/Qwen3-MOE-4x0.6B-2.4B-Writing-Thunder | qwen | about 1.5 B | 40,960 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| microsoft/phi-2 | phi | about 2.8 B | 2,048 | F16 (safetensors) | CPU, CUDA, Metal |
| microsoft/Phi-3-mini-4k-instruct | phi | about 3.8 B | 4,096 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| microsoft/Phi-3.5-mini-instruct | phi | about 3.8 B | 131,072 | IQ1_M (gguf), IQ1_S (gguf), IQ2_XS (gguf), IQ3_XS (gguf), IQ4_XS (gguf), Q2_K (gguf), Q3_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf) | CPU, CUDA, Vulkan |
| moonshotai/Kimi-Linear-48B-A3B-Instruct | kimi-linear | about 49.1 B | BF16 (safetensors) | CPU | |
| openai/gpt-oss-20b | gpt-oss | about 20.9 B | 131,072 | U8 (safetensors), MXFP4 (gguf) | CPU, CUDA, Vulkan |
| OuteAI/Lite-Oute-1-300M | mistral | about 300 M | 4,096 | F32 (safetensors) | CPU, CUDA, Metal |
| Qwen/Qwen2.5-0.5B-Instruct | qwen | about 494 M | 32,768 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| Qwen/Qwen3-0.6B | qwen | about 752 M | 40,960 | BF16 (gguf), F8_E4M3 (safetensors), IQ4_NL (gguf), IQ4_XS (gguf), Q2_K (gguf), Q2_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ1_M (gguf), UD-IQ1_S (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_XXS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf) | CPU, CUDA, Vulkan |
| Qwen/Qwen3-0.6B-Base | qwen | about 596 M | 32,768 | Qwen3-Embedding-0.6B-f16 (gguf), Q8_0 (gguf) | CPU, CUDA, Vulkan |
| SupraLabs/Supra2-100M-Base | qwen | about 101 M | 2,048 | F32 (safetensors) | CPU, CUDA, Vulkan |
| zai-org/GLM-4-9B-0414 | glm | about 9.4 B | 32,768 | BF16 (safetensors), IQ2_M (gguf), IQ3_M (gguf), IQ3_XS (gguf), IQ3_XXS (gguf), IQ4_NL (gguf), IQ4_XS (gguf), Q2_K (gguf), Q2_K_L (gguf), Q3_K_L (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q3_K_XL (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_L (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_L (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q6_K_L (gguf), Q8_0 (gguf), THUDM_GLM-4-9B-0414-bf16 (gguf) | CPU, CUDA, Vulkan, Metal |
| zai-org/GLM-4.5-Air | glm | about 110.5 B | 131,072 | BF16 (safetensors) | CPU |
Vision-language chat
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| baidu/Unlimited-OCR | deepseek_ocr | about 3.3 B | 32,768 | BF16 (safetensors) | CPU, CUDA, Metal |
| deepseek-ai/DeepSeek-OCR | deepseek_ocr | about 3.3 B | 8,192 | BF16 (safetensors) | CPU, CUDA, Metal |
| deepseek-ai/DeepSeek-OCR-2 | deepseek_ocr | about 3.4 B | 8,192 | BF16 (safetensors) | CPU, CUDA, Metal |
| deepseek-community/DeepSeek-OCR-2 | deepseek_ocr | about 3.4 B | 8,192 | BF16 (safetensors) | CPU, CUDA, Metal |
| florence-community/Florence-2-base | florence2 | about 232 M | 1,024 | F16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| google/gemma-4-26B-A4B | gemma | about 26.5 B | 262,144 | BF16 (safetensors) | CPU, CUDA |
| google/gemma-4-26B-A4B-it | gemma | about 25.8 B | 262,144 | BF16 (safetensors) | |
| google/gemma-4-26B-A4B-it-qat-q4_0-unquantized | gemma | about 26.5 B | 262,144 | BF16 (safetensors) | CUDA |
| google/gemma-4-31B | gemma | about 32.7 B | 262,144 | BF16 (safetensors) | CPU, CUDA |
| google/gemma-4-31B-it | gemma | about 31.3 B | 262,144 | BF16 (safetensors) | CPU, CUDA, Vulkan |
| google/gemma-4-31B-it-qat-q4_0-unquantized | gemma | about 32.7 B | 262,144 | BF16 (safetensors), I32 (safetensors), gemma-4-31B_q4_0-it (gguf) | CPU, CUDA, Vulkan |
| meta-models/Muse-Glimmer-30B | muse-glimmer | about 29.8 B | 131,072 | BF16 (safetensors) | CPU, CUDA, Vulkan |
| microsoft/Florence-2-base | florence2 | about 232 M | 1,024 | F16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| Qwen/Qwen2-VL-2B-Instruct | qwen-vl | about 2.2 B | 32,768 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| Qwen/Qwen3-VL-2B-Instruct | qwen-vl | about 2.1 B | 262,144 | BF16 (safetensors), F16 (gguf), Q4_K_M (gguf), Q8_0 (gguf) | CPU, CUDA, Vulkan, Metal |
| Qwen/Qwen3-VL-30B-A3B-Instruct | qwen-vl | about 31.1 B | 262,144 | F8_E4M3 (safetensors) | CPU, CUDA, Vulkan |
| Qwen/Qwen3.5-0.8B | qwen3-5 | about 873 M | 262,144 | BF16 (safetensors), IQ4_NL (gguf), IQ4_XS (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_0 (gguf), Q4_1 (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_XXS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf) | CPU, CUDA, Vulkan, Metal |
| Qwen/Qwen3.5-35B-A3B | qwen3-5-moe | about 36 B | 262,144 | BF16 (gguf), I32 (safetensors), MXFP4_MOE (gguf), Q3_K_M (gguf), Q3_K_S (gguf), Q4_K_M (gguf), Q4_K_S (gguf), Q5_K_M (gguf), Q5_K_S (gguf), Q6_K (gguf), Q8_0 (gguf), UD-IQ2_M (gguf), UD-IQ2_XXS (gguf), UD-IQ3_S (gguf), UD-IQ3_XXS (gguf), UD-IQ4_NL (gguf), UD-IQ4_XS (gguf), UD-Q2_K_XL (gguf), UD-Q3_K_XL (gguf), UD-Q4_K_L (gguf), UD-Q4_K_XL (gguf), UD-Q5_K_XL (gguf), UD-Q6_K_S (gguf), UD-Q6_K_XL (gguf), UD-Q8_K_XL (gguf) | CPU, CUDA, Vulkan |
| thinkingmachines/Inkling-Small | inkling | about 266 B | U8 (safetensors) | CPU |
Multimodal (any to any)
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| google/gemma-4-12B-it | gemma | about 12 B | 262,144 | BF16 (safetensors) | CPU, CUDA, Vulkan |
| google/gemma-4-12B-it-qat-q4_0-unquantized | gemma | about 12 B | 262,144 | gemma-4-12b-it-qat-q4_0 (gguf) | CPU, CUDA, Vulkan |
| google/gemma-4-E2B-it | gemma | about 5.1 B | 131,072 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| google/gemma-4-E2B-it-qat-q4_0-unquantized | gemma | about 5.1 B | 131,072 | gemma-4-E2B_q4_0-it (gguf) | CPU, CUDA, Vulkan |
| google/gemma-4-E4B-it | gemma | about 8 B | 131,072 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
| google/gemma-4-E4B-it-qat-q4_0-unquantized | gemma | about 7.9 B | 131,072 | gemma-4-E4B_q4_0-it (gguf) | CPU, CUDA, Vulkan |
Speech to text
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| nvidia/canary-1b-v2 | canary | about 979 M | F32 (safetensors) | ||
| nvidia/parakeet-ctc-0.6b | parakeet | about 609 M | F32 (safetensors) | ||
| nvidia/parakeet-rnnt-0.6b | parakeet | about 617 M | F32 (safetensors) | ||
| nvidia/parakeet-tdt-0.6b-v3 | parakeet | about 627 M | F32 (safetensors) | ||
| openai/whisper-large-v3-turbo | whisper | about 809 M | whisper-large-v3-turbo-q4_0 (gguf), whisper-large-v3-turbo-q4_1 (gguf), whisper-large-v3-turbo-q8_0 (gguf) | CPU, CUDA | |
| openai/whisper-tiny | whisper | about 38 M | 448 | F32 (safetensors), F16 (gguf), Q4_K_M (gguf), Q5_K_M (gguf), Q6_K (gguf), Q8_0 (gguf) | CPU, CUDA, Vulkan, Metal |
Text to speech
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| sesame/csm-1b | csm | about 1.6 B | 2,048 | F32 (safetensors) | CPU, CUDA, Metal |
Translation
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| tencent/Hy-MT2-1.8B | hunyuan | about 2 B | 262,144 | BF16 (safetensors), Q4_K_M (gguf), Q6_K (gguf), Q8_0 (gguf) | CPU, CUDA, Vulkan, Metal |
| tencent/Hy-MT2-30B-A3B | hy-v3 | about 30.1 B | 262,144 | F8_E4M3 (safetensors) | CPU, CUDA, Vulkan |
Text embedding
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| TaylorAI/bge-micro-v2 | bert | about 17 M | 512 | F16 (safetensors) | Metal |
Feature extraction
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| Qwen/Qwen3-Embedding-0.6B | qwen-embedding | about 596 M | 32,768 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
Image embedding
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| facebook/dinov2-small | dinov2 | about 22 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Reranking
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| Qwen/Qwen3-Reranker-0.6B | qwen-reranker | about 596 M | 40,960 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
Object detection
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| facebook/detr-resnet-50 | detr | about 42 M | 1,024 | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
| microsoft/table-transformer-detection | detr | about 29 M | 1,024 | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
| PekingU/rtdetr_r18vd | rt-detr | about 20 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| PekingU/rtdetr_v2_r18vd | rt-detr | about 20 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| Roboflow/rf-detr-nano | rf-detr | about 30 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| ustc-community/dfine-nano-coco | d-fine | about 4 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Open-vocabulary object detection
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| google/owlv2-base-patch16 | owlvit | about 155 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| google/owlvit-base-patch32 | owlvit | about 153 M | 16 | F32 (safetensors) | CPU, CUDA, Metal |
| IDEA-Research/grounding-dino-tiny | grounding-dino | about 172 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| openmmlab-community/mm_grounding_dino_tiny_o365v1_goldg | grounding-dino | about 173 M | 512 | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Zero-shot image classification
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| google/siglip-base-patch16-224 | siglip | about 203 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| google/siglip2-base-patch16-naflex | siglip | about 375 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| wkcn/TinyCLIP-ViT-8M-16-Text-3M-YFCC15M | clip | about 23 M | 77 | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Image segmentation
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| facebook/detr-resnet-50-panoptic | detr-panoptic | 1,024 | torch (torch) | CPU, CUDA, Vulkan, Metal | |
| Roboflow/rf-detr-seg-nano | rf-detr-seg | about 34 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Depth estimation
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| depth-anything/DA3-Small | da3 | about 34 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| depth-anything/Depth-Anything-V2-Small-hf | depth-anything | about 25 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal | |
| Intel/dpt-large | dpt | about 342 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Image to text
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| zai-org/GLM-OCR | glm | about 1.3 B | 131,072 | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
Text to image
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| Qwen/Qwen-Image | qwenimage | about 20.4 B | BF16 (safetensors) | CPU, CUDA, Vulkan | |
| Tongyi-MAI/Z-Image-Turbo | z-image | about 6.2 B | F32 (safetensors) | CPU, CUDA, Vulkan |
Audio-language chat
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| Qwen/Qwen2-Audio-7B-Instruct | qwen2-audio | about 8.4 B | BF16 (safetensors) | CPU, CUDA, Vulkan, Metal |
Image to 3D
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| depth-anything/DA3-BASE | da3 | about 135 M | F32 (safetensors) | CPU, CUDA, Vulkan, Metal |
Text to video
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| Wan-AI/Wan2.1-T2V-1.3B-Diffusers | wan | about 1.4 B | F32 (safetensors) | CPU, CUDA, Vulkan | |
| Wan-AI/Wan2.2-TI2V-5B-Diffusers | wan | about 5 B | F32 (safetensors) | CPU, CUDA, Vulkan |
Other
| Model | Family | Parameters | Context | Weights | Proven on |
|---|---|---|---|---|---|
| facebook/m2m100_418M | m2m-100 | 1,024 | torch (torch) | CPU, CUDA, Vulkan, Metal | |
| fromziro/ZeroS-v0.1-150M | qwen3-5 | about 152 M | 2,048 | F32 (safetensors) | CPU, CUDA, Metal |
| rpatel622/mamba2-130m-hf-Q8_0-GGUF | mamba2 | about 168 M | mamba2-130m-hf-q8_0 (gguf), mamba2-130m-q8_0 (gguf) | CPU, CUDA, Vulkan |