---
title: Supported models
description: "The models the release serves, read off its supported-models document: one row per model with its weight variants and the backends the release's run record covers."
---

<!-- GENERATED at build time by tools/supported_models_page.py from inputs/clikart/supported_models.csv; never edit or commit. -->

The release ships its supported-models document beside the archives: `supported_models.json`, and `supported_models.csv`, the same document flat. It lists 109 checkpoints over 56 model families with 312 weight variants. This page lists the 80 of them whose weights carry a permissive license, the document's own class; the model card is where a checkpoint's license is read, and the runtime prints it at fetch, at load and in `info`. The whole document is a download of the platform with the release ([Get ClikaRT](/clikart/getting-started/get-clikart)).

A model is named by its Hugging Face repository id, which `clikart-cli fetch <id>` and `AutoModel.from_pretrained(id)` take as they are. **Parameters** and **Context** are read off the checkpoint. **Weights** lists the weight variants the release serves for the model, base first, with their format. **Proven on** names the backends the release's own run record covers for the model; an empty cell is a model the release lists without a recorded run.

## Text generation and chat

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [arianraje/qwen3-4b-gdn-hybrid-stage1-align-dtfix](https://huggingface.co/arianraje/qwen3-4b-gdn-hybrid-stage1-align-dtfix) | `qwen3-next` | about 4.5 B | 40,960 | `BF16` (safetensors) | CPU, CUDA, Metal |
| [DavidAU/Qwen3-MOE-4x0.6B-2.4B-Writing-Thunder](https://huggingface.co/DavidAU/Qwen3-MOE-4x0.6B-2.4B-Writing-Thunder) | `qwen` | about 1.5 B | 40,960 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [microsoft/phi-2](https://huggingface.co/microsoft/phi-2) | `phi` | about 2.8 B | 2,048 | `F16` (safetensors) | CPU, CUDA, Metal |
| [microsoft/Phi-3-mini-4k-instruct](https://huggingface.co/microsoft/Phi-3-mini-4k-instruct) | `phi` | about 3.8 B | 4,096 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [microsoft/Phi-3.5-mini-instruct](https://huggingface.co/microsoft/Phi-3.5-mini-instruct) | `phi` | about 3.8 B | 131,072 | `IQ1_M` (gguf), `IQ1_S` (gguf), `IQ2_XS` (gguf), `IQ3_XS` (gguf), `IQ4_XS` (gguf), `Q2_K` (gguf), `Q3_K_L` (gguf), `Q3_K_M` (gguf), `Q3_K_S` (gguf), `Q4_K_M` (gguf), `Q4_K_S` (gguf), `Q5_K_M` (gguf), `Q5_K_S` (gguf), `Q6_K` (gguf), `Q8_0` (gguf) | CPU, CUDA, Vulkan |
| [moonshotai/Kimi-Linear-48B-A3B-Instruct](https://huggingface.co/moonshotai/Kimi-Linear-48B-A3B-Instruct) | `kimi-linear` | about 49.1 B |  | `BF16` (safetensors) | CPU |
| [openai/gpt-oss-20b](https://huggingface.co/openai/gpt-oss-20b) | `gpt-oss` | about 20.9 B | 131,072 | `U8` (safetensors), `MXFP4` (gguf) | CPU, CUDA, Vulkan |
| [OuteAI/Lite-Oute-1-300M](https://huggingface.co/OuteAI/Lite-Oute-1-300M) | `mistral` | about 300 M | 4,096 | `F32` (safetensors) | CPU, CUDA, Metal |
| [Qwen/Qwen2.5-0.5B-Instruct](https://huggingface.co/Qwen/Qwen2.5-0.5B-Instruct) | `qwen` | about 494 M | 32,768 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [Qwen/Qwen3-0.6B](https://huggingface.co/Qwen/Qwen3-0.6B) | `qwen` | about 752 M | 40,960 | `BF16` (gguf), `F8_E4M3` (safetensors), `IQ4_NL` (gguf), `IQ4_XS` (gguf), `Q2_K` (gguf), `Q2_K_L` (gguf), `Q3_K_M` (gguf), `Q3_K_S` (gguf), `Q4_0` (gguf), `Q4_1` (gguf), `Q4_K_M` (gguf), `Q4_K_S` (gguf), `Q5_K_M` (gguf), `Q5_K_S` (gguf), `Q6_K` (gguf), `Q8_0` (gguf), `UD-IQ1_M` (gguf), `UD-IQ1_S` (gguf), `UD-IQ2_M` (gguf), `UD-IQ2_XXS` (gguf), `UD-IQ3_XXS` (gguf), `UD-Q2_K_XL` (gguf), `UD-Q3_K_XL` (gguf), `UD-Q4_K_XL` (gguf), `UD-Q5_K_XL` (gguf), `UD-Q6_K_XL` (gguf), `UD-Q8_K_XL` (gguf) | CPU, CUDA, Vulkan |
| [Qwen/Qwen3-0.6B-Base](https://huggingface.co/Qwen/Qwen3-0.6B-Base) | `qwen` | about 596 M | 32,768 | `Qwen3-Embedding-0.6B-f16` (gguf), `Q8_0` (gguf) | CPU, CUDA, Vulkan |
| [SupraLabs/Supra2-100M-Base](https://huggingface.co/SupraLabs/Supra2-100M-Base) | `qwen` | about 101 M | 2,048 | `F32` (safetensors) | CPU, CUDA, Vulkan |
| [zai-org/GLM-4-9B-0414](https://huggingface.co/zai-org/GLM-4-9B-0414) | `glm` | about 9.4 B | 32,768 | `BF16` (safetensors), `IQ2_M` (gguf), `IQ3_M` (gguf), `IQ3_XS` (gguf), `IQ3_XXS` (gguf), `IQ4_NL` (gguf), `IQ4_XS` (gguf), `Q2_K` (gguf), `Q2_K_L` (gguf), `Q3_K_L` (gguf), `Q3_K_M` (gguf), `Q3_K_S` (gguf), `Q3_K_XL` (gguf), `Q4_0` (gguf), `Q4_1` (gguf), `Q4_K_L` (gguf), `Q4_K_M` (gguf), `Q4_K_S` (gguf), `Q5_K_L` (gguf), `Q5_K_M` (gguf), `Q5_K_S` (gguf), `Q6_K` (gguf), `Q6_K_L` (gguf), `Q8_0` (gguf), `THUDM_GLM-4-9B-0414-bf16` (gguf) | CPU, CUDA, Vulkan, Metal |
| [zai-org/GLM-4.5-Air](https://huggingface.co/zai-org/GLM-4.5-Air) | `glm` | about 110.5 B | 131,072 | `BF16` (safetensors) | CPU |

## Vision-language chat

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [baidu/Unlimited-OCR](https://huggingface.co/baidu/Unlimited-OCR) | `deepseek_ocr` | about 3.3 B | 32,768 | `BF16` (safetensors) | CPU, CUDA, Metal |
| [deepseek-ai/DeepSeek-OCR](https://huggingface.co/deepseek-ai/DeepSeek-OCR) | `deepseek_ocr` | about 3.3 B | 8,192 | `BF16` (safetensors) | CPU, CUDA, Metal |
| [deepseek-ai/DeepSeek-OCR-2](https://huggingface.co/deepseek-ai/DeepSeek-OCR-2) | `deepseek_ocr` | about 3.4 B | 8,192 | `BF16` (safetensors) | CPU, CUDA, Metal |
| [deepseek-community/DeepSeek-OCR-2](https://huggingface.co/deepseek-community/DeepSeek-OCR-2) | `deepseek_ocr` | about 3.4 B | 8,192 | `BF16` (safetensors) | CPU, CUDA, Metal |
| [florence-community/Florence-2-base](https://huggingface.co/florence-community/Florence-2-base) | `florence2` | about 232 M | 1,024 | `F16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [google/gemma-4-26B-A4B](https://huggingface.co/google/gemma-4-26B-A4B) | `gemma` | about 26.5 B | 262,144 | `BF16` (safetensors) | CPU, CUDA |
| [google/gemma-4-26B-A4B-it](https://huggingface.co/google/gemma-4-26B-A4B-it) | `gemma` | about 25.8 B | 262,144 | `BF16` (safetensors) |  |
| [google/gemma-4-26B-A4B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-26B-A4B-it-qat-q4_0-unquantized) | `gemma` | about 26.5 B | 262,144 | `BF16` (safetensors) | CUDA |
| [google/gemma-4-31B](https://huggingface.co/google/gemma-4-31B) | `gemma` | about 32.7 B | 262,144 | `BF16` (safetensors) | CPU, CUDA |
| [google/gemma-4-31B-it](https://huggingface.co/google/gemma-4-31B-it) | `gemma` | about 31.3 B | 262,144 | `BF16` (safetensors) | CPU, CUDA, Vulkan |
| [google/gemma-4-31B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-31B-it-qat-q4_0-unquantized) | `gemma` | about 32.7 B | 262,144 | `BF16` (safetensors), `I32` (safetensors), `gemma-4-31B_q4_0-it` (gguf) | CPU, CUDA, Vulkan |
| [meta-models/Muse-Glimmer-30B](https://huggingface.co/meta-models/Muse-Glimmer-30B) | `muse-glimmer` | about 29.8 B | 131,072 | `BF16` (safetensors) | CPU, CUDA, Vulkan |
| [microsoft/Florence-2-base](https://huggingface.co/microsoft/Florence-2-base) | `florence2` | about 232 M | 1,024 | `F16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [Qwen/Qwen2-VL-2B-Instruct](https://huggingface.co/Qwen/Qwen2-VL-2B-Instruct) | `qwen-vl` | about 2.2 B | 32,768 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [Qwen/Qwen3-VL-2B-Instruct](https://huggingface.co/Qwen/Qwen3-VL-2B-Instruct) | `qwen-vl` | about 2.1 B | 262,144 | `BF16` (safetensors), `F16` (gguf), `Q4_K_M` (gguf), `Q8_0` (gguf) | CPU, CUDA, Vulkan, Metal |
| [Qwen/Qwen3-VL-30B-A3B-Instruct](https://huggingface.co/Qwen/Qwen3-VL-30B-A3B-Instruct) | `qwen-vl` | about 31.1 B | 262,144 | `F8_E4M3` (safetensors) | CPU, CUDA, Vulkan |
| [Qwen/Qwen3.5-0.8B](https://huggingface.co/Qwen/Qwen3.5-0.8B) | `qwen3-5` | about 873 M | 262,144 | `BF16` (safetensors), `IQ4_NL` (gguf), `IQ4_XS` (gguf), `Q3_K_M` (gguf), `Q3_K_S` (gguf), `Q4_0` (gguf), `Q4_1` (gguf), `Q4_K_M` (gguf), `Q4_K_S` (gguf), `Q5_K_M` (gguf), `Q5_K_S` (gguf), `Q6_K` (gguf), `Q8_0` (gguf), `UD-IQ2_M` (gguf), `UD-IQ2_XXS` (gguf), `UD-IQ3_XXS` (gguf), `UD-Q2_K_XL` (gguf), `UD-Q3_K_XL` (gguf), `UD-Q4_K_XL` (gguf), `UD-Q5_K_XL` (gguf), `UD-Q6_K_XL` (gguf), `UD-Q8_K_XL` (gguf) | CPU, CUDA, Vulkan, Metal |
| [Qwen/Qwen3.5-35B-A3B](https://huggingface.co/Qwen/Qwen3.5-35B-A3B) | `qwen3-5-moe` | about 36 B | 262,144 | `BF16` (gguf), `I32` (safetensors), `MXFP4_MOE` (gguf), `Q3_K_M` (gguf), `Q3_K_S` (gguf), `Q4_K_M` (gguf), `Q4_K_S` (gguf), `Q5_K_M` (gguf), `Q5_K_S` (gguf), `Q6_K` (gguf), `Q8_0` (gguf), `UD-IQ2_M` (gguf), `UD-IQ2_XXS` (gguf), `UD-IQ3_S` (gguf), `UD-IQ3_XXS` (gguf), `UD-IQ4_NL` (gguf), `UD-IQ4_XS` (gguf), `UD-Q2_K_XL` (gguf), `UD-Q3_K_XL` (gguf), `UD-Q4_K_L` (gguf), `UD-Q4_K_XL` (gguf), `UD-Q5_K_XL` (gguf), `UD-Q6_K_S` (gguf), `UD-Q6_K_XL` (gguf), `UD-Q8_K_XL` (gguf) | CPU, CUDA, Vulkan |
| [thinkingmachines/Inkling-Small](https://huggingface.co/thinkingmachines/Inkling-Small) | `inkling` | about 266 B |  | `U8` (safetensors) | CPU |

## Multimodal (any to any)

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [google/gemma-4-12B-it](https://huggingface.co/google/gemma-4-12B-it) | `gemma` | about 12 B | 262,144 | `BF16` (safetensors) | CPU, CUDA, Vulkan |
| [google/gemma-4-12B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized) | `gemma` | about 12 B | 262,144 | `gemma-4-12b-it-qat-q4_0` (gguf) | CPU, CUDA, Vulkan |
| [google/gemma-4-E2B-it](https://huggingface.co/google/gemma-4-E2B-it) | `gemma` | about 5.1 B | 131,072 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [google/gemma-4-E2B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-E2B-it-qat-q4_0-unquantized) | `gemma` | about 5.1 B | 131,072 | `gemma-4-E2B_q4_0-it` (gguf) | CPU, CUDA, Vulkan |
| [google/gemma-4-E4B-it](https://huggingface.co/google/gemma-4-E4B-it) | `gemma` | about 8 B | 131,072 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [google/gemma-4-E4B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-E4B-it-qat-q4_0-unquantized) | `gemma` | about 7.9 B | 131,072 | `gemma-4-E4B_q4_0-it` (gguf) | CPU, CUDA, Vulkan |

## Speech to text

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [nvidia/canary-1b-v2](https://huggingface.co/nvidia/canary-1b-v2) | `canary` | about 979 M |  | `F32` (safetensors) |  |
| [nvidia/parakeet-ctc-0.6b](https://huggingface.co/nvidia/parakeet-ctc-0.6b) | `parakeet` | about 609 M |  | `F32` (safetensors) |  |
| [nvidia/parakeet-rnnt-0.6b](https://huggingface.co/nvidia/parakeet-rnnt-0.6b) | `parakeet` | about 617 M |  | `F32` (safetensors) |  |
| [nvidia/parakeet-tdt-0.6b-v3](https://huggingface.co/nvidia/parakeet-tdt-0.6b-v3) | `parakeet` | about 627 M |  | `F32` (safetensors) |  |
| [openai/whisper-large-v3-turbo](https://huggingface.co/openai/whisper-large-v3-turbo) | `whisper` | about 809 M |  | `whisper-large-v3-turbo-q4_0` (gguf), `whisper-large-v3-turbo-q4_1` (gguf), `whisper-large-v3-turbo-q8_0` (gguf) | CPU, CUDA |
| [openai/whisper-tiny](https://huggingface.co/openai/whisper-tiny) | `whisper` | about 38 M | 448 | `F32` (safetensors), `F16` (gguf), `Q4_K_M` (gguf), `Q5_K_M` (gguf), `Q6_K` (gguf), `Q8_0` (gguf) | CPU, CUDA, Vulkan, Metal |

## Text to speech

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [sesame/csm-1b](https://huggingface.co/sesame/csm-1b) | `csm` | about 1.6 B | 2,048 | `F32` (safetensors) | CPU, CUDA, Metal |

## Translation

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [tencent/Hy-MT2-1.8B](https://huggingface.co/tencent/Hy-MT2-1.8B) | `hunyuan` | about 2 B | 262,144 | `BF16` (safetensors), `Q4_K_M` (gguf), `Q6_K` (gguf), `Q8_0` (gguf) | CPU, CUDA, Vulkan, Metal |
| [tencent/Hy-MT2-30B-A3B](https://huggingface.co/tencent/Hy-MT2-30B-A3B) | `hy-v3` | about 30.1 B | 262,144 | `F8_E4M3` (safetensors) | CPU, CUDA, Vulkan |

## Text embedding

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [TaylorAI/bge-micro-v2](https://huggingface.co/TaylorAI/bge-micro-v2) | `bert` | about 17 M | 512 | `F16` (safetensors) | Metal |

## Feature extraction

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [Qwen/Qwen3-Embedding-0.6B](https://huggingface.co/Qwen/Qwen3-Embedding-0.6B) | `qwen-embedding` | about 596 M | 32,768 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Image embedding

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [facebook/dinov2-small](https://huggingface.co/facebook/dinov2-small) | `dinov2` | about 22 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Reranking

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [Qwen/Qwen3-Reranker-0.6B](https://huggingface.co/Qwen/Qwen3-Reranker-0.6B) | `qwen-reranker` | about 596 M | 40,960 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Object detection

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [facebook/detr-resnet-50](https://huggingface.co/facebook/detr-resnet-50) | `detr` | about 42 M | 1,024 | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [microsoft/table-transformer-detection](https://huggingface.co/microsoft/table-transformer-detection) | `detr` | about 29 M | 1,024 | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [PekingU/rtdetr_r18vd](https://huggingface.co/PekingU/rtdetr_r18vd) | `rt-detr` | about 20 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [PekingU/rtdetr_v2_r18vd](https://huggingface.co/PekingU/rtdetr_v2_r18vd) | `rt-detr` | about 20 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [Roboflow/rf-detr-nano](https://huggingface.co/Roboflow/rf-detr-nano) | `rf-detr` | about 30 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [ustc-community/dfine-nano-coco](https://huggingface.co/ustc-community/dfine-nano-coco) | `d-fine` | about 4 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Open-vocabulary object detection

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [google/owlv2-base-patch16](https://huggingface.co/google/owlv2-base-patch16) | `owlvit` | about 155 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [google/owlvit-base-patch32](https://huggingface.co/google/owlvit-base-patch32) | `owlvit` | about 153 M | 16 | `F32` (safetensors) | CPU, CUDA, Metal |
| [IDEA-Research/grounding-dino-tiny](https://huggingface.co/IDEA-Research/grounding-dino-tiny) | `grounding-dino` | about 172 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [openmmlab-community/mm_grounding_dino_tiny_o365v1_goldg](https://huggingface.co/openmmlab-community/mm_grounding_dino_tiny_o365v1_goldg) | `grounding-dino` | about 173 M | 512 | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Zero-shot image classification

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [google/siglip-base-patch16-224](https://huggingface.co/google/siglip-base-patch16-224) | `siglip` | about 203 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [google/siglip2-base-patch16-naflex](https://huggingface.co/google/siglip2-base-patch16-naflex) | `siglip` | about 375 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [wkcn/TinyCLIP-ViT-8M-16-Text-3M-YFCC15M](https://huggingface.co/wkcn/TinyCLIP-ViT-8M-16-Text-3M-YFCC15M) | `clip` | about 23 M | 77 | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Image segmentation

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [facebook/detr-resnet-50-panoptic](https://huggingface.co/facebook/detr-resnet-50-panoptic) | `detr-panoptic` |  | 1,024 | `torch` (torch) | CPU, CUDA, Vulkan, Metal |
| [Roboflow/rf-detr-seg-nano](https://huggingface.co/Roboflow/rf-detr-seg-nano) | `rf-detr-seg` | about 34 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Depth estimation

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [depth-anything/DA3-Small](https://huggingface.co/depth-anything/DA3-Small) | `da3` | about 34 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [depth-anything/Depth-Anything-V2-Small-hf](https://huggingface.co/depth-anything/Depth-Anything-V2-Small-hf) | `depth-anything` | about 25 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |
| [Intel/dpt-large](https://huggingface.co/Intel/dpt-large) | `dpt` | about 342 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Image to text

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [zai-org/GLM-OCR](https://huggingface.co/zai-org/GLM-OCR) | `glm` | about 1.3 B | 131,072 | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Text to image

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [Qwen/Qwen-Image](https://huggingface.co/Qwen/Qwen-Image) | `qwenimage` | about 20.4 B |  | `BF16` (safetensors) | CPU, CUDA, Vulkan |
| [Tongyi-MAI/Z-Image-Turbo](https://huggingface.co/Tongyi-MAI/Z-Image-Turbo) | `z-image` | about 6.2 B |  | `F32` (safetensors) | CPU, CUDA, Vulkan |

## Audio-language chat

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [Qwen/Qwen2-Audio-7B-Instruct](https://huggingface.co/Qwen/Qwen2-Audio-7B-Instruct) | `qwen2-audio` | about 8.4 B |  | `BF16` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Image to 3D

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [depth-anything/DA3-BASE](https://huggingface.co/depth-anything/DA3-BASE) | `da3` | about 135 M |  | `F32` (safetensors) | CPU, CUDA, Vulkan, Metal |

## Text to video

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [Wan-AI/Wan2.1-T2V-1.3B-Diffusers](https://huggingface.co/Wan-AI/Wan2.1-T2V-1.3B-Diffusers) | `wan` | about 1.4 B |  | `F32` (safetensors) | CPU, CUDA, Vulkan |
| [Wan-AI/Wan2.2-TI2V-5B-Diffusers](https://huggingface.co/Wan-AI/Wan2.2-TI2V-5B-Diffusers) | `wan` | about 5 B |  | `F32` (safetensors) | CPU, CUDA, Vulkan |

## Other

| Model | Family | Parameters | Context | Weights | Proven on |
| --- | --- | --- | --- | --- | --- |
| [facebook/m2m100_418M](https://huggingface.co/facebook/m2m100_418M) | `m2m-100` |  | 1,024 | `torch` (torch) | CPU, CUDA, Vulkan, Metal |
| [fromziro/ZeroS-v0.1-150M](https://huggingface.co/fromziro/ZeroS-v0.1-150M) | `qwen3-5` | about 152 M | 2,048 | `F32` (safetensors) | CPU, CUDA, Metal |
| [rpatel622/mamba2-130m-hf-Q8_0-GGUF](https://huggingface.co/rpatel622/mamba2-130m-hf-Q8_0-GGUF) | `mamba2` | about 168 M |  | `mamba2-130m-hf-q8_0` (gguf), `mamba2-130m-q8_0` (gguf) | CPU, CUDA, Vulkan |
