Model Format Support
InvokeAI recognizes a model’s format from the file itself, so you install every build the same way. This page lists which formats load for which model. What each format is, and how it behaves in memory, is explained under Quantized formats. Sizes, starter models and measured speeds are on each family’s own page.
Legend:
- ✓ loads
- L installs, but loading it for generation fails with a message naming the reason
- R the install fails with a message naming the reason
- ✗ not supported
- – no such build exists
How an unsupported file fails depends on the file. Some are not recognized at all and install as an unknown model, with the reason only in the server log: the Diffusers layout of LTX-2 and the GGUF builds of LTX-2, MiniMax H3 and the Ministral 3B encoder. Others are recognized by their layer names and fail, or misbehave, only when they load.
Main models
Section titled “Main models”| Family | Diffusers | Single file (bf16) | fp8 scaled | int8 | nvfp4 | MXFP8 | GGUF | SDNQ | NF4 |
|---|---|---|---|---|---|---|---|---|---|
| SD 1.x / 2.x, SDXL | ✓ | ✓ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ |
| SD 3 | ✓ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ |
| FLUX.1 | ✓ | ✓ | ✓ | ✓ | ✗ | ✓ | ✓ | ✓ | ✓ |
| FLUX.2 | ✓ | ✓ | ✓ | ✓ | ✗ | ✓ | ✓ | ✓ | inside the Diffusers folder (FLUX.2 [dev]) |
| CogView4 | ✓ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ |
| Z-Image | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✗ |
| Qwen-Image | ✓ | ✓ | ✓ | ✗ | ✓ | ✓ | ✓ | ✗ | ✗ |
| Anima | – | ✓ | ✓ | ✓ | ✗ | ✓ | ✗ | ✗ | ✗ |
| Krea-2 | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✗ | ✗ |
| Ideogram 4 | ✓ | ✓ | ✓ | ✓ | R | ✓ | ✓ | ✗ | inside the Diffusers folder |
| ERNIE-Image | ✓ | ✓ | L | L | L | ✗ | ✗ | ✗ | ✗ |
| Wan 2.2 | ✓ | ✓ | ✓ | L | L | L | ✓ | ✗ | ✗ |
| LTX-2 (2.5 only) | official layout ✓, Diffusers layout ✗ | ✓ | L | ✓ | ✓ | ✗ | ✗ | ✗ | ✗ |
| MiniMax H3 | ✓ | ✓ | L | ✓ | L | ✗ | ✗ | ✗ | ✗ |
- int8 means ComfyUI’s
int8_convrot(int8_tensorwise) builds. nvfp4 and fp8 scaled mean ComfyUI’s checkpoint layouts; other exports of the same schemes (AWQ, ModelOpt) fail when the model loads. - GGUF covers the llama.cpp types (Q4_0 through Q8_0, the K-quants, BF16). ComfyUI-GGUF’s
Q8_CRfiles load for Krea-2 only; for every other model with GGUF support the install fails. - LTX-2.0/2.3 files fail to install with a message. Wan 2.1, Wan Animate, S2V, Fun-Control and VACE single files and GGUFs are not recognized as Wan 2.2 and install as unknown models; a Wan 2.1 Diffusers folder is not detected and is not supported either.
Text encoders
Section titled “Text encoders”| Encoder | Used by | Folder | Single file (bf16) | fp8 scaled | int8 | nvfp4 | GGUF | SDNQ |
|---|---|---|---|---|---|---|---|---|
| T5-XXL | FLUX.1, SD 3 | ✓ | ✗ | ✗ | ✗ | ✗ | ✓ | ✓ (and bitsandbytes int8) |
| CLIP-L / CLIP-G | SD, SDXL, FLUX.1, SD 3 | ✓ | ✗ | ✗ | ✗ | ✗ | ✗ | CLIP-L inside a FLUX.1 SDNQ pipeline |
| Qwen3 0.6B / 4B / 8B | Z-Image, FLUX.2 Klein, Anima | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ | ✓ |
| Qwen3-VL 4B / 8B | Krea-2, Ideogram 4 | ✓ | ✓ | ✓ | ✓ | ✗ | ✓ | ✗ |
| Qwen3-VL 32B (truncated) | MiniMax H3 | in the Diffusers folder | ✓ | L | ✓ | L | ✗ | ✗ |
| Qwen3.5 4B | Anima 3.8B | ✗ | ✓ | L | L | L | ✗ | ✗ |
| Qwen2.5-VL | Qwen-Image | ✓ | ✓ | ✓ | L | ✓ | ✗ | ✗ |
| Mistral Small 3 | FLUX.2 dev | ✓ | ✓ | ✓ | L | ✓ | ✓ | ✗ |
| Ministral 3B | ERNIE-Image | ✓ | ✓ | ✓ | L | L | ✗ | ✗ |
| Gemma-2 2B | PiD decoders | ✓ | ✗ | ✗ | ✗ | ✗ | ✓ | ✗ |
| Gemma-4 12B | LTX-2 | ✓ (bf16 or int8) | ✗ | L | in the folder | L | ✗ | ✗ |
| UMT5 | Wan 2.2 | ✓ | ✗ | ✗ | ✗ | ✗ | ✗ | ✗ |
The Qwen3-VL GGUF encoder is the language model only; a separate mmproj vision file is refused, and split
multi-part GGUFs have to be joined before installing. The Gemma-4 encoder installs only as a folder: its
config.json and tokenizer files next to exactly one weight file, bf16 or int8. A standalone .safetensors file is
not recognized.
| VAE | Single file | Diffusers | Used by |
|---|---|---|---|
| SD 1.x, SDXL, SD 3, FLUX.1, FLUX.2 | ✓ | ✓ | its own family; FLUX.1’s also by Z-Image, FLUX.2’s also by Ideogram 4 and ERNIE-Image |
| SD 2.x | ✓ | ✗ | SD 2.x |
| Wan (16 and 48 channels) | ✓ | ✓ | Wan 2.2; the Wan 2.1 VAE also by Anima |
| Qwen-Image | ✓ | ✗ | Qwen-Image, Krea-2 |
| CogView4, LTX-2, MiniMax H3 | – | bundled | taken from the model’s own folder |
Quantized VAE files (fp8, int8, nvfp4) are not supported; use the unquantized VAE. The SD 3, FLUX.2, Qwen-Image, Wan, Anima and Ideogram 4 VAEs fail with a message when they load. The FLUX.1 and SD 1.x / SDXL single-file VAEs do not check, so a quantized file there fails with an unclear error or decodes incorrectly.
LoRAs and adapters
Section titled “LoRAs and adapters”| Family | LoRA | ControlNet | Other adapters |
|---|---|---|---|
| SD 1.x / 2.x | ✓ (LyCORIS, Diffusers) | ✓ | IP-Adapter, T2I-Adapter (SD 1.x), textual inversion |
| SDXL | ✓ (LyCORIS, Diffusers, OMI) | ✓ | IP-Adapter, T2I-Adapter, textual inversion |
| FLUX.1 | ✓ (LyCORIS, Diffusers, OMI) | ✓ | Control LoRA, IP-Adapter, Redux |
| FLUX.2, Z-Image | ✓ (LyCORIS, Diffusers) | Z-Image only | |
| Qwen-Image, Krea-2, Wan 2.2, Anima | ✓ (LyCORIS/kohya, including LoKR where noted on the family page) | Anima (ControlNet-LLLite) | |
| LTX-2, MiniMax H3 | ✓ (plain low-rank; LoKR, LoHA and DoRA are refused) | ✗ | |
| SD 3, CogView4, Ideogram 4, ERNIE-Image | ✗ | ✗ |
PiD decoders load for FLUX.1, FLUX.2, SD 3, SDXL and Qwen-Image (v1 and v1.5, bf16 and int8 builds).