Repository navigation
docs(models): add a format support matrix for every model type - #9726
Pfannkuchensack wants to merge 6 commits into
Conversation
New Model Format Support page lists which formats load, are refused or do not exist for every main model, text encoder, VAE and adapter family. Model Families links to it from its quantized-formats section, and the new-model checklist asks to keep it current.
joshistoast
left a comment
There was a problem hiding this comment.
[P2] Mark standalone Gemma-4 files as unsupported. The matrix advertises single-file bf16 support, but Gemma-4 requires a folder containing configuration, tokenizer, and weight files. Clarify that bf16 and int8 weights require those accompanying files.
[P2] Distinguish installation refusal from loading refusal. The legend says R formats are refused during installation, but some—including ERNIE-Image quantization and Qwen3.5 cases—are rejected only when loading for generation. Remove the installation-specific wording or distinguish these stages.
…mat matrix Mark formats that install but fail on load as L, keep R for install-time refusals, and mark unrecognized files ✗. Gemma-4 installs only as a folder with its config and tokenizer; quantized VAEs are not refused everywhere.
|
Thanks, both points were right. Fixed in 73172b0. Gemma-4: it is only recognized as a folder: Install vs. load: I went through every R cell against the configs and loaders. The legend now separates three outcomes:
While checking this I also corrected the VAE note. Quantized VAEs are refused when they load for SD 3, FLUX.2, Qwen-Image, Wan/Anima and Ideogram 4. The FLUX.1 and SD 1.x/SDXL single-file VAEs do not check at all. |
joshistoast
left a comment
There was a problem hiding this comment.
[P2] Include the supported FLUX.2 dev NF4 pipeline. Starter Models exposes FLUX.2 [dev] (Diffusers, NF4) from diffusers/FLUX.2-dev-bnb-4bit, and the existing family page documents it. Mark NF4 as supported inside the Diffusers folder, consistent with the Ideogram 4 row.
|
Right, thanks. FLUX.2 [dev]'s NF4 Diffusers pipeline ( |
Summary
Which file format loads for which model was spread over the Model Families prose and the individual family pages, and went stale whenever a format landed (Ideogram 4 and ERNIE-Image GGUF most recently). This adds one Model Format Support page under Users Guide → Models with four tables:
Cells are ✓ (loads), ✗ (not supported), R (refused at install with a message) or – (no such build). Model Families links to the page from its quantized-formats section. The new-model integration checklist asks to update it.
The tables reflect
mainas of #9708/#9709. ERNIE-Image GGUF is still ✗ here; #9725 flips that cell when it merges.Related Issues / Discussions
Follow-up to #9708, #9709 and #9725.
QA Instructions
pnpm -C docs build: completes.docs/dist), and the#quantized-formatsanchor exists.invokeai/backend/model_manager/.Review
Docs only. No material findings.
Checklist
What's Newcopy (if doing a release after this PR)