Is there an existing issue for this?
What should this feature add?
Load community GGUF builds of the Anima transformer. Anima loads as bf16, scaled fp8, int8 and MXFP8 single files,
but there is no GGUF config, so a .gguf Anima file does not install as Anima.
Alternatives
Use the fp8 or int8 single file.
Additional Content
The full matrix of what loads today is in #9726 (Model Format Support page).
Is there an existing issue for this?
What should this feature add?
Load community GGUF builds of the Anima transformer. Anima loads as bf16, scaled fp8, int8 and MXFP8 single files,
but there is no GGUF config, so a
.ggufAnima file does not install as Anima.Alternatives
Use the fp8 or int8 single file.
Additional Content
Bedovyy/Anima-GGUF,Abiray/Anima-base-v1.0-GGUF,vanes430/Anima-Turbo-V1.1-GGUF.Main_GGUF_*_Configfingerprinted by keys (GGUF headers often name another architecture), a loader that keeps quantized Linears packed and unpacks the rest withunpack_ggml_at_load, and the denoise node'speak_dequant_transient_bytes, which already counts GGUF layers (fix(model-cache): reserve the GGUF dequantization transient in every node #9709).The full matrix of what loads today is in #9726 (Model Format Support page).