Repository navigation
feat(ltx2): load LTX-2.5 transformer GGUF builds - #9814
Open
Pfannkuchensack wants to merge 1 commit into
Open
Pfannkuchensack wants to merge 1 commit into
Pfannkuchensack wants to merge 1 commit into
Conversation
Add Main_GGUF_LTX2_Config and LTX2GGUFModel: Linears stay packed, the bundled connectors are skipped at read. Refuse LTX-2.3 GGUFs and ComfyUI-GGUF Q8_CR files with a reason; pin both against real stripped headers. Add the vantagewithai Dev/Distilled Q4_K_M starters, webv2 support and measured docs.
4 of 7 tasks
This branch has not been deployed
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Community GGUF builds of the LTX-2.5 transformer now install and run. Until now every LTX-2 GGUF installed as an unknown model, with the reason only in the server log.
Main_GGUF_LTX2_Configclaims LTX-2.5 transformer GGUFs from their keys. That covers bare ormodel.diffusion_model.-prefixed keys under anygeneral.architecture. The generation comes from the header'smodel_versionwhen there is one, and from the structure otherwise. Dev or Distilled is read from the file name.Q8_CRfiles are refused as such. The factory now checksDECODES_GGUF_Q8_CRbefore parsing the markers, so molbal's build no longer fails on "names 'keyframes_abs_pos_embedding', not a weight".LTX2GGUFModelkeeps the quantized Linears packed asGGMLTensor. Everything read outside a matmul, and everything unquantized, is unpacked once at load. The text connectors that every build bundles are skipped before their data is copied, through a newkeepfilter ongguf_sd_loader; the component folder supplies them, as it does for the official files. Key renaming and the meta build are shared with the safetensors path.gguf_quantizedLTX-2 mains, which run like the single files: component folder plus Gemma-4 encoder.Related Issues / Discussions
Closes #9728
QA Instructions
Automated checks
uv tool run ruff@0.11.2 checkandformat --checkon all 17 changed Python files: clean.pytest -n 4 tests/backend/model_manager tests/backend/quantization tests/app/invocations/test_ltx2_nodes.py tests/app/invocations/test_text_encoders_with_packed_layers.py tests/model_identification: 2758 passed, 156 skipped, 1 xfailed.src/features/video/core): vitest 508 passed;lint(format, oxlint,tsc --noEmit, architecture) passed.node_moduleslocally). The docs change is prose only.New tests
tests/model_identification: vantagewithai distilled Q4_K_M installs; unsloth LTX-2.3 dev Q4_K_M is refused. The identification harness gains an optionalexpected_refusalfield.keepfilter, hard-coding float32, leaving quantized non-Linears packed, the old dating, and a disabled refusal branch.E2E on an RTX 4090
Setup: distilled, 8 steps, same seed. A/B runs are interleaved, with 1 warm-up and 3 runs each, and the model cache is emptied before every run.
device_working_mem_gb: 5Review
Material findings resolved:
keepsemantics are simplified: Q8_CR is refused file-wide.Remaining limitations:
main.Compatibility / Rollout
AnyModelConfig;openapi.jsonandschema.tsare regenerated. There is no migration.Checklist
What's Newcopy (if doing a release after this PR)Need help on this PR? Tag
@codesmith-botwith what you need. Autofix is disabled.