Problem
Kimi K3 applies quantization-aware training from the SFT stage onward, using MXFP4 weights with MXFP8 activations for broad hardware compatibility.
Proposed solution
We need the loading of the Original Kimi K3 Weights, without any changes. The Engine should work out of the box with the Original Quantization. Because any Changes to the Original Weights will degrade the Quality. Its not the Situation like in GLM 5.2 where Quantization makes sense, but not in Kimi K3.
Alternatives considered
No response
Scope and compatibility
No response
Problem
Kimi K3 applies quantization-aware training from the SFT stage onward, using MXFP4 weights with MXFP8 activations for broad hardware compatibility.
Proposed solution
We need the loading of the Original Kimi K3 Weights, without any changes. The Engine should work out of the box with the Original Quantization. Because any Changes to the Original Weights will degrade the Quality. Its not the Situation like in GLM 5.2 where Quantization makes sense, but not in Kimi K3.
Alternatives considered
No response
Scope and compatibility
No response