Skip to content

Convert: Support conversion to gguf for compressed-tensor mixed-precision NVFP4 checkpoint - #28636

Draft
ynankani wants to merge 2 commits into
ggml-org:masterfrom
ynankani:ynankani/CT_support_mixed_precision_convert
Draft

Convert: Support conversion to gguf for compressed-tensor mixed-precision NVFP4 checkpoint#28636
ynankani wants to merge 2 commits into
ggml-org:masterfrom
ynankani:ynankani/CT_support_mixed_precision_convert

Conversation

@ynankani

@ynankani ynankani commented Sep 9, 2026

Copy link
Copy Markdown
Contributor

Overview

Support conversion to gguf for compressed-tensor mixed-precision NVFP4 checkpoint, observed this issue while converting https://huggingface.co/unsloth/Qwen3.8-27B-NVFP4. This is observed due to lack of handling of FP8 layers during conversion of HF CT checkpoint to gguf

Additional information

Validated the change by converting https://huggingface.co/unsloth/Qwen3.8-27B-NVFP4 to gguf

Requirements

Signed-off-by: ynankani <ynankani@nvidia.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant