docs: correct Qwen3.8-Flash-Next generation claims - #3186
Open
localai-org-maint-bot wants to merge 3 commits into
Open
docs: correct Qwen3.8-Flash-Next generation claims#3186localai-org-maint-bot wants to merge 3 commits into
localai-org-maint-bot wants to merge 3 commits into
Conversation
The feature entry contradicts later generation evidence and repeats obsolete attempts. Scope a CPU-only correction with source checks and independent review. Tracks ISSUE-LOCAL-01M2EXYPHDZ03EX01VRTB5FECQ. FOLLOWING_AGENTS_PROTOCOL Following-Agents-Protocol: true AI-Assisted: true Assisted-by: Codex:GPT-6 [Codex]
The feature row contradicted retained CPU and CUDA generation evidence. Replace its attempt history with current behavior and unresolved gates. Announce real-checkpoint CPU and ROCm generation in the README. Preserve the original row verbatim for historical reference. Issue: ISSUE-LOCAL-01M2EXYPHDZ03EX01VRTB5FECQ FOLLOWING_AGENTS_PROTOCOL Following-Agents-Protocol: true AI-Assisted: true Assisted-by: Codex:GPT-6 [Codex]
Record independent review and the operator verification. Full preflight remains incomplete because the local tool environment cannot run its build and packaging checks. FOLLOWING_AGENTS_PROTOCOL Following-Agents-Protocol: true AI-Assisted: true Assisted-by: Codex:GPT-6 [Codex]
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
The Qwen3.8-Flash-Next feature entry mixes current generation results with obsolete claims that decode and CUDA output are unavailable. Replace that history with the tested UD-IQ1_S text path, the one-sequence limit, and the unresolved correctness gate. Preserve the original entry verbatim in an archive.
Add README news for real-checkpoint CPU and ROCm generation. Source registration, backend dispatch, and retained CPU, CUDA, and ROCm evidence support the text. No runtime behavior, benchmark measurement, or model lifecycle changes.
Tracks local issue
ISSUE-LOCAL-01M2EXYPHDZ03EX01VRTB5FECQ. The committed spec is.agents/specs/qwen4-exp-public-doc-refresh.md.CPU-only validation:
python3 scripts/check-readme-structure.pypython3 scripts/check-supported-models.pypython3 scripts/check-site.pypython3 scripts/check-agent-record.pygit diff --checkpass.Independent scoped review passed with no findings. The operator reran the focused gates.
Full preflight was attempted on the baseline and candidate. It is not green in this Alpine tool environment: unrelated registration and packaging checks lack tools or fail in subprocesses that remove the temporary Python library path. The incomplete sweeps were stopped after retaining their logs. The focused documentation checks pass. No GPU tests or new benchmarks were run.
FOLLOWING_AGENTS_PROTOCOL
Following-Agents-Protocol: true
AI-Assisted: true
Assisted-by: Codex:GPT-6 [Codex]