Three Qwen3.8 checkpoints benchmarked locally on RTX PRO 6000
10/03/2026 — 10/03, 06:11·1 sources·1 reports
Story overview
On October 3, 2026, a post appeared on Reddit's LocalLLaMA community describing a local test of three Qwen3.8 checkpoints on a single RTX PRO 6000. According to the author, the comparison spanned three weeks and covered the same 10 tests. All three models received the same prompts and ran with the sampler settings listed on their own model cards, with one attempt per task. The results were published alongside a video of the runs. The author also notes that an earlier test had covered Qwen3.8-Flash-Next on its own.
The three checkpoints compared are RadixArk/Qwen3.8-27B-NVFP4 (dense), orcarouter/Qwen3.8-27B-Uncensored-NVFP4 (a dense, uncensored fine-tune), and RadixArk/Qwen3.8-Flash-Next-NVFP4 (MoE). The post's title frames the work as a comparison of new SGLang/vLLM recipes.
What the summary and excerpt do not provide is the substance of the individual tests, any scores, latency figures, or throughput numbers, nor any indication of which of the three checkpoints came out ahead. As it stands, the verifiable details are the test setup, the model list, and the conditions under which the runs were made. The account stops at the author publishing the write-up and the accompanying video; the benchmark conclusions themselves are not yet reflected in the coverage. Only this one report is available so far, so there are no competing figures from other sources to set against it.
AI-generated from 1 reports · updated 1 hour ago
Latest turnOver three weeks the author ran the same 10 tests on a single RTX PRO 6000 across three Qwen3.8 checkpoints: RadixArk 27B NVFP4, orcarouter 27B Uncensored NVFP4 (an uncensored fine-tune) and RadixArk Flash-Next NVFP4 (MoE). Each model used the sampler from its own model card, with one attempt per task, and the write-up links a video of the runs.

Reports on this story headlines open the original
Over three weeks the author ran the same 10 tests on a single RTX PRO 6000 across three Qwen3.8 checkpoints: RadixArk 27B NVFP4, orcarouter 27B Uncensored NVFP4 (an uncensored fine-tune) and RadixArk Flash-Next NVFP4 (MoE). Each model used the sampler from its own model card, with one attempt per task, and the write-up links a video of the runs.
Reddit · LocalLLaMAAI score 72
Other stories people are talking about
- 659RisingNVIDIA launches 64GB DGX Spark desktop AI computer at $4,9998 sources
- 539NewApple tightens macOS Full Disk Access over AI agent risks6 sources
- 476NewMeta open-sources Muse Gadgets firmware and SDK5 sources
- 303Google releases Gemini 4 Argon as next-gen frontier AI model8 sources
- 270NewAnthropic launches Claude Frontier Academy with $100M3 sources
- 248SurgeOpenAI DevDay 2026 launches Dots, Decisions API, and more3 sources
How is heat calculated?About the methodHide
Heat counts how many independent sources covered a story in the last 48 hours: one source counts once no matter how many posts it published, decaying with a 24-hour half-life. What ranks first is what many people are talking about.
This page aggregates public feeds. Headlines and summaries are machine-organized and remain the property of the original authors; verify important facts at the source.
- Surge
- Discussion rising fast
- New
- First report within 6 hours
- Rising
- Still gathering discussion
