Qwen/Qwen3.8-Flash-Next
Also known as: Qwen3.8-Flash-Next
3stories this week
6last 30 days
6all time
Timeline
- Oct 11, 2026 · Opinion / analysis · 1 sourcevllm-ascend updates (GLM5.3-flash) on dual 310p Ascend cardsAlright, I learend a few valuable lessons since I posted about a week ago, and I've made more strides with the DDR4x Ascend 96GB 310P cards I purchased, so here we go, meet the new friendly and less verbose me.
- Oct 10, 2026 · Opinion / analysis · 1 sourceAgentic Coding with Qwen3.8 Flash Next GGUFs and Notes from COLMUpdates on Qwen3.8 Flash Next GGUF evaluations for agentic coding
- Oct 6, 2026 · Benchmark result · 1 sourceQwen3.8 Flash Next Reasoning Modes: Off vs Low vs Medium vs XhighLike Qwen3.8 27B, Qwen3.8 Flash Next has three reasoning efforts: low, medium, and xhigh.
- Oct 3, 2026 · Opinion / analysis · 1 sourceQwen3.8 Flash Next, Swift 1.5, and MiMo V2.6: What I’m Testing NowThe Kaitchup – AI on a Budget is a reader-supported publication.
- Sep 30, 2026 · Benchmark result · 1 sourceQwen3.8 Flash Next GGUF Benchmark: Q4 to Q1 Accuracy and Token EfficiencyQuantizing Qwen3.8 Flash Next is almost unavoidable if you want to run it locally.
- Sep 29, 2026 · Opinion / analysis · 1 sourcesergqwer/strata-nvfp4: Strata fork: Qwen3.8-Flash-Next 125B MoE in NVFP4 on one RTX 20-50 card (12 GB+) and 64 GB of RAM or more. Setup installs our GPTQ quants (hStrata fork: Qwen3.8-Flash-Next 125B MoE in NVFP4 on one RTX 20-50 card (12 GB+) and 64 GB of RAM or more.