Modelofficial release

Qwen/Qwen3.8-Flash-Next-FP8

Parameters180B
Context256K
Input → outputimage + text → text
Licenseother
PrecisionF8_E4M3 · fp8
ArchitectureQwen4ExpForConditionalGeneration
ReleasedAug 24, 2026
Downloads275.8K

Based on Qwen/Qwen3.8-Flash-Next

Other sizes and formats in the family

ModelParamsPrecisionDownloads
Qwen/Qwen3.8-Flash-Next180BBF161.8M

Compare side by side →

This release in the news

No coverage of this exact release yet.

Recent news about Qwen

  1. Oct 11, 2026 · Opinion / analysis
    What is Decision 3.0? vLLM Semantic Router’s New Open Decision Models
  2. Oct 11, 2026 · Opinion / analysis
    I tested different Qwen 3.8 27B quants
  3. Oct 11, 2026 · Opinion / analysis
    Every Model That Can Be Run On 10-16GB VRAM Ranked
  4. Oct 11, 2026 · Opinion / analysis
    Reverse Engineering w/ Local?
  5. Oct 11, 2026 · Benchmark result
    Agents That Own Their Inference — Du'an Lightfoot & Khaja Omer, Akamai Technologies
  6. Oct 11, 2026 · Opinion / analysis
    UPDATE: Qwen 3.8 27B 140 tok/s on single RTX 3090 Megakernel: KL divergence 0.0009 vs llama.cpp
  7. Oct 11, 2026 · Model release
    Qwen/Qwen-Image-2.1-Turbo
  8. Oct 11, 2026 · Opinion / analysis
    Building prompts with LLMs