Modelofficial release

nvidia/Qwen3.8-27B-NVFP4

Parameters18.2B
Context256K
Input → outputtext → text
Licenseapache-2.0
PrecisionU8 · nvfp4
ArchitectureQwen3_5ForConditionalGeneration
ReleasedSep 4, 2026
Downloads621.6K

Based on Qwen/Qwen3.8-27B

Other sizes and formats in the family

ModelParamsPrecisionDownloads
nvidia/Qwen3.8-2.4T-A95B-NVFP4nvfp41261BU810.6K

Compare side by side →

This release in the news

No coverage of this exact release yet.

Recent news about Qwen

  1. Oct 11, 2026 · Opinion / analysis
    What is Decision 3.0? vLLM Semantic Router’s New Open Decision Models
  2. Oct 11, 2026 · Opinion / analysis
    I tested different Qwen 3.8 27B quants
  3. Oct 11, 2026 · Opinion / analysis
    Every Model That Can Be Run On 10-16GB VRAM Ranked
  4. Oct 11, 2026 · Opinion / analysis
    Reverse Engineering w/ Local?
  5. Oct 11, 2026 · Benchmark result
    Agents That Own Their Inference — Du'an Lightfoot & Khaja Omer, Akamai Technologies
  6. Oct 11, 2026 · Opinion / analysis
    UPDATE: Qwen 3.8 27B 140 tok/s on single RTX 3090 Megakernel: KL divergence 0.0009 vs llama.cpp
  7. Oct 11, 2026 · Model release
    Qwen/Qwen-Image-2.1-Turbo
  8. Oct 11, 2026 · Opinion / analysis
    Building prompts with LLMs