Modelofficial release

Qwen/Qwen3.8-2.4T-A95B-FP8

Parameters2446B (95B active)
Context256K
Input → outputtext → text
Licenseother
PrecisionF8_E4M3 · fp8
ArchitectureQwen3_5MoeForCausalLM
ReleasedAug 8, 2026
Downloads8.1K

Based on Qwen/Qwen3.8-2.4T-A95B

Other sizes and formats in the family

ModelParamsPrecisionDownloads
Qwen/Qwen3.8-2.4T-A95B2446BBF1632.5K
Qwen/Qwen3.8-27B27.8BBF166.8M
Qwen/Qwen3.8-27B-FP8fp827.8BF8_E4M34.5M

Compare side by side →

This release in the news

No coverage of this exact release yet.

Recent news about Qwen

  1. Oct 11, 2026 · Opinion / analysis
    Every Model That Can Be Run On 10-16GB VRAM Ranked
  2. Oct 11, 2026 · Opinion / analysis
    Reverse Engineering w/ Local?
  3. Oct 11, 2026 · Benchmark result
    Agents That Own Their Inference — Du'an Lightfoot & Khaja Omer, Akamai Technologies
  4. Oct 11, 2026 · Opinion / analysis
    UPDATE: Qwen 3.8 27B 140 tok/s on single RTX 3090 Megakernel: KL divergence 0.0009 vs llama.cpp
  5. Oct 11, 2026 · Model release
    Qwen/Qwen-Image-2.1-Turbo
  6. Oct 11, 2026 · Opinion / analysis
    Building prompts with LLMs
  7. Oct 11, 2026 · Opinion / analysis
    [P] Pecision models that score every allowed label from the logits: Jebadiah v2.1 (27B, 9B), open weights and self-run benchmark results [P]
  8. Oct 11, 2026 · Opinion / analysis
    PSA: DeepSeek V4.1 Flash habitually exfiltrates API keys. It is dangerously misaligned and may be hazardous to use