AION
Opinion / analysisEfficiency & Inference · Agents & Tool Use · Large Language Models1 source · Oct 10, 2026

Is anyone running Qwen3.8 Flash Next with a 1M context?

So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.

Key points

  • That's the first I have heared of both, the 1M context, and YaRN.
  • I want to use this model for a Hermes agent, so a long context could be good...I think.
  • But I am still waiting for my mobo to ship, so for now all I can do is ask. x)

Sources (1)

  • [1]Is anyone running Qwen3.8 Flash Next with a 1M context?
    r/LocalLLaMA (top, daily) · Oct 10, 05:49 PM
    So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.
    That's the first I have heared of both, the 1M context, and YaRN.

Extractive summary: sentences quoted from the sources.