Is anyone running Qwen3.8 Flash Next with a 1M context?
So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.
Key points
- That's the first I have heared of both, the 1M context, and YaRN.
- I want to use this model for a Hermes agent, so a long context could be good...I think.
- But I am still waiting for my mobo to ship, so for now all I can do is ask. x)
Sources (1)
- [1]Is anyone running Qwen3.8 Flash Next with a 1M context?r/LocalLLaMA (top, daily) · Oct 10, 05:49 PM
So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.
That's the first I have heared of both, the 1M context, and YaRN.
Extractive summary: sentences quoted from the sources.