Unsloth
10stories this week
12last 30 days
14all time
Timeline
- Oct 11, 2026 · Opinion / analysis · 1 sourceUPDATE: Qwen 3.8 27B 140 tok/s on single RTX 3090 Megakernel: KL divergence 0.0009 vs llama.cppRECAP: The megakernel is a CUDA engine for Qwen3.8-27B that runs 1.4-1.9x faster than llama.cpp on a single 3090.
- Oct 11, 2026 · Model release · 4 sourcesQwen/Qwen-Image-2.1-TurboQwen published the model Qwen-Image-2.1-Turbo on Hugging Face.
- Oct 10, 2026 · Opinion / analysis · 1 sourceEngineer / developer observations of Gemma4-31B, Qwen3.8-27B, and 6.1-Sol for software engineering workModels: Gemma4-31B vs Qwen3.8-27B at the same quantization (an Unsloth flavor of Q4).
- Oct 10, 2026 · Opinion / analysis · 1 sourceQwen3.8-27B on a single 3090: 140 tok/s on code with a custom megakernelI've been using Claude Opus 5.5 to speed up Qwen3.8-27B on my PC (rtx 3090), it wrote a CUDA megakernel that is 1.4-1.9x faster than llama.cpp depending on the task/context length.
- Oct 10, 2026 · Opinion / analysis · 1 sourceStrata with Qwen3.8 Flash Next UD-Q4_K_XLMost of the benchmarks I've seen are using IQ2 or IQ3 quants, so I wanted to see how Unsloth's UD-Q4KXL performs instead.
- Oct 10, 2026 · Opinion / analysis · 1 source[AINews] TypeSafe/Jev at >$100M ARR, $7.5B valuation 3 weeks after launchAs you can see in the AINews X recap section below, everyone on earth has cloned the Jev API, but only one company can ever create the category.
- Oct 8, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.905-beta: Sandboxing is here!We're introducing Windows, Mac and Linux sandboxing in Unsloth!
- Oct 7, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.904-beta: Train your own Decision modelTurn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
- Oct 6, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.903-beta: New Browser + Voice CloningThis release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
- Oct 6, 2026 · Model release · 1 sourceunsloth/embeddinggemma-2-GGUFunsloth published the model embeddinggemma-2-GGUF on Hugging Face.
- Oct 1, 2026 · Product / feature launch · 1 sourceunslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UXThis release brings faster navigation, shareable run settings, and clearer errors to Unsloth Desktop.
- Sep 28, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.900-beta: Laya Decision Models + LibraryRun and serve Decision Models like Laya (open-source Jev) locally
- Jun 10, 2026 · Model release · 1 sourceDiffusionGemma: 4x faster text generationToday, we’re introducing DiffusionGemma, an experimental open model that explores text diffusion, an exceptionally fast approach to text generation.
- Jun 9, 2026 · Model release · 1 sourceIntroducing Gemma 4 12B: a unified, encoder-free multimodal modelIntroducing Gemma 4 12B: a unified, encoder-free multimodal model