Open source
Libraries, tools and repos: releases you depend on and projects gaining stars.

llm-openai-decisions 0.1a0
OpenAI released their new Jev-style Decisions API, as previously announced at last week's DevDay.
langchain-ai/langchain langchain-openai==1.7.0
chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/openai (#41096)
Asana cuts model costs 76x in browser tests with GPT-6.1 Sol
Using GPT-6 Astra in Codex, Asana made its browser agent 76x cheaper and 5x faster in tests to offer customers more capable models.
huggingface/transformers v5.19.0: Release v5.19.0
EmbeddingGemma 2 is a multimodal embedding model from Google built on the Gemma 4 architecture.
VideoParameter Golf with AutoResearch — Vayum Arora, Zhengyao Jiang, Dixing Xu & Dhruv Srikanth, Weco AI
Zhengyao Jiang introduces autoresearch as repeated proposals and evaluations, and Dixing Xu explains the team's Aiden system and its contributions to OpenAI's Parameter Golf challenge.
Converting dense models into Mixture-of-Experts
For the past few weeks I've been trying out converting existing dense models to sparse Mixture-of-Experts models, with no pretraining from scratch.
unslothai/unsloth v0.1.905-beta: Sandboxing is here!
We're introducing Windows, Mac and Linux sandboxing in Unsloth!
Mia-AiLab/GLM-5.3-Flash-EXL3-4bpw-TensorFold-Ablit
Mia-AiLab published the model GLM-5.3-Flash-EXL3-4bpw-TensorFold-Ablit on Hugging Face.
huggingface/trl v1.14.2
Patch release fixing two cases of silently wrong training and three crashes.
langchain-ai/langgraph cli==0.4.33: langgraph-cli==0.4.33
feat(cli): add 'langgraph deploy listeners list' (#9221)
sgl-project/sglang v0.5.21
| Model | Type | Cookbook |
huggingface/peft v0.21.2
This is a PEFT release fixes an issue that prevented encoder-decoder models to work when using Transformers ≥ 5.18.0.
ollama/ollama v0.35.0
Decision models return choices, probabilities, and scores instead of text.
Dreamer-Toby/STEPQuant: STEPQuant: When and Where Errors Matter in Delta-Rule Recurrent State Quantization
STEPQuant: When and Where Errors Matter in Delta-Rule Recurrent State Quantization Language: Python.
Infatoshi/GLM-5.3-UNCENSORED-EXL3-3.0bpw
Infatoshi published the model GLM-5.3-UNCENSORED-EXL3-3.0bpw on Hugging Face.
unslothai/unsloth v0.1.904-beta: Train your own Decision model
Turn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
langchain-ai/langchain langchain-huggingface==1.2.3
chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/huggingface (#41098)
langchain-ai/langchain langchain==1.4.4
chore(deps): bump langgraph-sdk from 0.4.2 to 0.4.4 in /libs/langchainv1 (#41074)
langchain-ai/langchain langchain-core==1.6.9
feat(core): accept a callable in withretry (#41158)
Cloudflare/clef-flash
Cloudflare published the model clef-flash on Hugging Face.
ttok 0.4
ttok is my CLI tool for counting tokens, using OpenAI's open source tiktoken library.
langchain-ai/langchain langchain-fireworks==1.7.1
chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/fireworks (#41100)
langchain-ai/langgraph sdk==0.4.6: langgraph-sdk==0.4.6
fix(sdk-py): percent-encode threadid and assistantid in thread stream requests (#9213)
huggingface/transformers v5.18.0: Release 5.18.0
Nemotron 3 Diarization is an open-weight streaming speaker diarization model designed to determine "who spoke when" in real-world audio.
langchain-ai/langgraph 1.2.13: langgraph==1.2.13
fix(langgraph): fork before replaying an update checkpoint the thread moved past (#9170)
unslothai/unsloth v0.1.903-beta: New Browser + Voice Cloning
This release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
langchain-ai/langchain langchain-text-splitters==1.1.3
chore(deps): bump tornado from 6.5.9 to 6.5.10 in /libs/text-splitters (#40971)
ollama/ollama v0.35.1
Clef (27B) and Clef Flash (9B) are multimodal: requests can now include images alongside the text state, shared by all questions and scored jointly with it.
langchain-ai/langchain langchain-core==1.6.6
fix(anthropic): support Claude Sonnet 5.5 compatibility (#40882)
langchain-ai/langchain langchain==1.4.3
feat(langchain): support Bedrock Mantle chat models in initchatmodel (#40837)
StayLameBro/backburner: Your iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable
Your iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable Language: Python.
Edwardxlai/easyread: 把英文论文读成舒服的中文:本地 PDF 论文翻译、原文对照、边读边问 AI、文献管理。Read English papers in comfortable Chinese.
把英文论文读成舒服的中文:本地 PDF 论文翻译、原文对照、边读边问 AI、文献管理。Read English papers in comfortable Chinese.
nanaism/yomiyasu: AI生成の日本語を自然な日本語へ推敲するAgent Skill / Agent Skill for Refining AI-Generated Japanese into Natural Japanese
AI生成の日本語を自然な日本語へ推敲するAgent Skill / Agent Skill for Refining AI-Generated Japanese into Natural Japanese Language: Python.
ollama/ollama v0.34.4
Qwen 3.8 prompt processing is faster on Apple Silicon.
sgl-project/sglang v0.5.20
| Model | Type | PRs | Cookbook |