Open source

Libraries, tools and repos: releases you depend on and projects gaining stars.

Simon Willison's Weblog2 sources2d ago

llm-openai-decisions 0.1a0

OpenAI released their new Jev-style Decisions API, as previously announced at last week's DevDay.

2 outlets
GitHub: BerriAI/litellm10h ago

BerriAI/litellm v1.105.0

Verify using the pinned commit hash (recommended):

Code
r/LocalLLaMA (top, daily)14h ago

Converting dense models into Mixture-of-Experts

For the past few weeks I've been trying out converting existing dense models to sparse Mixture-of-Experts models, with no pretraining from scratch.

GitHub: pydantic/pydantic-ai2d ago

pydantic/pydantic-ai v2.55.0: v2.55.0 (2026-10-09)

<!-- Release notes generated using configuration in .github/release.yml at main -->

Code
GitHub: open-webui/open-webui22h ago

open-webui/open-webui v0.12.0

Approvals and questions from tools still have to be answered in the chat, calls need the Allow Call permission and end after an hour, models can set their own Realtime Voice in the model editor, and the settings can also be given with the "AUDIOREALTIMEENABLED", "AUDIOREALTIMEOPENAIAPIBASEURL", "AUDIOREALTIMEOPENAIAPIKEY", "AUDIOREALTIMEMODEL", "AUDIOREALTIMEVOICE", "AUDIOREALTIMETRANSCRIPTIONMODEL" and "REALTIMECALLPROMPTTEMPLATE" environment variables.

Code
GitHub: crewAIInc/crewAI1d ago

crewAIInc/crewAI 1.15.27

Add deepinfra as an OpenAI-compatible provider

Code
GitHub: anthropics/anthropic-sdk-python2d ago

anthropics/anthropic-sdk-python v1.13.0

api: add types for the Chat and Cowork unified analytics metrics

Code
OpenAI News2d ago

Asana cuts model costs 76x in browser tests with GPT-6.1 Sol

Using GPT-6 Astra in Codex, Asana made its browser agent 76x cheaper and 5x faster in tests to offer customers more capable models.

GitHub: langchain-ai/langchain3d ago

langchain-ai/langchain langchain-openai==1.7.0

chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/openai (#41096)

Code
GitHub: huggingface/trl3d ago

huggingface/trl v1.15.0

SFT, DPO, KTO, GRPO, RLOO and Distillation now score tokens with a fused LM head: a Triton kernel projects the hidden states through the LM head in tiles and reduces to per-token log-probs and entropy directly, so the [batch, seq, vocab] logits tensor is never built.

Code
GitHub: unslothai/unsloth3d ago

unslothai/unsloth v0.1.905-beta: Sandboxing is here!

We're introducing Windows, Mac and Linux sandboxing in Unsloth!

Code
GitHub: ollama/ollama3d ago

ollama/ollama v0.40.2

Models downloaded with earlier versions of Ollama are upgraded in the background the first time you run them, for better performance and compatibility when running on llama.cpp.

Code
GitHub: anthropics/claude-code3d ago

anthropics/claude-code v2.1.294

Fixed prompt and agent hooks written as instructions (such as "Block commands that...") allowing what they should block

Code
NVIDIA Technical Blog4d ago

Faster Scientific Image Analysis with NVIDIA cuPhoton

Observatories and telescopes, lasers and X-ray light sources, and other high-throughput instruments generate image data faster than CPU-bound pipelines can...Observatories and telescopes, lasers and X-ray light sources, and other high-throughput instruments generate image data faster than CPU-bound pipelines can process it to support timely decisions.

GitHub: huggingface/transformers5d ago

huggingface/transformers v5.19.0: Release v5.19.0

EmbeddingGemma 2 is a multimodal embedding model from Google built on the Gemma 4 architecture.

Code
GitHub: browser-use/browser-use4d ago

browser-use/browser-use 0.13.11

This release includes Browser Use toolsets for Claude, available at browseruse.integrations.toolsetsforclaude.

Code
GitHub: langchain-ai/langgraph4d ago

langchain-ai/langgraph cli==0.4.33: langgraph-cli==0.4.33

feat(cli): add 'langgraph deploy listeners list' (#9221)

Code
GitHub: huggingface/diffusers5d ago

huggingface/diffusers v0.41.0: Diffusers 0.41.0: QwenImage 2.1 pipeline and more

> This release brings Qwen-Image 2.1 to Diffusers, with text-to-image generation, image editing, native transparency, and LoRA training.

Code
Latent Space4d ago

Can a Cloud-Native Harness Make Agents Reliable Beyond the Desktop?

Since that’s easiest to implement locally, many coding agents started out as terminal tools and then desktop apps (Claude Code famously began as a CLI tool).

Product Hunt: AI launches1d ago

Farol

The macOS terminal that shows which coding agent needs you

Hugging Face trending models6d ago

speridlabs/iris-3b

speridlabs published the model iris-3b on Hugging Face.

Weights
GitHub: vllm-project/vllm6d ago

vllm-project/vllm v0.31.0

Fast restart: the new vllm preload CLI launches the weight-cache daemon that keeps post-quantized weights resident in GPU memory across engine restarts (#56680), now with data parallelism (#57386), MTP draft models (#57312), a /health endpoint (#58552) and a readiness wait (#58370).

Code
Repo
Rising AI repositories on GitHub6d ago

GTKottman/mortiflix-oss: A motion design studio on your own machine: Claude makes the video step by step, you approve every stage. Bring your own Claude Code or API

A motion design studio on your own machine: Claude makes the video step by step, you approve every stage.

Code
GitHub: modelcontextprotocol/typescript-sdk6d ago

modelcontextprotocol/typescript-sdk v2.3.1: 2.3.1

requireBearerAuth in @modelcontextprotocol/server-legacy takes the optional expectedResource that @modelcontextprotocol/server 2.3.0 added: it accepts only tokens issued for this server (the token's audience).

Code
Together AI Blog6d ago

Together Link: open models in the harness you already use. Start with one command today.

Together Link brings frontier open models like GLM 5.3 and Kimi K3 into the coding agent your team already uses, cutting model spend by over 50%.

Simon Willison's Weblog20h ago

Python 3.15.0 added to actions/python-versions

Now that this has landed, you can add "3.15" to a GitHub Actions testing matrix to run tests against the new Python 3.15.0 release.

GitHub: BerriAI/litellm9h ago

BerriAI/litellm v1.104.3

Verify using the pinned commit hash (recommended):

Code
OpenAI News3d ago

LegalOn halves Codex costs while maintaining development speed

LegalOn cut estimated daily Codex costs by 65% while maintaining development speed.

GitHub: langchain-ai/langchain2d ago

langchain-ai/langchain langchain==1.4.4

chore(deps): bump langgraph-sdk from 0.4.2 to 0.4.4 in /libs/langchainv1 (#41074)

Code
GitHub: langchain-ai/langchain2d ago

langchain-ai/langchain langchain-core==1.6.9

feat(core): accept a callable in withretry (#41158)

Code
Simon Willison's Weblog2d ago

ttok 0.4

ttok is my CLI tool for counting tokens, using OpenAI's open source tiktoken library.

GitHub: langchain-ai/langchain3d ago

langchain-ai/langchain langchain-core==1.6.8

fix(core): harden scoped and compatible IPv6 SSRF checks (#41151)

Code
Simon Willison's Weblog1d ago

Deno is joining Cloudflare

The Deno team released the first version of celld back in August - their open source implementation of the Durable Objects pattern from Cloudflare Workers.

GitHub: anthropics/anthropic-sdk-python3d ago

anthropics/anthropic-sdk-python v1.12.1

docs: note that listing Claude Console spend limits is in early access

Code
GitHub: langchain-ai/langchain3d ago

langchain-ai/langchain langchain-huggingface==1.2.3

chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/huggingface (#41098)

Code
GitHub: crewAIInc/crewAI2d ago

crewAIInc/crewAI 1.15.26

Fix ownership retention of rwlock when acquisition is interrupted.

Code
GitHub: unslothai/unsloth4d ago

unslothai/unsloth v0.1.904-beta: Train your own Decision model

Turn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.

Code
GitHub: langchain-ai/langchain3d ago

langchain-ai/langchain langchain-fireworks==1.7.1

chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/fireworks (#41100)

Code
GitHub: BerriAI/litellm3d ago

BerriAI/litellm v1.104.2

Verify using the pinned commit hash (recommended):

Code
NVIDIA Technical Blog5d ago

How DOCA GPUNetIO Unifies GPU-Initiated Networking Across the NVIDIA Software Stack

GPU applications increasingly need networking and data movement to behave like first-class GPU-controlled operations rather than host-driven services.

GitHub: BerriAI/litellm3d ago

BerriAI/litellm v1.102.4

Verify using the pinned commit hash (recommended):

Code
GitHub: BerriAI/litellm3d ago

BerriAI/litellm v1.101.6

Verify using the pinned commit hash (recommended):

Code
GitHub: huggingface/trl4d ago

huggingface/trl v1.14.2

Patch release fixing two cases of silently wrong training and three crashes.

Code
GitHub: BerriAI/litellm3d ago

BerriAI/litellm v1.104.1

Verify using the pinned commit hash (recommended):

Code
GitHub: BerriAI/litellm4d ago

BerriAI/litellm v1.103.4

Verify using the pinned commit hash (recommended):

Code
GitHub: BerriAI/litellm4d ago

BerriAI/litellm v1.102.3

Verify using the pinned commit hash (recommended):

Code
GitHub: BerriAI/litellm4d ago

BerriAI/litellm v1.100.5

Verify using the pinned commit hash (recommended):

Code
GitHub: BerriAI/litellm4d ago

BerriAI/litellm v1.101.5

Verify using the pinned commit hash (recommended):

Code