Explore
Everything AION read, in seven sections. Pick one, a topic or a time window.
VideoAgent Speedrun — Elizabeth Fuentes Leone & Sandhya Subramani, AWS
A customer service agent looks up an account, checks an order, and handles a refund.

llm-openai-decisions 0.1a0
OpenAI released their new Jev-style Decisions API, as previously announced at last week's DevDay.
BerriAI/litellm v1.105.0
Verify using the pinned commit hash (recommended):
"Strata" for GLM5.3 Flash is here for some! Project Maya
I stumbled across this as I was currently having glm5.3 flash run only around 10tok/s basically unusable.

I trained a 414k-parameter transformer to fly a boids flock, then tested whether the rules a probe can read are the ones it uses [P]
I wrote a small boid simulator (12 birds), recorded it flying, and trained a transformer to predict each bird's next move without it knowing about any boid rules.
openai/codex rust-v0.162.1: 0.162.1
Fixed a TUI crash when asynchronous questions contain multiple lines, preserving line breaks and complete hyperlink destinations. (#51866)
crewAIInc/crewAI 1.15.27
Add deepinfra as an OpenAI-compatible provider

Ukraine’s drones knock out AI data center belonging to "Russia’s Google"
Ukrainian drone strikes have knocked out two of five data centers belonging to the Russian tech giant Yandex.
ICYMI: What landed for AI builders in September 2026
A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026
anthropics/anthropic-sdk-python v1.13.0
api: add types for the Chat and Cowork unified analytics metrics
langchain-ai/langchain langchain-openai==1.7.0
chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/openai (#41096)
anthropics/claude-code v2.1.295
Added onFailure: "block" for command and HTTP hooks: a hook that can't start, times out, or exits with an unexpected code blocks the action instead of letting it through
How Oracle turns days of work into minutes with ChatGPT and Codex
Across recruiting, engineering, and operations, Oracle turns specialist knowledge into fast, repeatable workflows with ChatGPT Work and Codex.
IBM connects enterprise AI orchestration to production readiness ahead of TechXchange
The post IBM connects enterprise AI orchestration to production readiness ahead of TechXchange appeared first on SiliconANGLE.

Launching an opt-in vulnerability-finding service for open-source software
We’re launching OSS Scanner, an opt-in vulnerability scanner for the open-source ecosystem informed by our experience using Claude to find vulnerabilities during Project Glasswing.
ollama/ollama v0.40.2
Models downloaded with earlier versions of Ollama are upgraded in the background the first time you run them, for better performance and compatibility when running on llama.cpp.
5 Steps to Create SimReady Assets for Robotics with Frontier AI Models
Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision...Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision geometry, joints, and other physics properties before testing robot behavior.
Krea 2 is a bit too... Help me please.
Can anyone recommend a Krea 2 checkpoint that does naughty stuff but isn't overly so?
Maildun for Mac
Design on-brand emails by hand, with AI, or over MCP
huggingface/diffusers v0.41.0: Diffusers 0.41.0: QwenImage 2.1 pipeline and more
> This release brings Qwen-Image 2.1 to Diffusers, with text-to-image generation, image editing, native transparency, and LoRA training.
COSMIC shuts the door on AI code as GNOME debates letting bug reports in
System76 is banning AI-generated content from contributions to the COSMIC desktop.

Can a Cloud-Native Harness Make Agents Reliable Beyond the Desktop?
Since that’s easiest to implement locally, many coding agents started out as terminal tools and then desktop apps (Claude Code famously began as a CLI tool).
Secret protection must scale with software
Today, one in three pull requests on GitHub involves an AI agent.
GeoNatureAgent (GNA): A Framework and Benchmark for Pre-Production Evaluation of Tool-Using Agents on Geospatial and Environmental Tasks
We introduce GeoNatureAgent (GNA), a framework for pre-production evaluation of tool-using agents: a fixed sixteen-tool geospatial interface published as a Model Context Protocol (MCP) server, so the agent under test is the only variable, scored against an identical tool layer, task suite, and deterministic scorer.
VideoDeepSeek-V4.1 Flash: The Most Insane Optimization So Far
DeepSeek-V4.1-Flash is probably the craziest architecture revamp to date.
langchain-ai/langgraph sdk==0.4.6: langgraph-sdk==0.4.6
fix(sdk-py): percent-encode threadid and assistantid in thread stream requests (#9213)
Best Books/Courses/Channels to Leapfrog on AI/ML Material
Assuming I have been on hibernation since 2018-19 time frame, which materials - books, MOOC courses and YT channels would be recommended for me to get started with understanding all the current progress on AI/ML.

Everything we launched during Birthday Week 2026
We celebrated our 16th birthday last week by sharing how we’re building a better Internet for today’s world.
modelcontextprotocol/typescript-sdk v2.3.1: 2.3.1
requireBearerAuth in @modelcontextprotocol/server-legacy takes the optional expectedResource that @modelcontextprotocol/server 2.3.0 added: it accepts only tokens issued for this server (the token's audience).
Together Link: open models in the harness you already use. Start with one command today.
Together Link brings frontier open models like GLM 5.3 and Kimi K3 into the coding agent your team already uses, cutting model spend by over 50%.
Qwen3.8 Flash Next fixed my GNOME extension
I love Dash2Dock Lite, but Icedman is always a week or two before updates.
VideoAI Apps in a Flash: Ship to GPUs Without Docker — Dean Quiñanola, Runpod
Dean Quiñanola, staff engineer at Runpod, introduces Runpod Flash, which lets you run Python on cloud GPUs as if the GPU were local, with no Dockerfiles, image builds or registry pushes.
Engineer / developer observations of Gemma4-31B, Qwen3.8-27B, and 6.1-Sol for software engineering work
Models: Gemma4-31B vs Qwen3.8-27B at the same quantization (an Unsloth flavor of Q4).
VideoBuild a Platform and Watch It Burn — Michael Forrester, Accenture & Whitney Lee, Datadog
A burrito ordering assistant deploys an external container image and exposes a secret recipe from its Kubernetes environment.
Python 3.15.0 added to actions/python-versions
Now that this has landed, you can add "3.15" to a GitHub Actions testing matrix to run tests against the new Python 3.15.0 release.
[Model] Support MiniCPM-V 4.7 by tc-mb · Pull Request #29416 · ggml-org/llama.cpp
Let me remind you that MiniCPM-V-4.7-35B-A3B was spotted on r/LocalLLaMA a few days ago (but the model was later hidden on HF).
BerriAI/litellm v1.104.3
Verify using the pinned commit hash (recommended):
OpenMed 3.0 is out: Apache-2.0 clinical AI that runs fully local and never falls back to the cloud. 422 open issues if you want in on 3.1
Quick recap: it's an open-source (Apache-2.0) medical AI toolkit with one rule we never break: patient data stays on your machine.
PSA: DeepSeek V4.1 Flash habitually exfiltrates API keys. It is dangerously misaligned and may be hazardous to use
EDIT: since people keep calling it out, this is API key abuse but not exfiltration.
openai/codex rust-v0.162.0: 0.162.0
Add tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. (#50148)
ttok 1.0
I released ttok 0.4, ran uv tool upgrade ttok, piped a file into the new version... and realized that it was defaulting to the GPT-4 tokenizer when it should very clearly default to GPT-5/GPT-6 instead!
VideoThe SDLC You Build Yourself — Andy Wong, Andrei Bocan & Shane Wolf, Atlassian
An engineering standard becomes a reusable code review rule, applied across selected repositories instead of enforced by hand on every pull request.
Dwarf Fortress uses version control now
I remember hearing a while ago that Dwarf Fortress didn't use version control.
Automate remediation post AWS DevOps Agent investigation
In this post, we demonstrate how to use AWS Lambda Durable Functions, a capability of AWS Lambda, Amazon EventBridge, and Amazon Bedrock to create an automated remediation workflow that complements AWS DevOps Agent to complete the issue resolution step.
Share GPU clusters across teams with isolation and fairness using Amazon SageMaker HyperPod
Multiple teams within the same company increasingly need shared access to expensive GPU clusters for their generative AI operations, while maintaining isolation boundaries, resource fairness, and operational independence.
LegalOn halves Codex costs while maintaining development speed
LegalOn cut estimated daily Codex costs by 65% while maintaining development speed.
langchain-ai/langchain langchain==1.4.4
chore(deps): bump langgraph-sdk from 0.4.2 to 0.4.4 in /libs/langchainv1 (#41074)
langchain-ai/langchain langchain-core==1.6.9
feat(core): accept a callable in withretry (#41158)