Explore
Everything AION read, in seven sections. Pick one, a topic or a time window.

Microsoft's Decision-1 model enters the fast-growing AI decision model race
With Decision-1, Microsoft enters the growing decision model space.

Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide
An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) tipline, according to a report from 6abc.
VideoAgent Speedrun — Elizabeth Fuentes Leone & Sandhya Subramani, AWS
A customer service agent looks up an account, checks an order, and handles a refund.
What's up with google scholar citations ? [D]
These papers have been there for months, but Google Scholar still has not updated its citations, and it has occurred before, but I ignored it and thought it was a one-time error.

Understanding Jev, the new model everyone is talking about
On September 15, the startup TypeSafe AI came out of stealth and released a new AI model.

When the Safety Test Became the Threat: The Machine That Found Its Own Way Out
OpenAI built a room with no doors – or so it thought.
"Strata" for GLM5.3 Flash is here for some! Project Maya
I stumbled across this as I was currently having glm5.3 flash run only around 10tok/s basically unusable.
pydantic/pydantic-ai v2.55.0: v2.55.0 (2026-10-09)
<!-- Release notes generated using configuration in .github/release.yml at main -->

Building AI for Reliable Execution: Lessons From Industrial Robotics
Standard Bots claims to be “America’s largest AI-native industrial robot manufacturer.” It recently raised $200 million at a $1 billion valuation, in a series C round led by General Catalyst and RoboStrategy, a fund focused on robotics.
Quoting The New York Times
— The New York Times, Anthropic Agents Tried to Fill Out Visa Forms on State Dept.
Memento 3: Model-Based Recursive Self-Improvement through Reflective Rulebooks
We introduce Memento 3, building on the Memento series to enable frozen LLM agents to continually learn explicit world models through external memory.
anthropics/claude-code v2.1.296
Added a code key to the Claude apps gateway's managed.policies[]: the same settings as cli, also applied in Claude Desktop's Code tab; beside desktop, it turns on Claude Desktop's gateway mode

AI agent makers are promising privacy — will they deliver?
At this year's OpenAI DevDay, CEO Sam Altman unveiled the company's new AI agent Dots - and told the crowd that the company wants to "set a new standard for privacy in frontier AI.
ICYMI: What landed for AI builders in September 2026
A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026

Microsoft event debuts new AI-friendly hardware and Windows changes
In its first live event in two years, Microsoft today announced its newest Surface laptop, outlined a host of changes coming to Windows 11, and shared a vision of how local-AI and agentic workflows could reshape personal computing for those in the dev community as well as home enthusiasts.

Introducing Mistral Large 4
Today, we’re launching a public preview of Mistral Large 4.
langchain-ai/langchain langchain-openai==1.7.0
chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/openai (#41096)
VideoMicrosoft Joins the Local AI Push
Microsoft is clearly going all in on local AI running on your machine and working with cloud models only when it needs to use them.
How Oracle turns days of work into minutes with ChatGPT and Codex
Across recruiting, engineering, and operations, Oracle turns specialist knowledge into fast, repeatable workflows with ChatGPT Work and Codex.
From Reactive Containment to Proactive Assurance: Lessons from OpenAI, Anthropic, and Google Agent Security Incidents
In 2026, cybersecurity evaluations involving OpenAI, Anthropic, and Google agents reached real systems outside their authorized test scope.
OpenPilot
Open-source desktop AI agent for any model you choose

The Epoch Brief - October 8, 2026
Welcome back to the Epoch Brief.
Show HN: Jevman – AI decision models play Pac-Man
Openai just launched their decisions endpoint, cloudflare launched clef the other week, and many more jev alternatives are out there.

Text is so 2023
Hark (from Brett Adcock - Figure robots) launched which suggests tasks as one-tap action buttons.
Nvidia, Samsung back $90M round for AI agent startup Nous Research
Nous Research Inc., the developer of the popular Hermes artificial intelligence agent, today announced that it has raised $90 million in funding.
5 Steps to Create SimReady Assets for Robotics with Frontier AI Models
Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision...Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision geometry, joints, and other physics properties before testing robot behavior.
Show HN: I Put an AI Agent on a Nokia 110
Recently got the idea to put ai agent in it.
Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses
Harnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework.
PaperRISED: Rubrics for Agentic Multi-Environment Selection and Self-Distillation
RISED: Rubrics for Agentic Multi-Environment Selection and Self-Distillation
browser-use/browser-use 0.13.11
This release includes Browser Use toolsets for Claude, available at browseruse.integrations.toolsetsforclaude.
COSMIC shuts the door on AI code as GNOME debates letting bug reports in
System76 is banning AI-generated content from contributions to the COSMIC desktop.

Why agent swarms could be the next “scaling law”
One of the most surprising aspects of July’s news that OpenAI agents attacked Hugging Face was how the agents had worked together.
Secret protection must scale with software
Today, one in three pull requests on GitHub involves an AI agent.
Open and Emergent Problems in Agentic Privacy and Security: A Contextual Angle
Inspired by the theory of Contextual Integrity, our new workshop report outlines key open research directions across system, model, and user levels to build AI agents users can trust.
GTKottman/mortiflix-oss: A motion design studio on your own machine: Claude makes the video step by step, you approve every stage. Bring your own Claude Code or API
A motion design studio on your own machine: Claude makes the video step by step, you approve every stage.
Connecting AI agents to enterprise knowledge
For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge.

Anthropic is cutting off its internal evaluations from the internet
After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations.
VideoSonarQube + OpenAI: Agentic Development — Killian Carlsen-Phelan, Sonar
Killian Carlsen-Phelan runs it through SonarQube, brings the findings into Codex and asks the agent to fix the vulnerable query.
VideoCan Your Agent Hear You Now? Building Live Voice Agents with Gemini — Thor Schaeff
An agent that hears you, sees you, and answers in your language, in real time.
Google brings agentic AI to Gemini, starting with businesses
Google is turning Gemini into an AI agent that can plan, execute tasks, and work across business apps and systems.
VideoAI Security Engineer Foundations + Certificate — Micah Silverman, Snyk
Micah Silverman uses that capture the flag exercise to connect AI risks with familiar application security controls.
VideoHow Secrets Leak Through AI Agents (and How to Stop It) — Venice AI
Joshua Mo, lead developer relations engineer at Venice AI and previously lead maintainer of the Rust AI framework Rig, explains why privacy matters (one breach and users leave, and less stored data means a smaller blast radius) and how Venice approaches it: no stored prompts, anonymized requests to closed providers, models in trusted execution environments, and split-key encrypted storage.

AI agent teams waste massive tokens for barely measurable quality gains, research finds
Teams of AI agents barely outperform solo agents but cost up to 5.1x more, according to Vals AI.

AI agents overstate their results and remain far from autonomous research, study finds
Epoch AI and Anthropic independently found the same thing: current AI models like GPT-5.6 Sol and Claude Fable 5 can run experiments but lack scientific self-criticism and genuine creative thinking.
Is anyone running Qwen3.8 Flash Next with a 1M context?
So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.
Embodied Turing Machines: Stateful Code for Robot Recursive Self-Improvement
We propose a different view: the embodied world is an Embodied Turing Machine, whose tape is the robot and environment state and rules are the policy.
OneSearch-VL: Unified Multimodal Deep Research Agent for Image and Video
We introduce OneSearch-VL, a unified agent centered on the Visually Grounded Evidence Graph (VGEG), which encodes these dependencies as a shared task-level reference for data construction, process supervision, and operation-level evaluation.
SuperNav: An Agentic Navigation System for Any Task in Any Scene
General-purpose service robots need navigation systems that can handle diverse human requests in unfamiliar environments, combining task generality with scene generality.