Explore

Everything AION read, in seven sections. Pick one, a topic or a time window.

The Decoder4 sources1d ago

Microsoft's Decision-1 model enters the fast-growing AI decision model race

With Decision-1, Microsoft enters the growing decision model space.

3 outlets
TechCrunch: AI4 sources1d ago

Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide

An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) tipline, according to a report from 6abc.

2 outlets
Video
AI Engineer (YouTube)5h ago

Agent Speedrun — Elizabeth Fuentes Leone & Sandhya Subramani, AWS

A customer service agent looks up an account, checks an order, and handles a refund.

Newr/MachineLearning (top, daily)2h ago

What's up with google scholar citations ? [D]

These papers have been there for months, but Google Scholar still has not updated its citations, and it has occurred before, but I ignored it and thought it was a one-time error.

Understanding AI (Timothy B. Lee)3 sources1d ago

Understanding Jev, the new model everyone is talking about

On September 15, the startup TypeSafe AI came out of stealth and released a new AI model.

3 outlets
MarkTechPost23h ago

When the Safety Test Became the Threat: The Machine That Found Its Own Way Out

OpenAI built a room with no doors – or so it thought.

Newr/LocalLLaMA (top, daily)2h ago

"Strata" for GLM5.3 Flash is here for some! Project Maya

I stumbled across this as I was currently having glm5.3 flash run only around 10tok/s basically unusable.

GitHub: pydantic/pydantic-ai2d ago

pydantic/pydantic-ai v2.55.0: v2.55.0 (2026-10-09)

<!-- Release notes generated using configuration in .github/release.yml at main -->

Code
Latent Space1d ago

Building AI for Reliable Execution: Lessons From Industrial Robotics

Standard Bots claims to be “America’s largest AI-native industrial robot manufacturer.” It recently raised $200 million at a $1 billion valuation, in a series C round led by General Catalyst and RoboStrategy, a fund focused on robotics.

Simon Willison's Weblog1d ago

Quoting The New York Times

— The New York Times, Anthropic Agents Tried to Fill Out Visa Forms on State Dept.

Paper
Hugging Face Daily Papers2 sources3d ago

Memento 3: Model-Based Recursive Self-Improvement through Reflective Rulebooks

We introduce Memento 3, building on the Memento series to enable frozen LLM agents to continually learn explicit world models through external memory.

Paper
GitHub: anthropics/claude-code2d ago

anthropics/claude-code v2.1.296

Added a code key to the Claude apps gateway's managed.policies[]: the same settings as cli, also applied in Claude Desktop's Code tab; beside desktop, it turns on Claude Desktop's gateway mode

Code
The Verge: AI1d ago

AI agent makers are promising privacy — will they deliver?

At this year's OpenAI DevDay, CEO Sam Altman unveiled the company's new AI agent Dots - and told the crowd that the company wants to "set a new standard for privacy in frontier AI.

AWS Machine Learning Blog2d ago

ICYMI: What landed for AI builders in September 2026

A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026

Ars Technica: AI2 sources3d ago

Microsoft event debuts new AI-friendly hardware and Windows changes

In its first live event in two years, Microsoft today announced its newest Surface laptop, outlined a host of changes coming to Windows 11, and shared a vision of how local-AI and agentic workflows could reshape personal computing for those in the dev community as well as home enthusiasts.

2 outlets
Mistral AI News3 sources4d ago

Introducing Mistral Large 4

Today, we’re launching a public preview of Mistral Large 4.

2 outlets
GitHub: langchain-ai/langchain3d ago

langchain-ai/langchain langchain-openai==1.7.0

chore(deps): bump langgraph-sdk from 0.4.4 to 0.4.6 in /libs/partners/openai (#41096)

Code
Video
Sam Witteveen (YouTube)3d ago

Microsoft Joins the Local AI Push

Microsoft is clearly going all in on local AI running on your machine and working with cloud models only when it needs to use them.

OpenAI News3d ago

How Oracle turns days of work into minutes with ChatGPT and Codex

Across recruiting, engineering, and operations, Oracle turns specialist knowledge into fast, repeatable workflows with ChatGPT Work and Codex.

Paper
arXiv (AI, ML, NLP, CV, robotics, multi-agent)3d ago

From Reactive Containment to Proactive Assurance: Lessons from OpenAI, Anthropic, and Google Agent Security Incidents

In 2026, cybersecurity evaluations involving OpenAI, Anthropic, and Google agents reached real systems outside their authorized test scope.

Paper
Product Hunt: AI launches2d ago

OpenPilot

Open-source desktop AI agent for any model you choose

Epoch AI: Gradient Updates3d ago

The Epoch Brief - October 8, 2026

Welcome back to the Epoch Brief.

Hacker News: LLM threads (40+ points)2 sources3d ago

Show HN: Jevman – AI decision models play Pac-Man

Openai just launched their decisions endpoint, cloudflare launched clef the other week, and many more jev alternatives are out there.

Ben's Bites3d ago

Text is so 2023

Hark (from Brett Adcock - Figure robots) launched which suggests tasks as one-tap action buttons.

SiliconANGLE: AI3d ago

Nvidia, Samsung back $90M round for AI agent startup Nous Research

Nous Research Inc., the developer of the popular Hermes artificial intelligence agent, today announced that it has raised $90 million in funding.

NVIDIA Technical Blog2d ago

5 Steps to Create SimReady Assets for Robotics with Frontier AI Models

Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision...Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision geometry, joints, and other physics properties before testing robot behavior.

Show HN: AI projects (15+ points)3d ago

Show HN: I Put an AI Agent on a Nokia 110

Recently got the idea to put ai agent in it.

Code
Microsoft Research Blog4d ago

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Harnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework.

Vendor claim only
Paper
Apple Machine Learning Research2 sources5d ago

RISED: Rubrics for Agentic Multi-Environment Selection and Self-Distillation

RISED: Rubrics for Agentic Multi-Environment Selection and Self-Distillation

Paper
GitHub: browser-use/browser-use4d ago

browser-use/browser-use 0.13.11

This release includes Browser Use toolsets for Claude, available at browseruse.integrations.toolsetsforclaude.

Code
The Register: AI/ML4d ago

COSMIC shuts the door on AI code as GNOME debates letting bug reports in

System76 is banning AI-generated content from contributions to the COSMIC desktop.

Import AI (Jack Clark)2 sources6d ago

Why agent swarms could be the next “scaling law”

One of the most surprising aspects of July’s news that OpenAI agents attacked Hugging Face was how the agents had worked together.

2 outlets
GitHub Blog: AI and ML4d ago

Secret protection must scale with software

Today, one in three pull requests on GitHub involves an AI agent.

Google Research Blog5d ago

Open and Emergent Problems in Agentic Privacy and Security: A Contextual Angle

Inspired by the theory of Contextual Integrity, our new workshop report outlines key open research directions across system, model, and user levels to build AI agents users can trust.

Repo
Rising AI repositories on GitHub6d ago

GTKottman/mortiflix-oss: A motion design studio on your own machine: Claude makes the video step by step, you approve every stage. Bring your own Claude Code or API

A motion design studio on your own machine: Claude makes the video step by step, you approve every stage.

Code
MIT Technology Review (AI)6d ago

Connecting AI agents to enterprise knowledge

For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge.

TechCrunch: AI2 sources1d ago

Anthropic is cutting off its internal evaluations from the internet

After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations.

2 outlets
Video
NewAI Engineer (YouTube)1h ago

SonarQube + OpenAI: Agentic Development — Killian Carlsen-Phelan, Sonar

Killian Carlsen-Phelan runs it through SonarQube, brings the findings into Codex and asks the agent to fix the vulnerable query.

Video
AI Engineer (YouTube)5h ago

Can Your Agent Hear You Now? Building Live Voice Agents with Gemini — Thor Schaeff

An agent that hears you, sees you, and answers in your language, in real time.

TechCrunch: AI3 sources2d ago

Google brings agentic AI to Gemini, starting with businesses

Google is turning Gemini into an AI agent that can plan, execute tasks, and work across business apps and systems.

3 outlets
Video
AI Engineer (YouTube)7h ago

AI Security Engineer Foundations + Certificate — Micah Silverman, Snyk

Micah Silverman uses that capture the flag exercise to connect AI risks with familiar application security controls.

Video
AI Engineer (YouTube)4h ago

How Secrets Leak Through AI Agents (and How to Stop It) — Venice AI

Joshua Mo, lead developer relations engineer at Venice AI and previously lead maintainer of the Rust AI framework Rig, explains why privacy matters (one breach and users leave, and less stored data means a smaller blast radius) and how Venice approaches it: no stored prompts, anonymized requests to closed providers, models in trusted execution environments, and split-key encrypted storage.

The Decoder5h ago

AI agent teams waste massive tokens for barely measurable quality gains, research finds

Teams of AI agents barely outperform solo agents but cost up to 5.1x more, according to Vals AI.

The Decoder7h ago

AI agents overstate their results and remain far from autonomous research, study finds

Epoch AI and Anthropic independently found the same thing: current AI models like GPT-5.6 Sol and Claude Fable 5 can run experiments but lack scientific self-criticism and genuine creative thinking.

r/LocalLLaMA (top, daily)1d ago

Is anyone running Qwen3.8 Flash Next with a 1M context?

So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN.

Paper
Hugging Face Daily Papers2 sources3d ago

Embodied Turing Machines: Stateful Code for Robot Recursive Self-Improvement

We propose a different view: the embodied world is an Embodied Turing Machine, whose tape is the robot and environment state and rules are the policy.

Paper
Paper
Hugging Face Daily Papers2 sources3d ago

OneSearch-VL: Unified Multimodal Deep Research Agent for Image and Video

We introduce OneSearch-VL, a unified agent centered on the Visually Grounded Evidence Graph (VGEG), which encodes these dependencies as a shared task-level reference for data construction, process supervision, and operation-level evaluation.

Paper
Paper
Hugging Face Daily Papers2 sources3d ago

SuperNav: An Agentic Navigation System for Any Task in Any Scene

General-purpose service robots need navigation systems that can handle diverse human requests in unfamiliar environments, combining task generality with scene generality.

Paper