AI agents
Also known as: AI agent, agentic AI, autonomous agents
52stories this week
61last 30 days
71all time
Timeline
- Oct 11, 2026 · Opinion / analysis · 1 source500B Tokens Later: Letting AI Agents Decompile a First-Person Shooter
- Oct 10, 2026 · Opinion / analysis · 1 sourceTalorys – A self-hosted personal AI agent on Cloudflare's free tier
- Oct 9, 2026 · Opinion / analysis · 1 sourceI expect rapid progress but not towards general superintelligenceI’ve often been surprised when I hear from top researchers in industry that they think AI will be better than them at their job in a few years, and I didn’t really know why I doubted it.
- Oct 9, 2026 · Opinion / analysis · 1 sourceICYMI: What landed for AI builders in September 2026A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026
- Oct 9, 2026 · Opinion / analysis · 1 sourceHow Postman runs Agent Mode for 40 million developers on Amazon BedrockPostman set out to build Agent Mode, an AI-native way to work across API testing, documentation, discovery, and implementation.
- Oct 9, 2026 · Opinion / analysis · 1 sourceShow HN: Let your AI agents paint big arrows, boxes and text on your screen
- Oct 8, 2026 · Opinion / analysis · 1 sourcePay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore paymentsIn this post, we look at how Incarna used AgentCore payments to let its agents pay BlockRun for model inference one request at a time.
- Oct 8, 2026 · Research paper · 1 sourceEcology of AI Agents: Collaboration Creates a Population Threshold for TakeoffHere, we develop an ecological theory of AI-agent populations based on a population growth equation in which fitness (growth rate) depends on cybersecurity capability.
- Oct 8, 2026 · Research paper · 1 sourceCan AI Agents Learn Their Way to the Top? Evaluating Heuristic Learning in a Long-Running Game Agent CompetitionAdversarial games have driven advances from heuristic search to reinforcement learning, yet learning and adapting strategies from limited samples remain challenging.
- Oct 8, 2026 · Research paper · 1 sourceDataSense-Bench: The First Step Toward an AI ScientistWe introduce DataSense-Bench to study this capability through the fundamental problem of data selection and performance forecasting in machine learning.
- Oct 8, 2026 · Research paper · 1 sourceForms of LLM-Integrated Applications from LLM-Chats to Autonomous AI Agent SystemLarge language models (LLMs) are increasingly embedded as components in software systems, marketed under labels such as chatbot, copilot, retrieval-augmented generation, workflow, coding agent and AI agent.
- Oct 8, 2026 · Opinion / analysis · 1 sourceBuilding a safer path to autonomous industrial AIIndustrial AI is entering a new phase.
- Oct 7, 2026 · Research paper · 1 sourceiAm.md: Robot Skill Self-Assessment through Agentic Introspection for Unknown Open-Vocabulary DomainsAgentic AI based on Large Language Model generalization capabilities offers a wide range of potential applications, including planning for embodied tasks.
- Oct 7, 2026 · Research paper · 1 sourceRFChipAgent: Multi-Agentic AI Flow for Analog/RF Chip DesignWe present RFChipAgent, a first-of-its-kind multi-agent flow of large language model (LLM) agents for end-to-end analog/RF circuit design automation, in which AI agents collaboratively orchestrate the complete design flow under human supervision.
- Oct 7, 2026 · Research paper · 1 sourceOn the Clock: Towards Punctual and Productive Time-Budgeted AI AgentsWe study whether small LLM agents can operate effectively under explicit wall-clock time budgets by both respecting the allocated runtime and using available time productively.
- Oct 7, 2026 · Research paper · 1 sourceSciExam for ENSO: Can AI Agents Build Climate Models?The AI Science Exam for El Nino-Southern Oscillation (SciExam for ENSO) is a benchmark in which agents build low-order stochastic models of ENSO, the dominant mode of interannual climate variability, from real observations.
- Oct 7, 2026 · Opinion / analysis · 1 sourceValidate AI Factory Changes with Digital Twins and AI AgentsAI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration...AI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration services, security controls, and a rapidly changing software stack.
- Oct 7, 2026 · Product / feature launch · 1 sourceAgent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real HarnessesHarnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework.
- Oct 7, 2026 · Product / feature launch · 1 sourceBeyond hours saved: Building the business case for agentic automationIn this post, we introduce a framework AI CoE leaders can use to build a business case that captures the full value of agentic automation.
- Oct 7, 2026 · Tutorial / explainer · 1 sourceBuilding AI builders: Playbook for closing the AI knowledge-capability gapThe biggest barrier to AI adoption isn’t awareness.
- Oct 7, 2026 · Research paper · 1 sourceAgentic AI-Assisted Modeling for Production Scheduling: Assessment in Constraint ProgrammingDeveloping optimization models for production scheduling requires substantial expert effort.
- Oct 7, 2026 · Research paper · 1 sourceExperienceIndex: Artifact-Grounded MemoryWe introduce ExperienceIndex, a novel experience layer for AI agents that captures and reuses knowledge about artifacts based on prior reasoning traces.
- Oct 7, 2026 · Research paper · 1 sourceThe Harness as the Only Mutable Surface: Compliance-Bounded Self-Evolution of LLM Agents in Credit Pipelines, with a Measured Admission GateSelf-improving LLM agents can adapt a credit pipeline to a changed rule, but an agent that rewrites itself destroys the artefact a supervisor reviews: a named change, a recorded test, an approval.
- Oct 7, 2026 · Research paper · 1 sourceAgentTime: Can Agents Estimate and Control Their Own Runtime?We present AgentTime, a benchmark for testing whether agents can work for a requested duration, predict their runtime, and estimate elapsed time afterward.
- Oct 7, 2026 · Research paper · 1 sourceEnd-to-End Autonomous Generation of Human Assembly PlansIn this work, we encode long-established design for assembly (DfA) principles into a contained, end-to-end approach for generating assembly plans.
- Oct 7, 2026 · Research paper · 1 sourceOn-Demand Robotic Assembly via Differentiable Geometric Part RepairThis paper presents an end-to-end, autonomous pipeline for the design and physical construction of bespoke wooden assemblies.
- Oct 7, 2026 · Research paper · 1 sourceShared and structured inputs undermine collective random choice by reasoning AI agentsRandom selection is widely used in resource allocation and auditing, making reliable implementation essential for AI-agent systems.
- Oct 7, 2026 · Research paper · 1 sourceLearning Situation-Conditioned Thinking Policies for Long-Term LLM AgentsLong-running autonomous agents must reuse accumulated reasoning experience without allowing explicit historical memory and LLM context to grow indefinitely.
- Oct 7, 2026 · Research paper · 1 sourceDrugTargetWorld: A Synthetic Biobank for Training and Benchmarking AI ScientistsWe introduce DrugTargetWorld, a framework that procedurally generates simulated biobanks, or "worlds," with known but concealed causal structure.
- Oct 7, 2026 · Research paper · 1 sourceVerification and Self-Improvement in Agentic AI: Foundations and LimitsAgentic AI systems can improve by searching longer, receiving additional support, or modifying how they propose and verify outputs.