Claude Code
16stories this week
16last 30 days
17all time
Timeline
- Oct 11, 2026 · Open-source release · 1 sourceBerriAI/litellm v1.105.0Verify using the pinned commit hash (recommended):
- Oct 11, 2026 · Tutorial / explainer · 1 sourceRunning Next Flash IQ3_XXS at ~70 tok/s with 100k context or 2 instances of Qwen 3.6 35B A3B IQ4 at ~145 tok/s with 256k all on $500 of ex mining BC-250 boardsThis will be my third update on the bc-250 cluster and for my first forray into local ai I have been having a blast.
- Oct 8, 2026 · Opinion / analysis · 1 sourcePay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore paymentsIn this post, we look at how Incarna used AgentCore payments to let its agents pay BlockRun for model inference one request at a time.
- Oct 8, 2026 · Research paper · 1 sourceCan AI Agents Learn Their Way to the Top? Evaluating Heuristic Learning in a Long-Running Game Agent CompetitionAdversarial games have driven advances from heuristic search to reinforcement learning, yet learning and adapting strategies from limited samples remain challenging.
- Oct 8, 2026 · Research paper · 1 sourceSWE-Journey: Towards More Realistic Evaluation of Coding Assistants through Long-Horizon, Multi-Turn InteractionTo address these gaps, we introduce SWE-Journey, a benchmark for more realistic evaluation of coding assistants.
- Oct 8, 2026 · Opinion / analysis · 1 sourceI think I found a planet nobody knew existed. I used Claude Code to find it
- Oct 8, 2026 · Research paper · 1 sourceSkill Constellations: Tracing the Supply Chain of Agent Skills on GitHubAgent skills are SKILL.md instructions and scripts that AI coding agents such as Claude Code and Codex run with the permissions of their user.
- Oct 7, 2026 · Research paper · 1 sourceStoreBench: A Live-Commerce Environment for Evaluating and Training Autonomous Operator AgentsWe introduce StoreBench, a live-commerce environment in which an agent runs a mid-size online apparel store on a production-grade commerce backend, testing long-horizon planning and economic judgment under uncertainty.
- Oct 7, 2026 · Product / feature launch · 1 sourceAgent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real HarnessesHarnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework.
- Oct 7, 2026 · Research paper · 1 sourceQuSema: Detecting Silent Bugs in Quantum Libraries via Quantum-knowledge-enhanced AgentsHere we present QuSema, an autonomous testing agent for finding silent bugs in quantum libraries.
- Oct 7, 2026 · Research paper · 1 sourceAgentTime: Can Agents Estimate and Control Their Own Runtime?We present AgentTime, a benchmark for testing whether agents can work for a requested duration, predict their runtime, and estimate elapsed time afterward.
- Oct 6, 2026 · Opinion / analysis · 1 sourceClaude Code’s suggested message feature: I think the real customer is the model
- Oct 6, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.903-beta: New Browser + Voice CloningThis release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
- Oct 6, 2026 · Research paper · 1 sourceSquidAgent: Parallelize Wisely, Coordinate EfficientlyLLM-based agents solve complex multi-step tasks, but sequential execution incurs substantial latency.
- Oct 5, 2026 · Product / feature launch · 1 sourceSupercharge regulated workloads with Claude Code and Amazon BedrockThe availability of Anthropic Claude Opus 5.5 and Claude Sonnet 5.5 in the AWS GovCloud (US) Regions introduces an on-ramp for AI-assisted development for workloads with regulatory or compliance requirements, including International Traffic in Arms Regulations (ITAR).
- Oct 5, 2026 · Product / feature launch · 1 sourceNew agent skill: Amazon SageMaker optimized generative AI inference for your coding agentToday, Amazon SageMaker AI optimized generative AI inference introduces the aws-ai-ml skill, available through the Agent Toolkit for AWS.
- Aug 21, 2026 · Open-source release · 1 sourceollama/ollama v0.33.0Developers can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider.