AION
Product

Codex

22stories this week
26last 30 days
26all time

Timeline

  1. Oct 11, 2026 · Open-source release · 1 source
    BerriAI/litellm v1.105.0
    Verify using the pinned commit hash (recommended):
  2. Oct 10, 2026 · Opinion / analysis · 1 source
    Engineer / developer observations of Gemma4-31B, Qwen3.8-27B, and 6.1-Sol for software engineering work
    Models: Gemma4-31B vs Qwen3.8-27B at the same quantization (an Unsloth flavor of Q4).
  3. Oct 10, 2026 · Opinion / analysis · 1 source
    Are .ipynb notebooks already outdated in the agentic era? [D]
    Back then, Jupyter Notebooks were a perfect fit for the classical DS pipeline: EDA -> data prep -> fit -> eval -> tune -> save model artefact and notebook.
  4. Oct 9, 2026 · Open-source release · 1 source
    pydantic/pydantic-ai v2.55.0: v2.55.0 (2026-10-09)
    <!-- Release notes generated using configuration in .github/release.yml at main -->
  5. Oct 9, 2026 · Product / feature launch · 1 source
    A new feature for my blog, built using my voice
    I used the ChatGPT desktop app for this, in the Codex tab, using the voice conversation mode, running against a local development environment.
  6. Oct 9, 2026 · Opinion / analysis · 1 source
    Asana cuts model costs 76x in browser tests with GPT-6.1 Sol
    Using GPT-6 Astra in Codex, Asana made its browser agent 76x cheaper and 5x faster in tests to offer customers more capable models.
  7. Oct 8, 2026 · Opinion / analysis · 1 source
    Pay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore payments
    In this post, we look at how Incarna used AgentCore payments to let its agents pay BlockRun for model inference one request at a time.
  8. Oct 8, 2026 · Opinion / analysis · 1 source
    How Oracle turns days of work into minutes with ChatGPT and Codex
    Across recruiting, engineering, and operations, Oracle turns specialist knowledge into fast, repeatable workflows with ChatGPT Work and Codex.
  9. Oct 8, 2026 · Research paper · 2 sources
    A Closer Look at Agentic BBO: Benchmarking LLM Agents for Black-Box Optimization
    We therefore introduce AgenticBBO-Bench, a cross-domain benchmark for agentic BBO spanning synthetic functions, hyperparameter optimization, database tuning, chip design, and molecular design under a unified finite-budget evaluation protocol.
  10. Oct 8, 2026 · Opinion / analysis · 1 source
    LegalOn halves Codex costs while maintaining development speed
    LegalOn cut estimated daily Codex costs by 65% while maintaining development speed.
  11. Oct 8, 2026 · Research paper · 1 source
    SWE-Journey: Towards More Realistic Evaluation of Coding Assistants through Long-Horizon, Multi-Turn Interaction
    To address these gaps, we introduce SWE-Journey, a benchmark for more realistic evaluation of coding assistants.
  12. Oct 8, 2026 · Research paper · 1 source
    Skill Constellations: Tracing the Supply Chain of Agent Skills on GitHub
    Agent skills are SKILL.md instructions and scripts that AI coding agents such as Claude Code and Codex run with the permissions of their user.
  13. Oct 7, 2026 · Research paper · 1 source
    Cross-Provider Review as a Runtime Contract for Coding Agents: A Controlled Pilot and Fault-Injection Study
    We describe an advisory cross-provider review contract: distinct resource pools, bounded execution, restricted reviewer capabilities, complete input delivery, usable semantic output, explicit failure states and durable per-attempt evidence.
  14. Oct 7, 2026 · Product / feature launch · 2 sources
    Claude Haiku 5.5
    As previously promised, here's Anthropic's new fast, low cost model: Introducing Claude Haiku 5.5.
  15. Oct 7, 2026 · Product / feature launch · 1 source
    Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses
    Harnessed Agentic RL: Microsoft Research Asia introduces a training paradigm in which the same agent harness used in deployment participates directly in reinforcement learning, removing the need to reimplement the agent inside the training framework.
  16. Oct 7, 2026 · Research paper · 1 source
    QuSema: Detecting Silent Bugs in Quantum Libraries via Quantum-knowledge-enhanced Agents
    Here we present QuSema, an autonomous testing agent for finding silent bugs in quantum libraries.
  17. Oct 7, 2026 · Research paper · 1 source
    AgentTime: Can Agents Estimate and Control Their Own Runtime?
    We present AgentTime, a benchmark for testing whether agents can work for a requested duration, predict their runtime, and estimate elapsed time afterward.
  18. Oct 6, 2026 · Tutorial / explainer · 1 source
    Using Parseable with Datasette for OpenTelemetry traces
    TIL: Using Parseable with Datasette for OpenTelemetry traces
  19. Oct 6, 2026 · Research paper · 2 sources
    Can AI Agents Make Open-Ended Scientific Discovery? Evidence from Station
    Recent AI systems have made rapid progress in scientific discovery when given well-defined metrics, but whether they can autonomously undertake open-ended scientific discovery remains unclear.
  20. Oct 6, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.903-beta: New Browser + Voice Cloning
    This release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
  21. Oct 5, 2026 · Product / feature launch · 1 source
    New agent skill: Amazon SageMaker optimized generative AI inference for your coding agent
    Today, Amazon SageMaker AI optimized generative AI inference introduces the aws-ai-ml skill, available through the Agent Toolkit for AWS.
  22. Oct 4, 2026 · Tutorial / explainer · 1 source
    Qwen3.8 27B addition in words
    Research: Qwen3.8 27B addition in words
  23. Oct 2, 2026 · Product / feature launch · 1 source
    Chatham scales its capital markets expertise with OpenAI
    Chatham Financial uses Codex and GPT-5.6 to build technology and redesign workflows, cutting trade validation from 30 minutes to under 4.
  24. Sep 29, 2026 · Product / feature launch · 1 source
    DevDay 2026 Recap
    Explore more than 20 announcements from OpenAI DevDay 2026, including GPT-6 Astra, ChatGPT, Codex, APIs, security, and new tools for builders.
  25. Sep 28, 2026 · Opinion / analysis · 1 source
    Are you a Codex Original?
    We’re collecting real stories of builders, tinkerers, researchers, and creators who are using Codex to do incredible things.
  26. Sep 23, 2026 · Open-source release · 1 source
    ollama/ollama v0.34.4
    Qwen 3.8 prompt processing is faster on Apple Silicon.

Often appears with