AnalysisOpinion / analysisAgents & Tool Use · MLOps, Tooling & Infrastructure · Reinforcement Learning1 source · Sep 29, 2026

[AINews] Opus 5.5 is good at explainer videos

Opus 5.5 shipped this week but the vibes are overwhelmingly positive:

Proof1 independent outlet

Key points

  • On vision evals, @skalskip92 ranks it Anthropic’s best vision model to date: better than Fable 5 and GPT-6 Sol, worse than GPT-6 Astra, at about 60% lower cost than Fable 5.1.
  • Terminal-Bench-Science leaders: GPT-6 Astra and Opus 5.5 lead Fable 5.1 by about 20 points.
  • DOOM agent matches show Astra at 82.5% win rate, Sol fastest, Luna best wins/$.
  • Xiaomi MiMo-V2.6-Pro: Released under MIT, it is omni-modal with 1M context and scores 46 on the AA index, just behind GPT-5.6 Sol at 47.

Sources (1)

  • [1][AINews] Opus 5.5 is good at explainer videos
    Latent Space · Sep 29, 02:44 AM
    Opus 5.5 shipped this week but the vibes are overwhelmingly positive:
    On vision evals, @skalskip92 ranks it Anthropic’s best vision model to date: better than Fable 5 and GPT-6 Sol, worse than GPT-6 Astra, at about 60% lower cost than Fable 5.1.

Extractive summary: sentences quoted from the sources.

Before this

  1. Sep 29, 2026Claude Code’s Next Era — Thariq Shihipar, Anthropic
  2. Sep 28, 2026edenfunf/reelmimic: Show it a video you love. Get a new video in the same style. An AI crew (Claude Code or Codex) plans, builds and reviews it with you.
  3. Sep 28, 2026Holo4: powering generalist computer-use agents
  4. Sep 28, 2026Basis completes a tax workbook 2x faster with GPT-6 Astra
  5. Sep 28, 2026Welcome RL Environments to the hub
  6. Sep 27, 2026Louis-CFM/coucou: A tiny friend in your Mac's notch and on your iPhone that keeps an eye on your AI coding agents: Claude Code, Codex, Cursor, Gemini CLI, Ant

Related