AION
Tutorial / explainerLarge Language Models · Reasoning & Planning · Evaluation & Benchmarks1 source · Oct 4, 2026

Qwen3.8 27B addition in words

Research: Qwen3.8 27B addition in words

Key points

  • Colin Frasier posted on Bluesky about an experiment he ran over two years ago using GPT-4o to see how well it could "compute the sum but return the answer in words" across increasingly large numbers.
  • I'm confident GPT-4o didn't cheat and use a calculator, especially since it got so many of the calculations wrong, but I was inspired to run the experiment again on local hardware (a DGX Spark) to explore the effect in a fully controlled environment.
  • I pasted his image into a Codex Remote session (GPT-6 Astra) and had it run the same experiment using Qwen3.8-27B-Q4KM.gguf.
  • Here's a version of the report that includes the reasoning traces from some of those larger calculations, which include text like this:

Sources (1)

  • [1]Qwen3.8 27B addition in words
    Simon Willison's Weblog · Oct 4, 11:34 PM
    Research: Qwen3.8 27B addition in words
    Colin Frasier posted on Bluesky about an experiment he ran over two years ago using GPT-4o to see how well it could "compute the sum but return the answer in words" across increasingly large numbers.

Extractive summary: sentences quoted from the sources.