Analysis
Opinion, explainers and guides from people worth reading.

Satya Nadella says we should assume all AI models are ‘compromised’
In a lengthy post on X, Microsoft's CEO laid out his views on the dangers posed by highly advanced AI models and how to confront those risks.
OpenAI’s revenue is reportedly $20 billion less than previously projected
It had previously been reported that the AI lab's annualized revenue was some $70 billion, but a new report claims it's a whole lot less than that.
VideoSonarQube + OpenAI: Agentic Development — Killian Carlsen-Phelan, Sonar
Killian Carlsen-Phelan runs it through SonarQube, brings the findings into Codex and asks the agent to fix the vulnerable query.
Quoting The New York Times
— The New York Times, Anthropic Agents Tried to Fill Out Visa Forms on State Dept.

Why AlphaFold Didn't Solve Protein Folding — Pushmeet Kohli, Google DeepMind & Sal Candido, Biohub
From the Bitter Lesson of AI scaling to the unsolved mysteries of protein folding, Google DeepMind’s Pushmeet Kohli and Biohub’s Sal Candido are rethinking what it takes to build AI that truly understands biology.
Sophos cuts threat investigation time by 96% with OpenAI Daybreak
Discover how Sophos uses OpenAI’s Daybreak to cut cyber-threat investigation time by 96% and automate 52% of MDR cases while preserving human oversight.
I Made Terrible Games With Google’s AI Playground
A long day’s haul in the video game slop mines.
ICYMI: What landed for AI builders in September 2026
A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026
What's up with google scholar citations ? [D]
These papers have been there for months, but Google Scholar still has not updated its citations, and it has occurred before, but I ignored it and thought it was a one-time error.
VideoWhy DeepSeek Wants AI To Forget
📝 The DeepSeek OCR paper is available here:

OpenAI says a misaligned model deliberately destroyed its own environment hoping for a fresh start with better data
OpenAI has documented new cases of misaligned model behavior.

Why agent swarms could be the next “scaling law”
One of the most surprising aspects of July’s news that OpenAI agents attacked Hugging Face was how the agents had worked together.
Every Model That Can Be Run On 10-16GB VRAM Ranked
It's been 4 years since c.ai first hallucination model, and yet, we're nowhere good enough at LLMs in terms of spontaneity/interesting hallucination features.

Ukraine’s drones knock out AI data center belonging to "Russia’s Google"
Ukrainian drone strikes have knocked out two of five data centers belonging to the Russian tech giant Yandex.
Byte Language Models: Scaling, Emergent Abstractions, and Information Allocation
The paper challenges the assumption that language models need explicit tokenizers to be efficient demonstrating that standard flat Transformers can process raw byte sequences and actually outperform traditional subword models as parameter sizes scale.

The State of AI Report 2026
After months of research and revisions right up to the last minute, I’m thrilled to bring you the 9th annual State of AI Report.

Text is so 2023
Hark (from Brett Adcock - Figure robots) launched which suggests tasks as one-tap action buttons.
Unlocking Earth AI’s planetary geospatial foundation models for global public health
In our latest work, we present five partner-driven case studies demonstrating how this model exemplifies the planetary geospatial foundation model paradigm for global public health.
OpenAI revenue falls short, models play hopscotch and Trump cracks down on tech green cards
OpenAI told investors this week that it actually had $18 billion less revenue than the $68 [...]

The Epoch Brief - October 8, 2026
Welcome back to the Epoch Brief.
What AI gets wrong and what failure teaches us
Jennifer Neville is a partner research manager at Microsoft who’s built a career around understanding and advancing AI for real-world use, and much like the human-AI interactions she’s been studying, her early-career path was multiturn: math, then physics; cognitive science, then work; and finally computer science—despite her best efforts to avoid the field.
Validate AI Factory Changes with Digital Twins and AI Agents
AI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration...AI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration services, security controls, and a rapidly changing software stack.

Language Models for Text Classification: From Bag-of-Words to Jev
The recently released Jev AI model has been quite a cultural phenomenon in technical communities in the past 2 weeks.

The missing map of the sky
Here, Brice Ménard, an astrophysicist at Johns Hopkins University and a researcher at Anthropic, explains how he worked with Claude Science to produce the first complete map of the sky in UV light.

Attempts to Keep Humans in the AI Loop May Actually Push Them Out
A crucial safeguard against AI agents going rogue—keeping humans in the loop to review and approve their decisions—will fail unless designers and users change their current practices, a trio of leading AI ethics researchers argue.
VideoHolo4: A Model That Clicks, Codes and Calls Tools
How one open-weight model click through a GUI, write and run code, and call MCP tools, and also work out which one to use at each step.
Connecting AI agents to enterprise knowledge
For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge.

LWiAI Podcast #258 - Opus 5.5, Sol and Luna, Muse, DeepSeek-V4.1-Flash, Xi
Our 258th episode with a summary and discussion of last week’s big AI news!

🔮 Quick weekend reads: the big AI questions
AI is really forcing some big questions into the open.
VideoOpenAI Security: Controlling Models is Now ‘Hell’
A cracked cipher, an OpenAI security warning, Gemini 4 Argon, RSI paper (co-authored by a who’s who of AI), Lab White House commitments, new hacks emerging, ‘deep personas’, biology Kasparov competitions, and so much more, ending with an epic Opus outro.
Show HN: Made an open-source Lego AI generator
So, the idea I had was: if I manage for maybe ChatGPT or Claude to generate high-quality LDraw source files... then, they would actually be generating high-quality LEGO CAD models, right?

Cloudflare OS: your company’s agent workspace, managed for you
Cloudflare OS gives everyone in your organization an agent workspace that knows how your company works and connects to its data and systems.
AI is changing developer work. Here are three skills to strengthen.
AI is changing how developers work and apply their skills.
Graph-centric agentic intelligence
In a world where AI agents need to reason about complex systems, not just retrieve information, graph structure provides the scaffolding for causal inference, dependency tracking, and compositional reasoning.

Limits of Confidence in Diffusion
Discrete diffusion, including remasking and uniform-state samplers, generate a sequence by writing multiple token positions per step, drawing each from a per-position distribution and choosing which positions to write from those same distributions.
Show colleagues how much you care by sending an AI avatar to your Google Meet call
If you've always wanted to say you have "people" to attend your meetings, there's an app in development for that, as long as you're comfortable with the idea of eventually sending an AI bot in your place and potentially offending your human colleagues.
Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents
Our latest paper, ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents (read it on Hugging Face, or on arXiv in the meantime), targets that gap.

Notes on NVIDIA Nemotron
Today, many key details of frontier large language models (LLMs) remain proprietary, but open-weights model families—such as DeepSeek, Kimi, and MiMo—continue to provide a valuable window into the development process for modern LLMs. Among these resources, the NVIDIA Nemotron model series is especially useful due to its transparency.
Ollama now supports Jev-style decision models
Ollama now supports decision models, based on TypeSafe's Jev API for fast, typed decisions.

How GLM5.3 Sparse Attention Affects HBM Memory Usage
How Sparse Attention Affects DRAM/NAND Memory

The Era of Personal Super-Intelligent Agents
Meta Muse is now the SOTA in Personal AI agents.

Team Bots: shared AI teammates that learn as they work
Today we’re launching Team Bots, Grok Bots that work and learn alongside your team.
Google Beam expands with new regions, partners, and customers
We’re expanding Google Beam to five new countries, and partnering with Industrious for an extended network.
VideoTeaching Agents to Search with NVIDIA Data Designer — Dhruv Nathawani, NVIDIA
Dhruv Nathawani uses that funnel to explain how NVIDIA builds synthetic data that teaches a model to search, rather than answer from memory.
VideoCan Your Agent Hear You Now? Building Live Voice Agents with Gemini — Thor Schaeff
An agent that hears you, sees you, and answers in your language, in real time.

These execs think voice AI hasn’t reached its ChatGPT moment yet
Voice AI's often misses important points for its context layer, and causes the whole pipeline to break
Pollo AI turns creative ideas into campaigns with OpenAI
With GPT-5.6, GPT-6 Astra, and GPT‐Image‐2.5, Pollo AI helps creators turn bold ideas into detailed images and cinematic video ads.
ttok 1.0
I released ttok 0.4, ran uv tool upgrade ttok, piped a file into the new version... and realized that it was defaulting to the GPT-4 tokenizer when it should very clearly default to GPT-5/GPT-6 instead!