Analysis

Opinion, explainers and guides from people worth reading.

The Verge: AI4 sources12h ago

Satya Nadella says we should assume all AI models are ‘compromised’

In a lengthy post on X, Microsoft's CEO laid out his views on the dangers posed by highly advanced AI models and how to confront those risks.

3 outlets
TechCrunch: AI2 sources2d ago

OpenAI’s revenue is reportedly $20 billion less than previously projected

It had previously been reported that the AI lab's annualized revenue was some $70 billion, but a new report claims it's a whole lot less than that.

2 outlets
Video
NewAI Engineer (YouTube)2h ago

SonarQube + OpenAI: Agentic Development — Killian Carlsen-Phelan, Sonar

Killian Carlsen-Phelan runs it through SonarQube, brings the findings into Codex and asks the agent to fix the vulnerable query.

Simon Willison's Weblog1d ago

Quoting The New York Times

— The New York Times, Anthropic Agents Tried to Fill Out Visa Forms on State Dept.

Latent Space1d ago

Why AlphaFold Didn't Solve Protein Folding — Pushmeet Kohli, Google DeepMind & Sal Candido, Biohub

From the Bitter Lesson of AI scaling to the unsolved mysteries of protein folding, Google DeepMind’s Pushmeet Kohli and Biohub’s Sal Candido are rethinking what it takes to build AI that truly understands biology.

OpenAI News2d ago

Sophos cuts threat investigation time by 96% with OpenAI Daybreak

Discover how Sophos uses OpenAI’s Daybreak to cut cyber-threat investigation time by 96% and automate 52% of MDR cases while preserving human oversight.

WIRED: AI10h ago

I Made Terrible Games With Google’s AI Playground

A long day’s haul in the video game slop mines.

AWS Machine Learning Blog2d ago

ICYMI: What landed for AI builders in September 2026

A recap of the latest Amazon Bedrock, Amazon Bedrock AgentCore, and Strands updates from September 2026

r/MachineLearning (top, daily)3h ago

What's up with google scholar citations ? [D]

These papers have been there for months, but Google Scholar still has not updated its citations, and it has occurred before, but I ignored it and thought it was a one-time error.

Video
Two Minute Papers (YouTube)4h ago

Why DeepSeek Wants AI To Forget

📝 The DeepSeek OCR paper is available here:

The Decoder1d ago

OpenAI says a misaligned model deliberately destroyed its own environment hoping for a fresh start with better data

OpenAI has documented new cases of misaligned model behavior.

Import AI (Jack Clark)2 sources6d ago

Why agent swarms could be the next “scaling law”

One of the most surprising aspects of July’s news that OpenAI agents attacked Hugging Face was how the agents had worked together.

2 outlets
Newr/LocalLLaMA (top, daily)2h ago

Every Model That Can Be Run On 10-16GB VRAM Ranked

It's been 4 years since c.ai first hallucination model, and yet, we're nowhere good enough at LLMs in terms of spontaneity/interesting hallucination features.

Ars Technica: AI1d ago

Ukraine’s drones knock out AI data center belonging to "Russia’s Google"

Ukrainian drone strikes have knocked out two of five data centers belonging to the Russian tech giant Yandex.

Lobsters: ai23h ago

Byte Language Models: Scaling, Emergent Abstractions, and Information Allocation

The paper challenges the assumption that language models need explicit tokenizers to be efficient demonstrating that standard flat Transformers can process raw byte sequences and actually outperform traditional subword models as parameter sizes scale.

Paper
Guide to AI (Nathan Benaich)3d ago

The State of AI Report 2026

After months of research and revisions right up to the last minute, I’m thrilled to bring you the 9th annual State of AI Report.

Ben's Bites3d ago

Text is so 2023

Hark (from Brett Adcock - Figure robots) launched which suggests tasks as one-tap action buttons.

Google Research Blog5d ago

Unlocking Earth AI’s planetary geospatial foundation models for global public health

In our latest work, we present five partner-driven case studies demonstrating how this model exemplifies the planetary geospatial foundation model paradigm for global public health.

SiliconANGLE: AI2d ago

OpenAI revenue falls short, models play hopscotch and Trump cracks down on tech green cards

OpenAI told investors this week that it actually had $18 billion less revenue than the $68 [...]

Epoch AI: Gradient Updates3d ago

The Epoch Brief - October 8, 2026

Welcome back to the Epoch Brief.

Microsoft Research Blog5d ago

What AI gets wrong and what failure teaches us

Jennifer Neville is a partner research manager at Microsoft who’s built a career around understanding and advancing AI for real-world use, and much like the human-AI interactions she’s been studying, her early-career path was multiturn: math, then physics; cognitive science, then work; and finally computer science—despite her best efforts to avoid the field.

NVIDIA Technical Blog4d ago

Validate AI Factory Changes with Digital Twins and AI Agents

AI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration...AI factories are some of the most complex operations in the world, combining GPUs, CPUs, switches, DPUs, and SuperNICs alongside schedulers, orchestration services, security controls, and a rapidly changing software stack.

Ahead of AI magazine (Sebastian Raschka)3 sources9d ago

Language Models for Text Classification: From Bag-of-Words to Jev

The recently released Jev AI model has been quite a cultural phenomenon in technical communities in the past 2 weeks.

3 outlets
Anthropic Research3d ago

The missing map of the sky

Here, Brice Ménard, an astrophysicist at Johns Hopkins University and a researcher at Anthropic, explains how he worked with Claude Science to produce the first complete map of the sky in UV light.

IEEE Spectrum: AI6d ago

Attempts to Keep Humans in the AI Loop May Actually Push Them Out

A crucial safeguard against AI agents going rogue—keeping humans in the loop to review and approve their decisions—will fail unless designers and users change their current practices, a trio of leading AI ethics researchers argue.

Video
Sam Witteveen (YouTube)5d ago

Holo4: A Model That Clicks, Codes and Calls Tools

How one open-weight model click through a GUI, write and run code, and call MCP tools, and also work out which one to use at each step.

MIT Technology Review (AI)6d ago

Connecting AI agents to enterprise knowledge

For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge.

Last Week in AI8d ago

LWiAI Podcast #258 - Opus 5.5, Sol and Luna, Muse, DeepSeek-V4.1-Flash, Xi

Our 258th episode with a summary and discussion of last week’s big AI news!

Exponential View (Azeem Azhar)8d ago

🔮 Quick weekend reads: the big AI questions

AI is really forcing some big questions into the open.

Video
AI Explained (YouTube)10d ago

OpenAI Security: Controlling Models is Now ‘Hell’

A cracked cipher, an OpenAI security warning, Gemini 4 Argon, RSI paper (co-authored by a who’s who of AI), Lab White House commitments, new hacks emerging, ‘deep personas’, biology Kasparov competitions, and so much more, ending with an epic Opus outro.

Show HN: AI projects (15+ points)9d ago

Show HN: Made an open-source Lego AI generator

So, the idea I had was: if I manage for maybe ChatGPT or Claude to generate high-quality LDraw source files... then, they would actually be generating high-quality LEGO CAD models, right?

Code
Cloudflare Blog: AI10d ago

Cloudflare OS: your company’s agent workspace, managed for you

Cloudflare OS gives everyone in your organization an agent workspace that knows how your company works and connects to its data and systems.

GitHub Blog: AI and ML9d ago

AI is changing developer work. Here are three skills to strengthen.

AI is changing how developers work and apply their skills.

Amazon Science10d ago

Graph-centric agentic intelligence

In a world where AI agents need to reason about complex systems, not just retrieve information, graph structure provides the scaffolding for causal inference, dependency tracking, and compositional reasoning.

Apple Machine Learning Research9d ago

Limits of Confidence in Diffusion

Discrete diffusion, including remasking and uniform-state samplers, generate a sequence by writing multiple token positions per step, drawing each from a per-position distribution and choosing which positions to write from those same distributions.

The Register: AI/ML11d ago

Show colleagues how much you care by sending an AI avatar to your Google Meet call

If you've always wanted to say you have "people" to attend your meetings, there's an app in development for that, as long as you're comfortable with the idea of eventually sending an AI bot in your place and potentially offending your human colleagues.

Hugging Face Blog12d ago

Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents

Our latest paper, ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents (read it on Hugging Face, or on arXiv in the meantime), targets that gap.

Deep (Learning) Focus (Cameron Wolfe)13d ago

Notes on NVIDIA Nemotron

Today, many key details of frontier large language models (LLMs) remain proprietary, but open-weights model families—such as DeepSeek, Kimi, and MiMo—continue to provide a valuable window into the development process for modern LLMs. Among these resources, the NVIDIA Nemotron model series is especially useful due to its transparency.

Ollama Blog12d ago

Ollama now supports Jev-style decision models

Ollama now supports decision models, based on TypeSafe's Jev API for fast, typed decisions.

SemiAnalysis13d ago

How GLM5.3 Sparse Attention Affects HBM Memory Usage

How Sparse Attention Affects DRAM/NAND Memory

AI Supremacy12d ago

The Era of Personal Super-Intelligent Agents

Meta Muse is now the SOTA in Personal AI agents.

xAI News13d ago

Team Bots: shared AI teammates that learn as they work

Today we’re launching Team Bots, Grok Bots that work and learn alongside your team.

Google Blog (AI)18d ago

Google Beam expands with new regions, partners, and customers

We’re expanding Google Beam to five new countries, and partnering with Industrious for an extended network.

Video
AI Engineer (YouTube)5h ago

Teaching Agents to Search with NVIDIA Data Designer — Dhruv Nathawani, NVIDIA

Dhruv Nathawani uses that funnel to explain how NVIDIA builds synthetic data that teaches a model to search, rather than answer from memory.

Video
AI Engineer (YouTube)6h ago

Can Your Agent Hear You Now? Building Live Voice Agents with Gemini — Thor Schaeff

An agent that hears you, sees you, and answers in your language, in real time.

TechCrunch: AI7h ago

These execs think voice AI hasn’t reached its ChatGPT moment yet

Voice AI's often misses important points for its context layer, and causes the whole pipeline to break

OpenAI News3d ago

Pollo AI turns creative ideas into campaigns with OpenAI

With GPT-5.6, GPT-6 Astra, and GPT‐Image‐2.5, Pollo AI helps creators turn bold ideas into detailed images and cinematic video ads.

Simon Willison's Weblog2d ago

ttok 1.0

I released ttok 0.4, ran uv tool upgrade ttok, piped a file into the new version... and realized that it was defaulting to the GPT-4 tokenizer when it should very clearly default to GPT-5/GPT-6 instead!