Apple
4stories this week
10last 30 days
18all time
Timeline
- Oct 10, 2026 · Opinion / analysis · 1 sourceOpen-source Mac app that runs EmbeddingGemma 2 locally to search your files by what’s in themDigUp is a free Mac app that runs Google DeepMind’s new EmbeddingGemma 2 locally over your own files.
- Oct 7, 2026 · Research paper · 1 sourceDecoupling Logic from Persona: Structural Immunity of Edge LLM Agents to Context PollutionWe study what happens to the logical part of such an agent when that history is long, misleading and persona-heavy (persona-logic interference), and present a Decoupling Architecture (AO-DA) that separates logical inference ("What") from persona expression ("How") into two inference paths on one INT4 base model with hot-swappable LoRA adapters.
- Oct 6, 2026 · Research paper · 1 sourceBreaking the Space Barrier and its Application to Language Model InferenceLanguage models are more and more often asked for structured output: JSON that follows a schema, or a tool call with typed arguments.
- Oct 6, 2026 · Opinion / analysis · 1 sourceWhat AI gets wrong and what failure teaches usJennifer Neville is a partner research manager at Microsoft who’s built a career around understanding and advancing AI for real-world use, and much like the human-AI interactions she’s been studying, her early-career path was multiturn: math, then physics; cognitive science, then work; and finally computer science—despite her best efforts to avoid the field.
- Oct 1, 2026 · Product / feature launch · 1 sourceunslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UXThis release brings faster navigation, shareable run settings, and clearer errors to Unsloth Desktop.
- Sep 28, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.900-beta: Laya Decision Models + LibraryRun and serve Decision Models like Laya (open-source Jev) locally
- Sep 25, 2026 · Open-source release · 1 sourceollama/ollama v0.40.0Models run on MLX on Apple Silicon by default
- Sep 23, 2026 · Open-source release · 1 sourceollama/ollama v0.34.4Qwen 3.8 prompt processing is faster on Apple Silicon.
- Sep 19, 2026 · Open-source release · 1 sourceollama/ollama v0.34.3GET /api/show now advertises each model's thinking controls and default:
- Sep 14, 2026 · Open-source release · 1 sourceollama/ollama v0.34.1GGUF model creation now requires using llama.cpp tooling for safetensor conversion and quantization.
- Sep 5, 2026 · Open-source release · 1 sourceollama/ollama v0.34.0Use Ollama models in ChatGPT Desktop
- Sep 2, 2026 · Open-source release · 1 sourcepytorch/pytorch v2.14.0: PyTorch 2.14.0 Release<tr><td><strong>A preview of our rewritten NCCL backend for PyTorch</strong>, ported from torchcomms, implementing the full collective contract with nonblocking communicators and eager communicator splitting and advanced features such as fault tolerance and windows designed as a drop-in replacement of existing NCCL c10d backend</td></tr>
- Aug 14, 2026 · Open-source release · 1 sourceollama/ollama v0.32.12Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks.
- Jul 15, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.14.0: Release v5.14.0Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and
- Jul 8, 2026 · Open-source release · 1 sourcepytorch/pytorch v2.13.0: PyTorch 2.13.0 Release<tr><td><strong>torchcomms</strong>, a new communications backend for PyTorch Distributed, improves fault tolerance, scalability, and debuggability for large-cluster training.</td></tr>
- Apr 6, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.10Piecewise CUDA Graph Enabled by Default: Piecewise CUDA graph capture is now the default execution mode, reducing memory overhead and improving throughput for models with complex control flow patterns: #16331
- Mar 28, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.10rc0Piecewise CUDA Graph Enabled by Default: Piecewise CUDA graph capture is now the default execution mode, reducing memory overhead and improving throughput for models with complex control flow patterns: #16331
- Mar 23, 2026 · Open-source release · 1 sourcepytorch/pytorch v2.11.0: PyTorch 2.11.0 Release<strong>FlexAttention</strong> now has a <strong>FlashAttention-4</strong> backend on <strong>Hopper</strong> and <strong>Blackwell</strong> GPUs