AION
Organization

Apple

4stories this week
10last 30 days
18all time

Timeline

  1. Oct 10, 2026 · Opinion / analysis · 1 source
    Open-source Mac app that runs EmbeddingGemma 2 locally to search your files by what’s in them
    DigUp is a free Mac app that runs Google DeepMind’s new EmbeddingGemma 2 locally over your own files.
  2. Oct 7, 2026 · Research paper · 1 source
    Decoupling Logic from Persona: Structural Immunity of Edge LLM Agents to Context Pollution
    We study what happens to the logical part of such an agent when that history is long, misleading and persona-heavy (persona-logic interference), and present a Decoupling Architecture (AO-DA) that separates logical inference ("What") from persona expression ("How") into two inference paths on one INT4 base model with hot-swappable LoRA adapters.
  3. Oct 6, 2026 · Research paper · 1 source
    Breaking the Space Barrier and its Application to Language Model Inference
    Language models are more and more often asked for structured output: JSON that follows a schema, or a tool call with typed arguments.
  4. Oct 6, 2026 · Opinion / analysis · 1 source
    What AI gets wrong and what failure teaches us
    Jennifer Neville is a partner research manager at Microsoft who’s built a career around understanding and advancing AI for real-world use, and much like the human-AI interactions she’s been studying, her early-career path was multiturn: math, then physics; cognitive science, then work; and finally computer science—despite her best efforts to avoid the field.
  5. Oct 1, 2026 · Product / feature launch · 1 source
    unslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UX
    This release brings faster navigation, shareable run settings, and clearer errors to Unsloth Desktop.
  6. Sep 28, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.900-beta: Laya Decision Models + Library
    Run and serve Decision Models like Laya (open-source Jev) locally
  7. Sep 25, 2026 · Open-source release · 1 source
    ollama/ollama v0.40.0
    Models run on MLX on Apple Silicon by default
  8. Sep 23, 2026 · Open-source release · 1 source
    ollama/ollama v0.34.4
    Qwen 3.8 prompt processing is faster on Apple Silicon.
  9. Sep 19, 2026 · Open-source release · 1 source
    ollama/ollama v0.34.3
    GET /api/show now advertises each model's thinking controls and default:
  10. Sep 14, 2026 · Open-source release · 1 source
    ollama/ollama v0.34.1
    GGUF model creation now requires using llama.cpp tooling for safetensor conversion and quantization.
  11. Sep 5, 2026 · Open-source release · 1 source
    ollama/ollama v0.34.0
    Use Ollama models in ChatGPT Desktop
  12. Sep 2, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.14.0: PyTorch 2.14.0 Release
    <tr><td><strong>A preview of our rewritten NCCL backend for PyTorch</strong>, ported from torchcomms, implementing the full collective contract with nonblocking communicators and eager communicator splitting and advanced features such as fault tolerance and windows designed as a drop-in replacement of existing NCCL c10d backend</td></tr>
  13. Aug 14, 2026 · Open-source release · 1 source
    ollama/ollama v0.32.12
    Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks.
  14. Jul 15, 2026 · Open-source release · 1 source
    huggingface/transformers v5.14.0: Release v5.14.0
    Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and
  15. Jul 8, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.13.0: PyTorch 2.13.0 Release
    <tr><td><strong>torchcomms</strong>, a new communications backend for PyTorch Distributed, improves fault tolerance, scalability, and debuggability for large-cluster training.</td></tr>
  16. Apr 6, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.10
    Piecewise CUDA Graph Enabled by Default: Piecewise CUDA graph capture is now the default execution mode, reducing memory overhead and improving throughput for models with complex control flow patterns: #16331
  17. Mar 28, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.10rc0
    Piecewise CUDA Graph Enabled by Default: Piecewise CUDA graph capture is now the default execution mode, reducing memory overhead and improving throughput for models with complex control flow patterns: #16331
  18. Mar 23, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.11.0: PyTorch 2.11.0 Release
    <strong>FlexAttention</strong> now has a <strong>FlashAttention-4</strong> backend on <strong>Hopper</strong> and <strong>Blackwell</strong> GPUs

Often appears with