AION
Model

Llama

Also known as: LLaMA

6stories this week
6last 30 days
6all time

Timeline

  1. Oct 11, 2026 · Tutorial / explainer · 1 source
    Running Next Flash IQ3_XXS at ~70 tok/s with 100k context or 2 instances of Qwen 3.6 35B A3B IQ4 at ~145 tok/s with 256k all on $500 of ex mining BC-250 boards
    This will be my third update on the bc-250 cluster and for my first forray into local ai I have been having a blast.
  2. Oct 8, 2026 · Research paper · 1 source
    Spectral Weight Decay: Inducing Low-Rank Structure in Neural Network Weights
    We introduce spectral weight decay, a post-step decoupled nuclear-norm update that applies additive rather than multiplicative spectral shrinkage.
  3. Oct 8, 2026 · Research paper · 1 source
    LadderEdit: Edit-Level Residual Compression for Memory-Efficient Lifelong Editing of LLMs
    Lifelong editing of LLMs requires storing thousands of edits after acquisition.
  4. Oct 7, 2026 · Research paper · 1 source
    Document-Level Text Simplification in Estonian Using Large Language Models
    Despite advances in sentence-level simplification for high-resource languages, document-level simplification in morphologically rich, low-resource languages such as Estonian remains largely unexplored.
  5. Oct 6, 2026 · Research paper · 1 source
    Few Bits, One Law: Toward W2A4KV2
    We introduce CanonQ, a unified quantization-aware training framework that addresses these challenges by separating source canonicalization from task-aware adaptation.
  6. Oct 6, 2026 · Research paper · 1 source
    How Fragile Is On-Device Language Model Safety? Localizing Safety-Critical Parameters for Sparse Fault Analysis
    As small language models (SLMs) are increasingly deployed on resource-constrained and on-device platforms, including as components of agentic systems, the integrity of locally stored model parameters becomes an important safety concern.

Often appears with