ResearchResearch paperEfficiency & Inference1 source · Oct 7, 2026

Decoupling Logic from Persona: Structural Immunity of Edge LLM Agents to Context Pollution

We study what happens to the logical part of such an agent when that history is long, misleading and persona-heavy (persona-logic interference), and present a Decoupling Architecture (AO-DA) that separates logical inference ("What") from persona expression ("How") into two inference paths on one INT4 base model with hot-swappable LoRA adapters.

Key points

  • Small language-model agents on edge devices must hold a persona and reason correctly at once, inside one context window that fills with conversational history and persona instructions.
  • The logic path receives only the core turn and emits a verifiable structured state (Micro-State); the persona path renders it in character with the full history.
  • Separation costs one extra decode on a topic's first turn (28.2 s vs 18.2 s on Llama) and buys persona hot-swapping in 1.7 ms without re-running the logic path.
  • Code, rubric, fixtures, adapters and logs are released.

Sources (1)

  • [1]Decoupling Logic from Persona: Structural Immunity of Edge LLM Agents to Context Pollution
    arXiv (AI, ML, NLP, CV, robotics, multi-agent) · Oct 7, 09:53 AM
    We study what happens to the logical part of such an agent when that history is long, misleading and persona-heavy (persona-logic interference), and present a Decoupling Architecture (AO-DA) that separates logical inference ("What") from persona expression ("How") into two inference paths on one INT4 base model with hot-swappable LoRA adapters.
    Small language-model agents on edge devices must hold a persona and reason correctly at once, inside one context window that fills with conversational history and persona instructions.

Extractive summary: sentences quoted from the sources.

Before this

  1. Oct 7, 2026Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute
  2. Oct 6, 2026EmbeddingGemma 2: an open, lightweight multimodal embedding model
  3. Oct 1, 2026unslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UX
  4. Sep 28, 2026unslothai/unsloth v0.1.900-beta: Laya Decision Models + Library
  5. Jul 11, 2026vllm-project/vllm v0.25.0
  6. Jun 29, 2026vllm-project/vllm v0.24.0

Related