AION
Technique

Retrieval-augmented generation

Also known as: RAG, retrieval augmented generation

35stories this week
36last 30 days
38all time

Timeline

  1. Oct 8, 2026 · Research paper · 1 source
    NativeScope: Relation-Localized Retrieval over Native Topology with a Correct Anchor
    We propose NativeScope, a scope-then-rank method for queries with a known anchor and relation.
  2. Oct 8, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.905-beta: Sandboxing is here!
    We're introducing Windows, Mac and Linux sandboxing in Unsloth!
  3. Oct 8, 2026 · Research paper · 1 source
    Is Memorization Context-Sensitive? Prefix-Based Extraction Beyond Isolated Prefixes
    Large language models (LLMs) can expose memorized training sequences under prefix-based extraction: given a prefix from a training example, the model may assign high probability to the original continuation.
  4. Oct 8, 2026 · Research paper · 1 source
    Forms of LLM-Integrated Applications from LLM-Chats to Autonomous AI Agent System
    Large language models (LLMs) are increasingly embedded as components in software systems, marketed under labels such as chatbot, copilot, retrieval-augmented generation, workflow, coding agent and AI agent.
  5. Oct 8, 2026 · Research paper · 1 source
    MAP4CS: A Multi-dimensional Data Pruning Framework for Efficient Code Retriever Fine-tuning
    To address these challenges, we propose MAP4CS (Multi-dimensional Awareness Pruning for Code Search), an adaptive data pruning framework.
  6. Oct 8, 2026 · Research paper · 1 source
    Personalization Matters: Long-Horizon Conversation Agent with User-Centric Information in Online Shopping Interactions
    We propose a multi-agent, multimodal Retrieval-Augmented Generation (RAG) framework that decomposes dialogue state tracking, recommendation retrieval, preference-aware reasoning, and response generation, while integrating product metadata, product reviews, image-derived descriptions, and user historical reviews.
  7. Oct 8, 2026 · Research paper · 1 source
    RIT-RAG: Navigating Document Corpora with Retrieval-Induced Trees
    Retrieval-augmented generation (RAG) grounds language models in external corpora.
  8. Oct 8, 2026 · Research paper · 1 source
    From Retrieval to Reconstruction: Constructing Evolvable Cognitive Memory for Long-Term Dialogue
    Large Language Models (LLMs) serving as long-term dialogue agents require memory systems that support reliable reasoning over extended interactions.
  9. Oct 8, 2026 · Research paper · 1 source
    RAG-Stress: Probing the Limits of Evidence Reliance in Retrieval-Augmented Generation
    We introduce RAG-Stress, a controlled diagnostic protocol for examining the limits of evidence reliance in retrieval-augmented generation.
  10. Oct 7, 2026 · Research paper · 1 source
    When Citations Mislead? A Claim-Level Benchmark for Legal Hallucination Detection
    We introduce PARCEL, a benchmark for checking whether a legal claim is supported by the underlying authority.
  11. Oct 7, 2026 · Research paper · 1 source
    RFChipAgent: Multi-Agentic AI Flow for Analog/RF Chip Design
    We present RFChipAgent, a first-of-its-kind multi-agent flow of large language model (LLM) agents for end-to-end analog/RF circuit design automation, in which AI agents collaboratively orchestrate the complete design flow under human supervision.
  12. Oct 7, 2026 · Opinion / analysis · 1 source
    Rethinking access control for RAG with Amazon Quick and Amazon Bedrock
    Enterprise organizations are adopting Retrieval Augmented Generation (RAG) to unlock insights from company knowledge sources like Microsoft SharePoint, Google Drive, and Atlassian Confluence.
  13. Oct 7, 2026 · Research paper · 1 source
    RECAST: Learning to Compute the Right Context through Adaptive Evidence Routing
    In this work, we introduce RECAST (Routing Evidence through Computation, Access, and Synthesized Tools), a learned framework that formulates evidence construction as a sequential decision process over heterogeneous retrieval and computation operations, allowing evidence to be actively derived rather than merely retrieved.
  14. Oct 7, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.904-beta: Train your own Decision model
    Turn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
  15. Oct 7, 2026 · Research paper · 1 source
    Does Document Structure Help Dense Retrieval? A Placebo-Controlled Ablation of Four Mechanisms Across Two Corpora
    Retrieval-augmented generation systems increasingly rely on document-structure treatments: structure-aligned chunking, LLM-generated chunk contexts, heading-path metadata, and hierarchical two-stage retrieval.
  16. Oct 7, 2026 · Research paper · 1 source
    EntroPrefill: Renyi-Guided Context Pruning with Conditional Stability Guarantees for Retrieval-Augmented Generation
    Mid-prefill pruning can reduce the sequence processed by deeper transformer layers, but attention concentration alone does not certify that discarded context is dispensable.
  17. Oct 7, 2026 · Research paper · 1 source
    Finding the Right Balance: Relevance and Diversity in LLM Retrieval
    Using controlled near-duplicate injection and production-style overlapping chunking, we find that diversification harms relevance, evidence coverage and answer quality on clean pools, but becomes beneficial on multi-evidence tasks when redundancy causes nearest-neighbor retrieval to select repeated passages.
  18. Oct 7, 2026 · Research paper · 1 source
    From Retrieval to Customer Context: Evaluating Frontier-Model Systems for Voice-of-Customer Analysis
    We define a customer context graph as a unified model of customer and business context.
  19. Oct 7, 2026 · Research paper · 1 source
    From Chunks to Functional Evidence: Function-Aware Retrieval for EDA Documentation QA
    Retrieval-Augmented Generation (RAG) is widely used to ground answers in documents.
  20. Oct 7, 2026 · Research paper · 1 source
    TopoGraphRAG-Bench: Evaluating Multimodal GraphRAG on Layout-Grounded Evidence Reasoning
    We introduce TOPOGRAPHRAG-BENCH, a layout-grounded benchmark for multimodal evidence reasoning in GraphRAG, comprising 2,024 questions over 201 long, visually rich documents.
  21. Oct 6, 2026 · Research paper · 1 source
    SPLATIFY: Reproduce, Discover, Innovate! From Papers and Ideas to Trainable 3DGS Code
    We introduce SPLATIFY, a multi-agent framework that converts 3DGS papers into trainable gsplat-based implementations, where generic paper-to-code methods and frontier models fail.
  22. Oct 6, 2026 · Product / feature launch · 1 source
    EmbeddingGemma 2: an open, lightweight multimodal embedding model
    EmbeddingGemma 2: an open, lightweight multimodal embedding model
  23. Oct 6, 2026 · Research paper · 1 source
    BEACON-SP: Ontology-Grounded GraphRAG Framework for Clinical Suicide Risk Assessment
    We present BEACON-SP, an ontology-grounded Graph Retrieval-Augmented Generation (GraphRAG) framework for clinician-facing decision support in behavioral health settings such as suicide prevention, where effective assessment requires integrating heterogeneous clinical, behavioral, social, and temporal evidence.
  24. Oct 6, 2026 · Research paper · 1 source
    EC-RAG: Event Chain Retrieval-Augmented Generation for Long Video Understanding
    In this paper, we propose Event Chain Retrieval-Augmented Generation (EC-RAG), a training-free framework that organizes video content into an explicit event chain before question answering.
  25. Oct 6, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.903-beta: New Browser + Voice Cloning
    This release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
  26. Oct 6, 2026 · Research paper · 1 source
    Trustworthy Domain-Specific AI for Structured Knowledge Retrieval and Reasoning
    This dissertation presents a scalable architecture for transforming unstructured, domain-specific text into structured knowledge for retrieval and reasoning.
  27. Oct 6, 2026 · Research paper · 1 source
    RAG-PIBench: A Leakage-Aware Benchmark for Prompt-Injection Detection in Trustworthy RAG Systems
    We introduce RAG-PIBench, a benchmark for RAG-style prompt-injection detection containing 4,876 contextual examples across frozen train, validation, and protected-test splits.
  28. Oct 6, 2026 · Opinion / analysis · 1 source
    Build a voice travel concierge with Amazon Bedrock AgentCore, Managed Knowledge Base and Nova Sonic
    Airlines already have apps and websites where travelers check flights, pick seats, and manage bookings, and adding a natural voice layer opens those tasks to spoken requests.
  29. Oct 6, 2026 · Research paper · 1 source
    UNREAL: Unifying Retrieval and Long-Context with a Single Model
    We introduce UNifying REtrieval And Long-Context with a Single Model (UNREAL), a model-native evidence selection framework to span corpus retrieval and long-context inference.
  30. Oct 6, 2026 · Research paper · 1 source
    Agentic AutoRAG: RAG Pipeline Optimization through Reasoning-Driven Agents
    We introduce Agentic AutoRAG, an LLM-agent optimizer for multi-objective RAG hyperparameter optimization with retrieval-versus-generation failure attribution.

Often appears with