Retrieval-augmented generation
Also known as: RAG, retrieval augmented generation
35stories this week
36last 30 days
38all time
Timeline
- Oct 8, 2026 · Research paper · 1 sourceNativeScope: Relation-Localized Retrieval over Native Topology with a Correct AnchorWe propose NativeScope, a scope-then-rank method for queries with a known anchor and relation.
- Oct 8, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.905-beta: Sandboxing is here!We're introducing Windows, Mac and Linux sandboxing in Unsloth!
- Oct 8, 2026 · Research paper · 1 sourceIs Memorization Context-Sensitive? Prefix-Based Extraction Beyond Isolated PrefixesLarge language models (LLMs) can expose memorized training sequences under prefix-based extraction: given a prefix from a training example, the model may assign high probability to the original continuation.
- Oct 8, 2026 · Research paper · 1 sourceForms of LLM-Integrated Applications from LLM-Chats to Autonomous AI Agent SystemLarge language models (LLMs) are increasingly embedded as components in software systems, marketed under labels such as chatbot, copilot, retrieval-augmented generation, workflow, coding agent and AI agent.
- Oct 8, 2026 · Research paper · 1 sourceMAP4CS: A Multi-dimensional Data Pruning Framework for Efficient Code Retriever Fine-tuningTo address these challenges, we propose MAP4CS (Multi-dimensional Awareness Pruning for Code Search), an adaptive data pruning framework.
- Oct 8, 2026 · Research paper · 1 sourcePersonalization Matters: Long-Horizon Conversation Agent with User-Centric Information in Online Shopping InteractionsWe propose a multi-agent, multimodal Retrieval-Augmented Generation (RAG) framework that decomposes dialogue state tracking, recommendation retrieval, preference-aware reasoning, and response generation, while integrating product metadata, product reviews, image-derived descriptions, and user historical reviews.
- Oct 8, 2026 · Research paper · 1 sourceRIT-RAG: Navigating Document Corpora with Retrieval-Induced TreesRetrieval-augmented generation (RAG) grounds language models in external corpora.
- Oct 8, 2026 · Research paper · 1 sourceFrom Retrieval to Reconstruction: Constructing Evolvable Cognitive Memory for Long-Term DialogueLarge Language Models (LLMs) serving as long-term dialogue agents require memory systems that support reliable reasoning over extended interactions.
- Oct 8, 2026 · Research paper · 1 sourceRAG-Stress: Probing the Limits of Evidence Reliance in Retrieval-Augmented GenerationWe introduce RAG-Stress, a controlled diagnostic protocol for examining the limits of evidence reliance in retrieval-augmented generation.
- Oct 7, 2026 · Research paper · 1 sourceWhen Citations Mislead? A Claim-Level Benchmark for Legal Hallucination DetectionWe introduce PARCEL, a benchmark for checking whether a legal claim is supported by the underlying authority.
- Oct 7, 2026 · Research paper · 1 sourceRFChipAgent: Multi-Agentic AI Flow for Analog/RF Chip DesignWe present RFChipAgent, a first-of-its-kind multi-agent flow of large language model (LLM) agents for end-to-end analog/RF circuit design automation, in which AI agents collaboratively orchestrate the complete design flow under human supervision.
- Oct 7, 2026 · Opinion / analysis · 1 sourceRethinking access control for RAG with Amazon Quick and Amazon BedrockEnterprise organizations are adopting Retrieval Augmented Generation (RAG) to unlock insights from company knowledge sources like Microsoft SharePoint, Google Drive, and Atlassian Confluence.
- Oct 7, 2026 · Research paper · 1 sourceRECAST: Learning to Compute the Right Context through Adaptive Evidence RoutingIn this work, we introduce RECAST (Routing Evidence through Computation, Access, and Synthesized Tools), a learned framework that formulates evidence construction as a sequential decision process over heterogeneous retrieval and computation operations, allowing evidence to be actively derived rather than merely retrieved.
- Oct 7, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.904-beta: Train your own Decision modelTurn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
- Oct 7, 2026 · Research paper · 1 sourceDoes Document Structure Help Dense Retrieval? A Placebo-Controlled Ablation of Four Mechanisms Across Two CorporaRetrieval-augmented generation systems increasingly rely on document-structure treatments: structure-aligned chunking, LLM-generated chunk contexts, heading-path metadata, and hierarchical two-stage retrieval.
- Oct 7, 2026 · Research paper · 1 sourceEntroPrefill: Renyi-Guided Context Pruning with Conditional Stability Guarantees for Retrieval-Augmented GenerationMid-prefill pruning can reduce the sequence processed by deeper transformer layers, but attention concentration alone does not certify that discarded context is dispensable.
- Oct 7, 2026 · Research paper · 1 sourceFinding the Right Balance: Relevance and Diversity in LLM RetrievalUsing controlled near-duplicate injection and production-style overlapping chunking, we find that diversification harms relevance, evidence coverage and answer quality on clean pools, but becomes beneficial on multi-evidence tasks when redundancy causes nearest-neighbor retrieval to select repeated passages.
- Oct 7, 2026 · Research paper · 1 sourceFrom Retrieval to Customer Context: Evaluating Frontier-Model Systems for Voice-of-Customer AnalysisWe define a customer context graph as a unified model of customer and business context.
- Oct 7, 2026 · Research paper · 1 sourceFrom Chunks to Functional Evidence: Function-Aware Retrieval for EDA Documentation QARetrieval-Augmented Generation (RAG) is widely used to ground answers in documents.
- Oct 7, 2026 · Research paper · 1 sourceTopoGraphRAG-Bench: Evaluating Multimodal GraphRAG on Layout-Grounded Evidence ReasoningWe introduce TOPOGRAPHRAG-BENCH, a layout-grounded benchmark for multimodal evidence reasoning in GraphRAG, comprising 2,024 questions over 201 long, visually rich documents.
- Oct 6, 2026 · Research paper · 1 sourceSPLATIFY: Reproduce, Discover, Innovate! From Papers and Ideas to Trainable 3DGS CodeWe introduce SPLATIFY, a multi-agent framework that converts 3DGS papers into trainable gsplat-based implementations, where generic paper-to-code methods and frontier models fail.
- Oct 6, 2026 · Product / feature launch · 1 sourceEmbeddingGemma 2: an open, lightweight multimodal embedding modelEmbeddingGemma 2: an open, lightweight multimodal embedding model
- Oct 6, 2026 · Research paper · 1 sourceBEACON-SP: Ontology-Grounded GraphRAG Framework for Clinical Suicide Risk AssessmentWe present BEACON-SP, an ontology-grounded Graph Retrieval-Augmented Generation (GraphRAG) framework for clinician-facing decision support in behavioral health settings such as suicide prevention, where effective assessment requires integrating heterogeneous clinical, behavioral, social, and temporal evidence.
- Oct 6, 2026 · Research paper · 1 sourceEC-RAG: Event Chain Retrieval-Augmented Generation for Long Video UnderstandingIn this paper, we propose Event Chain Retrieval-Augmented Generation (EC-RAG), a training-free framework that organizes video content into an explicit event chain before question answering.
- Oct 6, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.903-beta: New Browser + Voice CloningThis release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.
- Oct 6, 2026 · Research paper · 1 sourceTrustworthy Domain-Specific AI for Structured Knowledge Retrieval and ReasoningThis dissertation presents a scalable architecture for transforming unstructured, domain-specific text into structured knowledge for retrieval and reasoning.
- Oct 6, 2026 · Research paper · 1 sourceRAG-PIBench: A Leakage-Aware Benchmark for Prompt-Injection Detection in Trustworthy RAG SystemsWe introduce RAG-PIBench, a benchmark for RAG-style prompt-injection detection containing 4,876 contextual examples across frozen train, validation, and protected-test splits.
- Oct 6, 2026 · Opinion / analysis · 1 sourceBuild a voice travel concierge with Amazon Bedrock AgentCore, Managed Knowledge Base and Nova SonicAirlines already have apps and websites where travelers check flights, pick seats, and manage bookings, and adding a natural voice layer opens those tasks to spoken requests.
- Oct 6, 2026 · Research paper · 1 sourceUNREAL: Unifying Retrieval and Long-Context with a Single ModelWe introduce UNifying REtrieval And Long-Context with a Single Model (UNREAL), a model-native evidence selection framework to span corpus retrieval and long-context inference.
- Oct 6, 2026 · Research paper · 1 sourceAgentic AutoRAG: RAG Pipeline Optimization through Reasoning-Driven AgentsWe introduce Agentic AutoRAG, an LLM-agent optimizer for multi-objective RAG hyperparameter optimization with retrieval-versus-generation failure attribution.