研究库 论文知识库
Papers · organized/paper_cards

论文

192 张论文卡片 · RAG 检索增强 · OA 绿色

开放获取 全部 绿色 · 1640
HAKARI-Bench: A Lightweight Benchmark for Comparing Retrieval Architectures and Efficiency Settings under Unified Conditions
HAKARI-Bench: A Lightweight Benchmark for Comparing Retrieval Architectures and Efficiency Settings under Unified Conditions
arXiv:2606.22778 RAG 检索增强 评测集 OA · 绿色 被引 1 · S2

HAKARI-Bench 是一个轻量级基准,将现有检索套件重建为统一格式的小型数据集(Nano-sets),支持在同一条件下对五类检索方法及其效率变体进行与模型无关的对比。HAKARI-Bench is a lightweight benchmark that reconstructs existing retrieval suites into small datasets (Nano-sets) in a unified format, enabling same-condition, model-agnostic comparison of five retrieval families and their efficiency variants.

When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG
当置信度走上歧路:诊断 RAG 中的检索状态锁定
arXiv:2606.22728 RAG 检索增强 应用落地 OA · 绿色 被引 1 · S2

本文命名了"检索状态锁定"这一失败模式,通过分离单一置信度分数所混淆的三个对象——答案表面、检索到的证据以及检索状态本身——来诊断该问题,并直接衡量"一致性盲区"。This work names the failure retrieval-state lock-in and diagnose it by separating the three objects a single confidence score conflates: the answer surface, the retrieved evidence, and the retrieval state itself, and measures the agreement blind spot directly.

Qiskit Code Migration with LLMs
使用 LLM 进行 Qiskit 代码迁移
arXiv:2606.20173 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出一种将 LLM 与检索增强生成(RAG)相结合的混合方法,用于自动化跨版本 Qiskit 代码迁移,验证了这种以数据为中心的方法在促进技术独立性和提供缓解 API 过时问题的鲁棒智能助手方面的潜力。A hybrid approach integrating LLMs with Retrieval-Augmented Generation (RAG) to automate the migration of Qiskit code across versions and confirms the potential of this data-centric methodology to foster technological independence and provide robust, intelligent assistants that mitigate API obsolescence.

When Does Streaming Tool Use Help? Characterizing Tool-Intent Stabilization in Streaming Retrieval-Augmented Generation
流式工具使用何时有效?刻画流式 RAG 中的工具意图稳定化
arXiv:2606.20113 RAG 检索增强 评测集 OA · 绿色 被引 0 · S2 + OpenAlex

本文命名并测量了推测性查询的检索收敛到包含答案结果时的输入流位置——"工具意图稳定化":即推测性查询的检索收敛到包含答案结果的输入流位置。This work names and measures the point in the input stream at which a speculative query's retrieval converges on the answer-bearing result, tool-intent stabilization: the point in the input stream at which a speculative query's retrieval converges on the answer-bearing result.

Human-AI Coevolution Dynamics: A Formal Theory of Social Intelligence Emergence Through Long-Term Interaction
人-AI 协同进化动力学:通过长期交互涌现社交智能的形式化理论
arXiv:2606.19144 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

HACD-H 为建模自适应人-AI 社交交互与开发社交智能 AI 系统提供了统一的理论基础,并表明社交智能源自长期社交认知的协同进化,而非孤立的对话能力。The HACD-H provides a unified theoretical foundation for modeling adaptive human-AI social interaction and developing socially intelligent AI systems and suggests that social intelligence emerges from long-term social cognitive coevolution rather than isolated conversational capabilities.

MCompassRAG: Topic Metadata as a Semantic Compass for Paragraph-Level Retrieval
MCompassRAG:以主题元数据作为段落级检索的语义罗盘
arXiv:2606.18508 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出 MCompassRAG,一种由元数据引导的检索框架,将主题级信号作为语义罗盘以选择相关证据,在同一 embedding 空间中以主题元数据丰富 chunk 表示,并通过 LLM 教师蒸馏训练轻量级检索器。MCompassRAG is introduced, a metadata-guided retrieval framework that uses topic-level signals as a semantic compass for selecting relevant evidence and enriches chunk representations with topic metadata in the same embedding space and trains a lightweight retriever through LLM-teacher distillation.

HistoRAG: Embedding Historical Methodology in Retrieval-Augmented Generation Through Critical Technical Practice
HistoRAG:通过批判性技术实践将历史学方法论嵌入 RAG
arXiv:2606.18103 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出 HistoRAG,一个将史学原则转化为具体架构干预的框架,为特定领域认识论承诺如何转化为 RAG 设计决策提供模型,并可迁移至其他处理大规模语料的诠释性学科。HistoRAG is introduced, a framework that translates historiographical principles into concrete architectural interventions and offers a model for how domain-specific epistemological commitments can be translated into RAG design decisions, and may transfer to other interpretive disciplines working with large corpora.

A Unified Framework for Context-Aware and Relation-Aware Graph Retrieval-Augmented Generation
面向上下文感知与关系感知图 RAG 的统一框架
arXiv:2606.18075 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出 HyGRAG,一种分层图 RAG 框架,通过解决三个核心挑战超越源文档限制:构建真正融合上下文与关系信息的摘要、利用这些综合表示在检索阶段访问涌现知识、以及为动态语料高效更新分层结构。HyGRAG is proposed, a hierarchical graph RAG framework that transcends source documents by addressing three core challenges: constructing summaries that genuinely integrate contextual and relational information, leveraging these synthesized representations to access emergent knowledge during retrieval, and efficiently updating hierarchical structures for dynamic corpora.

When Rules Learn: A Self-Evolving Agent for Legal Case Retrieval
当规则学会学习:面向法律案例检索的自进化 Agent
arXiv:2606.17220 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出一个面向规则驱动查询改写的自进化框架,无需任何参数训练即可增强 BM25,并揭示 LLM 利用先前实验结果的能力以及其对规则消除的内在知识,在通过自进化精炼规则集方面起到关键作用。This work proposes a self-evolving framework for rule-driven query rewriting that enhances BM25 without any parameter training, and reveals that LLM's capabilities to leverage previous experimental results and its intrinsic knowledge of rule elimination play critical roles in refining the rule set via self-evolution.

RL-Index: Reinforcement Learning for Retrieval Index Reasoning
RL-Index:面向检索索引推理的强化学习
arXiv:2606.16316 RAG 检索增强 方法 OA · 绿色 被引 2 · S2

RL-Index 被提出,是一个将检索索引推理建模为强化学习问题的索引框架,能持续提升检索与下游问答性能,同时显著降低在线推理延迟。RL-Index is proposed, an indexing framework that formulates retrieval index reasoning as a reinforcement learning problem that consistently improves both retrieval and downstream question-answering performance, while significantly reducing online inference latency.

MAGE-RAG: Multigranular Adaptive Graph Evidence for Agentic Multimodal RAG in Long-Document QA
MAGE-RAG:面向长文档问答中 Agentic 多模态 RAG 的多粒度自适应图证据
arXiv:2606.15906 RAG 检索增强 方法 OA · 绿色 被引 2 · S2

本文提出 MAGE-RAG,一个面向长文档多模态问答的多粒度自适应图证据框架,并建立了涵盖 Direct MLLM、Text RAG、Page-level Visual RAG 与 Graph/Agentic RAG 的统一比较与分析协议。This paper proposes MAGE-RAG, a multigranular adaptive graph evidence framework for long-document multimodal QA, and establishes a unified comparison and analysis protocol covering Direct MLLM, Text RAG, Page-level Visual RAG, and Graph/Agentic RAG.

Ricci-Filtration: Boosting Retrieval-Augmented Generation Reranker to Query-Answer Tasks by Discrete Ricci Flow
Ricci-Filtration:通过离散 Ricci Flow 将 RAG 重排序器提升至 Query-Answer 任务
arXiv:2606.15482 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文理论上证明,归一化离散 Ricci Flow 可通过识别边权中的不同渐近行为来检测社区结构,并支持移除相对于 query 节点具有大权重与负 Ricci 曲率的"噪声"文档片段。It is theoretically prove that normalized discrete Ricci flow can detect community structures by identifying distinct asymptotic behaviors in edge weights, and supports the removal of ``noisy''document chunks characterized by large weights and negative Ricci curvature relative to the query node.

ScoreGate: Adaptive Chunk Selection for Retrieval-Augmented Generation via Dual-Score Statistical Fusion
ScoreGate:基于双分数统计融合的 RAG 自适应 Chunk 选择
arXiv:2606.14269 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

在 MS MARCO 与真实生产流量上的结果表明,自适应检索 cardinality 能够在不降低检索质量的前提下提升检索效率。Results on both MS MARCO and real-world production traffic suggest that adaptive retrieval cardinality can improve retrieval efficiency without degrading retrieval quality.

CQC-RAG: Robust Retrieval-Augmented Generation via Cross-Query Consistency
CQC-RAG:通过跨查询一致性实现鲁棒的检索增强生成
arXiv:2606.13438 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

CQC-RAG 框架被提出,它协同设计查询级多样性注入与跨查询一致性评估,无需外部监督即可实现自我评估,验证了跨查询一致性在过滤噪声引发幻觉方面的有效性。CQC-RAG, a framework that co-designs query-level diversity injection with cross-query consistency evaluation and enables self-evaluation without external supervision, is introduced, validating the effectiveness of cross-query consistency for filtering noise-induced hallucinations.

Large Behavior Model: A Promptable Digital Twin of the Retail Customer
Large Behavior Model:零售客户的可提示数字孪生
arXiv:2607.06993 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

结果表明,交易历史中编码的行为知识可以被语言模型有效学习,为客户数字孪生和行为模拟提供了可扩展的基础。The results demonstrate that behavioral knowledge encoded in transaction histories can be effectively learned by language models, providing a scalable foundation for customer digital twins and behavior simulation.

End-to-End LLM Flight Planning with RAG-based Memory and Multi-modal Coach Agent
基于 RAG 记忆与多模态教练智能体的端到端 LLM 飞行规划
arXiv:2607.06964 RAG 检索增强 应用落地 OA · 绿色 被引 1 · S2

FRAMe 展示了先进 LLM 如何被部署用于以人为本的任务规划,将自然语言指令转化为安全、高效且灵活的飞行路线。FRAMe signifies how advanced LLMs can be deployed for human-centric mission planning, translating natural language instructions into safe, efficient, and flexible flight routes.

Conversational Retrieval and On-the-Fly Knowledge Modeling of Historical Penitentiary Repression Records
历史监狱压迫记录的对话式检索与即时知识建模
arXiv:2607.08459 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出一种面向历史数字图书馆管理的文档分析系统,支持即时知识建模,并促进生成更丰富、更全面的信息。This article presents a document analysis system designed for the management of historical digital libraries that supports on-the-fly knowledge modeling and facilitates the generation of richer and more comprehensive information.

PolyUQuest: Verifiable Structure-Aware Web RAG over Heterogeneous Graphs
PolyUQuest:异构图上的可验证结构感知 Web RAG
arXiv:2607.08269 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 PolyUQuest,一个可验证、感知结构的 Web RAG 框架,基于异构图构建,统一了页面间超链接拓扑、页面内 DOM 层级以及跨页面实体关系知识。PolyUQuest is presented, a verifiable, structure-aware web RAG framework built on a heterogeneous graph that unifies hyperlink topology between pages, DOM hierarchy within pages, and entity-relation knowledge across pages.

Deceptive Grounding: Entity Attribution Failure in Clinical Retrieval-Augmented Generation
欺骗性 grounding:临床 RAG 中的实体归因失败
arXiv:2607.09349 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

一项受控消融实验揭示了机制:从检索到的文档中移除特定实体的临床证据,可彻底消除实体归因失败,使所有失败转移到虚构生成。A controlled ablation identifies the mechanism: removing entity-specific clinical evidence from retrieved documents eliminates entity-attribution failure entirely, shifting all failures to confabulation.

Augmenting Fundamental Analysis with Large Language Models: A RAG-Based System for Generating Investor Briefs
利用 LLM 增强基本面分析:基于 RAG 的投资者简报生成系统
arXiv:2607.09121 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

论文探讨了 LLM 为公司基本面分析各方面带来的机会,分析依据包括公司报告、描述宏观经济状况(如 GDP 和通胀变化)的数据与文件,以及提交至美国证券交易委员会(SEC)的文件。The opportunities brought by Large Language Models (LLMs) to various aspects of fundamental analysis of companies based on their reports as well as data and documents describing macroeconomic situation like GDP and inflation changes as well as documents filled to the U.S. Securities and Exchange Commission (SEC) are examined.

AgentKGV: Agentic LLM-RAG Framework with Two-Stage Training for the Fact Verification of Knowledge Graphs
AgentKGV:面向知识图谱事实核查的智能体 LLM-RAG 框架与两阶段训练
arXiv:2607.09092 RAG 检索增强 应用落地 OA · 绿色 被引 0 · S2 + OpenAlex

提出 AgentKGV,一种用于知识图谱事实核查的智能体 LLM-RAG 框架,集成动态路由与迭代查询改写,以应对文档级检索中的表层形式不匹配问题。AgentKGV, the Agentic LLM-RAG framework for KG fact Verification, is proposed, that integrates dynamic routing and iterative query rewriting, which handles surface-form mismatch in document-level retrieval.

How Temperature Shapes Ideological Discourse in Retrieval-Augmented Generation?
温度如何塑造 RAG 中的意识形态话语?
arXiv:2607.11783 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

通过考察包含意识形态话语的 RAG 框架对 LLM 生成答案的影响,发现 RAG 框架倾向于将意识形态话语传递到 LLM 响应中,且采样温度对这种传递的强度有可测量的影响。Examining the influence of the RAG framework, comprising ideological discourses, in LLM-generated answers shows that the RAG framework is prone to transferring ideological discourses into LLM responses, with sampling temperature having a measurable impact on the strength of this transfer.

EvoGraph-R1: Self-Evolving Multimodal Knowledge Hypergraphs for Agentic Retrieval
EvoGraph-R1:面向 Agentic 检索的自演化多模态知识超图
arXiv:2607.12764 RAG 检索增强 方法 OA · 绿色 被引 2 · S2

提出 EvoGraph-R1,一个自演化 GraphRAG 框架,将知识图谱重新概念化为由 Agent 交互塑造的动态环境,将自演化知识图谱确立为跨模态的基础范式。EvoGraph-R1 is introduced, a self-evolving GraphRAG framework that reconceptualizes knowledge graphs as dynamic environments shaped through agent interactions, establishing self-evolving knowledge graphs as a fundamental paradigm across modalities.

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education
Earthquaker-AI:面向小学地震教育的、采用评分量表评估的 RAG 框架
arXiv:2607.14046 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 Earthquaker-AI,一个混合式教育框架,在已有教育机器人项目基础上集成基于 RAG 的对话式 AI 助手,旨在提升小学生的地震应急准备与主动行动意识。该系统将曾获奖的 STEM 项目 Earthquaker 从 Lego WeDo2 的机械模拟拓展至认知与元认知层面:机器人组件利用 Lego WeDo2 自动化模拟地震响应,使学生能够与传感器和执行器进行交互。This paper presents Earthquaker-AI, a hybrid educational framework building upon a previously implemented educational robotics project by integrating a conversational AI assistant based on Retrieval-Augmented Generation. It aims to enhance earthquake preparedness and conscious action among primary-school students. The system extends the award-winning STEM project Earthquaker moving from mechanical simulation with Lego WeDo2 to cognitive and metacognitive processing. The robotics component uses Lego WeDo2 automation to simulate seismic response, letting students interact with sensors and actuat

GRASP: GRanularity-Aware Search Policy for Agentic RAG
GRASP:面向 Agentic RAG 的粒度感知搜索策略
arXiv:2607.10463 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出 GRASP,一个用于训练智能体在多步推理过程中自适应协调互补检索工具的强化学习(RL)框架,并指出学会协调检索信号与上下文粒度对智能体的正确推理至关重要。GRASP is introduced, a reinforcement learning (RL) framework for training agents to adaptively coordinate complementary retrieval tools during multi-step reasoning, and it is suggested that learning to coordinate retrieval signals and context granularity is critical for agent's correct reasoning.

AI Prototyper: A Figma Plugin for Decomposition-Based GUI Prototyping with LLMs
AI Prototyper:基于分解的 LLM GUI 原型设计 Figma 插件
arXiv:2607.14830 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

提出 AI Prototyper,一个开源 Figma 插件,通过分解与 RAG 流水线自动完成 GUI 原型设计,并引入人在回路编辑步骤,允许用户在渲染前审查、修改或扩展生成的功能列表。AI Prototyper is presented, an open-source Figma plugin that automates GUI prototyping through a decomposition and retrieval-augmented generation (RAG) pipeline, and introduces a human-in-the-loop editing step that lets users review, modify, or extend the generated feature list before rendering.

Is External Database Protection Static in Retrieval-Augmented Generation? Rethinking Privacy Preservation under Dynamic Queries
RAG 中的外部数据库保护是静态的吗?重新审视动态查询下的隐私保护
arXiv:2607.14811 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

提出 PA-HDP(Prompt-Aware Dynamic Hierarchical Differential Privacy)框架,通过 prompt 感知的风险分层动态评估不同查询下的隐私风险,并采用自适应敏感实体替换与基于指数机制的文本选择,在保留语义可用性的同时提供差异化的隐私保护。A Prompt-Aware Dynamic Hierarchical Differential Privacy framework (PA-HDP) is proposed, which performs a prompt-aware risk hierarchy to dynamically assess privacy risks under different queries and applies adaptive sensitive entity replacement and exponential mechanism-based text selection to provide differentiated privacy protection while preserving semantic utility.

Chat2Scenic: An Iterative RAG-Based Framework for Scenario Generation in Autonomous Driving
Chat2Scenic:面向自动驾驶场景生成的迭代式 RAG 框架
arXiv:2607.14387 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

提出 Chat2Scenic,是首个以领域特定语言 (DSL) 生成场景脚本的迭代式检索增强框架,并构建了一个涵盖 NHTSA、联合国车辆法规及其他来源共 123 个场景的开源场景生成基准。Chat2Scenic is presented, the first iterative retrieval-augmented framework to generate scenario scripts in Domain Specific Language (DSL) and proposes an open benchmark for scenario generation comprising 123 scenarios from various regulations, including NHTSA and United Nations Vehicle Regulations, as well as other sources.

Retrieval-Augmented Generation for Large Language Models: A Survey
Retrieval-Augmented Generation for Large Language Models: A Survey
arXiv:2312.10997 RAG 检索增强 综述 OA · 绿色 被引 4251 · S2

该综述细致梳理了 RAG 范式的演进,涵盖 Naive RAG、Advanced RAG 与 Modular RAG,并对 RAG 框架的三大基础支柱——检索、生成与增强技术——进行了深入审视。This comprehensive review paper offers a detailed examination of the progression of RAG paradigms, encompassing the Naive RAG, the Advanced RAG, and the Modular RAG, and meticulously scrutinizes the tripartite foundation of RAG frameworks, which includes the retrieval, the generation and the augmentation techniques.

RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM
RAGU:基于紧凑领域适配LLM的多步GraphRAG引擎
arXiv:2607.11683 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

RAGU是一个开源模块化GraphRAG引擎,通过将抽取与整合分离来解决抽取-整合问题:实体和关系经过两阶段类型化抽取、基于DBSCAN的去重、LLM摘要和Leiden社区检测。RAGU, an open-source modular GraphRAG engine, addresses extraction from consolidation by separating extraction from consolidation: entities and relations pass through two-stage typed extraction, DBSCAN-backed deduplication, LLM summarization, and Leiden community detection.

A Human-Centric Evaluation of a Retrieval-Augmented Generation System for Explaining Quebec Insurance Contracts
面向魁北克保险合同解释的RAG系统的人类中心评估
arXiv:2607.15963 RAG 检索增强 评测集 OA · 绿色 被引 0 · S2 + OpenAlex

对一个旨在使魁北克汽车保险合同更易理解的SOTA RAG系统进行以人为中心的外在评估,结果显示该系统被视为认知均衡器,用户对系统所提供的自主感的重视程度甚至超过知识本身。A human-centric, extrinsic evaluation of a state-of-the-art Retrieval-Augmented Generation system, designed to make Quebec automobile insurance contracts more understandable, shows the system is perceived as a cognitive equalizer, and users value the sense of autonomy the system provides even more than the knowledge itself.

NOWJ@COLIEE 2026: Adaptive Pipelines for Legal Retrieval and Reasoning
NOWJ@COLIEE 2026:面向法律检索与推理的自适应流水线
arXiv:2607.16603 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文介绍了 NOWJ 团队参加 COLIEE 2026 全部五项任务的方法与结果,采用基于稠密检索、注意力重排序和小样本提示 LLM 推理的检索增强生成框架。This paper presents the methodologies and results of the NOWJ team's participation across all five tasks of the COLIEE 2026 competition and adopts a retrieval-augmented generation framework with dense retrieval, attention-based reranking, and few-shot-prompted LLM reasoning.

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking
通过知识蒸馏将 LLM 转化为高效的 RAG 重排序 Cross-Encoder
arXiv:2607.11933 RAG 检索增强 应用落地 OA · 绿色 被引 0 · S2 + OpenAlex

Cross-encoder 在 RAG 流水线中具有较高的重排序准确率,但推理成本随序列长度呈二次增长,难以实时部署。本文通过两阶段流水线解决该问题:使用 Unsloth 框架与 LoRA 适配器,在自定义的查询-文档相关性数据集上对 LLaMA 3 (8B) 进行监督微调,随后进行 4-bit 量化以提升推理效率。该模型可替换双路检索 RAG 流水线中结合 BM25 与稠密向量检索的 cross-encoder,并在特定领域问答……Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-time deployment. We address this by fine-tuning LLaMA 3 (8B) as a drop-in reranker using a two-stage pipeline: supervised fine-tuning on a custom query-document relevance dataset via the Unsloth framework with LoRA adapters, followed by 4-bit quantization for efficient inference. The resulting model replaces the cross-encoder in a dual-retriever RAG pipeline combining BM25 and dense vector search. Evaluated on a domain-specific question-answering

Vector Search As Nearest Neighbor Matching: RAG-based Policy Learning in Causal Inference
向量搜索作为最近邻匹配:基于RAG的因果推断策略学习
arXiv:2607.18225 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

该工作将两步方法的遗憾分解为候选生成遗憾和候选内选择遗憾,并利用最近邻估计器和Transformer的预测误差保证对后者进行了界。This work decomposes the regret of the two-step method into candidate-generation regret and within-candidate choice regret, and bound the latter using prediction-error guarantees for nearest-neighbor estimators and transformers.

Testing Retrieval-Augmented Generation Systems with Chunk Coverage
Testing Retrieval-Augmented Generation Systems with Chunk Coverage
arXiv:2607.18155 RAG 检索增强 应用落地 OA · 绿色 被引 2 · S2

本文提出 Chunk Coverage (CC),一种独立于 oracle 的 RAG 系统检索组件测试充分性准则,结果表明 CC 在无需测试 oracle 的情况下捕获了与有效测试相关的检索多样性。Chunk Coverage (CC), an oracle-independent test adequacy criterion for testing the retrieval component of RAG systems, is introduced and results show that CC captures retrieval diversity relevant to effective testing without requiring test oracles.

AutoIndex: Learning Representation Programs for Retrieval
AutoIndex:为检索学习表征程序
arXiv:2607.18603 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

研究结果表明,文档表示不应被视为检索开始前一次性的固定预处理选择,而应作为一个明确的优化目标。The results suggest that document representation should not be treated as a fixed preprocessing choice made before retrieval begins, but as an explicit optimization target.