研究库 论文知识库
Papers · organized/paper_cards

论文

218 张论文卡片 · RAG 检索增强

开放获取 全部 绿色 · 1640
PlanSightRAG: A Visual-First Multimodal RAG for Automating Question Answering and Compliance Checking for Civil Standard Plans
PlanSightRAG:面向土木标准图自动化问答与合规审查的视觉优先多模态 RAG
arXiv:2608.26091 RAG 检索增强 评测集 OA · 绿色 被引 0 · S2 + OpenAlex

PlanSightRAG 是一种 Visual-First 多模态 RAG 框架,直接对图纸图像建立索引并进行推理,集成了 ColNomic-3B 多向量检索、Agentic Planner-Retriever-Auditor-Synthesizer,并以 MaxSim 热力图作为证据链。A Visual-First Multimodal Retrieval-Augmented Generation (RAG) framework called PlanSightRAG, which indexes and reasons directly over plan imagery, integrates a ColNomic-3B multi-vector retrieval, an agentic Planner-Retriever-Auditor-Synthesizer, and MaxSim heatmaps as an evidence trail.

Multi-Granularity Context-Enhanced RAG over Multimodal Knowledge Graphs
多模态知识图谱上的多粒度上下文增强 RAG
arXiv:2608.25986 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出一种构建 Context-Enhanced MMKG (CEMMKG) 的新框架,能够有效利用上下文信息提升基于 MMKG 的 RAG 性能,并在不同基于 MMKG 的 RAG 方法上的有效性验证了其广泛适用性。A novel framework for constructing a Context-Enhanced MMKG (CEMMKG) is proposed, effective in leveraging contextual information to improve MMKG-based RAG performance and its effectiveness across different MMKG-based RAG methods demonstrates its broad applicability.

RetrievalRouter: Joint Modality and Architecture Selection for Document Retrieval
RetrievalRouter:面向文档检索的模态与架构联合选择
arXiv:2608.25625 RAG 检索增强 方法 被引 0 · S2

RetrievalRouter 是一种轻量级 query-aware router,仅依据 query 文本即可学习最匹配的检索 pipeline,在面向准确率的设置下 nDCG@5 显著更高,而在面向延迟的设置下,nDCG@5 和延迟均匹配或数值上优于基线。RetrievalRouter is a lightweight query-aware router that learns, from the query text alone, which retrieval pipeline best fits each query, and achieves significantly higher nDCG@5 across accuracy-oriented settings, while matching or numerically outperforming them on both nDCG@5 and latency in latency-oriented settings.

CaSKG: Counterfactual-Causal Skill Graphs for Scalable Agent Skill Retrieval
CaSKG:用于可扩展 Agent 技能检索的反事实-因果技能图
arXiv:2608.25500 RAG 检索增强 方法 OA · 绿色 被引 3 · S2

提出 CaSKG,一种反事实因果 Skill 图谱框架,在检索前校准程序关系,将边置信度校准定位为大规模紧凑且可执行的 Skill 检索的有效路径。CaSKG, a counterfactual-causal skill graph framework that calibrates procedural relations before retrieval, is proposed, position edge-confidence calibration as an effective route to compact and executable skill retrieval at scale.

6. LLM Research Papers: The 2026 List (Jan–May) — Sebastian Raschka
LLM 研究论文:2026 年清单(1—5 月)— Sebastian Raschka
arXiv:2601.21204 RAG 检索增强 方法 Open MIND OA · 绿色 被引 14 · S2

本工作将 embedding 缩放作为正交于稀疏度缩放的强有力维度加以探索,并推出 LongCat-Flash-Lite,一个从零训练的 68.5B 参数、约 30 亿激活参数的模型,不仅超越参数等量级的 MoE 基线,还对同规模现有模型展现出卓越竞争力。This work explores embedding scaling as a potent, orthogonal dimension for scaling sparsity and introduces LongCat-Flash-Lite, a 68.5B parameter model with ~3B activated trained from scratch that not only surpasses parameter-equivalent MoE baselines but also exhibits exceptional competitiveness against existing models of comparable scale.

LINE Conversation History Retrieval for Personal Memory RAG: Evaluating Search Representations and Hybrid Retrieval
基于 LINE 对话历史的个人记忆 RAG 检索:搜索表示与混合检索评估
arXiv:2608.27809 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

该研究将 358,896 条消息切分为 22,329 个时间连贯的块,并构建三种搜索表示:raw_text、生成的 summary,以及将 summary 与 raw_text 片段及其他固定文本相结合的 embedding_text。This study segmented 358,896 messages into 22,329 temporally coherent chunks and constructed three search representations: raw_text, a generated summary, and embedding_text, which combines a summary with a raw-text excerpt and other fixed text.

CamoDocs: A Poisoning Attack Against Retrieval-Augmented Language Models Using Camouflaged Documents
CamoDocs:利用伪装文档针对检索增强语言模型的投毒攻击
arXiv:2608.28389 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

提出 CamoDocs,一种通过将对抗文档伪装在良性内容中来避免直接包含查询的投毒攻击,并表明 TrustRAG 等以擦除为主的聚类防御可降低 ASR,但会在 NeoQA 等依赖检索的基准上造成显著的效用下降。CamoDocs is proposed, a poisoning attack that avoids direct query inclusion by camouflaging adversarial documents among benign content, and shows that erasure-heavy clustering defenses such as TrustRAG can reduce ASR, but only with substantial utility drops on retrieval-dependent benchmarks such as NeoQA.

SymbolLKG: Towards Verifiable Logical Reasoning via Logical Knowledge Graph and Symbolic Solvers
SymbolLKG:通过逻辑知识图谱与符号求解器实现可验证的逻辑推理
arXiv:2608.26836 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

一种神经符号架构,将逻辑知识图谱(LKG)与动态求解器路由相结合,并引入基于本体的 LKG,将逻辑规则和约束视为一等拓扑节点,从而支持对从文本中抽取的依赖关系进行显式建模。A Neuro-Symbolic architecture that integrates a Logical Knowledge Graph (LKG) with dynamic solver routing, and introduces an ontology-based LKG that treats logical rules and constraints as first-class topological nodes, enabling explicit modeling of dependencies extracted from text.

Pointing the Way, Hiding the Destination: Practical Private Dense Retrieval at Scale
指路而隐目的地:面向大规模场景的实用化私有稠密检索
arXiv:2608.25735 RAG 检索增强 方法 OA · 绿色 被引 1 · S2

该候选名单可在不牺牲检索质量的前提下短路全语料库加密搜索:200–500 个候选即可在 5 个零样本语料库(规模从 25K 到 5.4M 文档)中与全语料库检索效果接近匹配。This shortlist short-circuits full-corpus cryptographic search without sacrificing retrieval quality: with 200-500 candidates, it closely matches full-corpus retrieval across five zero-shot corpora spanning 25K to 5.4M documents.

Configurable Semantic Chunking for Biomedical Information Extraction in Retrieval-Augmented Generation
RAG 中面向生物医学信息抽取的可配置语义分块
arXiv:2608.31139 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

跨数据集分析表明,语义切分在具有显式关系线索的抽取数据集(如 GM-CIHT 和 DDI)上表现更优,而固定切分在密集生化抽取和二分类场景(如 ChemProt 和 ADE)下仍具竞争力甚至更强。Cross-dataset analysis shows that semantic chunking improves extraction datasets with explicit relation cues, such as GM-CIHT and DDI, while fixed chunking remains competitive or stronger for dense biochemical extraction and binary classification settings such as ChemProt and ADE.

SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers
SMELT:面向计算匹配的 MoE 循环 Transformer 的扩展定律
arXiv:2609.01343 RAG 检索增强 方法 OA · 绿色 被引 13 · S2

结果表明,即便在算力预算匹配的前提下,循环(looping)仍可提升 Transformer,提供了一种将深度复用转化为可衡量增益的实用方案。Results show that looping can improve Transformers even under budget matching, offering a practical recipe that turns depth reuse into measurable gains.

Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation
自适应关键 token 感知的仓库级代码生成检索
arXiv:2609.01601 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

ACToR 在生成过程中识别关键 token,按需触发有针对性的检索,在这些决定性位置提供仓库上下文;并为稠密检索器设计了一种位置感知加权方法,以优先考虑对生成更具信息量的上下文。ACToR identifies critical tokens during generation and triggers targeted retrieval on demand to provide repository context at these decisive positions, and designs a position-aware weighting method for dense retrievers to prioritize context that is more informative for generation.

A Glance Is All You Need: Single-Pass Fine-Grained Image Captioning with SimLoss
一瞥即可:基于 SimLoss 的单遍细粒度图像描述生成
arXiv:2609.00591 RAG 检索增强 观点 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 SimLoss,一种面向单轮细粒度图像描述的无参考 embedding 空间目标;结果表明 embedding 空间监督能在单轮描述器延迟下恢复多阶段验证的质量。SimLoss is proposed, a reference-free embedding-space objective for single-pass fine-grained image captioning, and results show that embedding-space supervision can recover the quality of multi-stage verification at the latency of a single-pass captioner.

5️⃣ arXiv · Is Agentic RAG Worth It? An Experimental Comparison of RAG Approaches(⭐⭐⭐⭐ 高优先级)
5️⃣ arXiv · Agentic RAG 是否值得?RAG 方法的实验对比(⭐⭐⭐⭐ 高优先级)
arXiv:2601.07711 RAG 检索增强 评测集 OA · 绿色 被引 6 · S2

基于实证对 "Enhanced" 与 "Agentic" RAG 范式进行评估,为真实场景中选取最有效的 RAG 设计(兼顾性能与成本)提供指导。An empirically driven evaluation of the "Enhanced" and "Agentic" RAG paradigms is conducted, offering guidance on selecting the most effective RAG design for real-world applications, considering both performance and costs.

SnapBench: Benchmarking Snap-and-Ask Multimodal Retrieval for Mobile Interactions
SnapBench:面向移动端交互的即拍即问多模态检索 benchmark
arXiv:2608.29607 RAG 检索增强 评测集 OA · 绿色 被引 1 · S2

本文提出 SnapBench——首个面向鲁棒"拍照即问"多模态检索的配对基准,以及一种简单的自适应融合方法 MOOR(Modality-anchored, Outlier-aware, Optimal Reweighting),并指出在"拍照即问"检索中需要具备可靠性感知的模态校准。SnapBench is introduced, the first paired benchmark for robust snap-and-ask multimodal retrieval, and MOOR (Modality-anchored, Outlier-aware, Optimal Reweighting), a simple adaptive fusion approach, highlighting the need for reliability-aware modality calibration in snap-and-ask retrieval.

ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering
ViSAR:面向视觉文档问答的无训练自适应 $k$ 检索
arXiv:2609.02486 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 ViSAR(Visual Semantic Activation Retrieval),一种面向 late-interaction 视觉文档检索的无训练自适应 k 检索方法,并表明相似度矩阵结构与答案准确率相关,为面向检索质量感知的文档理解指明了未来方向。ViSAR (Visual Semantic Activation Retrieval), a training-free adaptive-$k$ retrieval method for late-interaction visual document retrieval, is introduced and it is shown that the similarity matrix structure correlates with answer accuracy, suggesting future directions for retrieval quality-aware document understanding.

Improving Health Literacy through Lay Summarization of Radiological Reports: An Evaluation of BioNER and Retrieval-Augmented Generation
通过放射学报告的大众化摘要提升健康素养:BioNER 与检索增强生成的评估
arXiv:2609.02396 RAG 检索增强 评测集 OA · 绿色 被引 0 · S2 + OpenAlex

本研究探讨了相较于标准 LLM 生成方式,RAG 与命名实体识别在多大程度上能提升自动生成通俗摘要的质量、事实一致性及可读性。This study investigates the extent to which Retrieval-Augmented Generation and Named Entity Recognition improve the quality, factual consistency, and readability of automatically generated lay summaries compared with standard LLM-based generation.

NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning
NE-R1:通过强化学习增强命名实体识别模型
arXiv:2609.02366 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 NE-R1,一种面向自适应检索增强 NER 的新框架,在多个基准上达到 SOTA 性能,域内评估平均 F1 提升 2.52%,零样本跨域评估平均 F1 提升 1.18%。This paper proposes NE-R1, a novel framework for adaptive retrieval-augmented NER, which achieves state-of-the-art performance on various benchmarks, with an average F1 score gain of 2.52% in in-domain evaluation and 1.18% in zero-shot cross-domain evaluation.

Beyond Visual Similarity: Entity-Aligned Retrieval for Knowledge-Based Visual Question Answering
超越视觉相似性:面向知识库视觉问答的实体对齐检索
arXiv:2608.21450 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 KBMR,首个面向 KB-VQA 的基于 MLLM 的 embedding retriever,并引入一个基于 MLLM 的语义判别器以生成连续的实体一致性权重,应对维基百科规模检索中的噪声监督挑战。KBMR is proposed, the first MLLM-based embedding retriever tailored for KB-VQA, and an MLLM-based semantic discriminator that generates continuous entity-consistency weights is introduced to tackle the challenge of noisy supervision in Wikipedia-scale retrieval.

Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens
Sparse Readout Prism:用特征而非 token 解释 Logit-Lens 分数
arXiv:2609.01936 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

Sparse Readout Prism (SRP) 仅使用 readout 的权重对其进行分解,将任意 token logit 或 logit 差表示为来自稀疏 readout 特征贡献之和,揭示了 readout 特征作为 lens 解读新单元的价值,暴露出 token 身份可能掩盖的结构,并支持跨 token、上下文、层与 lens 的比较。Sarse Readout Prism (SRP), which decomposes the readout using only its weights and expresses any token logit or logit difference as a sum of contributions from sparse readout features, reveals readout features as a new unit of analysis for lens readings, exposing structure that token identities can obscure and enabling comparisons across tokens, contexts, layers, and lenses.

Wasserstein-Barycentric Interaction Fields for Spatial Factor Models: Evidence from Language-Model Representations
空间因子模型的 Wasserstein-重心交互场:来自语言模型表征的证据
arXiv:2608.29669 RAG 检索增强 方法 被引 1 · S2

空间资产定价模型将企业间交互结构视为已知,并利用语言模型表征从企业的信息环境中推断该结构;语言模型表征充当资本市场中潜在企业间信息结构的测量工具。Spatial asset-pricing models take the structure of inter-firm interaction as given and infer that structure from firms' information environments using language-model representations, which serve as a measurement instrument for latent inter-firm information structure in capital markets.

Beyond Retrieval: Progressive Latent Memory Evolution for Streaming Video Understanding
超越检索:面向流式视频理解的渐进式潜在记忆演化
arXiv:2609.04131 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

该工作提出 LatentStream,一种渐进式 latent working memory 框架,将流式记忆从"存储-检索"转变为"检索-内化",在现有在线和离线视频 benchmark 上取得新的 SOTA 结果。This work introduces LatentStream, a progressive latent working memory framework that shifts streaming memory from store-and-retrieve to retrieve-and-internalize, and achieves new state-of-the-art results on existing online and offline video benchmarks.

PACE: Towards Surfacing Hidden Conflicts in User Requests
PACE:揭示用户请求中的隐性冲突
arXiv:2609.03293 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出 PACE 数据集,用于评估模型能否识别以自我中心知识或事件形式表达的潜在约束(这些约束使看似合理的用户请求变得不当),以及 PaceMaker 多 Agent 框架,其中专门 Agent 通过查询重构、多跳图遍历与冲突感知过滤进行协调,以检索上下文决定性证据。PACE is introduced, a dataset for evaluating whether models can identify latent constraints, expressed as egocentric knowledge or events, that render seemingly reasonable user requests inappropriate, and PaceMaker, a multi-agent framework in which specialized agents coordinate across query reformulation, multi-hop graph traversal, and conflict-aware filtering to retrieve contextually decisive evidence.

RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning
RoboTok:面向人类演示检索与灵巧操作学习的大规模互联网数据引擎
arXiv:2609.03199 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本工作提出 RoboTok,一个可扩展的数据引擎:利用人类操作视频作为查询,从互联网检索与操作相关的演示以训练灵巧机器人策略,并从以演员为中心的参考系下估计的 3D 手部轨迹中学习一个潜在运动空间This work introduces RoboTok, a scalable data engine that uses a query human manipulation video to retrieve manipulation-relevant internet demonstrations for training dexterous robot policies and learns a latent motion space from 3D hand trajectories expressed in estimated actor-centered reference frames.

Does Your Agent's Memory Survive a Model Upgrade? A Controlled Study of Memory Portability
你的 Agent 记忆能否在模型升级后幸存?一项关于记忆可移植性的受控研究。
arXiv:2609.05339 RAG 检索增强 方法 被引 1 · S2

研究结果强调了方向特定的迁移测试、严格的 embedding space 隔离,以及在 memory migrations 中为 memory repair 保留源历史的必要性。Findings highlight the necessity of direction-specific migration testing, strict embedding space isolation, and the retention of source histories for memory repair in memory migrations.

ENEAS: Embedding-guided Neural Ensemble for Adaptive Segmentation
ENEAS:嵌入引导的自适应分割神经集成
arXiv:2609.03756 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

我们提出 ENEAS,一种用于实例追踪与语义发现的统一且文本可提示的方法。包括 SAM 3 在内的文本可提示分割模型仍存在时间幻觉、空间碎片化与语义误分类问题:目标离开视野时无法报告目标缺失;极端特写下只分割局部纹理而非完整目标;将视觉特征置于本体事实之上,从而把雕像、绘画或反射等视觉相似的物体误分割为目标。We present ENEAS, a unified, text-promptable method for instance tracking and semantic discovery. Text-promptable segmentation models, including the latest foundation models such as SAM 3, still suffer from temporal hallucinations, spatial fragmentation, and semantic misclassification: they fail to report target absence when an object leaves the field of view, segment local textures instead of the complete object during extreme close-ups, and prioritize visual features over ontological reality, so that visually similar artifacts such as statues, paintings, or reflections are segmented as targe

Unifying Conformal Language Tasks with In-Context Ensembles
通过上下文集成统一保形语言任务
arXiv:2609.03005 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文提出 Conformal Relevance 框架,利用上下文学习的示例筛选与集成构造打分函数,在保持覆盖的同时以极低人工成本提升简洁性。The Conformal Relevance framework is introduced which uses in-context learning example curation and ensembling to create a score function which maintains coverage while improving conciseness with minimal manual input.

Generative Late-Interaction Embeddings For Visual Document Retrieval
面向视觉文档检索的生成式晚交互 embedding
arXiv:2609.11808 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

Generative Late-Interaction Embeddings(GLIE):从归一化质心中学习每个页面 k<<N 个向量,既作为轻量级索引,也作为重建页面完整嵌入集的基础,解码器是其主要设计面。Generative Late-Interaction Embeddings (GLIE): k<<N vectors per page learned from the normalized centroids to serve as both a lightweight index and a basis for regenerating the page's full embedding set, with the decoder as its main design surface.

Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking
先思后链:多语言实体链接中的稀有性、推理与检索
arXiv:2609.10745 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

一个简单、无训练的框架,其中具备推理能力的视觉语言模型在 Wikipedia 上迭代搜索与推理,动态收集证据,表明推理与检索在稀有实体上具有互补性。A simple, training-free framework in which a reasoning-capable vision-language model iteratively searches and reasons over Wikipedia, gathering evidence dynamically, shows that reasoning and retrieval are complementary on rare entities.

Affordance-Compiled Intelligence: Observable-Only Cognitive Impedance Matching for No-Meta LLM-Integrated Systems
Affordance-Compiled Intelligence: Observable-Only Cognitive Impedance Matching for No-Meta LLM-Integrated Systems
arXiv:2005.11401 RAG 检索增强 方法 OA · 绿色 被引 19105 · S2

提出面向检索增强生成(RAG)的通用微调方案——RAG 模型融合预训练参数化记忆与非参数化记忆进行语言生成;研究发现,相较 SOTA 的纯参数化 seq2seq 基线,RAG 模型生成的文本更具针对性、更多样且更符合事实。A general-purpose fine-tuning recipe for retrieval-augmented generation (RAG) -- models which combine pre-trained parametric and non-parametric memory for language generation, and finds that RAG models generate more specific, diverse and factual language than a state-of-the-art parametric-only seq2seq baseline.

ReMoMask-2: Latent Retrieval-Augmented Masked Motion Generation
ReMoMask-2:潜在检索增强的掩码运动生成
arXiv:2609.08365 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

提出结构感知的 RAG 框架 ReMoMask-2:耦合 Hierarchical Bidirectional Momentum 对比学习以对齐全局与部件级特征与文本;采用 Semantic Spatial-Temporal Attention (SSTA) 实现拓扑感知的融合;通过 Topology Structured Masking (TSM) 借助自适应掩码强化鲁棒的部件级 grounding。ReMoMask-2, a structure-aware RAG framework coupling Hierarchical Bidirectional Momentum contrastive learning to align global and part-level features with text; Semantic Spatial-Temporal Attention (SSTA) for topology-aware fusion; and Topology Structured Masking (TSM) to force robust part-level grounding via adaptive masking are introduced.

CiteGuard-RAG: A Validation-Centered AI System for Evidence-Grounded Question Answering
CiteGuard-RAG:一种以验证为中心的、基于证据的问答 AI 系统
arXiv:2609.15830 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

外部评估显示,尽管引用有效性保持稳健,但在领域偏移下证据利用、片段对齐与拒答校准变得更加困难,表明可信的 RAG 系统需要在检索与最终答案交付之间进行显式验证。External evaluation shows that while citation validity remains strong, evidence utilization, span alignment, and refusal calibration become harder under domain shift, indicating that trustworthy RAG systems require explicit validation between retrieval and final answer delivery.

Navigating Sparse Evidence: Agentic Visual RAG via Explicit Context Selection and Consolidation
驾驭稀疏证据:通过显式上下文选择与整合的 Agentic Visual RAG
arXiv:2609.15800 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

该工作提出了 SCoRE(Selection and Consolidation for Robust Evidence),一个用于显式证据选择与整合的统一 agent 循环,将最终推理与探索式试错解耦,并通过索引化的声明-图像关联确保严格的视觉锚定。This work proposes SCoRE (Selection and Consolidation for Robust Evidence), a unified agent loop for explicit evidence selection and consolidation, which decouples final reasoning from exploratory trial-and-error while ensuring strict visual grounding via indexed claim-to-image linkages.

ORDER: Task-Conditioned Routing for Retrieval-Augmented Generation
ORDER:面向检索增强生成的任务条件化路由。
arXiv:2609.17012 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

检索增强生成(RAG)管线通常依赖在预处理阶段确定的固定索引与检索配置。这种一刀切的设计难以适配领域专家场景,因为异构查询需要不同的分块粒度、元数据约束与来源选择策略。因此,针对某一类查询有效的配置,往往在其他类查询上表现欠佳。本文提出 ORDER(Optimal Routing for Dynamic Evidence Retrieval),一种查询条件化的 RAG 框架,可联合自适应地调整索引与……Retrieval-Augmented Generation (RAG) pipelines typically rely on a fixed indexing and retrieval configuration determined at preprocessing time. This one-size-fits-all design is ill-suited to domain-expert settings, where heterogeneous queries require different chunking granularities, metadata constraints, and source-selection strategies. As a result, configurations that are effective for one family of queries often perform poorly for others. In this paper, we introduce ORDER (Optimal Routing for Dynamic Evidence Retrieval), a query-conditioned RAG framework that jointly adapts indexing and ret

InceptionRAG: Stealthy Poisoning Attack Against Retrieval-Augmented Generation
InceptionRAG:针对检索增强生成的隐蔽投毒攻击
arXiv:2609.16818 RAG 检索增强 方法 被引 0 · S2

论文提出了 InceptionRAG,一种颠覆 RAG 标准攻击范式的隐蔽攻击机制,展现出更优的规避能力,能够有效绕过针对传统单文档注入的既有防御。This paper introduces InceptionRAG, a stealthy attack mechanism that subverts the standard attack paradigm of RAG, and shows superior evasion capabilities, effectively bypassing established defenses that mitigate traditional single-document injections.

FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation
FLAT:将图像和文本重采样为 1D 变长对齐跨模态 token 用于检索与生成
arXiv:2609.16591 RAG 检索增强 方法 OA · 绿色 被引 0 · S2 + OpenAlex

本文重新审视联合多模态表示学习与生成,旨在产生可直接被生成式解码器使用的线性可插值嵌入,并确保其表示同时充当判别性语义描述符和生成条件。This work revisits joint multimodal representation learning and generation to produce linearly interpolatable embeddings that are directly consumable by generative decoders and ensures its representations function as both discriminative semantic descriptors and generative conditions.