用于大规模搜索、抓取与交互网页的 API。🔥The context API to search, scrape, and interact with the web at scale. 🔥
仓库/Skill 库
294 个 · RAG 检索增强
RAGFlow 是一款领先的开源 RAG 引擎,融合前沿 RAG 与 Agent 能力,为 LLM 构建更优的上下文层。RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
Crawl4AI:开源、对 LLM 友好的网络爬虫与抓取工具。🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
Context7 Platform——为 LLM 与 AI 代码编辑器提供最新的代码文档Context7 Platform -- Up-to-date code documentation for LLMs and AI code editors
开箱即用的云模板,支持 RAG、AI 流水线与企业级搜索,接入实时数据。🐳Docker 友好。⚡始终与 Sharepoint、Google Drive、S3、Kafka、PostgreSQL 以及各类实时数据 API 保持同步。Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.
[EMNLP2025] "LightRAG:简单且快速的检索增强生成"[EMNLP2025] LightRAG: Simple and Fast Retrieval-Augmented Generation
模块化的、基于图结构的 Retrieval-Augmented Generation (RAG) 系统。A modular graph-based Retrieval-Augmented Generation (RAG) system
📑 PageIndex:面向无向量、基于推理的 RAG 的文档索引📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
FastGPT 是基于 LLM 构建的知识平台,提供开箱即用的全套能力,包括数据处理、RAG 检索与可视化 AI 工作流编排,让你无需复杂配置即可轻松开发并部署复杂的问答系统。FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data processing, RAG retrieval, and visual AI workflow orchestration, letting you easily develop and deploy complex question-answering systems without the need for extensive setup or configuration.
该仓库展示了多种面向 Retrieval-Augmented Generation (RAG) 系统的高级技术,每个技术均配有详细的 notebook 教程。This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial.
面向 AI 就绪数据的 PDF 解析器,自动化 PDF 无障碍处理,开源。PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
基于 RAG 的开源文档对话工具。An open-source RAG-based tool for chatting with your documents.
💡 面向语义搜索、LLM 编排与语言模型工作流的一体化 AI 框架💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
网页解析的终结,可扩展像素原生搜索的开端。链接:https://pixelrag.ai/The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/
根据你的 Zotero 库,每日推荐你感兴趣的新 arxiv 论文。Recommend new arxiv papers of your interest daily according to your Zotero libarary.
LLM 实战指南:从基础到使用 LLMOps 最佳实践将 LLM 和 RAG 应用部署到 AWSThe LLM's practical guide: From the fundamentals to deploying advanced LLM and RAG apps to AWS using LLMOps best practices
面向 monorepo 的终极 RAG。借助 AI 与知识图谱的能力,对多语言代码库进行查询、理解与编辑。The ultimate RAG for your monorepo. Query, understand, and edit multi-language codebases with the power of AI and knowledge graphs
免费学习如何使用 LLMOps 最佳实践构建端到端生产级 LLM & RAG 系统:源码 + 12 个实操课程。🤖 𝗟𝗲𝗮𝗿𝗻 for 𝗳𝗿𝗲𝗲 how to 𝗯𝘂𝗶𝗹𝗱 an end-to-end 𝗽𝗿𝗼𝗱𝘂𝗰𝘁𝗶𝗼𝗻-𝗿𝗲𝗮𝗱𝘆 𝗟𝗟𝗠 & 𝗥𝗔𝗚 𝘀𝘆𝘀𝘁𝗲𝗺 using 𝗟𝗟𝗠𝗢𝗽𝘀 best practices: ~ 𝘴𝘰𝘶𝘳𝘤𝘦 𝘤𝘰𝘥𝘦 + 12 𝘩𝘢𝘯𝘥𝘴-𝘰𝘯 𝘭𝘦𝘴𝘴𝘰𝘯𝘴
LeanCTX — 面向 AI 系统的 Context Intelligence。LeanCTX — Context Intelligence for AI systems.
⚡FlashRAG:面向高效 RAG 研究的 Python 工具包(WWW2025 Resource)⚡FlashRAG: A Python Toolkit for Efficient RAG Research (WWW2025 Resource)
漂亮的、AI 原生的 markdown IDE 与 LLM 知识库。Beautiful, AI-native markdown IDE and LLM wiki
内置知识库 · 2.4 亿文献源 · PDF 全文翻译零配置 · 无插件网页版 app.ai4paper.pro · AI 速搜文献 · AI 一键综述 · 小程序期刊订阅 · 原生支持 Zotero 7–10 | AI4Paper.pro
RAG 领域新 SOTA —— 一种全新的原创检索架构,以及面向人类与 Agent 的开源知识库。A new SOTA for RAG — an original retrieval architecture and an open-source knowledge base for humans and agents.
收录 AIGC 领域 RAG 的优秀论文。我们在论文 "Retrieval-Augmented Generation for AI-Generated Content: A Survey" 中提出了 RAG 基础、增强与应用分类法。Collecting awesome papers of RAG for AIGC. We propose a taxonomy of RAG foundations, enhancements, and applications in paper "Retrieval-Augmented Generation for AI-Generated Content: A Survey".
用于获取和校验 DOI 与 shortDOI 的 Zotero 扩展。Zotero extension to retrieve and validate DOIs and shortDOIs
基于 OpenSearch 构建的开源、自托管企业及站内搜索服务器,爬取网页、文件、数据库与云端数据源,支持 20+ 语言、REST API,以及 AI/RAG 与语义搜索。Apache-2.0。Open-source, self-hosted enterprise & site search server built on OpenSearch. Crawls web / file / DB / cloud sources, 20+ languages, REST API, and AI/RAG & semantic search. Apache-2.0.
Rust 库,用于本地生成向量 embedding 和重排序!Rust library for generating vector embeddings and reranking locally!
一份共享记忆,被 Claude Code、Codex、Cursor、Copilot CLI、OpenClaw 以及另外 29 个 coding agent 共同使用,基于磁盘上已有的会话历史构建。在任一 agent 中发现的修复会自动同步到其他所有 agent,包括安装前数月的会话。无需 LLM,无需 embeddings,单个本地 Go 二进制。One memory shared by Claude Code, Codex, Cursor, Copilot CLI, OpenClaw and 29 more coding agents, built from the session history already on disk. A fix found in one agent comes back in any of them, including months of sessions from before you installed it. No LLM, no embeddings, one local Go binary.
📚 AIGC 求职面经、必备基础知识、提示词工程、ChatGPT、Stable Diffusion、Prompt、Embedding、Fintune 等 AIGC 求职你所需要知道的一切~
使用开源 LLM 摄取文件用于检索增强生成(RAG),无需第三方,数据不出你的网络。Ingest files for retrieval augmented generation (RAG) with open-source Large Language Models (LLMs), all without 3rd parties or sensitive data leaving your network.
极简网络搜索平台,配有可在浏览器直接运行的 AI 助手。Demo:https://felladrin-minisearch.hf.spaceMinimalist web-searching platform with an AI assistant that runs directly from your browser. Demo: https://felladrin-minisearch.hf.space
自动化研究论文的发现与分析。例如,让计算机自动打开与任意 URL 论文最相似的文献,或找出在某数据集上报告结果的全部论文,重新定义文献综述Automating research publications discovery and analysis. For example, ever wish your computer could automatically open papers that are most similar to a paper at an arbitrary url? How about finding all papers that report results on some dataset? Let's re-imagine literature review.