面向 LLM、VLM、DiT 和 REC 模型的高性能推理引擎,针对多种 AI 加速器进行了优化。该项目托管于 OpenAtom 基金会。A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
仓库/Skill 库
487 个 · 应用 · AI 核心
高性能推理引擎,支持 LLM、VLM、DiT 和 REC 模型,针对多种 AI 加速器优化A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.
RTP-LLM:阿里巴巴面向多样化应用的高性能 LLM 推理引擎。RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.
基于 Rust 🦀 构建的高性能、模块化、内存安全、生产可用的推理、数据接入与索引系统Highly Performant, Modular, Memory Safe and Production-ready Inference, Ingestion and Indexing built in Rust 🦀
基于自主编码 Agent 的开源自动研究系统。运行 Claude Code、OpenCode 与 Codex,具备评分、共享知识与多 Agent 进化功能。被 COLM 2026 接收。Open-source autoresearch powered by autonomous coding agents. Run Claude Code, OpenCode, and Codex with grading, shared knowledge, and multi-agent evolution. Accepted at COLM 2026.
基于 Next.js、LangChain、PostgreSQL + pgvector 构建的 AI 驱动的初创企业加速引擎。可上传、组织并与文档对话。包含预测性缺失文档检测、基于角色的工作流以及页面级洞察提取。AI-powered StartUp Accelerator Engine built with Next.js, LangChain, PostgreSQL + pgvector. Upload, organize, and chat with documents. Includes predictive missing-document detection, role-based workflows, and page-level insight extraction.
从零开始构建 LLM 驱动的多 Agent 应用Building LLM-Enabled Multi Agent Applications from Scratch
Athena 是一个本地优先的 agentic PKM,借助你自己的上下文帮助你做出更优决策——持久记忆、结构化推理,以及跨任意 LLM 工作的受治理 AI agent。拥有状态,租用智能。Athena is a local-first agentic PKM that helps you make better decisions with your own context — persistent memory, structured reasoning, and governed AI agents that work across any LLM. Own the state. Rent the intelligence.
一款惊艳的、本地化的通用 AI Agent,适用于 Android、iOS 和 Mac。| Aether 扶摇A stunning, localized, general-purpose AI Agent for Android, iOS & Mac. | Aether 扶摇
繁体中文(zh-TW)的语言学 linter。A linguistic linter for Traditional Chinese (zh-TW)
Caura(前身 MemClaw)——面向 AI Agent 集群的受治理共享内存。多 Agent、多租户、原生 MCP。支持信任分级、keystone 策略、审计日志、知识图谱、自改进检索。Apache 2.0。Caura (formerly MemClaw) — governed shared memory for AI agent fleets. Multi-agent, multi-tenant, MCP-native. Trust tiers, keystone policies, audit trails, knowledge graph, self-improving retrieval. Apache 2.0.
一款 MCP,让 AI 工具安全连接基础设施、编写 IaaS 代码、调试问题并在故障中提供协助,且不危及生产稳定性。专为安全团队审批而设计,为基础设施团队带来如魔法般的体验。An MCP that lets AI tools securely connect to your infrastructure, write IaaS code, debug issues, and assist during incidents - without risking production stability. Built for security teams to approve and infrastructure teams to experience like magic.
Pi Coding Agent 扩展的 TypeScript monorepo,涵盖自动化、规划、语言工具、浏览器控制、网络研究、Git 工作流和配置同步。A TypeScript monorepo of Pi Coding Agent extensions for automation, planning, language tooling, browser control, web research, Git workflows, and configuration sync.
StarWhisper 天文 LLMs、StarWhisper Telescope、Virtual-GOTTA,以及面向 embodied observing workflow 的天文定制研究 skills。StarWhisper astronomy LLMs, StarWhisper Telescope, Virtual-GOTTA, and astronomy-adapted research skills for embodied observing workflows
将本地文件转化为 AI Agent 可检索的上下文。Turn local files into searchable context for AI agents.
你的主动式个人 AI 助手与日常生产力伙伴 🌎Your proactive personal AI assistant & companion for daily productivity 🌎
Claude Code 的完整 AI 开发工具包。包含 106 个 skill、36 个 agent、171 个 hook。安装 `ork` 获取稳定版(v9.x),或安装 `ork-alpha` 获取每日发布的 v10 版本。The Complete AI Development Toolkit for Claude Code. 106 skills, 36 agents, 171 hooks. Install `ork` for stable (v9.x), or `ork-alpha` for the v10 line, which ships daily.
一份配置即可统管所有 AI agents:跨项目跨会话可移植、内置精选写作/路由/skills、更安全(内置破坏性命令防护)One config to rule all your AI agents: portable (every project, every session), effective (curated writing, routing, skills), and safer (destructive-command guard).
[ICML 2026] effGen:让小型语言模型具备自主 Agent 能力[ICML 2026] effGen: Enabling Small Language Models as Capable Autonomous Agents
一门实战课程,用 PyTorch 从零构建现代 LLM,包含 26 个可运行的 Jupyter Notebook,涵盖 tokenizer、attention、MoE、RLHF、推理、评估和蒸馏。A hands-on course for building modern LLMs from scratch in PyTorch, with 26 runnable Jupyter Notebooks covering tokenizers, attention, MoE, RLHF, inference, evaluation, and distillation.
面向 SolCex Exchange 的自主 AI BD Agent:7×24 跨链代币评分与支付 2026。Autonomous AI BD Agent for SolCex Exchange: 24/7 Cross-Chain Token Scoring & Payments 2026
AKB — Agent Knowledgebase。AI agent 的组织记忆:以 URI 图统一 vault 范围的文档 / 表格 / 文件,通过 MCP 提供服务。AKB — Agent Knowledgebase. Organizational memory for AI agents: vault-scoped docs / tables / files unified by URI graph, served over MCP.
面向 macOS 的强约束 AI 编码 Agent 与开发环境自动化Opinionated AI coding agent and dev environment automation for macOS
2026 年在社交信息流中通过实时聊天机器人模拟 AI 人格Simulate AI Personas in Social Feeds with Real-Time Chat Bots 2026
面向实验的 GenAI/RAG 优化器与工具包,基于 Oracle Database AI Vector Search 与 NL2SQLGenAI/RAG Optimizer and Toolkit for experimentation using Oracle Database AI Vector Search and NL2SQL
为大型语言模型(LLM)和 RAG 系统转换并优化你的 markdown 文档,自动生成 llms.txt。Transform and optimize your markdown documentation for Large Language Models (LLMs) and RAG systems. Generate llms.txt automatically.
由 LLM 驱动的科研自动化 AgentLLM-powered agents for scientific research automation
可解释、本地优先的 AI Agent 记忆引擎。约 9 MB 的单一二进制在 VelesQL 下融合向量 + 图 + 列存;why() 返回每次召回背后的证据路径。无云依赖、无胶水代码 —— 支持服务器、浏览器、移动端与桌面端运行。The explainable, local-first memory engine for AI agents. One ~9 MB binary fuses vector + graph + columnar under VelesQL; why() returns the evidence path behind every recall. No cloud, no glue code — runs on server, browser, mobile and desktop.
面向法律与重型文档工作流的 AI-native IDE 工作空间:文件、Agent、插件、WPS 编辑、OCR、证据链。律师的 VS Code。AI-native IDE workspace for legal and document-heavy workflows: files, agents, plugins, WPS editing, OCR, evidence chains. VS Code for lawyers.
基于对象存储的快速搜索引擎,原生支持 Parquet 上的全文检索、向量检索与 SQL。Fast search engine on object storage, with full text search, vectors, and SQL, natively on Parquet.
通过 hooks 拦截 AI 编程 Agent 发起的 cargo/gcc 构建,并透明地将其路由到远程工作机器,像在本地编译一样返回产物。Intercepts cargo/gcc builds from AI coding agents via hooks and transparently routes them to remote worker machines, returning artifacts as if compiled locally
ICLR 2026 Oral 论文 "Q-RAG: Long Context Multi-Step Retrieval via Value-Based Embedder Training" 的官方仓库。Official repository for the ICLR 2026 Oral Paper🔥 “Q-RAG: Long Context Multi-Step Retrieval via Value-Based Embedder Training”
即插即用、成本高效且轻量化的长周期 LLM Agent 记忆插件A Plug-and-Play, Cost-Efficient, and Lightweight Memory Plugin for Long-Horizon LLM Agents
面向编程 Agent 的更快运行时。平均提速 25%、降本 30%,同时质量持平或更优。同等任务、同等质量,更快更便宜。Faster runtime for coding agents. Make coding agents 25% faster and 30% cheaper on average while keeping the quality same or more. Same Task, Same Quality, Faster and Cheaper.
本地 MCP server,用于有状态、fail-closed 的 Logic Pro 控制与项目实时回读。Local MCP server for stateful, fail-closed Logic Pro control and live project readback.